Cloudera
DS-200 · Question #61
What is the most common reason for a k-means clustering algorithm to returns a sub-optimal clustering of its input?
The correct answer is C. Non-normal distribution of the input data. See the full explanation below for the reasoning.
Question
What is the most common reason for a k-means clustering algorithm to returns a sub-optimal clustering of its input?
Options
- ANon-negative values for the distance function
- BInput data set is too large
- CNon-normal distribution of the input data
- DPoor selection of the initial controls
How the community answered
(34 responses)- A6% (2)
- B3% (1)
- C79% (27)
- D12% (4)
Community Discussion
No community discussion yet for this question.