nerdexam
Cloudera

DS-200 · Question #61

What is the most common reason for a k-means clustering algorithm to returns a sub-optimal clustering of its input?

The correct answer is C. Non-normal distribution of the input data. See the full explanation below for the reasoning.

Question

What is the most common reason for a k-means clustering algorithm to returns a sub-optimal clustering of its input?

Options

  • ANon-negative values for the distance function
  • BInput data set is too large
  • CNon-normal distribution of the input data
  • DPoor selection of the initial controls

How the community answered

(34 responses)
  • A
    6% (2)
  • B
    3% (1)
  • C
    79% (27)
  • D
    12% (4)

Community Discussion

No community discussion yet for this question.

Full DS-200 Practice