K-Means decides the groups by distance, and distance depends on the units of each column. On two readings of machine M001, load alone gives 1,632 of the 2,071 squared units of difference; vibration gives 0.8. For K-Means, the column that predicts faults does not exist. This lesson shows why every column must be scaled before the fit, what happens on the NorthPeak readings without scaling, and introduces hierarchical clustering, a method that needs no k up front.
Preview — the rest of the lesson is for enrolled readers.
Your access is tied to your account, not to this link. Sign in with the same email you used in class: your course is waiting, no need to enter the code again.
Sign inNo account yet? Create oneThe first modules of the course are open to everyone. For the rest you have three options: buy this course once and for all, subscribe, or enter the code handed out in class.