Week 4 — Lesson 2: Encoding categories and scaling numbers

3 min

A model reads numbers only. The NorthPeak machine type is a word, pump or chiller, and the sensor columns run on very different ranges: temperature around 46.2 with a spread of 14.7, vibration around 4.2 with a spread of 1.3. Both facts must be handled before some models can learn. This lesson covers one-hot encoding for the text category and standardisation for the numbers, with the exact output of both on the NorthPeak data.

Preview — the rest of the lesson is for enrolled readers.

Already enrolled with a code?

Your access is tied to your account, not to this link. Sign in with the same email you used in class: your course is waiting, no need to enter the code again.

Sign inNo account yet? Create one
This lesson is part of the “Week 4 — Data preparation and dataset separation” module

The first modules of the course are open to everyone. For the rest you have three options: buy this course once and for all, subscribe, or enter the code handed out in class.

Are you a student on this course?

The code is tied to your account: sign in or create an account and it will be applied automatically when you come back.

No account yet? Create one