Exercise 2 — Predict faultnext7d with the Week 4 features

Guided practice5 min
Time
20-30 min
You need
the kit, the venv active, data/features.csv from Week 4
Deliverable
your confusion matrix at threshold 0.5 and the four scores of Step 5

The lab kit of the course: https://github.com/hrhouma2/aiopsatlas-ml-data-diagnostics-labs-en

Goal

This is the main task of the course, done properly for the first time. You give the 15 Week 4 features to a logistic regression inside a pipeline. You train on January to September and test on October to December. You compare with the dummy that says "no fault" every time. Then you read the confusion matrix cell by cell, in NorthPeak words, and compute precision, recall and F1.

Preview — the rest of the lesson is for enrolled readers.

Already enrolled with a code?

Your access is tied to your account, not to this link. Sign in with the same email you used in class: your course is waiting, no need to enter the code again.

Sign inNo account yet? Create one
This lesson is part of the “Week 5 — Regression, classification and evaluation” module

The first modules of the course are open to everyone. For the rest you have three options: buy this course once and for all, subscribe, or enter the code handed out in class.

Are you a student on this course?

The code is tied to your account: sign in or create an account and it will be applied automatically when you come back.

No account yet? Create one