Week 8 — Lesson 3: Embeddings and semantic similarity

2 min

Two NorthPeak notes can say the same thing with different words: Oil leak under the gearbox. and A puddle of lubricant was found below the gear housing. TF-IDF gives them a similarity of 0.092, because they share only the word the. An embedding places each text at a position in a space where texts with the same meaning are close, whatever words they use. This lesson explains that space, the cosine similarity that measures closeness, and how to read the numbers it gives.

Preview — the rest of the lesson is for enrolled readers.

Already enrolled with a code?

Your access is tied to your account, not to this link. Sign in with the same email you used in class: your course is waiting, no need to enter the code again.

Sign inNo account yet? Create one
This lesson is part of the “Week 8 — NLP and text representations” module

The first modules of the course are open to everyone. For the rest you have three options: buy this course once and for all, subscribe, or enter the code handed out in class.

Are you a student on this course?

The code is tied to your account: sign in or create an account and it will be applied automatically when you come back.

No account yet? Create one