Submission timeline
2007–2026One slot for every year since HN launched. Height is that year's peak points; orange marks a 100+ point or 50+ comment breakout. Select a bar to open its strongest thread.
First comments on top threads
HN comment orderThere are clear places where “king” and “queen” are similar to each other and distinct from all the others. Could these be coding for a vague concept of royalty? This is a common misunderstanding and unfortunately strengthened by the example of 'personality embeddings'. It is easy to understand intuitively why this is normally not the case. If you rotate a vector/embedding space, all the cosine similarities between words are preserved. Suppose that component 20 encoded 'royalty', there is an infinite…
This is a great guide. Also - despite the fact that language model embedding [1] are currently the hot rage, good old embedding models are more than good enough for most tasks. With just a bit of tuning, they're generally as good at many sentence embedding tasks [2], and with good libraries [3] you're getting something like 400k sentence/sec on laptop CPU versus ~4k-15k sentences/sec on a v100 for LM embeddings. When you should use language model embeddings: - Multilingual…
The first top-level comment from each of the four biggest threads, in HN’s own order. Excerpts are shortened; open a comment for full context.
- Breakout years
- 2
- Total points
- 532
- Total comments
- 56
100+ points or 50+ comments
reference only — not used in Hall rules or ranking
reference only — not used in Hall rules or ranking
Every submission
| Date | Title as submitted | By | Points | Comments |
|---|---|---|---|---|
| 2019-03-27 | The Illustrated Word2vecFirst breakout · Best thread | jalammar | 348 | 37 |
| 2021-08-20 | The Illustrated Word2vec | ColinWright | 1 | 0 |
| 2022-11-01 | The Illustrated Word2vec (2019) | rrampage | 3 | 0 |
| 2024-04-18 | The Illustrated Word2Vec (2019)Latest 20+ point return | wcedmisten | 180 | 19 |
