The 128-Vector Trick: Why Perplexity Stopped Squeezing Documents Into One Vector
This story is from 2026-10-09. It is preserved in the archive; the latest stories are on the live feed.
On October 7, Perplexity dropped two embedding models on Hugging Face under an MIT license — and the size gap between them is the whole story. The small one, pplx-embed-v2-late , runs on 0.6 billion parameters and lives on edge hardware. The big one is 9 billion parameters and, paired with Gemini 3…
Read the full story at DEV Community — Machine Learning ↗
Timeline · 1 report
- 2026-10-09 16:04 · DEV Community — Machine Learning
The 128-Vector Trick: Why Perplexity Stopped Squeezing Documents Into One Vector