Making cosine similarities comparable across separately trained embedding spaces
For my open-source project, I am training PPMI + SVD embeddings separately on each book within a collection (~25 economics texts from Project Gutenberg), then trying to compare how similar terms are to a query word across books. However, each space has been trained independently, so the raw cosines…
Read the full story at r/LanguageTechnology ↗
Timeline · 1 report
- 2026-09-24 20:18 · r/LanguageTechnology
Making cosine similarities comparable across separately trained embedding spaces