I benchmarked Jev aginst gpt-5.6-luna!
This story is from 2026-09-19. It is preserved in the archive; the latest stories are on the live feed.
I got access to TypeSafe's Jev a few days ago. It's an odd kind of model that doesn't generate text at all. You send it some content plus typed questions (yes/no, pick one of these options, rate this on a scale) and it gives you back probabilities. Setup: 49 tasks, about 8,200 items, all from publi…
Read the full story at r/OpenAI ↗
Timeline · 1 report
- 2026-09-19 08:25 · r/OpenAI
I benchmarked Jev aginst gpt-5.6-luna!