I benchmarked Jev against gpt-5.6-luna!
I got access to TypeSafe's Jev a few days ago. It's an odd kind of model that doesn't generate text at all. You send it some content plus typed questions (yes/no, pick one of these options, rate this on a scale) and it gives you back probabilities. Setup: 49 tasks, about 8,200 items, all from publi…
Read the full story at r/ArtificialInteligence ↗
Timeline · 2 reports
- 2026-09-19 08:25 · r/OpenAI
I benchmarked Jev aginst gpt-5.6-luna! - 2026-09-19 08:23 · r/ArtificialInteligence
I benchmarked Jev against gpt-5.6-luna!