Just another purely open-weight models benchmark
Live results, coding benchmarks included, agentic benchmarks included, domains and other classifications filters included. (Spoiler: DeepSeek-V4-Vision-Exp rules, but other models have their rule areas): https://beta.locallm.top Evaluated by domain experts (my friends mostly; coding part is evaluat…
Read the full story at r/LocalLLaMA ↗
Timeline · 1 report
- 2026-10-09 18:43 · r/LocalLLaMA
Just another purely open-weight models benchmark