Tested 15 local models for agent/tool use.. Bonsai 27B was last!
Tested 15 local models for agent/tool use — Bonsai 27B was last I wanted to see which local models are actually usable for agent/tool-use tasks, so I ran 15 of them through the same benchmark instead of guessing from model hype. The benchmark was Toolery . It doesn't just check whether a model can…
Read the full story at r/LocalLLaMA ↗
Timeline · 1 report
- 2026-10-04 19:39 · r/LocalLLaMA
Tested 15 local models for agent/tool use.. Bonsai 27B was last!