18 open models on 33 real DevOps tasks, one RTX 5090: pass rate, tokens/s and measured VRAM
I wanted to know which local model can actually do infrastructure work, so I recorded 33 broken scenarios and ran 18 tool-calling models that fit a single RTX 5090 (32 GB) through them: the twelve my tool recommends plus six others I was curious about. Ollama 0.35.0, the default build of each model…
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-10-01 17:23 · r/LocalLLM
18 open models on 33 real DevOps tasks, one RTX 5090: pass rate, tokens/s and measured VRAM