AINewsnow

18 open models on 33 real DevOps tasks, one RTX 5090: pass rate, tokens/s and measured VRAM

I wanted to know which local model can actually do infrastructure work, so I recorded 33 broken scenarios and ran 18 tool-calling models that fit a single RTX 5090 (32 GB) through them: the twelve my tool recommends plus six others I was curious about. Ollama 0.35.0, the default build of each model…

Read the full story at r/LocalLLM ↗

Timeline · 1 report

  1. 2026-10-01 17:23 · r/LocalLLM
    18 open models on 33 real DevOps tasks, one RTX 5090: pass rate, tokens/s and measured VRAM

More stories

  1. Bring near-Astra intelligence to everyday work with GPT-6.1 Sol on Amazon Bedrock — AWS Machine Learning Blog
  2. Gemini 4 Argon: our next era of frontier intelligence — Google Gemini Blog
  3. OpenAI Says It Will Not Release Newest Astra A.I. Model Over Safety Concerns — New York Times Technology
  4. Google Releases New Gemini Model With Guardrails Amid A.I. Safety Debate — New York Times Technology
  5. OpenAI DevDay 2026 Keynote (FULL) — OpenAI YouTube
  6. Introducing Olmo-core 3: Open, scalable training infrastructure for large MoEs — Allen Institute for AI (Ai2)
  7. Introducing dots — OpenAI News
  8. Ollama now supports Jev-style decision models — Ollama Blog

Get the daily brief of stories like this at 6:30 every morning →