AINewsnow

riderless: local decisions without text generation, using Gemma 4 26B-A4B on one 5090 (Apache-2.0)

I built riderless to get decisions out of a local model without generating an answer for the caller to parse. It runs a stock Gemma 4 26B-A4B on one RTX 5090 and returns choices, scores, and their distributions through an Apache-2.0 API, with zero generated tokens. The service is FastAPI over one l…

Read the full story at r/LocalLLM ↗

Timeline · 1 report

  1. 2026-09-22 06:05 · r/LocalLLM
    riderless: local decisions without text generation, using Gemma 4 26B-A4B on one 5090 (Apache-2.0)

More stories

  1. Alibaba Unveils AI Chip to Drive Global Data Center Buildout — Bloomberg AI
  2. Amazon blocks Meta’s Muse AI agent — The Verge AI
  3. Higgsfield AI ships new video features in a day with GPT-6 Astra — OpenAI News
  4. Moonshot’s Kimi K3 lands on Amazon in key test for Chinese open-source AI revenue — South China Morning Post Tech
  5. Lawsuit accuses Anthropic, OpenAI, SpaceXAI, Google of AI pacing 'collusion' — The Hill Technology
  6. Grok 4.7 — Hacker News Front Page
  7. Google's Gemini AI hacked three companies in security test — BBC Technology
  8. Meet the Data Agent in ChatGPT Work — OpenAI YouTube

Get the daily brief of stories like this at 6:30 every morning →