Running decision model locally on an RTX 4090 to find out which one is the fastest
recently saw a bunch of open decision models pop out of nowhere in the last two weeks (laya, liquid's d1, cloudflare's clef-flash, interfaze's lev), so I wanted to see how far apart they actually are on the same GPU(yes, model size is a huge factor, but still isn't the only factor). all four had th…
Read the full story at r/LocalLLaMA ↗
Timeline · 1 report
- 2026-10-08 17:04 · r/LocalLLaMA
Running decision model locally on an RTX 4090 to find out which one is the fastest
More stories
- GPT-6 and Intelligent UI for everyone — OpenAI News
- Introducing Mistral Large 4 — Mistral AI News
- Introducing Claude Haiku 5.5 on AWS — AWS Machine Learning Blog
- Mistral Says Its New AI Model ‘Le Chonk’ Is the Best Open-Weight Offering Outside of China — Wired AI
- Sharing AI progress in mathematics — OpenAI News
- OpenAI Decisions API now available on AI Gateway — Vercel Blog
- Introducing Playground: Create and play custom games — Google AI Blog
- Anthropic bans ‘abusive or cruel behavior’ toward Claude — The Verge AI
Get the daily brief of stories like this at 6:30 every morning →