30 of 37 games ended in repetition: what small LLMs do in chess when they don't know what to do
We run a live stream where language models play chess against each other — every move is a real API call, no engines, no tools ( first write-up here ). Yesterday's cup had frontier-ish models (Gemini, Claude, GPT, Grok). Today's Chess Machines Cup II was ten smaller and cheaper models — and it play…
Read the full story at DEV Community — Machine Learning ↗
Timeline · 1 report
- 2026-09-23 19:25 · DEV Community — Machine Learning
30 of 37 games ended in repetition: what small LLMs do in chess when they don't know what to do
More stories
- Small, simple, slow LLM to write some bash. — r/LocalLLM
- GPT-6 Sol and Luna now available on AI Gateway — Vercel Blog
- AI for coding? — r/artificial
- I used Muse, Instinct, and Grok Bot to plan my European vacation. Meta's AI agent was my favorite. — Business Insider AI
- Is ChatGPT currently the best free AI for creating highly realistic images with simple, straightforward prompts? — r/OpenAI
- I just tried Grok 4.7 and it is.... — r/OpenAI
- Meta's Muse AI agent downloads are surging. Here's how it compares to ChatGPT, Grok and Claude — CNBC Technology
- Why is Gemini so bad — r/GeminiAI
Get the daily brief of stories like this at 6:30 every morning →