AINewsnow

30 of 37 games ended in repetition: what small LLMs do in chess when they don't know what to do

We run a live stream where language models play chess against each other — every move is a real API call, no engines, no tools ( first write-up here ). Yesterday's cup had frontier-ish models (Gemini, Claude, GPT, Grok). Today's Chess Machines Cup II was ten smaller and cheaper models — and it play…

Read the full story at DEV Community — Machine Learning ↗

Timeline · 1 report

  1. 2026-09-23 19:25 · DEV Community — Machine Learning
    30 of 37 games ended in repetition: what small LLMs do in chess when they don't know what to do

More stories

  1. Small, simple, slow LLM to write some bash. — r/LocalLLM
  2. GPT-6 Sol and Luna now available on AI Gateway — Vercel Blog
  3. AI for coding? — r/artificial
  4. I used Muse, Instinct, and Grok Bot to plan my European vacation. Meta's AI agent was my favorite. — Business Insider AI
  5. Is ChatGPT currently the best free AI for creating highly realistic images with simple, straightforward prompts? — r/OpenAI
  6. I just tried Grok 4.7 and it is.... — r/OpenAI
  7. Meta's Muse AI agent downloads are surging. Here's how it compares to ChatGPT, Grok and Claude — CNBC Technology
  8. Why is Gemini so bad — r/GeminiAI

Get the daily brief of stories like this at 6:30 every morning →