AINewsnow

Opus 5.5 (high) improves on Opus 5 (high) 3.5 → 3.8 on the Short-Story Creative Writing Benchmark, just behind Fable 5.1 (high) and Opus 5 (xhigh).

https://github.com/lechmazur/writing/ Grok 4.7 (high) makes a substantial jump over Grok 4.6 (high): −2.8 → 0.4. MiMo V2.6 Pro (thinking) improves sharply over V2.5 Pro: −0.7 → 1.6. Gemini 3.8 Flash (high) advances over Gemini 3.7 Flash (high): −0.7 → 0.2. DeepSeek V4.1 Flash (high) enters at −0.5.…

Read the full story at r/singularity ↗

Timeline · 1 report

  1. 2026-09-28 19:22 · r/singularity
    Opus 5.5 (high) improves on Opus 5 (high) 3.5 → 3.8 on the Short-Story Creative Writing Benchmark, just behind Fable 5.1 (high) and Opus 5 (xhigh).

More stories

  1. We added the "Vibe" into vibecoding - Introducing the first Malleable AI workstation — r/AI_Agents
  2. Why are Gemini's answers outdated? — r/Bard
  3. Which is best local AI tools that can access all like Chatgpt, DeepSeek, Grok etc — r/huggingface
  4. Gemini 3.8 flash VS DeepSeek V4.1 — r/GeminiAI
  5. One key for claude, gpt, gemini, and deepseek in my coding tools — r/ChatGPTCoding
  6. Gemini Flash 3.8 Better than 3.6 for RP/DND Solo? — r/GeminiAI
  7. Increased Hallucinations Lately? — r/GeminiAI
  8. PSA: Dual 3090 - Qwen Flash Next - 80tps/2k+ prefill — r/LocalLLM

Get the daily brief of stories like this at 6:30 every morning →