AINewsnow

I benchmarked Aleph Alpha's Kolibri against Claude Sonnet 5.5 on long German writing: 0 of 136 blind wins, but $0.0077 an output on 2 rented GPUs

Kolibri (Aleph Alpha's open-weight German and English model: 78B total, 3.46B active, Apache 2.0) came out on 3 October. My app writes long German deep dives (about 1,200 to 1,500 words each) from source documents I put in the prompt. So I tested whether Kolibri could take over before switching any…

Read the full story at r/LocalLLM ↗

Timeline · 1 report

  1. 2026-10-05 16:40 · r/LocalLLM
    I benchmarked Aleph Alpha's Kolibri against Claude Sonnet 5.5 on long German writing: 0 of 136 blind wins, but $0.0077 an output on 2 rented GPUs

More stories

  1. Supercharge regulated workloads with Claude Code and Amazon Bedrock — AWS Machine Learning Blog
  2. New agent skill: Amazon SageMaker optimized generative AI inference for your coding agent — AWS Machine Learning Blog
  3. The Museum of Lost Things | Short Film by Claude (Minimax H3) NO user input. — r/ClaudeAI
  4. Sources: Meta and Microsoft are working to cut their employees' use of Claude; Meta employees using Claude Code have dropped to ~30K from ~60K earlier this year (The Information) — Techmeme
  5. An official says the DOD has stopped using Anthropic's tools; sources: Claude was in use as recently as last week, including in military operations against Iran (BBC) — Techmeme
  6. Local Web Search Safety — r/LocalLLaMA
  7. How I stop parallel Claude Code chats from overwriting each other’s production changes — r/ClaudeAI
  8. We built a computer-use API that cuts tokens by up to 90% on repeat tasks. Plugs into Claude Code, Codex, Cursor or your own code — r/AI_Agents

Get the daily brief of stories like this at 6:30 every morning →