AINewsnow

De los exámenes al trabajo real: un modelo abierto de 180B lidera 10 clasificaciones oficiales de Hugging Face

This story is from 2026-10-06. It is preserved in the archive; the latest stories are on the live feed.

TL;DR Un modelo de código abierto, Darwin-180B-RSI , ocupa ahora el puesto #1 en 10 de los 48 benchmarks oficiales de Hugging Face , el mayor número entre las 95 organizaciones que compiten. La segunda posición es Z.ai, con 4. El décimo primer puesto es MDPBench (parsing de documentos multilingües,…

Read the full story at DEV Community — AI ↗

Timeline · 1 report

  1. 2026-10-06 00:43 · DEV Community — AI
    De los exámenes al trabajo real: un modelo abierto de 180B lidera 10 clasificaciones oficiales de Hugging Face

More stories

  1. Can an Open Model Do Security Research? Cantina's apex-flash-1 Solves 40 of 60 Held-Out Bug Tasks — MarkTechPost
  2. Hugging Face Pulls GLM-5.3 Build Made for Cyberattacks — r/ArtificialInteligence
  3. Introducing GLM 5.3 on Amazon Bedrock — AWS Machine Learning Blog
  4. Aleph-Alpha/Kolibri-1 · Hugging Face - 78B parameters. 3.46B active. Up to 1M tokens of context - Apache 2.0 — r/LocalLLaMA
  5. Aleph Alpha releases open-weight Kolibri with 1M context — TestingCatalog AI News
  6. Qwen3.8-Flash-Next 177B running at 11–15 tok/s on a single RTX 5070 12GB + 32GB RAM DDR4 — r/LocalLLaMA
  7. Analysis: Anthropic's subscriptions offer ~5x more API-equivalent value per month than OpenAI's for agentic workloads with Claude Opus 5.5 vs. GPT-6.1 Sol (SemiAnalysis) — Techmeme
  8. World Models: The Simulation Strikes Back — r/computervision

Get the daily brief of stories like this at 6:30 every morning →