AINewsnow

Agens Volundr 32B Preview: our small team's first model on our own hybrid architecture. Only 18 of 72 layers keep a KV cache (Apache-2.0)

Hi r/LocalLLaMA . I'm on the team at Blockway, a small team in Hong Kong (disclosure: this is our model). Today we released Agens Volundr 32B Preview, the first model built on our own hybrid architecture. We trained it on limited compute, it isn't perfect, and we'd rather tell you where it falls sh…

Read the full story at r/LocalLLaMA ↗

Timeline · 2 reports

  1. 2026-10-05 13:01 · r/LocalLLM
    New Model: Agens Volundr 32B Preview: our small team's first model on our own hybrid architecture. Only 18 of 72 layers keep a KV cache (Apache-2.0)
  2. 2026-10-05 12:58 · r/LocalLLaMA
    Agens Volundr 32B Preview: our small team's first model on our own hybrid architecture. Only 18 of 72 layers keep a KV cache (Apache-2.0)

More stories

  1. Trump’s big AI move: ‘Super Intelligence Force’ launched, Jay Clayton named AI czar — Mint AI
  2. OpenAI safety employee resigns, claiming the company’s ‘culture is broken’ — TechCrunch AI
  3. Sam Altman to Decoded: ‘The world should accept some bad things happening’ for the benefits of AI — Politico Technology
  4. Introducing Oscilloscope Diffusion — r/comfyui
  5. Aleph-Alpha/Kolibri-1 · Hugging Face - 78B parameters. 3.46B active. Up to 1M tokens of context - Apache 2.0 — r/LocalLLaMA
  6. Strata is seriously impressive, running Qwen 3.8 Flash Next on hermes at 512k context. — r/LocalLLM
  7. The Story of Qwen: Alibaba's AI Models From 7B to 2.4T — MarkTechPost
  8. Everything we launched during Birthday Week 2026 — Cloudflare Blog — AI

Get the daily brief of stories like this at 6:30 every morning →