AINewsnow

yandex/AliceAI-Foundation-80B-A3B-Base: Russian-developed competitor to Qwen 35B and DeepSeek V4 Flash

https://huggingface.co/yandex/AliceAI-Foundation-80B-A3B-Base It's not a Qwen3 finetune, it's actually its own fully custom architecture. No Llama.cpp support yet sadly (Also note that this model is NOT post-trained like Qwen3.5/3.6)

Read the full story at r/LocalLLaMA ↗

Timeline · 3 reports

  1. 2026-09-21 19:26 · r/artificial
    Benchmarks Grok 4.7, GPT 6 Astra Fable 4.1 and DeepSeek V4.1 Flash
  2. 2026-09-21 19:25 · r/LocalLLM
    yandex/AliceAI-Foundation-80B-A3B-Base: Russian-developed competitor to Qwen 35B and DeepSeek V4 Flash
  3. 2026-09-21 19:24 · r/LocalLLaMA
    yandex/AliceAI-Foundation-80B-A3B-Base: Russian-developed competitor to Qwen 35B and DeepSeek V4 Flash

More stories

  1. My contribution to the local AI community: 9 abliterated models, 99 GGUF quantizations in progress — r/huggingface
  2. I gave 6 different AIs the same 5 questions — r/AI_Agents
  3. XiaomiMiMo/MiMo-V2.6-Pro-RL · Hugging Face — r/LocalLLaMA
  4. GPT image 2.5 vs Nano Banana pro vs Nano Banana 2 vs Qwen image 3 vs Seedream 5.0 pro — r/GeminiAI
  5. Success running Qwen 3.8 27B EXL3 on RTX 3060 + 5060 Ti — r/LocalLLM
  6. The bear can dance: Qwen 3.8 27B on one 3090 for 3 weeks — r/LocalLLaMA
  7. Dual B60 24GB Performance — r/LocalLLM
  8. Is llama.cpp meant to be slow at long context, even when you aren't using that context? — r/LocalLLaMA

Get the daily brief of stories like this at 6:30 every morning →