AINewsnow

Update #5: Post training yandex/AliceAI-80B-A3B [instruct!] from scratch

Last update for those following: https://www.reddit.com/r/LocalLLaMA/comments/1wxlytt/comment/pdz728b/?screen_view_count=1 Project in a sentence: An instruct finetune of ALiceAI-80B-A3B-Base capable of agentic work and conversation. I'm creating a shallow distill of qwen 3.8 27b on medium to teach…

Read the full story at r/LocalLLaMA ↗

Timeline · 1 report

  1. 2026-10-05 22:10 · r/LocalLLaMA
    Update #5: Post training yandex/AliceAI-80B-A3B [instruct!] from scratch

More stories

  1. Qwen Flash Next on Single B200 or B300, any pointers ? — r/LocalLLM
  2. I mapped every major Qwen release from 2023 to 2026: 44 models, from Qwen-7B to the 2.4T open weights (with sources) — r/machinelearningnews
  3. Qwen3.8-Flash-Next-Q8_0 running on a V100 @ 130Watts 32GB Vram and 128GB System Ram — r/LocalLLM
  4. Qwen Image 2.1 Uncensored MCP — r/StableDiffusion
  5. A benchmark for LLMs playing Civilization V. GLM-5.3 is ahead of Opus-5.5, and Qwen-3.8-27B holds up surprisingly well. — r/LocalLLaMA
  6. VNCCS 3.2.0 Released with Qwen Image 2.1 and MiniMax H3 support! — r/StableDiffusion
  7. Reflection AI Is About to Release a US Open-Weight Model to Take On DeepSeek and Qwen — r/LocalLLaMA
  8. 8GB VRAM -> ??? — r/LocalLLM

Get the daily brief of stories like this at 6:30 every morning →