AINewsnow

Update #4: Post training yandex/AliceAI-80B-A3B [instruct!] from scratch

Last update for those following: https://www.reddit.com/r/LocalLLaMA/comments/1wvyc3e/update_3_post_training_yandexaliceai80ba3b/ Project in a sentence: An instruct finetune of ALiceAI-80B-A3B-Base capable of agentic work and conversation. I'm creating a shallow distill of qwen 3.8 27b on medium to…

Read the full story at r/LocalLLaMA ↗

Timeline · 1 report

  1. 2026-10-04 17:55 · r/LocalLLaMA
    Update #4: Post training yandex/AliceAI-80B-A3B [instruct!] from scratch

More stories

  1. Strata is seriously impressive, running Qwen 3.8 Flash Next on hermes at 512k context. — r/LocalLLM
  2. Qwen3.8-Flash-Next 177B running at 11–15 tok/s on a single RTX 5070 12GB + 32GB RAM DDR4 — r/LocalLLaMA
  3. One .char model, Consistent face, body & cloths, now works in Comfy(Custom node & workflows) MinimaxH3 & Flux2 — r/comfyui
  4. llama, server: add /v1/systemone API (models: laya, julia-1, lev, openjev, kev) by ngxson · Pull Request #29818 · ggml-org/llama.cpp — r/LocalLLaMA
  5. Is all the work that's being put into Qwen3.8 Flash Next going to set us up for a very quick uplift to Qwen4? — r/LocalLLaMA
  6. ComfyUI Qwen image 2.1 Enhancer (Two nodes) — r/StableDiffusion
  7. I built Ninfer 4080 for 16GB class GPUs — r/LocalLLaMA
  8. How I use a local Qwen 27B for real work on home hardware (unscripted workflow demo) — r/LocalLLM

Get the daily brief of stories like this at 6:30 every morning →