AINewsnow

Searching for 3.8 35B: Qwen3.6-35B-A3B (Testing 5 Finetunes vs. Base)

TL;DR -- You should probably just use base Qwen3.6-35B, as only Occamy-1.0 is competitive with it. Tiel is a major let-down, worse than Ornith. KAT surprises (good), Nex surprises (bad). This post is long. Sorry, lots to cover. I think we all want to see a next-generation small MoE from the Qwen te…

Read the full story at r/LocalLLaMA ↗

Timeline · 1 report

  1. 2026-09-28 21:52 · r/LocalLLaMA
    Searching for 3.8 35B: Qwen3.6-35B-A3B (Testing 5 Finetunes vs. Base)

More stories

  1. PSA: Dual 3090 - Qwen Flash Next - 80tps/2k+ prefill — r/LocalLLM
  2. Verzeta Studio: an open source desktop app where several local models work-together as a team in one conversation — r/LocalLLM
  3. Qwen 3.8 27B vs Qwen 3.8 Flash Next and time to complete a coding task. — r/LocalLLaMA
  4. Qwen-Image 2.1 Inpainting with LanPaint — alpha channel included — r/StableDiffusion
  5. Layer Extract & Layer Remove Loras For Qwen Image 2.1 — r/StableDiffusion
  6. 85 GB DeepSeek-V4-Flash at ~3 tok/s on a 12 GB RTX 3060 + 64 GB DDR5 RAM - Overspill for FreeToken, inspired by Colibri — r/LocalLLaMA
  7. A LoRA I made: AnyAngle LoRA for Qwen Image 2.1. Style-Aligned Arbitrary Camera Angles — r/StableDiffusion
  8. 2x Tesla P100, q6_k quant 50+tps. V2.0 — r/LocalLLM

Get the daily brief of stories like this at 6:30 every morning →