AINewsnow

Worth going from Qwen3.8 27B to flash next or maaybe deepseek v4 flash?

I am running qwen3.8 27b on my dual rtx 3090 (fp8 quant, unquantized cache, 129k context) and I think it works decently well with hermes, opencode etc. But! I am tempted by the new models coming out such as qwen3.8 flash next, deepseek v4 flash, glm 5.3 flash. However, there is a big jump in vram a…

Read the full story at r/LocalLLaMA ↗

Timeline · 2 reports

  1. 2026-09-27 13:44 · r/LocalLLM
    Worth going from Qwen3.8 27B to flash next or maaybe deepseek v4 flash?
  2. 2026-09-27 13:44 · r/LocalLLaMA
    Worth going from Qwen3.8 27B to flash next or maaybe deepseek v4 flash?

More stories

  1. Is Qwen Flash Next at like Q2 better than 27B at Q4? — r/LocalLLaMA
  2. Another "Harness matters" post (codex cli > pi and opencode) — r/LocalLLaMA
  3. How do you guys give your models web browsing capabilities? — r/LocalLLaMA
  4. Which is best local AI tools that can access all like Chatgpt, DeepSeek, Grok etc — r/huggingface
  5. Gemini 3.8 flash VS DeepSeek V4.1 — r/GeminiAI
  6. Nonobench v1.2: 43 LLMs on nonogram puzzles. Open-weight DeepSeek V4 Pro ties for 4th, and no open model solves the new 20×20 Hard mode — r/LocalLLaMA
  7. Which provider actually wins on pure affordability right now for gemma qwen gpt oss and deepseek under one roof — r/AI_Agents
  8. JiRackUltra_1b Runs AI Routing on Any Laptop Without a GPU — AlphaSignal

Get the daily brief of stories like this at 6:30 every morning →