AINewsnow

Strata is seriously impressive, running Qwen 3.8 Flash Next on hermes at 512k context.

I've been messing with Qwen 3.8 Flash Next for a while on my local inference server and stumbled onto Strata, saw a small headline/article, didn't make much of it, passed on reading it fully. Then again somewhere else, then once more... So I gave it a quick read, didn't quite believe it, today I fi…

Read the full story at r/LocalLLM ↗

Timeline · 5 reports

  1. 2026-10-04 02:48 · r/LocalLLM
    Anyone tried strata qwen.38 flash next on a RX6700XT?
  2. 2026-10-03 23:44 · r/LocalLLaMA
    Strata Qwen 3.8 flash next is the biggest thing since the release of Qwen 3.8 27b
  3. 2026-10-03 19:51 · r/LocalLLM
    I would like to run Qwen 3.8 Flash Next.
  4. 2026-10-03 15:54 · r/LocalLLM
    can i run qwen flash next with these specs, or am i out of luck?
  5. 2026-10-03 01:47 · r/LocalLLM
    Strata is seriously impressive, running Qwen 3.8 Flash Next on hermes at 512k context.

More stories

  1. Local model for Hermes on a 48GB M5 Pro Mac mini that’s also my daily dev machine? — r/LocalLLM
  2. Pi extension: Skip reasoning with local Qwen 27B and proceed to answer right now — r/LocalLLaMA
  3. Qwen4Exp: add MTP by am17an · Pull Request #29761 · ggml-org/llama.cpp — r/LocalLLaMA
  4. ComfyUI Tutorial Qwen Image 2 1 vs FLUX 2 Krea Turbo Which Makes Better Character Sheets — r/comfyui
  5. Is all the work that's being put into Qwen3.8 Flash Next going to set us up for a very quick uplift to Qwen4? — r/LocalLLaMA
  6. I built Ninfer 4080 for 16GB class GPUs — r/LocalLLaMA
  7. Direct weight surgery from Qwen-4B to 0.8B on an 8GB RX 580: why editing all layers breaks everything, and how 4 anchor blocks fixed it — r/machinelearningnews
  8. I made my iPhone a second GPU for my 24 GB MacBook: Qwen 3.8 27B prefills 29–44% faster & my holds part of the CTX window. — r/LocalLLaMA

Get the daily brief of stories like this at 6:30 every morning →