AINewsnow

Released: Qwen3.8 Flash-Next REAP-384 oQ4e with the full native 512-expert MTP embedded

This story is from 2026-09-15. It is preserved in the archive; the latest stories are on the live feed.

I’ve been working on a slightly different approach to Qwen3.8 Flash-Next REAP builds and finally pushed the model to Hugging Face: https://huggingface.co/mensaprodigy/Qwen3.8-Flash-Next-REAP-384-mlx-mtp The basic idea: Prune the expensive 48-layer target model, but leave the native MTP predictor in…

Read the full story at r/LocalLLM ↗

Timeline · 1 report

  1. 2026-09-15 21:33 · r/LocalLLM
    Released: Qwen3.8 Flash-Next REAP-384 oQ4e with the full native 512-expert MTP embedded

More stories

  1. We’re Not Losing Control of A.I. We’re Giving It Away. — New York Times AI
  2. Deploy Hugging Face models on Amazon SageMaker AI with coding agents — AWS Machine Learning Blog
  3. A quick Minimax H3 news round-up - 17th September 2026 — r/comfyui
  4. Hugging Face Hack Shows Humans Can Keep AI In Check — AI Now Institute
  5. ‘Godfather of AI’ Geoffrey Hinton warns humans running out of time to control Artificial Intelligence – ‘maybe a year’ — Mint AI
  6. Your AI agents are isolated. Your infrastructure isn’t — InfoWorld AI
  7. this looks promising: stepfun-ai/Step-5-Preview-BF16 · Hugging Face — r/LocalLLaMA
  8. What is actually going on with all the recent AI safety / “rogue agent” stories? — r/ArtificialInteligence

Get the daily brief of stories like this at 6:30 every morning →