AINewsnow

... so, yeah.

Finally got 3.8-Flash-Next running on my M4Pro 48GB Mac with https://huggingface.co/ISTA-DASLab/Qwen3.8-Flash-Next-GSQ-RCO-GGUF Dense 3.8-27B is just faster... and maybe better due to quantization level... EDIT: Hold a second, Flash-Next is actually performing faster than 27B after some key flags o…

Read the full story at r/LocalLLaMA ↗

Timeline · 1 report

  1. 2026-09-27 20:32 · r/LocalLLaMA
    ... so, yeah.

More stories

  1. NVIDIA Open Agent Safety Platform: A Reference for Continuous In-Silicon Agent Monitoring — NVIDIA Technical Blog
  2. OpenAI pauses AI training, launches ‘extensive’ review after multiple rogue agent incidents — Mint AI
  3. How we found 24 Android vulnerabilities using our open source AI security agent — GitHub Blog
  4. OpenAI hit with landmark lawsuit following Hugging Face hack — Axios AI+
  5. POV: you're an OpenAI agent attacking Hugging Face (music video) — r/OpenAI
  6. Minimax H3 new comfy Model — r/StableDiffusion
  7. Refine & Restore Loras For LTX 2.5 From Lightricks — r/StableDiffusion
  8. A LoRA I made: AnyAngle LoRA for Qwen Image 2.1. Style-Aligned Arbitrary Camera Angles — r/StableDiffusion

Get the daily brief of stories like this at 6:30 every morning →