AINewsnow

we made Qwen 3.8 27b MLX vision quants and compared them against other popular community publishers (lm-studio, lukaskremla, mlx-community and etc) New Model

This story is from 2026-08-26. It is preserved in the archive; the latest stories are on the live feed.

we made vision mlx quants of qwen3.8 27b (9 builds from 8bit at 29.5 GB down to 3.23bpw DWQ at 11.8 GB) and compared them against other community vision mlx quants from hf (we only compared vision builds) the layout comes out of a clipping search we wrote for mlx and on top of that we wanted to try…

Read the full story at r/huggingface ↗

Timeline · 1 report

  1. 2026-08-26 11:47 · r/huggingface
    we made Qwen 3.8 27b MLX vision quants and compared them against other popular community publishers (lm-studio, lukaskremla, mlx-community and etc) New Model

More stories

  1. Alibaba ships Qwen3.8-Omni-Flash to watch, listen and call tools — r/LocalLLM
  2. Qwen 3.8 27B Running for 63 hours on a RTX 3090 to solve the Riemann hypothesis — r/LocalLLM
  3. Qwen q4 3.8 27b 16 tok/s 32k RTX 3060 :D — r/LocalLLM
  4. 10 hours left fo Qwen Image 2.1 Public Open Source Release — r/StableDiffusion
  5. US government website used Chinese model the FBI called "malicious" — Ars Technica AI
  6. Deployed Qwen 3.6 35B A3B on a single DGX Spark supporting 12 concurrent users at 262K context. Are there better ways to optimize this? — r/LocalLLM
  7. Qwen Developers on X: "Qwen-Image 2.1 is going open source" — r/StableDiffusion
  8. Ternary Bonsai 2 27B — r/LocalLLaMA

Get the daily brief of stories like this at 6:30 every morning →