AINewsnow

Qwen 3.8 27B 4bit, prompting H3 MiniMax

Qwen3.8-27B-oQ4e-mtp running on oMLX on a Mac Mini M5 Pro. Avg ~30 tok/sec. H3 MiniMax approximately 15 minutes on a RTX 3090 MiniMaxAI/MiniMax-H3 t2va lightx2v/Minimax-h3-Turbo Input prompt : "make a video of two pieces of paper ona desk that turn into origami figures that then fight eaach other,…

Read the full story at r/StableDiffusion ↗

Timeline · 1 report

  1. 2026-10-10 00:07 · r/StableDiffusion
    Qwen 3.8 27B 4bit, prompting H3 MiniMax

More stories

  1. A Very Strange GPU Stall Situation — r/comfyui
  2. Qwen Image 2.1 Turbo Released -- Hugging Face — r/StableDiffusion
  3. Release: Qwen-2B-RCOL Dynamic Low-Bit Quantization (IQ1_M, IQ2_M, IQ3_M) — r/LocalLLaMA
  4. Qwen 3.8 Flash-Next on 64GB RAM + 16GB VRAM, worth it over a fully-loaded 3.6 35B-A3B? — r/LocalLLaMA
  5. China’s open-weight AI models are winning global users. Who is capturing the value? — South China Morning Post Tech
  6. Best uncensored version of Qwen3.8-Flash-Next-GSQ-RCO-IQ3_S ? — r/LocalLLM
  7. Hunyuan image 3 vs qwen 2.1 — r/StableDiffusion
  8. feat: add GLM5Next MTP, optimize by pwilkin · Pull Request #29928 · ggml-org/llama.cpp — r/LocalLLaMA

Get the daily brief of stories like this at 6:30 every morning →