AINewsnow

MiniMax H3 native 720p→1440p on one RTX 4090: 112s / 223s / 334s with auto-scheduled sparse attention

This story is from 2026-08-29. It is preserved in the archive; the latest stories are on the live feed.

Hi everyone — I’m an independent developer experimenting with making MiniMax H3 more practical on consumer NVIDIA GPUs. I built an automatically scheduled sparse-attention system for X-MinimaxH3 and tested native H3 second sampling from 720p to 1440p on a single RTX 4090. Measured second-sampling t…

Read the full story at r/StableDiffusion ↗

Timeline · 1 report

  1. 2026-08-29 17:55 · r/StableDiffusion
    MiniMax H3 native 720p→1440p on one RTX 4090: 112s / 223s / 334s with auto-scheduled sparse attention

More stories

  1. Will this make YouTube compression reducing the quality of the video? — r/comfyui
  2. NVIDIA CEO Jensen Huang rejects ‘AI will end the world’ claim, yet cautions ‘we should go as fast as we can but...’ — Mint AI
  3. Building an open-source 500+ language Sparse MoE translation model from scratch (Apache 2.0) — r/huggingface
  4. Follow-up: making a quieter TNG scene with MiniMax H3 in ComfyUI, and why I had to regenerate the whole thing at 1MP — r/StableDiffusion
  5. what's the state of the art recipe for running Qwen3.8-Flash-Next with a pair of 3090s and a ton of system RAM rn? — r/LocalLLaMA
  6. Huawei details AI accelerator roadmap, pulls in next-generation Ascend NPUs by several quarters — FP4 performance of the Ascend 960PR doubles expectations — Tom's Hardware
  7. Flyweight: open-source C++/CUDA engine for running MoE models bigger than your VRAM on one GPU + system RAM. First PyPI release, looking for contributors. — r/LocalLLaMA
  8. A quick Minimax H3 news round-up - 18th September 2026 — r/comfyui

Get the daily brief of stories like this at 6:30 every morning →