Video DeltaNet: Hybrid Attention to Speed Up Video Models with Near-Lossless Quality
This story is from 2026-09-04. It is preserved in the archive; the latest stories are on the live feed.
We release VDN-Minimax-H3 ( VDN-H3 ), a hybrid-attention model that generates video faster than it plays, powered by MiniMax H3 . It offers these key features: Fast inference: On 8 B200 GPUs, VDN-H3 generates a 14.4-second clip in 11.23 seconds using 8 denoising steps. Hybrid Architecture: We propo…
Read the full story at r/StableDiffusion ↗
Timeline · 1 report
- 2026-09-04 16:17 · r/StableDiffusion
Video DeltaNet: Hybrid Attention to Speed Up Video Models with Near-Lossless Quality