Here's a little short tester I made using MiniMax H3, Yue2 for the music. Three 10s clips, first was t2va then rest are ref2va using clips and the voice from previous gens. A little post work for making the dogs barks sound the same, removing the original generated music and isolating vocals.
0.6mp, lcm/beta57, 4 steps with the DMAD 4 step lora
Read the full story at r/StableDiffusion ↗
Timeline · 1 report
- 2026-10-07 02:51 · r/StableDiffusion
Here's a little short tester I made using MiniMax H3, Yue2 for the music. Three 10s clips, first was t2va then rest are ref2va using clips and the voice from previous gens. A little post work for making the dogs barks sound the same, removing the original generated music and isolating vocals.
More stories
- I made a free all-in-one LoRA trainer for consumer GPUs (Windows + Linux): Qwen-Image 2.1, FLUX.2 Klein 9B, Krea 2, Z-Image, Ideogram 4, Anima, SDXL/Pony/Illustrious, LTX 2.3 and MiniMax-H3 (video + audio) — from 4-8 GB VRAM — r/StableDiffusion
- Anthropic Subscriptions Offer 5x+ More Value Than OpenAI — SemiAnalysis
- [MiniMax H3 / ComfyUI] How long does it take you to generate a 17-second video at 720p / 14 steps? — r/comfyui
- A quick Minimax H3 news round-up - 4th October 2026 — r/comfyui
- FastVideo’s FastH3 now runs on a single consumer machine — r/StableDiffusion
- AND HOW DOES THAT MAKE YOU FEEL? | An AI Short Comedy Film Made by Claude in Minimax H3 and my Video builder in ComfyUI. — r/StableDiffusion
- MiniMax-M2 (230B) running from disk on a 32 GB laptop, CPU only — r/StableDiffusion
- Prism (Tencent Hunyuan + Fudan) just dropped a preview: native 2K video + audio, MIT license — r/StableDiffusion
Get the daily brief of stories like this at 6:30 every morning →