AINewsnow

Qwen-Audio-3.1-TTS-Next: one call per sound scene, and a silent 120-second cut

This story is from 2026-09-28. It is preserved in the archive; the latest stories are on the live feed.

One call turned a 298-character script into 32 seconds of finished radio drama: three speaking characters, rain on an awning, a kettle in the kitchen, sequenced in the order I wrote them. That part is real, and on headphones it holds up. What no demo shows you is a wall at 120 seconds. Go past it a…

Read the full story at DEV Community — AI ↗

Timeline · 1 report

  1. 2026-09-28 07:08 · DEV Community — AI
    Qwen-Audio-3.1-TTS-Next: one call per sound scene, and a silent 120-second cut

More stories

  1. PSA: Dual 3090 - Qwen Flash Next - 80tps/2k+ prefill — r/LocalLLM
  2. Verzeta Studio: an open source desktop app where several local models work-together as a team in one conversation — r/LocalLLM
  3. 2x Tesla P100, q6_k quant 50+tps. V2.0 — r/LocalLLM
  4. Another "Harness matters" post (codex cli > pi and opencode) — r/LocalLLaMA
  5. Qwen, where's the small stuff? (1B/2B/4B) — r/LocalLLaMA
  6. I added Qwen-Image 2.1 + LoRA support to TensorSharp (GGUF, local inference) — r/LocalLLaMA
  7. Qwen 2.1 Might Be Just TOO Good at Face Swap... [Free Workflow] — r/StableDiffusion
  8. Character Design Sheet V2.0 Update: A Practical Approach to Character Sheet Generation. — r/StableDiffusion

Get the daily brief of stories like this at 6:30 every morning →