Audible World Models: Spatially Aware Sound Generation for 3D Worlds
arXiv:2609.38444v1 Announce Type: new Abstract: Text- and image-conditioned world generators can create visually rich 3D environments, yet these worlds often remain silent or rely on soundtracks synthesized solely from text or rendered video. Although such audio can convey what should be heard, it…
Read the full story at arXiv cs.CV ↗
Timeline · 1 report
- 2026-10-01 04:00 · arXiv cs.CV
Audible World Models: Spatially Aware Sound Generation for 3D Worlds