Image-to-video in three parameters: how a 5-second clip costs twice what it should
This story is from 2026-09-22. It is preserved in the archive; the latest stories are on the live feed.
一张 1024×768 的室内效果图,一句运动提示词,五秒后拿到一段可以发出去的镜头。这是图生视频里最省事的一档,也是我们六月份给渲见(RenVi)做营销素材时走的那条路。 结论先放这里: 模型本身没坑,坑在两个参数的默认值上。 我们实际跑出来的东西 项目 实测值 模型 wan2.6-i2v-flash 输入 一张效果图,以 data URI 内联 请求参数 resolution: 720P 、 duration: 5 输出 1108×830、5.007 秒、H.264、约 5.5 MB 附带音轨 有,AAC 立体声 44.1 kHz 单条成本 ¥1.50 (本该是 ¥0.75) 那个 AA…
Read the full story at DEV Community — AI ↗
Timeline · 1 report
- 2026-09-22 10:09 · DEV Community — AI
Image-to-video in three parameters: how a 5-second clip costs twice what it should