Models & tools · 7 Oct 2026 · 20:46 CEST
Vidu Releases Q4 Preview of Next-Generation Flagship AI Video Model

Publisher preview · OZZZER analysis pending editorial review.
ShengShu Technology launched Vidu Q4 Preview on October 7, 2026, the first public preview of its next-generation flagship model for expressive audio and video. Launch pricing starts at $0.014 per second, and the preview is available through Vidu’s web product and API platform. The model supports up to 15 image references and up to three audio references, with output at 2K or 4K resolution and 10-bit color depth.
The company said the preview targets independent creators, small studios and production teams spanning both narrative and commercial work. ShengShu Technology said Q4 Preview is built to better coordinate facial expression, emotion, body movement and voice, yielding subtler expressions, clearer emotional reads and more natural movement. As many as three reference audio clips help keep a character’s voice consistent and its delivery emotionally aligned, while up to 15 reference images can pin down characters, wardrobe, props, products and environments within a single creative setup.
The company said these controls are meant to hold visual and vocal choices together across a scene and to give narrative projects more deliberate control over staging and…
Excerpt supplied by the publisher.
Source
Unite.AI · 7 Oct 2026 · 20:46 CEST
Open the original at Unite.AI ↗