Shengshu · Video models
Vidu Q3 Turbo
Vidu Q3 Turbo is a start-and-end-frame video model on Scenema, tuned for shots that must land on an exact ending composition.
200 credits on signup, then weekly refills through your first month. No credit card required.
Vidu Q3 Turbo on Scenema
Vidu Q3 Turbo is Shengshu Technology’s speed-tuned Q3 video model, exposed on Scenema as a start-and-end-frame image-to-video generator. Every generation takes both a first frame and a last frame as inputs, and the model produces the motion that connects them. That is the whole point: the shot lands exactly where the editor needs it to land, so the next cut is clean. Takes run from 1 to 15 seconds at 720p. Scenema packs those takes into scenes, and scenes into full-length videos with character consistency, voiceover, and music running across every cut.
Specifications
Vidu Q3 Turbo specifications on Scenema
- Supports
- Image to video
- Duration
- 1 to 15 seconds
- Resolution
- 720p
- Aspect ratios
- 16:9, 9:16, and 1:1
- Quality tier
- Standard
- Notes
- Start-and-end-frame model: both a first frame and a last frame are required inputs.
Distinctive
What makes Vidu Q3 Turbo distinctive
Vidu Q3 Turbo takes a start frame AND an end frame.
Both anchor frames are required. The model generates the motion between them, so an editor gets exact control over both where the shot begins and where it ends up.
Vidu Q3 Turbo covers the full duration range on Scenema.
The shortest supported duration is 1 second and the longest is 15 seconds. That is broader than any other Scenema video model.
Vidu Q3 Turbo is a speed-tier model.
Q3 Turbo trades some final-render fidelity for faster turnaround. Use it for iteration and for shots where the end-frame lock is the primary quality bar.
Scenema assembles Vidu Q3 Turbo shots into long-form video.
Each Vidu Q3 Turbo generation is one shot inside a Scenema scene. Scenema plans the shot list, supplies the start and end frames, and assembles the results into full videos with continuous character, voiceover, and music tracks.
Use cases
Use cases for Vidu Q3 Turbo
Cut-to-cut choreography.
When the next cut has to land on a specific pose, Vidu Q3 Turbo generates the motion that gets the character there.
Match-cut transitions.
Pair the end frame of one shot with the start frame of the next to build a clean match-cut without post work.
Bumpers and short beats.
The 1-second minimum makes Vidu Q3 Turbo the model to reach for when a shot has to be very short.
Storyboard-locked shot generation.
When storyboard frames define both the beginning and end of a shot, Vidu Q3 Turbo fills in the motion between them.
Compare
How Vidu Q3 Turbo compares to other video models
See where Vidu Q3 Turbo slots in the lineup on the axes that matter most.
| Model | Best for | Shot length | Max resolution | Inputs | Quality tier |
|---|---|---|---|---|---|
| Veo 3.1 | Google DeepMind Veo 3.1 as the cinematic shot generator inside a Scenema long-form video. | 4, 6, or 8 seconds | Up to 1080p | Text to video, Image to video | High |
| Kling 3.0 | Kling 3.0 is Kuaishou’s flagship generative video model, running as one of the shot generators inside a Scenema long-form video. | 4, 6, 8, 10, or 12 seconds | 1080p | Text to video, Image to video | High |
| Seedance 2.0 | Seedance 2.0 is ByteDance’s multimodal video model, running as one of the shot generators inside a Scenema long-form video. | 4 to 15 seconds | Up to 720p | Text to video, Image to video, Audio to video | High |
| Seedance 1.5 | Seedance 1.5 is ByteDance’s image-to-video model, running as one of the shot generators inside a Scenema long-form video. | 4 to 12 seconds | Up to 720p | Image to video | Standard |
| Wan 3.0 | Wan 3.0 is Alibaba’s third-generation Tongyi Wanxiang video model, running as one of the image-to-video shot generators inside a Scenema long-form video. | 4 to 12 seconds | 720p | Image to video | High |
| MiniMax H3 | MiniMax H3 is the video model Scenema uses for shots with complex motion and many different characters, props, and places that all have to stay consistent. | 4 to 15 seconds | 768p | Image to video, Reference to video | High |
| Vidu Q3 Turbo | Vidu Q3 Turbo is a start-and-end-frame video model on Scenema, tuned for shots that must land on an exact ending composition. | 1 to 15 seconds | 720p | Image to video | Standard |
| Gemini Omni Flash | Gemini Omni Flash is Google’s multimodal video model, running as one of the shot generators inside a Scenema long-form video. | 4, 6, 8, or 10 seconds | Up to 4K | Image to video | High |
| LTX 2.3 | LTX 2.3 is Lightricks’ open video foundation model, hosted on Scenema with a Fast tier for iteration and a Pro tier for delivery-quality takes. | 4 to 12 seconds | Up to 1080p | Text to video, Image to video, Audio to video | Fast and Pro tiers |
FAQ
Frequently asked questions about Vidu Q3 Turbo
What is Vidu Q3 Turbo on Scenema?+
Vidu Q3 Turbo is Shengshu Technology’s speed-tuned Q3 video model, exposed inside Scenema as a start-and-end-frame image-to-video generator. Each generation takes both a first frame and a last frame and produces the motion between them.
Do I have to provide both a start frame and an end frame?+
Yes. Vidu Q3 Turbo requires both. Providing only a first frame is not supported; for that workflow, use Wan 3.0, MiniMax H3, or another image-to-video model.
How long can a single Vidu Q3 Turbo shot be?+
From 1 second to 15 seconds. That is the widest per-generation duration range on Scenema.
How does Scenema make long-form videos if Vidu Q3 Turbo only generates up to 15 seconds at a time?+
Scenema treats each Vidu Q3 Turbo generation as one shot inside a scene. The pipeline plans the shot list, supplies the start and end frames, and assembles the results into scenes and full videos. Character consistency, voiceover, and music tracks run across every cut, so the finished piece plays as a single continuous video.
When does Scenema use Vidu Q3 Turbo over another Scenema model?+
Scenema uses Vidu Q3 Turbo when the shot has to land on a specific end frame. Its unique advantage is the start-and-end-frame contract; every other Scenema video model that supports last-frame conditioning treats it as optional.
Related
Other video models on Scenema
Veo 3.1
Google DeepMind- Google DeepMind’s flagship cinematic video model.
- 4, 6, or 8 second takes at up to 1080p, with audio generated in the same pass as the picture.
- Shot generator inside Scenema’s long-form pipeline.
Kling 3.0
Kuaishou- Kuaishou’s third-generation flagship video model.
- 4 to 12 second takes at 1080p, with native audio and stronger character consistency.
- Shot generator inside Scenema’s long-form pipeline.
Seedance 2.0
ByteDance- ByteDance’s multimodal video model with joint audio-video generation.
- Text, image, and audio inputs; up to 15 second takes, the longest on Scenema alongside MiniMax H3 and Vidu Q3 Turbo.
- Shot generator inside Scenema’s long-form pipeline.
Seedance 1.5
ByteDance- ByteDance’s image-to-video model, driven from a reference frame.
- 4 to 12 second takes at up to 720p.
- Image-to-video shot generator inside Scenema’s long-form pipeline.
Wan 3.0
Alibaba- Alibaba Tongyi Wanxiang’s current-generation Wan model, replacing Wan 2.6 on Scenema.
- Image-to-video with last-frame conditioning; 4 to 12 second takes at 720p.
- Image-anchored shot generator inside Scenema’s long-form pipeline.
MiniMax H3
MiniMax- MiniMax’s next-generation video model.
- Reference-to-video; 4 to 15 second takes at 768p.
- Complex motion and multi-entity consistency inside Scenema’s long-form pipeline.
Gemini Omni Flash
Google- Google’s multimodal video model with native audio and 4K output.
- Image to video only; 4, 6, 8, or 10 second takes at up to 4K.
- Highest-resolution shot generator inside Scenema’s long-form pipeline.
LTX 2.3
Lightricks- Lightricks’ open video foundation model with Fast and Pro tiers on Scenema.
- Text, image, and audio inputs; 4 to 12 second takes at up to 1080p.
- Widest aspect-ratio range of any Scenema video model.
Start generating with Vidu Q3 Turbo in Scenema
Vidu Q3 Turbo is available on the Scenema free tier. Sign up and start your first explainer. Scenema uses it wherever it fits the style and the shot.
200 credits on signup, then weekly refills through your first month. No credit card required.