
Model input
Select an available input mode
Image3 exposes text-to-video, image-to-video, and reference-to-video when the selected provider supports them. Frame animation and reference guidance are distinct inputs.
- Prompt-led scenes, animating source frames, and reference-guided visual direction.
- The prompt assigns one action and a camera move, while leaving room to review whether clothing and subject details stay coherent.