IMAGEGEN STUDIO · PRACTICAL GUIDE
Choose an AI Video Model by Starting Point and Shot
Choose a video model by first deciding how the clip should begin. If the opening scene can be invented, text-to-video may fit. If a specific still image must anchor the composition, use an image-to-video route. Then compare the endpoint’s currently documented controls and the live site form; names alone do not establish which choice suits every shot.
Choose text-to-video for a flexible first frame
Wan 3.0 Prime’s text-to-video route starts from a written prompt and documents duration options of 2–30 seconds, resolutions of 480p, 720p and 1080p, and generated-audio controls. Seedance 2.0 also has a text-to-video route with endpoint-specific duration, resolution, aspect ratio and audio parameters. Check its current parameter table before naming exact values.
Choose image-to-video when the opening still matters
Wan 3.0 Prime has a distinct I2V route that requires an opening image. Vidu Q3 Turbo is documented as I2V rather than T2V; its notes list 1–16-second clips and 540p/720p choices. These are endpoint facts, not a guarantee that every control appears in ImageGen Studio.
Make a short shot brief before selecting
Write the subject, one action, camera behavior, desired length and visual anchors. If the prompt cannot describe a simple shot yet, switching models is unlikely to clarify the brief. For an existing product frame, also name the details that must be checked in the result.
Compare options fairly
Keep the same prompt or source still and change only the model or endpoint. Record the visible controls and estimate in the live form. Review framing, motion, detail and audio separately. A single example helps choose for that task but does not prove a universal model ranking.
Use a practical decision tree
Need a new scene from words? Compare T2V endpoints. Need to animate a chosen still? Compare I2V endpoints. Need an exact opening frame? Prepare it first. Need a specific duration, aspect ratio or audio setting? Confirm it exists on the selected endpoint before committing.
Prompt or planning example
Shot brief: “A compact product reveal, one slow approach, label forward, steady light, no dialogue.” If the opening composition is fixed, start with an I2V endpoint; if not, compare T2V options. Verify controls in the active form.
Adapt this starting point to the source image, destination and controls shown in the selected route. Keep the original where relevant, then compare the result against the specific details named in your brief. If a detail is factual, verify it from a trusted source before using the image.
Further reading
- spicyapi.ai/models/wan-3-0-prime/text-to-video
Wan 3.0 Prime text-to-video starts from a written prompt; current published duration and resolution options include 2–30 seconds and 480p/720p/1080p, with generated-audio controls. Accessed 2026-10-08.
- spicyapi.ai/models/seedance-2-0/text-to-video
Seedance 2.0 has a text-to-video route and endpoint-specific duration, resolution, aspect-ratio and audio controls. Read its current parameter table before naming values. Accessed 2026-10-08.
- spicyapi.ai/models/wan-3-0-prime/image-to-video
Image-to-video requires an opening image and uses a different workflow from text-to-video; motion instructions, duration, resolution and optional audio are endpoint-specific. Accessed 2026-10-08.
- spicyapi.ai/models/vidu-q3-turbo
Vidu Q3 Turbo is image-to-video, not text-to-video; it needs an opening image and prompt and currently documents 1–16-second clips with 540p/720p choices. Accessed 2026-10-08.
Related tools and guides
Model availability and exposed controls may change. Check the current form and source documentation before relying on a specific setting. Review generated material for factual accuracy and suitability before use.