Text to video
Build a short scene from a prompt describing subjects, dialogue, action, framing, light, pacing, and sound.
Create short videos from a written scene, animate one still image, or guide a transition with first and last frames using the Vidu Q3 workflow.
Choose a generation mode and Q3 variant, then set duration, resolution, audio, and optional off-peak processing before reviewing the point estimate.
Generated content will appear here
Facts reviewed:
Vidu Q3 is a video generation model series from Vidu. Its official model map documents synchronized audio and video, smart scene cuts, up to 1080p output, and durations up to 16 seconds. Q3 Pro and Q3 Turbo support text-to-video, image-to-video, and first-and-last-frame generation, while Turbo prioritizes faster generation.
Nano Banana Pro provides a focused Vidu Q3 workspace with Pro, Turbo, and Pro Fast routing. Pro Fast is shown only in image-to-video mode because that is the workflow listed for it in current Vidu documentation. The 2–16 second range, point estimates, and off-peak control shown here are platform settings and may differ from using the Vidu API directly.
Choose the input method and processing route that best fit the shot you need to produce.
Build a short scene from a prompt describing subjects, dialogue, action, framing, light, pacing, and sound.
Animate a permitted still image while directing camera movement, subject motion, scene timing, and audio.
Supply two visually compatible frames to define the opening and destination of a continuous transition.
Enable generated dialogue and sound effects, then describe audible events in the same timeline as the visuals.
Use Pro for quality-focused generation, Turbo for faster processing, or Pro Fast for supported image-to-video tasks.
Choose the lower-cost off-peak route when turnaround is flexible; Vidu notes that these tasks may take up to 48 hours.
Choose the input method and processing route that best fit the shot you need to produce.
Prototype dialogue, reaction shots, scene cuts, and sound-led storytelling for creative review.
Animate approved stills to test camera movement, highlights, texture, and commercial pacing.
Create short horizontal, square, or vertical concepts for feeds, stories, and ad iterations.
Connect two prepared key frames for transformations, time changes, entrances, and visual reveals.
Prototype short exchanges, reactions, shot changes, and synchronized room ambience for story review.
Use faster routing to compare openings, product beats, and vertical ad structures before choosing a final direction.
Each example loads its recommended generation mode and platform variant so you can revise it in context.
A Pro text-to-video brief built around staging, dialogue, reaction, and ambient sound.
Medium two-shot in a quiet late-night café. A tired designer slides a folded sketch across the table and says softly, “I think this is the one.” Her colleague pauses, smiles, and answers, “Then let’s build it.” Natural eye lines, restrained hand movement, warm practical lighting, rain against the window, low room ambience, one continuous cinematic shot.
A Turbo prompt for a compact social video with clear movement and sound cues.
Vertical close-up of a pastry chef placing the final glossy berry on a small tart. Quick but stable push-in, crisp texture, soft morning window light, a light dusting of sugar falling in slow motion, subtle kitchen ambience and one clean plate tap, no text or logos.
Upload a product still and use the image-only Pro Fast route for controlled motion.
Preserve the exact product shape, material, label placement, and background composition from the uploaded image. Add a smooth 20-degree camera orbit, slow moving highlights across the surface, subtle condensation, and a soft shadow shift. Premium commercial pacing, realistic physics, quiet studio ambience.
A Turbo brief with a fast scene change, product detail, and synchronized impact sound.
Fast 9:16 sports ad with two clean beats. First: a runner launches from wet pavement at dawn, low tracking camera, strong but realistic stride. Second: match cut to a close-up of the shoe landing, water droplets scattering in slow motion. Cool blue light with one coral accent, synchronized foot impact and breath, no logos or text.
A first-and-last-frame transition focused on spatial continuity.
Connect the empty studio first frame to the furnished final frame in one continuous transformation. Preserve walls, windows, floor lines, camera angle, and perspective. Furniture assembles in a physically believable sequence as daylight warms. Smooth pacing, subtle room ambience, no people, no camera cuts.
Start from text, one still image, or a compatible pair of first and last frames.
Use Pro or Turbo in any listed mode; Pro Fast appears only for image-to-video.
Write the action in sequence and align dialogue, effects, ambience, and camera movement with it.
Set duration and resolution, decide whether off-peak processing fits, then inspect the point estimate.
AI video can produce unstable identities, unintended cuts, incorrect dialogue, inconsistent hands or objects, and imperfect audio synchronization. Review the full visual and audio track before publishing or sending the result to a client.
Vidu’s official API documentation lists 1–16 second generation for core Q3 routes. Nano Banana Pro currently exposes 2–16 seconds and its own point system. Available variants, resolutions, point estimates, and the off-peak option on this page describe this integration; they are not a quote of Vidu API billing.
Off-peak processing trades speed for lower cost. Vidu states that off-peak tasks may be generated within 48 hours and unfinished tasks may be cancelled and refunded under its API terms. Only upload images you are authorized to use, and avoid deceptive, infringing, or privacy-invasive content.
Model capabilities, routes, and pricing can change. Official Vidu documentation and Nano Banana Pro platform information are listed separately below.
Return to the workspace, choose the right route for the shot, and review points and timing before generating.