Video Generator
Generate AI videos from a text prompt, image, or reference clip. Cinematic shots, product reveals, lifestyle ads, with synchronized audio and natural motion.
Runs on
-
ByteDance
-
Google
-
MiniMax
-
xAI
Examples
Laundry ball completes one orbit (Seedance 2.0 Fast)
Seedance 2.0 Fast. A generated Seoul laundromat frame becomes a silent five-second 720p study of one rigid ball completing a clear drum orbit.
How to use Video Generator
A compact decision guide for placing this pipeline in a workflow.
Best for
- Creating short video clips from text descriptions or images
- Product reveals and scene transitions with real camera movement
- Cinematic shots that need real motion (not just pan/zoom on a still)
- Text-to-video with native audio on models that support it (others are silent, add narration in video-reel)
Tips
- Provide detailed text prompts describing the scene, motion, and desired audio
- Use a start image for consistent framing, the model animates from it
- Works best for product reveals, scene changes, and dramatic transformations
- For simple pan/zoom on stills, image-motion is 10x faster and cheaper
- Native audio is generated alongside the video on models that support it, no separate voiceover step. Other models are silent; add narration or music in video-reel.
Video Generator API
Call this pipeline from your own code. One request dispatches a run; the model is an input field, not a separate endpoint.
Get an API token- inputs
- 13
- Pricing
- from 9.5 credits
| Input | Type | default | Description |
|---|---|---|---|
end_image | string | default— | Optional: URL of an image to use as the last frame. AI interpolates between start and end. |
prompt | string | default— | Text description of the video. The AI enhances this automatically with cinematic details. |
reference_audios | array | default— | Optional audio references for music, speech, rhythm, or sound. Model-specific limits are applied automatically. |
reference_images | array | default— | Optional images that guide identity, objects, style, or composition. They are not timeline frames. |
reference_videos | array | default— | Optional video references for composition, motion, camera work, or timing. Model-specific limits are applied automatically. |
start_image | string | default— | Optional: URL of an image to use as the first frame. |
aspect_ratio | enum | defaultauto | Auto infers framing and preserves attached reference framing when the selected model supports it; fixed ratios force the output canvas.auto · 16:9 · 9:16 +4auto16:99:161:14:33:421:9 |
audio | boolean | defaulttrue | Keep the generated synchronized audio track when the model returns one. |
duration | enum | defaultauto | Auto right-sizes the clip. Model-specific limits are applied automatically.auto · 4 · 5 +25auto456789101112131415161718192021222324252627282930 |
enhance_prompt | boolean | defaulttrue | Improve the prompt for the selected model before generation. Turn off to send your prompt unchanged. |
model | enum | defaultauto | Auto selects a compatible engine, or pin a model for its exact input and output controls.auto · seedance-2-0-mini · seedance-2-0-fast +8autoseedance-2-0-miniseedance-2-0-fastseedance-2-0-proseedance-2-5veo-3-1-liteveo-3-1-fastveo-3-1-standardminimax-h3gemini-omni-1-1-flashgrok-imagine-video |
resolution | enum | default720p | Model-aware output resolution.480p · 720p · 768p +3480p720p768p1080p2k4k |
seed | integer | default— | Optional: pin the random seed for a reproducible Seedance result. Leave empty for a fresh random seed each run. |
Point any MCP client at pipe2 and your agent gets these tools. It finds this pipeline, reads the same schema above, then runs it.
list_pipelinesget_pipeline_schemarun_pipelineget_pipeline_run_statusrequest_upload
https://mcp.pipe2.ai/mcp {
"mcpServers": {
"pipe2ai": {
"url": "https://mcp.pipe2.ai/mcp",
"headers": { "Authorization": "Bearer YOUR_TOKEN" }
}
}
}Recipes using this pipeline
Show all recipes-
One prompt → a vertical AI dance reel
Generate a labeled dance-move grid, use it as the choreography reference for a music-synced dance, then combine both into a vertical reel.
-
An ordinary box opens onto a tiny impossible world
Create a silent reveal of a tiny moonlit ocean inside a hinged wooden box: build the open scene, close its lid with an image edit, then animate the discovery.
Frequently Asked Questions
How long are generated videos?
Does it generate audio?
What aspect ratios are supported?
Can I use references?
How is pricing determined?
Can I choose the model?
AI Video Generator
Generate short videos from a text prompt, image, or reference clip. Cinematic narrative shots, product ads, social-ready vertical content, with synchronized audio, real camera movement, and natural motion.
What you can make
- Text-to-video scenes: describe a moment and get a finished clip with audio
- Image-to-video: start from a reference image and animate it into a cinematic shot
- Product and lifestyle ads: bring a reference video for camera move and a reference audio for the music bed
- Vertical shorts: 9:16 for TikTok, Reels, and Shorts in one call
- Cinematic explainers: golden hour, film noir, anime, watercolor, 35mm film, any style on demand
How it works
- Describe the scene: subject, action, camera, lighting, style
- Optionally attach references: images, a reference video for camera movement, or audio for the soundtrack
- Pick aspect and length: 16:9, 9:16, 1:1, 4:3, 3:4, or 21:9, 4–15 seconds
- Generate: the right AI engine is picked automatically based on what you uploaded
Prompt tips
- Include camera angles: close-up, aerial, tracking shot, dolly in
- Describe lighting: golden hour, neon glow, film noir shadows, volumetric fog
- Set the style: cinematic, anime, watercolor, 35mm film look
- Mention audio in a separate sentence, "rain pattering on concrete, distant thunder"
- For longer clips with reference video and audio, the multimodal path keeps your soundtrack and camera move in sync