
Vidu Q3 Pro Fast / 图生视频
High-efficiency variant optimized for cost and generation speed, ideal for high-volume creative pipelines. Supports 1–16 second image-to-video at 720p or 1080p, offering the most affordable per-second pricing in the Q3 series
Upload Images
JPG, JPEG, PNG (Max 10MB)
Upload Wm Url
JPG, JPEG, PNG (Max 10MB)
Video Playground Ready
在左侧参数面板中输入提示词,配置参数后点击 Generate。
Vidu Q3 Pro FastFast Image-to-Video at 720p–1080p
Vidu Q3 Pro Fast is a latency-optimized image-to-video model at 720p or 1080p with optional native audio, audio type selection, and watermark position control.

At a Glance
Vidu Pro Fast Stack
Low-latency I2V with audio controls.

I2V-First Design
Every request starts from a source still — ideal for catalog motion and social packshots.

Optional Native Audio
Enable audio generation with type selection — all, speech only, or sound effects only.

720p to 1080p
Draft at 720p and promote heroes to 1080p.

Watermark Position Control
Optional watermark with selectable corner positions (1–4) and custom watermark URL.
How It Works
Four steps from still to motion.
Upload a Still
Every generation starts from your source image.
Optional Motion Prompt
Describe the movement when the still alone is not enough.
Pick Size & Audio
720p or 1080p; optional audio type.
Watermark & Deliver
Optionally place a watermark and pull the finished clip.
Pro Fast Domains
Where stills become motion at speed.
Catalog Motion
Bulk packshot animation from product stills.
Social Drafts
Rapid 5s variants for testing creative.
Event Stills → Clips
Same-day recap motion from photography.
A/B Motion Tests
Multiple motion styles per still for creative tests.
Prompt Tips
Keep Fast I2V predictable.
Describe movement and camera, not a full redesign.
Sharp, well-lit product photos animate more reliably.
See viduq3-turbo for multi-mode generation.
Pro Fast Quickstart
Low-latency image-to-video.
curl -X POST "https://api.powertokens.ai/v1/videos" \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "viduq3-pro-fast",
"prompt": "Cinematic tracking shot through a neon cyberpunk city in heavy rain.",
"seconds": "5",
"size": "1080p",
"ratio": "16:9",
"resolution": "1080p",
"duration": 5,
"audio": true
}'Technical Specifications
Confirmed parameters and runtime execution protocols.
价格详情
此模型的实际计费根据您在 API 请求中传入的具体参数动态计算。以下是具体的组合及其对应的价格:
| 模态 | 积分 | 价格 (USD) |
|---|---|---|
| Image to Video/720P/Peak Shifting | 48/ Second | $0.048 |
| Image to Video/720P | 95/ Second | $0.095 |
| Image to Video/1080P/Peak Shifting | 62/ Second | $0.062 |
| Image to Video/1080P | 119/ Second | $0.119 |
Models from the Same Channel
Explore complementary models and alternative versions from the same provider channel.


Vidu Q3 Pro
viduq3-pro
Flagship video model supporting text-to-video, image-to-video, and start/end-frame workflows. Generates up to 16-second clips with native audio sync and storyboard capabilities. Available in 540p–1080p resolutions with premium motion dynamics


Vidu Q3 Turbo
viduq3-turbo
Optimized for speed, generating 1–16 second video clips from text or images with faster inference than the Pro variant. Supports 540p–1080p output, balancing generation quality with reduced latency for rapid iteration


Vidu Q3
viduq3
Base reference-to-video model built for narrative video creation. Supports native audio-video generation and multi-character dialogue. Delivers robust character consistency and scene coherence across up to 16-second clips
Recommended Related Models
Explore complementary video and multimodal models with your unified API key.


Seedance 2.5
dreamina-seedance-2-5-260628
ByteDance's latest flagship video generation model, built for longer-form storytelling and production-ready output. Generates up to 30 seconds of continuous, cinematic video with native audio sync in a single pass . Accepts up to 50 multimodal references (images, videos, audio, character sheets, storyboards) for precise scene, character, and motion consistency . Features localized region editing to fix specific areas without full regeneration……


Wan 3.0
wan3.0-video
Wan3.0-Video is an all-in-one video generation model unified support for multiple creative capabilities, including reference, editing, replication, and driving. It generates videos up to 30 seconds with omni-modal reference, and can parse files, web pages and complex images. With production-grade character consistency and lifelike visuals and sound, it delivers an immersive audiovisual experience.


Seedance 2.0 Mini
dreamina-seedance-2-0-mini-260615
Lightweight, cost-efficient video model from ByteDance, optimized for speed and high-volume content creation. Supports text-to-video, image-to-video, and reference-based generation with up to 12 references (6 images, 3 audio, 3 video). Delivers faster generation and lower credit consumption than Seedance 2.0, with strong motion quality and character consistency. Ideal for social media content, product videos, AI short dramas, and rapid creative iteration


Seedance 2.0
dreamina-seedance-2-0-260128
Generate videos from reference images, videos, and audio; edit videos; extend videos; generate videos from start and end frames
Frequently Asked Questions
Everything you need to know before integrating this model.
Start Building with Vidu Q3 Pro Fast Today
Create an account in seconds to receive 100 free credits and start generating immediately. No credit card or upfront contract required.