新用户免费领取 100 积分,即刻探索与构建您的 AI 应用免费领取
viduq3-turbo

Vidu Q3 Turbo / 文生视频

Commercial
ID: viduq3-turbo

Optimized for speed, generating 1–16 second video clips from text or images with faster inference than the Pro variant. Supports 540p–1080p output, balancing generation quality with reduced latency for rapid iteration

模型类型:
价格$0.01/ Second(10 积分)
输入 Prompt
0 / 5000
5

Upload Wm Url

JPG, JPEG, PNG (Max 10MB)

视频生成

Video Playground Ready

在左侧参数面板中输入提示词,配置参数后点击 Generate。

Vidu Q3 Turbo Hero
Vidu Q3 Turbo • 4-Mode Fast Video

Vidu Q3 TurboT2V / I2V / Dual-Frame / Reference

Vidu Q3 Turbo exposes four conditioning modes at 540p–1080p with optional native audio, audio type selection, and watermark position control.

4 Conditioning Modes
540p / 720p / 1080p
Optional Native Audio
Audio Type Select
ByteDance Seed Foundation Architecture
Commercial License & Enterprise SLA

At a Glance

4
Modes
T2V · I2V · Dual · R2V
1080p
Max Resolution
Also 540p / 720p
Audio
Native Optional
3 audio types
Seed
Reproducible
Optional pin
Capability Highlights

Vidu Q3 Turbo Architecture

Audio-aware multimodal video at production speed.

Four Conditioning Modes

Text-to-video, image-to-video, start-and-end-frame, and reference-to-video cover most production needs.

T2VI2VDual FrameR2V
Modes

Optional Native Audio

Enable audio generation with type selection — all, speech only, or sound effects only.

audiospeech_onlysound_effect_only

540p to 1080p

Draft at 540p, ship social at 720p, and promote heroes to 1080p.

540p720p1080p

Watermark Position Control

Optional watermark with selectable corner positions (1–4) and custom watermark URL.

wm_positionwm_url
How It Works

How It Works

From brief to audio-ready clip.

01

Pick a Mode

Text, image, dual-frame, or reference conditioning.

02

Set Resolution

540p draft, 720p social, or 1080p hero.

03

Enable Audio

Choose all, speech only, or sound effects only.

04

Watermark & Deliver

Optionally place a watermark and pull the finished clip.

Vidu Q3 Turbo Domains

Where sound and picture ship together at speed.

Content

Social with Sound

Native audio for Reels and TikTok without post Foley.

Audio
Commerce

Product Film

I2V packshots with optional speech or SFX.

I2V
Film

Transition Shots

Start-and-end frames for match cuts.

Dual Frame
Growth

Reference Motion

Subject-reference clips for campaign variants.

R2V
Best Practices

Prompt Tips

Cleaner Vidu Q3 Turbo output.

Mention acoustic mood

If audio is on, name environmental cues so sound matches picture.

Audio type is a product choice

Use speech_only for VO-led cuts and sound_effect_only for ambient plates.

Draft at 540p

Explore motion cheaply, then promote to 1080p for finals.

Developer Quickstart

Vidu Turbo Quickstart

Four-mode audio-ready video.

curl -X POST "https://api.powertokens.ai/v1/videos" \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
  "model": "viduq3-turbo",
  "prompt": "Cinematic tracking shot through a neon cyberpunk city in heavy rain.",
  "seconds": "5",
  "size": "1080p",
  "ratio": "16:9",
  "resolution": "1080p",
  "duration": 5,
  "aspect_ratio": "16:9",
  "audio": true
}'

Technical Specifications

Confirmed parameters and runtime execution protocols.

Provider & Model ID
Vidu • viduq3-turbo
Modes
Text-to-video, Image-to-video, Start & end frame, Reference-to-video
Resolutions
540p, 720p, 1080p (default 540p)
Duration
Configurable via duration slider (default 5)
Aspect Ratios
16:9, 9:16, 3:4, 4:3, 1:1 (default 16:9 for T2V)
Audio
Optional; all / speech_only / sound_effect_only
Seed
Optional reproducible seed
Watermark
Optional with wm_position 1–4 and wm_url
Billing
See pricing matrix
API Endpoint
POST /v1/videos

价格详情

此模型的实际计费根据您在 API 请求中传入的具体参数动态计算。以下是具体的组合及其对应的价格:

Text to Video540PPeak Shifting
积分19/ Second
价格 (USD)$0.019
Text to Video540P
积分34/ Second
价格 (USD)$0.034
Text to Video720PPeak Shifting
积分29/ Second
价格 (USD)$0.029
Text to Video720P
积分53/ Second
价格 (USD)$0.053
Text to Video1080PPeak Shifting
积分34/ Second
价格 (USD)$0.034
Text to Video1080P
积分62/ Second
价格 (USD)$0.062
Image to Video540PPeak Shifting
积分19/ Second
价格 (USD)$0.019
Image to Video540P
积分34/ Second
价格 (USD)$0.034
Image to Video720PPeak Shifting
积分29/ Second
价格 (USD)$0.029
Image to Video720P
积分53/ Second
价格 (USD)$0.053
Image to Video1080PPeak Shifting
积分34/ Second
价格 (USD)$0.034
Image to Video1080P
积分62/ Second
价格 (USD)$0.062
Start&End Frame to Video540PPeak Shifting
积分19/ Second
价格 (USD)$0.019
Start&End Frame to Video540P
积分34/ Second
价格 (USD)$0.034
Start&End Frame to Video720PPeak Shifting
积分29/ Second
价格 (USD)$0.029
Start&End Frame to Video720P
积分53/ Second
价格 (USD)$0.053
Start&End Frame to Video1080PPeak Shifting
积分34/ Second
价格 (USD)$0.034
Start&End Frame to Video1080P
积分62/ Second
价格 (USD)$0.062
Reference to Video540PPeak Shifting
积分10/ Second
价格 (USD)$0.010
Reference to Video540P
积分19/ Second
价格 (USD)$0.019
Reference to Video720PPeak Shifting
积分24/ Second
价格 (USD)$0.024
Reference to Video720P
积分48/ Second
价格 (USD)$0.048
Reference to Video1080PPeak Shifting
积分34/ Second
价格 (USD)$0.034
Reference to Video1080P
积分62/ Second
价格 (USD)$0.062
Ecosystem Models

Recommended Related Models

Explore complementary video and multimodal models with your unified API key.

Browse All Models
Seedance 2.5
videoCommercial
Seedance 2.5

Seedance 2.5

dreamina-seedance-2-5-260628

ByteDance's latest flagship video generation model, built for longer-form storytelling and production-ready output. Generates up to 30 seconds of continuous, cinematic video with native audio sync in a single pass . Accepts up to 50 multimodal references (images, videos, audio, character sheets, storyboards) for precise scene, character, and motion consistency . Features localized region editing to fix specific areas without full regeneration……

Text to VideoImage to Video
Wan 3.0
videoCommercial
Wan 3.0

Wan 3.0

wan3.0-video

Wan3.0-Video is an all-in-one video generation model unified support for multiple creative capabilities, including reference, editing, replication, and driving. It generates videos up to 30 seconds with omni-modal reference, and can parse files, web pages and complex images. With production-grade character consistency and lifelike visuals and sound, it delivers an immersive audiovisual experience.

Text to VideoImage to Video
Seedance 2.0 Mini
videoCommercial
Seedance 2.0 Mini

Seedance 2.0 Mini

dreamina-seedance-2-0-mini-260615

Lightweight, cost-efficient video model from ByteDance, optimized for speed and high-volume content creation. Supports text-to-video, image-to-video, and reference-based generation with up to 12 references (6 images, 3 audio, 3 video). Delivers faster generation and lower credit consumption than Seedance 2.0, with strong motion quality and character consistency. Ideal for social media content, product videos, AI short dramas, and rapid creative iteration

Text to VideoImage to Video
Seedance 2.0
videoCommercial
Seedance 2.0

Seedance 2.0

dreamina-seedance-2-0-260128

Generate videos from reference images, videos, and audio; edit videos; extend videos; generate videos from start and end frames

Text to VideoImage to Video

Frequently Asked Questions

Everything you need to know before integrating this model.

Text-to-video, image-to-video, start-and-end-frame, and reference-to-video.

Start Building with Vidu Q3 Turbo Today

Create an account in seconds to receive 100 free credits and start generating immediately. No credit card or upfront contract required.