
Vidu Q2 · Image Generation / 참조로 이미지
Next-gen image-to-video model with "emotional acting" for nuanced facial expressions. Supports 2–8s clips at 1080p with cinematic camera control. Features multi-modal reference (2 videos + 4 images) enabling precise replication of expressions, actions, textures, and effects. Includes video editing: element add/delete/replace, style transfer, aspect adjustment. Lightning mode outputs 5s in 20s. Supports up to 7 consistent subjects, 3x faster than Q1
Upload Images
JPG, JPEG, PNG (Max 10MB)
Image Playground Ready
왼쪽 프롬프트 패널에서 내용을 작성하고 옵션을 조정하세요.
가격 상세
이 모델의 실제 요금은 API 요청에서 전달된 특정 매개변수를 기반으로 동적으로 계산됩니다. 아래는 구체적인 조합과 해당 가격입니다:
| 모달리티 | 크레딧 | 가격 (USD) |
|---|---|---|
| Text to Image/1080P | 29/ Image | $0.029 |
| Text to Image/2K | 38/ Image | $0.038 |
| Text to Image/4K | 48/ Image | $0.048 |
| Reference to Image/1080P/1-3 Images | 38/ Image | $0.038 |
| Reference to Image/1080P/4-7 Images | 48/ Image | $0.048 |
| Reference to Image/2K/1-3 Images | 57/ Image | $0.057 |
| Reference to Image/2K/4-7 Images | 76/ Image | $0.076 |
| Reference to Image/4K/1-3 Images | 95/ Image | $0.095 |
| Reference to Image/4K/4-7 Images | 143/ Image | $0.143 |
Recommended Related Models
Explore complementary video and multimodal models with your unified API key.


Seedream 5.0
seedream-5-0-260128
Supports text, single-image and multi-image inputs, and enables the generation of image sets


Wan 2.7 Image Pro
wan2.7-image-pro
Wan2.7–image-pro,supports text to image, text/image to sequential images, image editing, multi-image reference generation, and interactive editing. Delivers enhanced performance in text rendering, subject consistency, and complex instruction following.


Kling Image O1
kling-image-o1
Advanced image generation model from the Kling O1 family. Supports text-to-image and image-to-image workflows with deep semantic understanding. Features character/subject consistency across generations, multi-reference input (up to 7 images), and intelligent aspect ratio adaptation. Professional-grade output suitable for commercial creative pipelines


Kling V2
kling-v2
V2.0 foundational image model supporting text-to-image and image-to-image generation. Enables multi-type reference control—character, face, subject, scene, and style. Built for role consistency and thematic control across generated visuals. Multiple aspect ratio support