新規登録で 100 クレジットを無料プレゼント、今すぐAIアプリを探索・構築無料受取
audio1.0

Vidu Audio 1.0 / テキストからオーディオ

Commercial
ID: audio1.0

Vidu's text-to-audio generation model (model ID: audio1.0). Generates sound effects and background music from text prompts. Duration range: 2–10 seconds. Supports random seed configuration for reproducible outputs

モデルタイプ:
料金$0.048/ Audio(48 クレジット)
Input Prompt
0 / 1500
10
音声生成

Audio Playground Ready

効果音やBGMを説明し、パラメーターを設定してGenerateをクリックしてください。

料金詳細

このモデルの実際の課金は、API リクエストで渡される特定のパラメータに基づいて動的に計算されます。以下は具体的な組み合わせとそれに対応する料金です:

Text to Audio2-5s
クレジット48/ Audio
料金 (USD)$0.048
Text to Audio5-10s
クレジット95/ Audio
料金 (USD)$0.095
Timing to Audio2-5s
クレジット48/ Audio
料金 (USD)$0.048
Timing to Audio5-10s
クレジット95/ Audio
料金 (USD)$0.095
Ecosystem Models

Recommended Related Models

Explore complementary video and multimodal models with your unified API key.

Browse All Models
Qwen3 TTS Instruct Flash
audioCommercial
Qwen3 TTS Instruct Flash

Qwen3 TTS Instruct Flash

qwen3-tts-instruct-flash

Qwen3-TTS-Flash model is Tongyi's latest real-time speech synthesis model. The Instruct model processes the synthesis effect through natural language, ensuring highly appropriate emotional and expressive speech in different contexts. Currently, it supports 25 timbres for both Chinese and English Instruct adjustments.

Text to Speech
MiniMax Speech 2.8 HD
audioCommercial
MiniMax Speech 2.8 HD

MiniMax Speech 2.8 HD

speech-2.8-hd

speech-2.8-hd is a high-definition AI speech synthesis model tailored for individual users. It delivers studio-grade natural voice texture with ultra-realistic pronunciation and smooth intonation. It supports rich exclusive timbres and multilingual conversion, and is capable of simulating vivid emotions like laughter and sighs. It perfectly fits daily voice dubbing, audio creation, reading narration and personal voice customization, bringing you immersive and high-quality voice experience.

Text to Music
MiniMax Speech 2.8 Turbo
audioCommercial
MiniMax Speech 2.8 Turbo

MiniMax Speech 2.8 Turbo

speech-2.8-turbo

speech-2.8-turbo is a lightweight and ultra-fast AI speech synthesis model for all users. It features instant response, efficient generation and stable audio output. With natural and smooth timbre performance, it supports multilingual conversion and basic emotional intonation adjustment. Optimized for low-latency scenarios such as daily narration, short video dubbing and real-time voice interaction, it balances speed, quality and ease of use, delivering a fluent and convenient voice creation exp

Text to Music
MiniMax Speech 2.6 HD
audioCommercial
MiniMax Speech 2.6 HD

MiniMax Speech 2.6 HD

speech-2.6-hd

speech-2.6-hd is a high-definition AI voice synthesis model designed for general users. It delivers lifelike, studio-level vocal quality with natural pronunciation, smooth rhythm and rich emotional expression. It supports multiple languages and diverse premium voice tones, enabling vivid voice dubbing, audiobook narration and personalized voice creation. With stable sound quality and authentic intonation, it perfectly fits daily entertainment, content creation and daily voice playback needs, bri

Text to Music