新規登録で 100 クレジットを無料プレゼント、今すぐAIアプリを探索・構築無料受取
qwen3-tts-instruct-flash

Qwen3 TTS Instruct Flash / テキストから音声

Commercial
ID: qwen3-tts-instruct-flash

Qwen3-TTS-Flash model is Tongyi's latest real-time speech synthesis model. The Instruct model processes the synthesis effect through natural language, ensuring highly appropriate emotional and expressive speech in different contexts. Currently, it supports 25 timbres for both Chinese and English Instruct adjustments.

料金$0.011/ 1K Characters(11 クレジット)
24 total
音声合成 (TTS)
0
音声出力準備完了。「合成を開始」をクリックして音声を生成します。
0:00 / 0:00

料金詳細

このモデルの実際の課金は、API リクエストで渡される特定のパラメータに基づいて動的に計算されます。以下は具体的な組み合わせとそれに対応する料金です:

Standard
クレジット11/ 1K Characters
料金 (USD)$0.011
Ecosystem Models

Recommended Related Models

Explore complementary video and multimodal models with your unified API key.

Browse All Models
MiniMax Speech 2.8 HD
audioCommercial
MiniMax Speech 2.8 HD

MiniMax Speech 2.8 HD

speech-2.8-hd

speech-2.8-hd is a high-definition AI speech synthesis model tailored for individual users. It delivers studio-grade natural voice texture with ultra-realistic pronunciation and smooth intonation. It supports rich exclusive timbres and multilingual conversion, and is capable of simulating vivid emotions like laughter and sighs. It perfectly fits daily voice dubbing, audio creation, reading narration and personal voice customization, bringing you immersive and high-quality voice experience.

Text to Music
MiniMax Speech 2.8 Turbo
audioCommercial
MiniMax Speech 2.8 Turbo

MiniMax Speech 2.8 Turbo

speech-2.8-turbo

speech-2.8-turbo is a lightweight and ultra-fast AI speech synthesis model for all users. It features instant response, efficient generation and stable audio output. With natural and smooth timbre performance, it supports multilingual conversion and basic emotional intonation adjustment. Optimized for low-latency scenarios such as daily narration, short video dubbing and real-time voice interaction, it balances speed, quality and ease of use, delivering a fluent and convenient voice creation exp

Text to Music
MiniMax Speech 2.6 HD
audioCommercial
MiniMax Speech 2.6 HD

MiniMax Speech 2.6 HD

speech-2.6-hd

speech-2.6-hd is a high-definition AI voice synthesis model designed for general users. It delivers lifelike, studio-level vocal quality with natural pronunciation, smooth rhythm and rich emotional expression. It supports multiple languages and diverse premium voice tones, enabling vivid voice dubbing, audiobook narration and personalized voice creation. With stable sound quality and authentic intonation, it perfectly fits daily entertainment, content creation and daily voice playback needs, bri

Text to Music
MiniMax Speech 2.6 Turbo
audioCommercial
MiniMax Speech 2.6 Turbo

MiniMax Speech 2.6 Turbo

speech-2.6-turbo

A lightweight, ultra-fast AI speech synthesis model with premium sound quality and sub-250ms low latency. Supports 40+ languages, 7 emotions, and 300+ curated voices for real-time interaction, short video dubbing, and daily narration. Delivers natural, smooth audio with high cost-performance.

Text to Music