シリーズ

Qwen Audio AIモデルシリーズ

2 モデル更新日 Jul 2026
Qwen Audio AIモデルシリーズ

Qwen Audioモデルについて

Qwen-Audio is a unified audio-language model series by Alibaba Cloud that processes speech, natural sounds, music, and singing across multiple languages and tasks, enabling universal audio understanding and multimodal interaction.

すべてのQwen Audioモデル

alibaba/qwen-audio-3.0-tts-flash

alibaba/qwen-audio-3.0-tts-flash

text-to-speech

[Core Function] Qwen-Audio 3.0 TTS Flash is Alibaba's low-latency Qwen-Audio text-to-speech model on the same SpeechSynthesizer endpoint family as CosyVoice. [Strengths] Fast synthesis with voice, format, sample-rate, prosody, SSML, instruction, language_hint, and AIGC watermark controls; system voices include longanhuan_v3.6, longjielidou_v3.6, loongeva_v3.6, and loongjohn (see the Qwen-Audio-TTS voice list). [Best For] Voice assistants, interactive prompts, short announcements, multilingual product flows, and latency-sensitive batch TTS using Qwen-Audio voices. [Limitations] Do NOT mix Plus-only voices (e.g. longanlingxin) with Flash. Use Plus when maximum narration quality matters more than turnaround time. [Routing] Choose Flash when speed matters most. Choose qwen-audio-3.0-tts-plus for premium narration quality.

alibaba/qwen-audio-3.0-tts-plus

alibaba/qwen-audio-3.0-tts-plus

text-to-speech

[Core Function] Qwen-Audio 3.0 TTS Plus is Alibaba's high-quality Qwen-Audio text-to-speech model on the same SpeechSynthesizer endpoint family as CosyVoice. [Strengths] Natural speech synthesis with voice, format, sample-rate, prosody, SSML, instruction, language_hint, and AIGC watermark controls; system voices include longanlingxin and longanlufeng (see the Qwen-Audio-TTS voice list). [Best For] Premium narration, brand voiceovers, multilingual product audio, and quality-sensitive batch TTS when Qwen-Audio voices are preferred. [Limitations] Do NOT mix Flash-only voices (e.g. longanhuan_v3.6) with Plus. [Routing] Choose Plus when speech quality is the priority. Choose qwen-audio-3.0-tts-flash when lower latency matters more.

必要なモデルが見つかりませんか? ご要望をお聞かせください。

さらに探す

テキスト読み上げ
カテゴリー9 モデル
Voice Cloning
Voice Cloning
コレクション2 モデル
Fun
Fun
シリーズ2 モデル
CosyVoice
CosyVoice
シリーズ4 モデル
Grok Voice
Grok Voice
シリーズ2 モデル
音声認識
カテゴリー5 モデル