System and base voices for qwen-audio-3.0-tts-plus and qwen-audio-3.0-tts-flash
Qwen-Audio-TTS supports the system voices listed below. In addition to system voices, pre-generated base voices created through voice cloning are also available. For details, see Base voice list. For a more personalized voice, you can use the voice cloning feature to create a custom voice for free. For details, see Synthesize speech with a cloned voice.
When synthesizing speech:
In addition to the system voices above,
- Each
modelsupports only a specific set of voices (voice). Voices cannot be mixed across models. If you specify a voice that is not in the voice list of the selected model, the service returns anInvalidParametererror (for example,[cosyvoice:]Engine error [411]: TTS speak operation failed). In this case, refer to the voice list of the corresponding model below and make sure the voice is supported by the current model. - The
textto be synthesized must be in a language supported by the selected voice. Otherwise, pronunciation errors or unnatural speech may occur.
qwen-audio-3.0-tts-plus system voices
| Use case | Voice name | Voice parameter | Characteristic | Age | Gender | Language | Sample |
|---|---|---|---|---|---|---|---|
| Social companionship (flagship voice) | Long An Ling Xin | longanlingxin | Warm and empathetic | 25 | Female | Chinese (Mandarin), English | |
| Social companionship (flagship voice) | Long An Lu Feng | longanlufeng | Bright and cheerful | 25 | Male | Chinese (Mandarin), English |
qwen-audio-3.0-tts-flash system voices
| Use case | Voice name | Voice parameter | Characteristic | Age | Gender | Language | Sample |
|---|---|---|---|---|---|---|---|
| Social companionship (premium Chinese voice) | Long An Feng Yue | longanfengyue | Natural and friendly | 30 | Female | Chinese (Mandarin), English | |
| Social companionship (premium Chinese voice) | Long An Yuan Fei | longanyuanfei | Proud and regal | 30 | Female | Chinese (Mandarin), English | |
| Social companionship (premium Chinese voice) | Long An Ling Xi | longanlingxi | Cute and sweet | 25 | Female | Chinese (Mandarin), English | |
| Social companionship (premium Chinese voice) | Long An Xiao Xin | longanxiaoxin | Friendly and lively | 22 | Female | Chinese (Mandarin), English | |
| Social companionship (premium Chinese voice) | Long An Huan | longanhuan_v3.6 | — | 25 | Female | Chinese (Mandarin), English | |
| Child companion / smart toy (premium child voice) | Long Jie Li Dou | longjielidou_v3.6 | Innocent young boy | 5 | Male | Chinese (Mandarin), English | |
| Child companion / smart toy (premium child voice) | Long Pao Pao | longpaopao_v3.6 | Soft and adorable | 5 | Female | Chinese (Mandarin), English | |
| Character / gaming voice (premium Chinese voice) | Long Huo Huo | longhuohuo_v3.6 | Mischievous young boy | 8 | Male | Chinese (Mandarin), English | |
| Character / gaming voice (premium Chinese voice) | Long Chuan Shu | longchuanshu_v3.6 | Sichuan-accented middle-aged male | 40 | Male | Chinese (Mandarin), English | |
| Social companionship / voice assistant (premium English voice) | loongmary | loongmary | Warm British accent | 20 | Female | English | |
| Social companionship / voice assistant (premium English voice) | loongeva | loongeva_v3.6 | Intelligent-sounding American accent | 28 | Female | English | |
| Social companionship / voice assistant (premium English voice) | loongjohn | loongjohn | Calm and friendly American accent | 28 | Male | English |
Base voice list
In addition to the system voices above, qwen-audio-3.0-tts-plus and qwen-audio-3.0-tts-flash each provide over 500 base voices generated through voice cloning. They are called the same way as system voices. Base voices follow the naming pattern qwen-audio-3.0-tts-{plus|flash}-{voice suffix}, and the same suffix maps to the same set of preview audio files across both models. Download the Excel files below for the full voice list:
qwen-audio-3.0-tts-plusbase voice list (Excel): qwen-audio-3.0-tts-plus base voices.xlsxqwen-audio-3.0-tts-flashbase voice list (Excel): qwen-audio-3.0-tts-flash base voices.xlsx- Base voice preview audio package (shared by plus and flash): base voice preview audio package.zip
- Download the Excel files and the preview audio package, then extract the audio package to your local machine.
- In the Excel file, find the "Preview audio file name" column and copy the audio file name.
- Locate the corresponding file in the extracted folder and open it with a media player to preview.