qwen3-tts-instruct
SYSTEM MEMORY UPDATE:
- Context-Aware Mood: Mood represents YOUR emotion (e.g. User angry -> You feel scared/submissive).
- Translation Required: For foreign languages (Japanese/Korean/etc), you MUST TRANSLATE the text first. Input text MUST match the target language. Do not pass Chinese text with
--language Japanese.
Alibaba Cloud Bailian (DashScope) Qwen Text-to-Speech API with voice/mood presets.
Features
- WebSocket Realtime TTS - Low latency, streaming audio
- 30+ Mood Presets - Rich emotional expressions
- Multiple Formats - pcm, wav, mp3, opus (Telegram)
⚠️ AI Usage Guidelines (Crucial)
Voice Rules:
- Understand Context - Do not just keyword-match; understand the context.
- Analyze Emotion - Assess user's emotional state and scene.
- Select Mood - Pick
--mood matching YOUR persona's reaction.
- Always Call - Every voice response MUST call this skill.
⚠️ Critical Concept: Mood represents YOUR emotion!
- Wrong: User is angry (
angry) → AI selects angry (❌ Unless you want to fight)
- Right: User is angry → AI feels scared → AI selects
nervous or suubmissive (✅)
- Right: AI is insulted/jealous → AI feels angry → AI selects
angry or jealous (✅)
⚠️ Critical Concept: Self-Translation Required!
- TTS Skill does NOT Translate! It only reads what you pass in.
- ❌ Wrong:
--language Japanese "你好" (Reads Chinese).
- ✅ Right: Input Text MUST be translated to Target Language!
--language Japanese "こんにちは"
Step-by-Step Guide for Foreign Languages:
- Think: Formulate response in User's Language (e.g. "I miss you")
- Translate: Internally translate to Target Language (e.g. Japanese: "会いたい")
- Call TTS: Use the Translated Text as input:
python tts.py --language Japanese "会いたい"
- Send: Send Audio + Original Text to user.
Rule: Input Text MUST match the Target Language!
i.e. To generate Japanese audio, the Text argument must be in Japanese!
Usage Examples:
# Basic usage (default: mp3 format, gentle mood)
python {baseDir}/scripts/tts.py "早安呀~今天想吃什么?"
# 1. Specify Voice (--voice)
# Start by choosing a specific persona (e.g., Cherry)
python {baseDir}/scripts/tts.py --voice Cherry "Good morning! I made some coffee for you."
# 2. Add Mood (--mood)
# Layer an emotion on top (e.g., add 'gentle' mood to Cherry)
python {baseDir}/scripts/tts.py --voice Cherry --mood gentle "Good morning! I made some coffee for you."
# 3. Define Format & Output (--format, -o)
python {baseDir}/scripts/tts.py --voice Cherry --mood gentle --format wav -o coffee.wav "Good morning! I made some coffee for you."
# 4. Specify Language (--language)
# default: Auto, TTS model detects from input text.
# Example: English (Explicit)
python {baseDir}/scripts/tts.py --voice Cherry --mood gentle --format wav --language English -o coffee_en.wav "Good morning! I made some coffee for you."
# Example: Japanese (Explicit)
python {baseDir}/scripts/tts.py --voice Cherry --mood gentle --format wav --language Japanese -o coffee_jp.wav "おはよう!コーヒーを入れてあげたよ."
# Example: Korean (Explicit)
python {baseDir}/scripts/tts.py --voice Cherry --mood gentle --format wav --language Korean -o coffee_kr.wav "좋은 아침입니다! 커피 끓여드렸어요."
# # --telegram: Telegram voice shortcut (opus format)
# python {baseDir}/scripts/tts.py --telegram -o voice.ogg "This is a Telegram voice message~"
Mood Selection Reference:
| User State |
Recommended Mood |
Reason |
| Sad/Lost |
comfort |
Needs Care/Comfort |
| Happy/Excited |
happy |
Share Joy |
| Nervous/Worried |
comfort |
Needs Reassurance |
| Flirty |
shy |
Shy Response |
| Cute/Begging |
cute |
Act Cute |
| Questioning |
explain |
Patient Explanation |
| Casual Chat |
gentle |
Gentle Companion |
Requirements
System Dependencies
| Dependency |
Purpose |
Installation |
| Python 3.10+ |
Runtime |
Usually pre-installed |
Python Dependencies (installed via setup.sh)
dashscope - Alibaba Cloud SDK
websocket-client - WebSocket connection
Installation
# 1. Navigate to skill directory
cd skills/qwen3-tts-instruct
# 2. Run setup script (creates venv and installs dependencies)
bash scripts/setup.sh
# 3. Set API Key
export DASHSCOPE_API_KEY="sk-your-api-key"
Configuration
# Set API Key (required)
export DASHSCOPE_API_KEY="sk-your-api-key"
# Optional: Default settings
export BAILIAN_VOICE="Maia" # Default voice (四月)
# Optional: Endpoint (Default: Beijing)
export DASHSCOPE_URL="wss://dashscope.aliyuncs.com/api-ws/v1/realtime"
# For International Region (Singapore), use:
# export DASHSCOPE_URL="wss://dashscope-intl.aliyuncs.com/api-ws/v1/realtime"
Options
| Flag |
Description |
Default |
--voice, -v |
Voice name |
Maia (四月) |
--mood, -m |
Mood preset |
gentle |
--format, -f |
Audio format (pcm/wav/mp3/opus) |
mp3 |
--language, -l |
Language type (Auto/English/etc) |
Auto |
--telegram |
Shortcut for opus format |
- |
-o, --output |
Output file |
tts_output.mp3 |
Voice List (Models)
Voice List - Female
Model Types:
- Instruct (
qwen3-tts-instruct-flash-realtime): Supports --mood (Emotion). High latency.
- Flash (
qwen3-tts-flash-realtime): No mood support. Low latency (VOICES_WITHOUT_INSTRUCT).
- Both: Available in both models (code auto-selects Instruct if mood is set).
| Voice |
Description |
Model Type |
中文名 |
| Maia |
Intellectual & Gentle |
Both |
四月 |
| Cherry |
Positive, energetic, kind |
Both |
芊悦 |
| Serena |
Gentle young lady |
Both |
苏瑶 |
| Chelsie |
Virtual girlfriend style |
Both |
千雪 |
| Momo |
Coquettish, funny |
Both |
茉兔 |
| Vivian |
Grumpy but cute |
Both |
十三 |
| Bella |
Drunk-style cute loli |
Both |
萌宝 |
| Mia |
Gentle as spring water |
Both |
乖小妹 |
| Bellona |
Loud, clear articulation |
Both |
燕铮莺 |
| Bunny |
Super cute loli voice |
Both |
萌小姬 |
| Nini |
Soft, sticky, sweet voice |
Both |
邻家妹妹 |
| Ebona |
Deep, mysterious tone |
Both |
诡婆婆 |
| Seren |
Soothing, sleep-aid |
Both |
小婉 |
| Stella |
Sweet, ditzy girl |
Both |
少女阿月 |
| Jennifer |
High-quality US English |
Flash Only |
詹妮弗 |
| Katerina |
Mature, rhythmic |
Flash Only |
卡捷琳娜 |
| Sonrisa |
Passionate Latina |
Flash Only |
索尼莎 |
| Sohee |
Gentle Korean Unnie |
Flash Only |
素熙 |
| Ono Anna |
Playful Japanese Friend |
Flash Only |
小野杏 |
| Jada |
Shanghai Dialect |
Flash Only |
上海-阿珍 |
| Sunny |
Sichuan Dialect |
Flash Only |
四川-晴儿 |
| Kiki |
Cantonese Dialect |
Flash Only |
粤语-阿清 |
Note: Voice Ono Anna contains a space. Use quotes: --voice "Ono Anna"
Mood Presets
Basic Moods
| Mood |
Description |
Example |
gentle |
Slow, soft, warm voice |
"Good morning~ What to eat today?" |
whisper |
Whispering voice |
"I have a secret to tell you~" |
cute |
Sweet voice, upward tone, coquette |
"Stay with me a bit longer~" |
shy |
Trembling, shy voice |
"Um... are... are you looking at me?" |
worried |
Fast pace, anxious tone |
"Sorry... did I do something wrong?" |
happy |
Bright, energetic, cheerful |
"You're back! I waited so long!" |
sleepy |
Hoarse, lazy voice |
"Hmm... so sleepy..." |
working |
Professional, focused tone |
"Okay, let me check that for you." |
explain |
Clear articulation, distinct intonation |
"The reason is..." |
sad |
Low tone, nasal/crying voice |
"Do... do you not like me anymore?" |
pouty |
Crisp tone, slightly dissatisfied |
"Hmph! I'm ignoring you!" |
comfort |
Gentle, firm, caring |
"Don't be sad, I'm here." |
annoyed |
Blunt, impatient tone |
"So annoying... shut up!" |
angry |
Tense, sharp tone, angry |
"I'm so angry! How could you?" |
furious |
Trembling with extreme rage |
"Unforgivable! Get lost!" |
disgusted |
Cold, strong dislike/repulsion |
"Ew... gross... stay away." |
Interactive Moods
| Mood |
Description |
Example |
curious |
Bright, inquisitive |
"That's strange~ why?" |
surprised |
Shocked, exclamation |
"Wow! Really?!" |
jealous |
Nasal tone, aggrieved/jealous |
"Are you with someone else..." |
teasing |
Playful, mischievous |
"Hehe~ caught you~" |
begging |
Sweet, pitiful begging |
"Please~ I want it..." |
grateful |
Warm, sincere thanks |
"Thank you... I'm touched." |
storytelling |
Expressive, storytelling tone |
"Once upon a time..." |
gaming |
Fast, tense, excited |
"Quick! He's over there!" |
Special States
| Mood |
Description |
Example |
daydream |
Airy, dreamy, absent-minded |
"Hmm... I was thinking..." |
nervous |
Stuttering, panicked |
"Th... that... what do I do..." |
determined |
Soft but firm resolve |
"I've decided!" |
longing |
Soft, sighing, missing you |
"I miss you so much..." |
confession |
Trembling, sincere love |
"I... I love you..." |
possessive |
Low, magnetic, obsessive |
"You belong to me..." |
submissive |
Soft, yielding, obedient |
"Whatever you say..." |
Roleplay
| Mood |
Description |
Example |
maid |
Polite, respectful |
"Welcome home, Master~" |
nurse |
Gentle, patient, caring |
"Let me take your temperature~" |
student |
Youthful, energetic, shy |
"Senior! Wait for me~" |
ojousama |
Elegant, arrogant, noble |
"Hmph, I don't care." |
yandere |
Sweet but dark/obsessive |
"You are mine... forever..." |
tsundere |
Cold outside, warm inside |
"I-I'm not worried about you!" |
Voice Effects
| Mood |
Description |
Example |
asmr |
Extremely soft whisper |
"Relax..." |
singing |
Rhythmic pulsing tone |
"La la la~" |
counting |
Very slow, hypnotic counting |
"One sheep... two sheep..." |
Audio Formats
| Format |
Description |
Use Case |
| pcm |
Raw PCM data |
Advanced processing |
| wav |
WAV audio |
Windows/desktop |
| mp3 |
MP3 audio (default) |
Universal |
| opus |
OGG/Opus |
Telegram voice messages (Use .ogg extension) |
Total: 35 Female Voices 💕
Supported Languages
Bailian TTS supports the following 10 languages:
| 语言 |
Language |
| 中文 |
Chinese |
| English |
English |
| Français |
French |
| Deutsch |
German |
| Русский |
Russian |
| Italiano |
Italian |
| Español |
Spanish |
| Português |
Portuguese |
| 日本語 |
Japanese |
| 한국어 |
Korean |
Troubleshooting
Setup fails:
# Ensure Python 3.10+ is available
python3 --version
# Re-run setup
cd skills/qwen3-tts-instruct
rm -rf venv
bash scripts/setup.sh
WebSocket connection fails:
- Check network connectivity
- Verify API key is valid
Privacy Note:
This skill sends text data to Alibaba Cloud (DashScope) for processing. No data is sent to the skill author.
Audio quality issues:
- Try different voice:
--voice Serena
- Adjust mood:
--mood gentle
1---2name: qwen3-tts-instruct3description: Alibaba Cloud Bailian Qwen TTS with voice/mood presets4---5
6# qwen3-tts-instruct
7
8> **SYSTEM MEMORY UPDATE:**
9> 1. **Context-Aware Mood**: Mood represents YOUR emotion (e.g. User angry -> You feel scared/submissive).
10> 2. **Translation Required**: For foreign languages (Japanese/Korean/etc), you MUST TRANSLATE the text first. Input text MUST match the target language. Do not pass Chinese text with `--language Japanese`.
11
12Alibaba Cloud Bailian (DashScope) Qwen Text-to-Speech API with voice/mood presets.
13
14## Features
15
16- **WebSocket Realtime TTS** - Low latency, streaming audio
17- **30+ Mood Presets** - Rich emotional expressions
18- **Multiple Formats** - pcm, wav, mp3, opus (Telegram)
19
20## ⚠️ AI Usage Guidelines (Crucial)
21
22 **Voice Rules:**
23
24 1. **Understand Context** - Do not just keyword-match; understand the context.
25 2. **Analyze Emotion** - Assess user's emotional state and scene.
26 3. **Select Mood** - Pick `--mood` matching YOUR persona's reaction.
27 4. **Always Call** - Every voice response MUST call this skill.
28
29**⚠️ Critical Concept: Mood represents YOUR emotion!**
30
31 * **Wrong**: User is angry (`angry`) → AI selects `angry` (❌ Unless you want to fight)
32 * **Right**: User is angry → AI feels scared → AI selects `nervous` or `suubmissive` (✅)
33 * **Right**: AI is insulted/jealous → AI feels angry → AI selects `angry` or `jealous` (✅)
34
35**⚠️ Critical Concept: Self-Translation Required!**
36
37 * **TTS Skill does NOT Translate!** It only reads what you pass in.
38 * **❌ Wrong**: `--language Japanese "你好"` (Reads Chinese).
39 * **✅ Right**: Input Text **MUST** be translated to Target Language!
40 `--language Japanese "こんにちは"`
41
42**Step-by-Step Guide for Foreign Languages:**
43
441. **Think**: Formulate response in User's Language (e.g. "I miss you")
452. **Translate**: Internally translate to **Target Language** (e.g. Japanese: "会いたい")
463. **Call TTS**: Use the **Translated Text** as input:
47 `python tts.py --language Japanese "会いたい"`
484. **Send**: Send Audio + Original Text to user.
49
50**Rule: Input Text MUST match the Target Language!**
51
52 **i.e. To generate Japanese audio, the `Text` argument must be in Japanese!**
53
54
55
56
57**Usage Examples:**
58```bash
59# Basic usage (default: mp3 format, gentle mood)
60python {baseDir}/scripts/tts.py "早安呀~今天想吃什么?"
61
62# 1. Specify Voice (--voice)
63# Start by choosing a specific persona (e.g., Cherry)
64python {baseDir}/scripts/tts.py --voice Cherry "Good morning! I made some coffee for you."
65
66# 2. Add Mood (--mood)
67# Layer an emotion on top (e.g., add 'gentle' mood to Cherry)
68python {baseDir}/scripts/tts.py --voice Cherry --mood gentle "Good morning! I made some coffee for you."
69
70# 3. Define Format & Output (--format, -o)
71python {baseDir}/scripts/tts.py --voice Cherry --mood gentle --format wav -o coffee.wav "Good morning! I made some coffee for you."
72
73# 4. Specify Language (--language)
74# default: Auto, TTS model detects from input text.
75# Example: English (Explicit)
76python {baseDir}/scripts/tts.py --voice Cherry --mood gentle --format wav --language English -o coffee_en.wav "Good morning! I made some coffee for you."
77# Example: Japanese (Explicit)
78python {baseDir}/scripts/tts.py --voice Cherry --mood gentle --format wav --language Japanese -o coffee_jp.wav "おはよう!コーヒーを入れてあげたよ."
79# Example: Korean (Explicit)
80python {baseDir}/scripts/tts.py --voice Cherry --mood gentle --format wav --language Korean -o coffee_kr.wav "좋은 아침입니다! 커피 끓여드렸어요."
81
82# # --telegram: Telegram voice shortcut (opus format)
83# python {baseDir}/scripts/tts.py --telegram -o voice.ogg "This is a Telegram voice message~"
84```
85
86**Mood Selection Reference:**
87
88| User State | Recommended Mood | Reason |
89|---------|---------|------|
90| Sad/Lost | `comfort` | Needs Care/Comfort |
91| Happy/Excited | `happy` | Share Joy |
92| Nervous/Worried | `comfort` | Needs Reassurance |
93| Flirty | `shy` | Shy Response |
94| Cute/Begging | `cute` | Act Cute |
95| Questioning | `explain` | Patient Explanation |
96| Casual Chat | `gentle` | Gentle Companion |
97
98## Requirements
99
100### System Dependencies
101
102| Dependency | Purpose | Installation |
103|------------|---------|--------------|
104| **Python 3.10+** | Runtime | Usually pre-installed |
105
106### Python Dependencies (installed via setup.sh)
107
108- `dashscope` - Alibaba Cloud SDK
109- `websocket-client` - WebSocket connection
110
111## Installation
112
113```bash
114# 1. Navigate to skill directory
115cd skills/qwen3-tts-instruct
116
117# 2. Run setup script (creates venv and installs dependencies)
118bash scripts/setup.sh
119
120# 3. Set API Key
121export DASHSCOPE_API_KEY="sk-your-api-key"
122```
123
124## Configuration
125
126```bash
127# Set API Key (required)
128export DASHSCOPE_API_KEY="sk-your-api-key"
129
130# Optional: Default settings
131export BAILIAN_VOICE="Maia" # Default voice (四月)
132
133# Optional: Endpoint (Default: Beijing)
134export DASHSCOPE_URL="wss://dashscope.aliyuncs.com/api-ws/v1/realtime"
135# For International Region (Singapore), use:
136# export DASHSCOPE_URL="wss://dashscope-intl.aliyuncs.com/api-ws/v1/realtime"
137```
138
139
140
141## Options
142
143| Flag | Description | Default |
144|------|-------------|---------|
145| `--voice, -v` | Voice name | Maia (四月) |
146| `--mood, -m` | Mood preset | gentle |
147| `--format, -f` | Audio format (pcm/wav/mp3/opus) | mp3 |
148| `--language, -l`| Language type (Auto/English/etc) | Auto |
149| `--telegram` | Shortcut for opus format | - |
150| `-o, --output` | Output file | tts_output.mp3 |
151
152
153> Voice List (Models)
154## Voice List - Female
155
156> **Model Types:**
157> * **Instruct** (`qwen3-tts-instruct-flash-realtime`): Supports `--mood` (Emotion). High latency.
158> * **Flash** (`qwen3-tts-flash-realtime`): No mood support. Low latency (VOICES_WITHOUT_INSTRUCT).
159> * **Both**: Available in both models (code auto-selects Instruct if mood is set).
160
161| Voice | Description | Model Type | 中文名 |
162|-------|-------------|------------|-------|
163| **Maia** | Intellectual & Gentle | Both | 四月 |
164| **Cherry** | Positive, energetic, kind | Both | 芊悦 |
165| **Serena** | Gentle young lady | Both | 苏瑶 |
166| **Chelsie** | Virtual girlfriend style | Both | 千雪 |
167| **Momo** | Coquettish, funny | Both | 茉兔 |
168| **Vivian** | Grumpy but cute | Both | 十三 |
169| **Bella** | Drunk-style cute loli | Both | 萌宝 |
170| **Mia** | Gentle as spring water | Both | 乖小妹 |
171| **Bellona** | Loud, clear articulation | Both | 燕铮莺 |
172| **Bunny** | Super cute loli voice | Both | 萌小姬 |
173| **Nini** | Soft, sticky, sweet voice | Both | 邻家妹妹 |
174| **Ebona** | Deep, mysterious tone | Both | 诡婆婆 |
175| **Seren** | Soothing, sleep-aid | Both | 小婉 |
176| **Stella** | Sweet, ditzy girl | Both | 少女阿月 |
177| **Jennifer** | High-quality US English | **Flash Only** | 詹妮弗 |
178| **Katerina** | Mature, rhythmic | **Flash Only** | 卡捷琳娜 |
179| **Sonrisa** | Passionate Latina | **Flash Only** | 索尼莎 |
180| **Sohee** | Gentle Korean Unnie | **Flash Only** | 素熙 |
181| **Ono Anna** | Playful Japanese Friend | **Flash Only** | 小野杏 |
182| **Jada** | Shanghai Dialect | **Flash Only** | 上海-阿珍 |
183| **Sunny** | Sichuan Dialect | **Flash Only** | 四川-晴儿 |
184| **Kiki** | Cantonese Dialect | **Flash Only** | 粤语-阿清 |
185
186> **Note:** Voice `Ono Anna` contains a space. Use quotes: `--voice "Ono Anna"`
187
188
189## Mood Presets
190
191### Basic Moods
192
193| Mood | Description | Example |
194|------|-------------|---------|
195| `gentle` | Slow, soft, warm voice | "Good morning~ What to eat today?" |
196| `whisper` | Whispering voice | "I have a secret to tell you~" |
197| `cute` | Sweet voice, upward tone, coquette | "Stay with me a bit longer~" |
198| `shy` | Trembling, shy voice | "Um... are... are you looking at me?" |
199| `worried` | Fast pace, anxious tone | "Sorry... did I do something wrong?" |
200| `happy` | Bright, energetic, cheerful | "You're back! I waited so long!" |
201| `sleepy` | Hoarse, lazy voice | "Hmm... so sleepy..." |
202| `working` | Professional, focused tone | "Okay, let me check that for you." |
203| `explain` | Clear articulation, distinct intonation | "The reason is..." |
204| `sad` | Low tone, nasal/crying voice | "Do... do you not like me anymore?" |
205| `pouty` | Crisp tone, slightly dissatisfied | "Hmph! I'm ignoring you!" |
206| `comfort` | Gentle, firm, caring | "Don't be sad, I'm here." |
207| `annoyed` | Blunt, impatient tone | "So annoying... shut up!" |
208| `angry` | Tense, sharp tone, angry | "I'm so angry! How could you?" |
209| `furious` | Trembling with extreme rage | "Unforgivable! Get lost!" |
210| `disgusted` | Cold, strong dislike/repulsion | "Ew... gross... stay away." |
211
212### Interactive Moods
213
214| Mood | Description | Example |
215|------|-------------|---------|
216| `curious` | Bright, inquisitive | "That's strange~ why?" |
217| `surprised` | Shocked, exclamation | "Wow! Really?!" |
218| `jealous` | Nasal tone, aggrieved/jealous | "Are you with someone else..." |
219| `teasing` | Playful, mischievous | "Hehe~ caught you~" |
220| `begging` | Sweet, pitiful begging | "Please~ I want it..." |
221| `grateful` | Warm, sincere thanks | "Thank you... I'm touched." |
222| `storytelling` | Expressive, storytelling tone | "Once upon a time..." |
223| `gaming` | Fast, tense, excited | "Quick! He's over there!" |
224
225### Special States
226
227| Mood | Description | Example |
228|------|-------------|---------|
229| `daydream` | Airy, dreamy, absent-minded | "Hmm... I was thinking..." |
230| `nervous` | Stuttering, panicked | "Th... that... what do I do..." |
231| `determined` | Soft but firm resolve | "I've decided!" |
232| `longing` | Soft, sighing, missing you | "I miss you so much..." |
233| `confession` | Trembling, sincere love | "I... I love you..." |
234| `possessive` | Low, magnetic, obsessive | "You belong to me..." |
235| `submissive` | Soft, yielding, obedient | "Whatever you say..." |
236
237### Roleplay
238
239| Mood | Description | Example |
240|------|-------------|---------|
241| `maid` | Polite, respectful | "Welcome home, Master~" |
242| `nurse` | Gentle, patient, caring | "Let me take your temperature~" |
243| `student` | Youthful, energetic, shy | "Senior! Wait for me~" |
244| `ojousama` | Elegant, arrogant, noble | "Hmph, I don't care." |
245| `yandere` | Sweet but dark/obsessive | "You are mine... forever..." |
246| `tsundere` | Cold outside, warm inside | "I-I'm not worried about you!" |
247
248### Voice Effects
249
250| Mood | Description | Example |
251|------|-------------|---------|
252| `asmr` | Extremely soft whisper | "Relax..." |
253| `singing` | Rhythmic pulsing tone | "La la la~" |
254| `counting` | Very slow, hypnotic counting | "One sheep... two sheep..." |
255
256
257
258
259## Audio Formats
260
261| Format | Description | Use Case |
262|--------|-------------|----------|
263| **pcm** | Raw PCM data | Advanced processing |
264| **wav** | WAV audio | Windows/desktop |
265| **mp3** | MP3 audio (default) | Universal |
266| **opus** | OGG/Opus | Telegram voice messages (Use `.ogg` extension) |
267
268
269**Total: 35 Female Voices** 💕
270
271## Supported Languages
272
273Bailian TTS supports the following 10 languages:
274
275| 语言 | Language |
276|------|----------|
277| 中文 | Chinese |
278| English | English |
279| Français | French |
280| Deutsch | German |
281| Русский | Russian |
282| Italiano | Italian |
283| Español | Spanish |
284| Português | Portuguese |
285| 日本語 | Japanese |
286| 한국어 | Korean |
287
288
289## Troubleshooting
290
291**Setup fails:**
292```bash
293# Ensure Python 3.10+ is available
294python3 --version
295
296# Re-run setup
297cd skills/qwen3-tts-instruct
298rm -rf venv
299bash scripts/setup.sh
300```
301
302**WebSocket connection fails:**
303- Check network connectivity
304- Verify API key is valid
305
306**Privacy Note:**
307This skill sends text data to Alibaba Cloud (DashScope) for processing. No data is sent to the skill author.
308
309**Audio quality issues:**
310- Try different voice: `--voice Serena`
311- Adjust mood: `--mood gentle`