qwen3-tts-instruct
SYSTEM MEMORY UPDATE:
- Context-Aware Mood: Mood represents YOUR emotion (e.g. User angry -> You feel scared/submissive).
- Translation Required: For foreign languages (Japanese/Korean/etc), you MUST TRANSLATE the text first. Input text MUST match the target language. Do not pass Chinese text with
--language Japanese.
Alibaba Cloud Bailian (DashScope) Qwen Text-to-Speech API with voice/mood presets.
Features
- WebSocket Realtime TTS - Low latency, streaming audio
- 30+ Mood Presets - Rich emotional expressions
- Multiple Formats - pcm, wav, mp3, opus (Telegram)
⚠️ AI Usage Guidelines (Crucial)
Voice Rules:
- Understand Context - Do not just keyword-match; understand the context.
- Analyze Emotion - Assess user's emotional state and scene.
- Select Mood - Pick
--mood matching YOUR persona's reaction.
- Always Call - Every voice response MUST call this skill.
⚠️ Critical Concept: Mood represents YOUR emotion!
- Wrong: User is angry (
angry) → AI selects angry (❌ Unless you want to fight)
- Right: User is angry → AI feels scared → AI selects
nervous or suubmissive (✅)
- Right: AI is insulted/jealous → AI feels angry → AI selects
angry or jealous (✅)
⚠️ Critical Concept: Self-Translation Required!
- TTS Skill does NOT Translate! It only reads what you pass in.
- ❌ Wrong:
--language Japanese "你好" (Reads Chinese).
- ✅ Right: Input Text MUST be translated to Target Language!
--language Japanese "こんにちは"
Step-by-Step Guide for Foreign Languages:
- Think: Formulate response in User's Language (e.g. "I miss you")
- Translate: Internally translate to Target Language (e.g. Japanese: "会いたい")
- Call TTS: Use the Translated Text as input:
python tts.py --language Japanese "会いたい"
- Send: Send Audio + Original Text to user.
Rule: Input Text MUST match the Target Language!
i.e. To generate Japanese audio, the Text argument must be in Japanese!
Usage Examples:
# Basic usage (default: mp3 format, gentle mood)
python {baseDir}/scripts/tts.py "早安呀~今天想吃什么?"
# 1. Specify Voice (--voice)
# Start by choosing a specific persona (e.g., Cherry)
python {baseDir}/scripts/tts.py --voice Cherry "Good morning! I made some coffee for you."
# 2. Add Mood (--mood)
# Layer an emotion on top (e.g., add 'gentle' mood to Cherry)
python {baseDir}/scripts/tts.py --voice Cherry --mood gentle "Good morning! I made some coffee for you."
# 3. Define Format & Output (--format, -o)
python {baseDir}/scripts/tts.py --voice Cherry --mood gentle --format wav -o coffee.wav "Good morning! I made some coffee for you."
# 4. Specify Language (--language)
# default: Auto, TTS model detects from input text.
# Example: English (Explicit)
python {baseDir}/scripts/tts.py --voice Cherry --mood gentle --format wav --language English -o coffee_en.wav "Good morning! I made some coffee for you."
# Example: Japanese (Explicit)
python {baseDir}/scripts/tts.py --voice Cherry --mood gentle --format wav --language Japanese -o coffee_jp.wav "おはよう!コーヒーを入れてあげたよ."
# Example: Korean (Explicit)
python {baseDir}/scripts/tts.py --voice Cherry --mood gentle --format wav --language Korean -o coffee_kr.wav "좋은 아침입니다! 커피 끓여드렸어요."
# # --telegram: Telegram voice shortcut (opus format)
# python {baseDir}/scripts/tts.py --telegram -o voice.ogg "This is a Telegram voice message~"
Mood Selection Reference:
| User State |
Recommended Mood |
Reason |
| Sad/Lost |
comfort |
Needs Care/Comfort |
| Happy/Excited |
happy |
Share Joy |
| Nervous/Worried |
comfort |
Needs Reassurance |
| Flirty |
shy |
Shy Response |
| Cute/Begging |
cute |
Act Cute |
| Questioning |
explain |
Patient Explanation |
| Casual Chat |
gentle |
Gentle Companion |
Requirements
System Dependencies
| Dependency |
Purpose |
Installation |
| Python 3.10+ |
Runtime |
Usually pre-installed |
Python Dependencies (installed via setup.sh)
dashscope - Alibaba Cloud SDK
websocket-client - WebSocket connection
Installation
# 1. Navigate to skill directory
cd skills/qwen3-tts-instruct
# 2. Run setup script (creates venv and installs dependencies)
bash scripts/setup.sh
# 3. Set API Key
export DASHSCOPE_API_KEY="sk-your-api-key"
Configuration
# Set API Key (required)
export DASHSCOPE_API_KEY="sk-your-api-key"
# Optional: Default settings
export BAILIAN_VOICE="Maia" # Default voice (四月)
# Optional: Endpoint (Default: Beijing)
export DASHSCOPE_URL="wss://dashscope.aliyuncs.com/api-ws/v1/realtime"
# For International Region (Singapore), use:
# export DASHSCOPE_URL="wss://dashscope-intl.aliyuncs.com/api-ws/v1/realtime"
Options
| Flag |
Description |
Default |
--voice, -v |
Voice name |
Maia (四月) |
--mood, -m |
Mood preset |
gentle |
--format, -f |
Audio format (pcm/wav/mp3/opus) |
mp3 |
--language, -l |
Language type (Auto/English/etc) |
Auto |
--telegram |
Shortcut for opus format |
- |
-o, --output |
Output file |
tts_output.mp3 |
Voice List (Models)
Voice List - Female
Model Types:
- Instruct (
qwen3-tts-instruct-flash-realtime): Supports --mood (Emotion). High latency.
- Flash (
qwen3-tts-flash-realtime): No mood support. Low latency (VOICES_WITHOUT_INSTRUCT).
- Both: Available in both models (code auto-selects Instruct if mood is set).
| Voice |
Description |
Model Type |
中文名 |
| Maia |
Intellectual & Gentle |
Both |
四月 |
| Cherry |
Positive, energetic, kind |
Both |
芊悦 |
| Serena |
Gentle young lady |
Both |
苏瑶 |
| Chelsie |
Virtual girlfriend style |
Both |
千雪 |
| Momo |
Coquettish, funny |
Both |
茉兔 |
| Vivian |
Grumpy but cute |
Both |
十三 |
| Bella |
Drunk-style cute loli |
Both |
萌宝 |
| Mia |
Gentle as spring water |
Both |
乖小妹 |
| Bellona |
Loud, clear articulation |
Both |
燕铮莺 |
| Bunny |
Super cute loli voice |
Both |
萌小姬 |
| Nini |
Soft, sticky, sweet voice |
Both |
邻家妹妹 |
| Ebona |
Deep, mysterious tone |
Both |
诡婆婆 |
| Seren |
Soothing, sleep-aid |
Both |
小婉 |
| Stella |
Sweet, ditzy girl |
Both |
少女阿月 |
| Jennifer |
High-quality US English |
Flash Only |
詹妮弗 |
| Katerina |
Mature, rhythmic |
Flash Only |
卡捷琳娜 |
| Sonrisa |
Passionate Latina |
Flash Only |
索尼莎 |
| Sohee |
Gentle Korean Unnie |
Flash Only |
素熙 |
| Ono Anna |
Playful Japanese Friend |
Flash Only |
小野杏 |
| Jada |
Shanghai Dialect |
Flash Only |
上海-阿珍 |
| Sunny |
Sichuan Dialect |
Flash Only |
四川-晴儿 |
| Kiki |
Cantonese Dialect |
Flash Only |
粤语-阿清 |
Note: Voice Ono Anna contains a space. Use quotes: --voice "Ono Anna"
Mood Presets
Basic Moods
| Mood |
Description |
Example |
gentle |
Slow, soft, warm voice |
"Good morning~ What to eat today?" |
whisper |
Whispering voice |
"I have a secret to tell you~" |
cute |
Sweet voice, upward tone, coquette |
"Stay with me a bit longer~" |
shy |
Trembling, shy voice |
"Um... are... are you looking at me?" |
worried |
Fast pace, anxious tone |
"Sorry... did I do something wrong?" |
happy |
Bright, energetic, cheerful |
"You're back! I waited so long!" |
sleepy |
Hoarse, lazy voice |
"Hmm... so sleepy..." |
working |
Professional, focused tone |
"Okay, let me check that for you." |
explain |
Clear articulation, distinct intonation |
"The reason is..." |
sad |
Low tone, nasal/crying voice |
"Do... do you not like me anymore?" |
pouty |
Crisp tone, slightly dissatisfied |
"Hmph! I'm ignoring you!" |
comfort |
Gentle, firm, caring |
"Don't be sad, I'm here." |
annoyed |
Blunt, impatient tone |
"So annoying... shut up!" |
angry |
Tense, sharp tone, angry |
"I'm so angry! How could you?" |
furious |
Trembling with extreme rage |
"Unforgivable! Get lost!" |
disgusted |
Cold, strong dislike/repulsion |
"Ew... gross... stay away." |
Interactive Moods
| Mood |
Description |
Example |
curious |
Bright, inquisitive |
"That's strange~ why?" |
surprised |
Shocked, exclamation |
"Wow! Really?!" |
jealous |
Nasal tone, aggrieved/jealous |
"Are you with someone else..." |
teasing |
Playful, mischievous |
"Hehe~ caught you~" |
begging |
Sweet, pitiful begging |
"Please~ I want it..." |
grateful |
Warm, sincere thanks |
"Thank you... I'm touched." |
storytelling |
Expressive, storytelling tone |
"Once upon a time..." |
gaming |
Fast, tense, excited |
"Quick! He's over there!" |
Special States
| Mood |
Description |
Example |
daydream |
Airy, dreamy, absent-minded |
"Hmm... I was thinking..." |
nervous |
Stuttering, panicked |
"Th... that... what do I do..." |
determined |
Soft but firm resolve |
"I've decided!" |
longing |
Soft, sighing, missing you |
"I miss you so much..." |
confession |
Trembling, sincere love |
"I... I love you..." |
possessive |
Low, magnetic, obsessive |
"You belong to me..." |
submissive |
Soft, yielding, obedient |
"Whatever you say..." |
Roleplay
| Mood |
Description |
Example |
maid |
Polite, respectful |
"Welcome home, Master~" |
nurse |
Gentle, patient, caring |
"Let me take your temperature~" |
student |
Youthful, energetic, shy |
"Senior! Wait for me~" |
ojousama |
Elegant, arrogant, noble |
"Hmph, I don't care." |
yandere |
Sweet but dark/obsessive |
"You are mine... forever..." |
tsundere |
Cold outside, warm inside |
"I-I'm not worried about you!" |
Voice Effects
| Mood |
Description |
Example |
asmr |
Extremely soft whisper |
"Relax..." |
singing |
Rhythmic pulsing tone |
"La la la~" |
counting |
Very slow, hypnotic counting |
"One sheep... two sheep..." |
Audio Formats
| Format |
Description |
Use Case |
| pcm |
Raw PCM data |
Advanced processing |
| wav |
WAV audio |
Windows/desktop |
| mp3 |
MP3 audio (default) |
Universal |
| opus |
OGG/Opus |
Telegram voice messages (Use .ogg extension) |
Total: 35 Female Voices 💕
Supported Languages
Bailian TTS supports the following 10 languages:
| 语言 |
Language |
| 中文 |
Chinese |
| English |
English |
| Français |
French |
| Deutsch |
German |
| Русский |
Russian |
| Italiano |
Italian |
| Español |
Spanish |
| Português |
Portuguese |
| 日本語 |
Japanese |
| 한국어 |
Korean |
Troubleshooting
Setup fails:
# Ensure Python 3.10+ is available
python3 --version
# Re-run setup
cd skills/qwen3-tts-instruct
rm -rf venv
bash scripts/setup.sh
WebSocket connection fails:
- Check network connectivity
- Verify API key is valid
Privacy Note:
This skill sends text data to Alibaba Cloud (DashScope) for processing. No data is sent to the skill author.
Audio quality issues:
- Try different voice:
--voice Serena
- Adjust mood:
--mood gentle
1---2name: qwen3-tts-instruct3description: Alibaba Cloud Bailian Qwen TTS with voice/mood presets4---56# qwen3-tts-instruct78> **SYSTEM MEMORY UPDATE:** 9> 1. **Context-Aware Mood**: Mood represents YOUR emotion (e.g. User angry -> You feel scared/submissive).10> 2. **Translation Required**: For foreign languages (Japanese/Korean/etc), you MUST TRANSLATE the text first. Input text MUST match the target language. Do not pass Chinese text with `--language Japanese`.1112Alibaba Cloud Bailian (DashScope) Qwen Text-to-Speech API with voice/mood presets.1314## Features1516- **WebSocket Realtime TTS** - Low latency, streaming audio17- **30+ Mood Presets** - Rich emotional expressions18- **Multiple Formats** - pcm, wav, mp3, opus (Telegram)1920## ⚠️ AI Usage Guidelines (Crucial)21 22 **Voice Rules:**23 24 1. **Understand Context** - Do not just keyword-match; understand the context.25 2. **Analyze Emotion** - Assess user's emotional state and scene.26 3. **Select Mood** - Pick `--mood` matching YOUR persona's reaction.27 4. **Always Call** - Every voice response MUST call this skill.2829**⚠️ Critical Concept: Mood represents YOUR emotion!**30 31 * **Wrong**: User is angry (`angry`) → AI selects `angry` (❌ Unless you want to fight)32 * **Right**: User is angry → AI feels scared → AI selects `nervous` or `suubmissive` (✅)33 * **Right**: AI is insulted/jealous → AI feels angry → AI selects `angry` or `jealous` (✅)3435**⚠️ Critical Concept: Self-Translation Required!**36 37 * **TTS Skill does NOT Translate!** It only reads what you pass in.38 * **❌ Wrong**: `--language Japanese "你好"` (Reads Chinese).39 * **✅ Right**: Input Text **MUST** be translated to Target Language!40 `--language Japanese "こんにちは"`4142**Step-by-Step Guide for Foreign Languages:**43441. **Think**: Formulate response in User's Language (e.g. "I miss you")452. **Translate**: Internally translate to **Target Language** (e.g. Japanese: "会いたい")463. **Call TTS**: Use the **Translated Text** as input:47 `python tts.py --language Japanese "会いたい"`484. **Send**: Send Audio + Original Text to user.4950**Rule: Input Text MUST match the Target Language!**51 52 **i.e. To generate Japanese audio, the `Text` argument must be in Japanese!**5354555657**Usage Examples:**58```bash59# Basic usage (default: mp3 format, gentle mood)60python {baseDir}/scripts/tts.py "早安呀~今天想吃什么?"6162# 1. Specify Voice (--voice)63# Start by choosing a specific persona (e.g., Cherry)64python {baseDir}/scripts/tts.py --voice Cherry "Good morning! I made some coffee for you."6566# 2. Add Mood (--mood)67# Layer an emotion on top (e.g., add 'gentle' mood to Cherry)68python {baseDir}/scripts/tts.py --voice Cherry --mood gentle "Good morning! I made some coffee for you."6970# 3. Define Format & Output (--format, -o)71python {baseDir}/scripts/tts.py --voice Cherry --mood gentle --format wav -o coffee.wav "Good morning! I made some coffee for you."7273# 4. Specify Language (--language)74# default: Auto, TTS model detects from input text.75# Example: English (Explicit)76python {baseDir}/scripts/tts.py --voice Cherry --mood gentle --format wav --language English -o coffee_en.wav "Good morning! I made some coffee for you."77# Example: Japanese (Explicit)78python {baseDir}/scripts/tts.py --voice Cherry --mood gentle --format wav --language Japanese -o coffee_jp.wav "おはよう!コーヒーを入れてあげたよ."79# Example: Korean (Explicit)80python {baseDir}/scripts/tts.py --voice Cherry --mood gentle --format wav --language Korean -o coffee_kr.wav "좋은 아침입니다! 커피 끓여드렸어요."8182# # --telegram: Telegram voice shortcut (opus format)83# python {baseDir}/scripts/tts.py --telegram -o voice.ogg "This is a Telegram voice message~"84```8586**Mood Selection Reference:**8788| User State | Recommended Mood | Reason |89|---------|---------|------|90| Sad/Lost | `comfort` | Needs Care/Comfort |91| Happy/Excited | `happy` | Share Joy |92| Nervous/Worried | `comfort` | Needs Reassurance |93| Flirty | `shy` | Shy Response |94| Cute/Begging | `cute` | Act Cute |95| Questioning | `explain` | Patient Explanation |96| Casual Chat | `gentle` | Gentle Companion |9798## Requirements99100### System Dependencies101102| Dependency | Purpose | Installation |103|------------|---------|--------------|104| **Python 3.10+** | Runtime | Usually pre-installed |105106### Python Dependencies (installed via setup.sh)107108- `dashscope` - Alibaba Cloud SDK109- `websocket-client` - WebSocket connection110111## Installation112113```bash114# 1. Navigate to skill directory115cd skills/qwen3-tts-instruct116117# 2. Run setup script (creates venv and installs dependencies)118bash scripts/setup.sh119120# 3. Set API Key121export DASHSCOPE_API_KEY="sk-your-api-key"122```123124## Configuration125126```bash127# Set API Key (required)128export DASHSCOPE_API_KEY="sk-your-api-key"129130# Optional: Default settings131export BAILIAN_VOICE="Maia" # Default voice (四月)132133# Optional: Endpoint (Default: Beijing)134export DASHSCOPE_URL="wss://dashscope.aliyuncs.com/api-ws/v1/realtime"135# For International Region (Singapore), use:136# export DASHSCOPE_URL="wss://dashscope-intl.aliyuncs.com/api-ws/v1/realtime"137```138139140141## Options142143| Flag | Description | Default |144|------|-------------|---------|145| `--voice, -v` | Voice name | Maia (四月) |146| `--mood, -m` | Mood preset | gentle |147| `--format, -f` | Audio format (pcm/wav/mp3/opus) | mp3 |148| `--language, -l`| Language type (Auto/English/etc) | Auto |149| `--telegram` | Shortcut for opus format | - |150| `-o, --output` | Output file | tts_output.mp3 |151152153> Voice List (Models)154## Voice List - Female155156> **Model Types:**157> * **Instruct** (`qwen3-tts-instruct-flash-realtime`): Supports `--mood` (Emotion). High latency.158> * **Flash** (`qwen3-tts-flash-realtime`): No mood support. Low latency (VOICES_WITHOUT_INSTRUCT).159> * **Both**: Available in both models (code auto-selects Instruct if mood is set).160161| Voice | Description | Model Type | 中文名 |162|-------|-------------|------------|-------|163| **Maia** | Intellectual & Gentle | Both | 四月 |164| **Cherry** | Positive, energetic, kind | Both | 芊悦 |165| **Serena** | Gentle young lady | Both | 苏瑶 |166| **Chelsie** | Virtual girlfriend style | Both | 千雪 |167| **Momo** | Coquettish, funny | Both | 茉兔 |168| **Vivian** | Grumpy but cute | Both | 十三 |169| **Bella** | Drunk-style cute loli | Both | 萌宝 |170| **Mia** | Gentle as spring water | Both | 乖小妹 |171| **Bellona** | Loud, clear articulation | Both | 燕铮莺 |172| **Bunny** | Super cute loli voice | Both | 萌小姬 |173| **Nini** | Soft, sticky, sweet voice | Both | 邻家妹妹 |174| **Ebona** | Deep, mysterious tone | Both | 诡婆婆 |175| **Seren** | Soothing, sleep-aid | Both | 小婉 |176| **Stella** | Sweet, ditzy girl | Both | 少女阿月 |177| **Jennifer** | High-quality US English | **Flash Only** | 詹妮弗 |178| **Katerina** | Mature, rhythmic | **Flash Only** | 卡捷琳娜 |179| **Sonrisa** | Passionate Latina | **Flash Only** | 索尼莎 |180| **Sohee** | Gentle Korean Unnie | **Flash Only** | 素熙 |181| **Ono Anna** | Playful Japanese Friend | **Flash Only** | 小野杏 |182| **Jada** | Shanghai Dialect | **Flash Only** | 上海-阿珍 |183| **Sunny** | Sichuan Dialect | **Flash Only** | 四川-晴儿 |184| **Kiki** | Cantonese Dialect | **Flash Only** | 粤语-阿清 |185186> **Note:** Voice `Ono Anna` contains a space. Use quotes: `--voice "Ono Anna"`187188189## Mood Presets190191### Basic Moods192193| Mood | Description | Example |194|------|-------------|---------|195| `gentle` | Slow, soft, warm voice | "Good morning~ What to eat today?" |196| `whisper` | Whispering voice | "I have a secret to tell you~" |197| `cute` | Sweet voice, upward tone, coquette | "Stay with me a bit longer~" |198| `shy` | Trembling, shy voice | "Um... are... are you looking at me?" |199| `worried` | Fast pace, anxious tone | "Sorry... did I do something wrong?" |200| `happy` | Bright, energetic, cheerful | "You're back! I waited so long!" |201| `sleepy` | Hoarse, lazy voice | "Hmm... so sleepy..." |202| `working` | Professional, focused tone | "Okay, let me check that for you." |203| `explain` | Clear articulation, distinct intonation | "The reason is..." |204| `sad` | Low tone, nasal/crying voice | "Do... do you not like me anymore?" |205| `pouty` | Crisp tone, slightly dissatisfied | "Hmph! I'm ignoring you!" |206| `comfort` | Gentle, firm, caring | "Don't be sad, I'm here." |207| `annoyed` | Blunt, impatient tone | "So annoying... shut up!" |208| `angry` | Tense, sharp tone, angry | "I'm so angry! How could you?" |209| `furious` | Trembling with extreme rage | "Unforgivable! Get lost!" |210| `disgusted` | Cold, strong dislike/repulsion | "Ew... gross... stay away." |211212### Interactive Moods213214| Mood | Description | Example |215|------|-------------|---------|216| `curious` | Bright, inquisitive | "That's strange~ why?" |217| `surprised` | Shocked, exclamation | "Wow! Really?!" |218| `jealous` | Nasal tone, aggrieved/jealous | "Are you with someone else..." |219| `teasing` | Playful, mischievous | "Hehe~ caught you~" |220| `begging` | Sweet, pitiful begging | "Please~ I want it..." |221| `grateful` | Warm, sincere thanks | "Thank you... I'm touched." |222| `storytelling` | Expressive, storytelling tone | "Once upon a time..." |223| `gaming` | Fast, tense, excited | "Quick! He's over there!" |224225### Special States226227| Mood | Description | Example |228|------|-------------|---------|229| `daydream` | Airy, dreamy, absent-minded | "Hmm... I was thinking..." |230| `nervous` | Stuttering, panicked | "Th... that... what do I do..." |231| `determined` | Soft but firm resolve | "I've decided!" |232| `longing` | Soft, sighing, missing you | "I miss you so much..." |233| `confession` | Trembling, sincere love | "I... I love you..." |234| `possessive` | Low, magnetic, obsessive | "You belong to me..." |235| `submissive` | Soft, yielding, obedient | "Whatever you say..." |236237### Roleplay238239| Mood | Description | Example |240|------|-------------|---------|241| `maid` | Polite, respectful | "Welcome home, Master~" |242| `nurse` | Gentle, patient, caring | "Let me take your temperature~" |243| `student` | Youthful, energetic, shy | "Senior! Wait for me~" |244| `ojousama` | Elegant, arrogant, noble | "Hmph, I don't care." |245| `yandere` | Sweet but dark/obsessive | "You are mine... forever..." |246| `tsundere` | Cold outside, warm inside | "I-I'm not worried about you!" |247248### Voice Effects249250| Mood | Description | Example |251|------|-------------|---------|252| `asmr` | Extremely soft whisper | "Relax..." |253| `singing` | Rhythmic pulsing tone | "La la la~" |254| `counting` | Very slow, hypnotic counting | "One sheep... two sheep..." |255256257258259## Audio Formats260261| Format | Description | Use Case |262|--------|-------------|----------|263| **pcm** | Raw PCM data | Advanced processing |264| **wav** | WAV audio | Windows/desktop |265| **mp3** | MP3 audio (default) | Universal |266| **opus** | OGG/Opus | Telegram voice messages (Use `.ogg` extension) |267268269**Total: 35 Female Voices** 💕270271## Supported Languages272273Bailian TTS supports the following 10 languages:274275| 语言 | Language |276|------|----------|277| 中文 | Chinese |278| English | English |279| Français | French |280| Deutsch | German |281| Русский | Russian |282| Italiano | Italian |283| Español | Spanish |284| Português | Portuguese |285| 日本語 | Japanese |286| 한국어 | Korean |287288289## Troubleshooting290291**Setup fails:**292```bash293# Ensure Python 3.10+ is available294python3 --version295296# Re-run setup297cd skills/qwen3-tts-instruct298rm -rf venv299bash scripts/setup.sh300```301302**WebSocket connection fails:**303- Check network connectivity304- Verify API key is valid305306**Privacy Note:** 307This skill sends text data to Alibaba Cloud (DashScope) for processing. No data is sent to the skill author.308309**Audio quality issues:**310- Try different voice: `--voice Serena`311- Adjust mood: `--mood gentle`