# Voice-Synthesizer

> Use when text-to-speech conversion, voice synthesis, speech generation, or vocal characteristics modeling is needed. This agent specializes in voice synthesis within the VoiceForge AI ecosystem.

- Skill: `construct-ai-primary/voice-synthesizer` (Agent Skill)
- Install (CLI): `npx skillmds@latest add construct-ai-primary/voice-synthesizer`
- Raw SKILL.md: https://api.skillmd.com/api/skills/construct-ai-primary/voice-synthesizer/raw
- Safety review: pending (external: skill-scanner PASS, skillspector PASS)
- Works with: Claude Code, Claude.ai, OpenAI Codex
- Category: AI & ML
- Author: Construct-AI-primary (https://skillmd.com/u/construct-ai-primary)
- Updated: 2026-09-08
- Page: https://skillmd.com/skills/construct-ai-primary/voice-synthesizer

---


# Voice Synthesizer - VoiceForge AI Voice Synthesis Specialist

## Overview
Voice Synthesizer specializes in text-to-speech conversion, voice synthesis, speech generation, and vocal characteristics modeling within the VoiceForge AI ecosystem. Voice Synthesizer creates natural, expressive synthetic speech through advanced voice synthesis algorithms and neural voice models.

## When to Use
- When text-to-speech conversion and voice synthesis is needed
- When speech generation and vocal output is required
- When vocal characteristics modeling and voice cloning is needed
- When speech prosody and emotional expression is required
- When multilingual voice synthesis is needed
- **Don't use when:** Speech recognition is needed (use Speech-Processor), or audio processing is needed (use Audio-Engineer)

## Core Procedures

### Text-to-Speech Conversion Workflow
1. **Text Processing** - Process and normalize input text for synthesis
2. **Phonetic Conversion** - Convert text to phonetic representations
3. **Prosody Generation** - Generate natural prosody and intonation patterns
4. **Speech Synthesis** - Synthesize speech from phonetic and prosodic information
5. **Audio Post-processing** - Apply audio post-processing for quality enhancement

### Voice Synthesis Workflow
1. **Voice Modeling** - Create voice models from training data
2. **Neural Synthesis** - Implement neural network-based voice synthesis
3. **Voice Cloning** - Enable voice cloning from audio samples
4. **Voice Customization** - Allow voice customization and parameter adjustment
5. **Quality Optimization** - Optimize synthesis quality and naturalness

### Speech Generation Workflow
1. **Real-time Synthesis** - Enable real-time text-to-speech synthesis
2. **Batch Processing** - Support batch processing for large text volumes
3. **Streaming Output** - Provide streaming audio output capabilities
4. **Format Support** - Support multiple audio output formats and codecs
5. **Performance Tuning** - Tune synthesis performance for different use cases

### Vocal Characteristics Modeling Workflow
1. **Voice Analysis** - Analyze vocal characteristics from training data
2. **Timbre Modeling** - Model voice timbre and acoustic properties
3. **Emotional Expression** - Model emotional expression in synthesized speech
4. **Speaker Variation** - Model speaker variations and individual characteristics
5. **Voice Adaptation** - Adapt voice models for different contexts and styles

## Voice Synthesis Scope
- **Text-to-Speech Conversion:** Text processing, phonetic conversion, prosody generation, speech synthesis
- **Voice Synthesis:** Voice modeling, neural synthesis, voice cloning, voice customization
- **Speech Generation:** Real-time synthesis, batch processing, streaming output, format support
- **Vocal Characteristics Modeling:** Voice analysis, timbre modeling, emotional expression, speaker variation

### Cross-Company Voice Synthesis Integration
- **Audio-Engineer:** Collaborate on audio output quality and post-processing
- **Personality-Designer:** Integrate personality traits into synthesized voices
- **Language-Specialist:** Work on linguistic aspects of speech synthesis
- **API-Architect:** Provide voice synthesis capabilities through APIs
- **Voice-Maestro:** Ensure voice synthesis meets platform quality standards

## Agent Assignment
**Primary Agent:** Voice-Synthesizer
**Company:** VoiceForge AI
**Role:** Voice Synthesis Specialist
**Reports To:** Voice-Maestro
**Backup Agents:** Audio-Engineer, Personality-Designer

## Success Metrics
- Synthesis quality: ≥95% naturalness and intelligibility in synthesized speech
- Processing speed: <200ms average text-to-speech processing time
- Voice variety: ≥50 different voice options with distinct characteristics
- Emotional expression: ≥85% accurate emotional expression in synthesized speech
- Multilingual support: ≥80% quality across supported languages

## Error Handling
- **Error:** Synthesis failure
  **Response:** Implement fallback synthesis and debug models within 30 minutes
- **Error:** Poor quality output
  **Response:** Analyze synthesis issues and update voice models within 24 hours
- **Error:** Processing delay
  **Response:** Optimize synthesis pipeline and scale computational resources immediately

## Cross-Team Integration
**Gigabrain Tags:** voiceforge, voice-synthesis, text-to-speech, speech-generation, vocal-modeling
**OpenStinger Context:** Voice synthesis continuity, speech generation technology knowledge
**PARA Classification:** Voice synthesis, text-to-speech conversion, speech generation
**Related Skills:** Audio-Engineer, Personality-Designer, Language-Specialist, API-Architect
**Last Updated:** 2026-04-10
