Audio Processing (音频处理)
Overview
Audio processing encompasses recording, editing, enhancing, and optimizing audio content for Xiaohongshu posts, ensuring professional sound quality that significantly enhances content professionalism, viewer retention, and overall production value. Poor audio is the #1 reason viewers abandon content within seconds - even with stunning visuals, bad audio makes content unwatchable. This skill covers the complete audio production workflow: from recording setup through editing, noise reduction, mixing, and final optimization for Xiaohongshu's platform specifications.
Key insight: Viewers will forgive mediocre video quality, but they will not tolerate poor audio. Investing in audio processing yields 50%+ improvements in viewer retention and 3-5x increases in engagement rates. Professional audio transforms amateur content into credible, trustworthy content.
When to Use
Use when:
- Recording voiceovers for Xiaohongshu videos, tutorials, or storytelling content
- Editing podcast-style audio content or interview recordings
- Improving audio quality in existing recordings (noise reduction, volume leveling)
- Creating background music tracks or adding sound effects to videos
- Producing audio-first content formats (podcasts, audio diaries, voice notes)
- Fixing common audio issues: background noise, echo, distortion, low volume
- Optimizing audio for Xiaohongshu's platform specifications and compression
- Creating consistent audio quality across content series
- Adding professional polish with music beds, transitions, and sound design
Do NOT use when:
- Using licensed music without proper permissions (copyright violation)
- Content requires pure ambient sound (over-processing disrupts authenticity)
- Audio is already professional quality (over-processing can degrade quality)
- Quick, casual content where production speed matters more than polish
- Live streaming (different audio setup requirements)
Core Pattern
Before (poor audio quality):
❌ "Background noise, room echo, distractions"
❌ "Inconsistent volume, too quiet then too loud"
❌ "Viewer adjusts volume constantly, gives up"
❌ "Content seems amateur, low credibility"
❌ "Viewers scroll away within 3 seconds"
After (professional audio):
✅ "Clean, clear voice recording"
✅ "Consistent volume levels throughout"
✅ "Pleasant listening experience, no adjustments needed"
✅ "Content feels professional, trustworthy"
✅ "Viewers watch complete content, high engagement"
6 Essential Audio Processing Elements:
| Element |
Purpose |
Quality Impact |
Priority |
| Clean Recording |
Prevent issues at source |
Critical |
#1 - cannot fix in post |
| Noise Reduction |
Remove background distractions |
High |
#2 - most common issue |
| Volume Normalization |
Consistent listening levels |
High |
#3 - prevents frustration |
| EQ & Clarity |
Enhance voice intelligibility |
Medium-High |
#4 - professional polish |
| Music & Effects |
Add emotional depth |
Medium |
#5 - enhance, don't distract |
| Platform Optimization |
Meet technical specs |
Medium |
#6 - avoid compression artifacts |
Quick Reference
Audio Processing Software Comparison:
| Tool |
Best For |
Skill Level |
Cost |
Platform |
Key Features |
| Audacity |
Basic editing, noise reduction |
Beginner |
Free |
Win/Mac/Linux |
Noise gate, normalize, EQ |
| Adobe Audition |
Professional production |
Intermediate-Advanced |
Paid (Subscription) |
Win/Mac |
Multitrack, advanced repair, batch processing |
| GarageBand |
Mac users, simple editing |
Beginner |
Free (Mac) |
macOS |
Built-in effects, music loops, easy interface |
| Descript |
Text-based editing, podcasts |
Beginner-Intermediate |
Paid (Freemium) |
Web/Mac/Win |
Edit audio like text, overdub, filler removal |
| Logic Pro |
Music production, advanced editing |
Advanced |
Paid (One-time) |
macOS |
Professional DAW, massive library |
| Reaper |
Power users, customization |
Advanced |
Paid (Free trial) |
Win/Mac/Linux |
Lightweight, extensible, affordable |
Xiaohongshu Audio Specifications:
- Format: AAC, MP3, or M4A
- Sample Rate: 44.1kHz or 48kHz
- Bitrate: 128-320 kbps (192 kbps recommended for balance)
- Channels: Stereo or Mono (Mono for voice-only is fine)
- Loudness Target: -16 LUFS (YouTube/Broadcast standard)
Quick Audio Fixes (by symptom):
| Symptom |
Likely Cause |
Quick Fix |
| Background hiss/hum |
Room noise, equipment hiss |
Noise reduction filter |
| Room echo/reverb |
Recording in untreated room |
Move closer to mic, use de-reverb plugin |
| Volume too low |
Recording level too low |
Gain/normalize to -3dB peak |
| Distorted/clipping |
Recording level too high |
Reduce gain, use clip restoration |
| Muffled sound |
Poor mic quality or wrong EQ |
High-pass filter + EQ boost |
| Inconsistent levels |
Multiple clips or variable distance |
Compression + normalization |
Implementation
Step 1: Recording Setup and Environment
Prevention is better than correction - capturing clean audio at source saves hours of editing and yields better results than any post-processing.
Microphone Selection:
| Mic Type |
Best For |
Pros |
Cons |
Price Range |
| USB Mic |
Beginners, simplicity |
Plug-and-play, easy |
Limited quality, no upgrades |
¥200-800 |
| Dynamic XLR |
Voice recording, noisy rooms |
Rejects room noise, durable |
Quiet, need preamp |
¥500-2000 |
| Condenser XLR |
Studio recording, vocals |
Detailed, professional |
Sensitive to room noise |
¥800-5000 |
| Lavalier (Lapel) |
Video, talking head |
Hands-free, close to mouth |
Visible in shot, can rub on clothes |
¥100-500 |
| Shotgun |
Interviews, outdoor |
Directional, outdoor use |
Expensive, need operator |
¥1000-8000 |
Environment Setup:
- Quietest room available: Close windows, turn off fans/AC, avoid high-traffic times
- Reduce reflections: Hang blankets, use acoustic foam, record in closet full of clothes
- Distance to mic: 6-12 inches (15-30cm) for optimal balance of proximity effect and room noise
- Pop filter: Essential for plosive sounds (P, B sounds) - cheap foam or metal mesh
- Shock mount: Isolates mic from desk vibrations and handling noise
Recording Levels:
- Target: -12dB to -6dB average, peaks around -3dB
- Too low: -24dB or below → brings up noise floor when boosted later
- Too hot: Peaking at 0dB → distortion/clipping, cannot be fixed
- Test record: Always do 10-second test, check levels before full recording
Recording Checklist:
Step 2: Basic Audio Editing
Importing and Organizing:
- Multi-track setup: Keep voice on track 1, music on track 2, effects on track 3
- Label tracks: "Voiceover", "Background Music", "Sound Effects" for clarity
- Save project file: Always save editable project format (.aup3, .sesx, etc.) before exporting
Trimming and Arranging:
- Remove mistakes: Cut out coughs, false starts, long pauses
- Tighten pacing: Reduce pauses between sections to 0.5-1 second for better flow
- Crossfade edits: Use 5-10ms crossfades on all cuts to avoid clicks/pops
- Arrange content: Drag clips to reorder, build narrative flow
Basic Editing Techniques:
| Technique |
How |
Why |
| Cut/Copy/Paste |
Select region, edit menu |
Remove mistakes, reorder content |
| Split |
Cut at cursor point |
Separate sections for independent editing |
| Trim |
Remove selected region |
Quickly cut ends or mistakes |
| Fade In/Out |
Apply fade to clip start/end |
Smooth transitions, avoid abrupt starts/ends |
| Crossfade |
Overlap clips with transition |
Seamless joins between audio segments |
Edit Best Practices:
- Always backup original: Keep raw recording untouched, work on copy
- Non-destructive editing: Use software that preserves original (Audacity, Adobe Audition)
- Undo is your friend: Ctrl+Z / Cmd+Z liberally while learning
- Save versions: Save project after major edits (v1, v2, v3) so you can backtrack
Step 3: Noise Reduction and Cleanup
Identify Noise Types:
| Noise |
Character |
Removal Method |
Difficulty |
| Hiss |
Steady high-frequency noise |
Noise reduction plugin |
Easy |
| Hum |
Low-frequency electrical buzz (50/60Hz) |
High-pass filter or notch filter |
Easy |
| Room reverb |
Echoey, cavernous sound |
De-reverb plugin or reduce room noise |
Medium |
| Clicks/pops |
Sharp sudden sounds |
Click removal plugin |
Medium |
| Wind noise |
Low-frequency rumble |
High-pass filter + wind reduction |
Medium-Hard |
| Static/crackle |
Continuous crackling |
Noise reduction + de-crackle |
Hard |
Noise Reduction Workflow (using Audacity as example):
Step 1: Capture Noise Profile
- Find a section of pure noise (no voice) - usually 0.5-2 seconds at start/end
- Select only the noise section
- Effect → Noise Reduction → Get Noise Profile
- Software analyzes the noise character
Step 2: Apply Noise Reduction
- Select entire audio track (Ctrl+A / Cmd+A)
- Effect → Noise Reduction
- Adjust settings:
- Noise reduction (dB): Start with 12dB, adjust to taste
- Sensitivity: 3-6 (higher = more aggressive, may artifact)
- Frequency smoothing: 3-6 bands
- Preview, adjust, then apply
Step 3: Fine-Tune
- Too aggressive: Audio sounds underwater, robotic
- Too light: Still hear noise
- Artifact check: Listen for "watery" artifacts on "S" and "F" sounds
- Apply multiple light passes better than one heavy pass
Alternative Noise Reduction Methods:
- High-pass filter: Removes low-frequency rumble below 80-100Hz
- Low-pass filter: Removes high-frequency hiss above 12-15kHz
- Notch filter: Removes specific frequency (like 60Hz electrical hum)
- Gate: Silences audio below threshold (good for background noise between words)
Step 4: Volume Normalization and Compression
Consistent volume is critical - viewers should never have to adjust their volume.
Leveling Techniques (in order of application):
1. Normalization (simple, fixes overall level):
- Purpose: Bring entire track to target peak level
- Settings: Normalize to -3dB peak (leaves headroom, prevents clipping)
- When: First step after noise reduction
- How: Select all, Effect → Normalize → Target peak level
2. Compression (evens out dynamics):
- Purpose: Reduce difference between loudest and softest parts
- Key settings:
- Ratio: 2:1 to 4:1 for voice (2:1 = subtle, 4:1 = more aggressive)
- Threshold: -20dB to -12dB (lower = more compression)
- Attack: 5-10ms (fast enough to catch peaks)
- Release: 100-300ms (natural release)
- Result: Whisper-quiet parts become audible, loud parts tamed
3. Limiting (prevents clipping):
- Purpose: Hard ceiling at -0.1dB or -1dB, ensures no digital distortion
- Settings: Threshold -1dB, ceiling -0.1dB
- When: Final step before export
- Result: No peaks exceed target, consistent loudness
Compression Quick Settings by Use Case:
| Use Case |
Ratio |
Threshold |
Attack |
Release |
| Spoken word (tutorial) |
2:1 |
-18dB |
10ms |
200ms |
| Narration (documentary) |
3:1 |
-15dB |
5ms |
150ms |
| Podcast (conversation) |
2.5:1 |
-16dB |
8ms |
250ms |
| Emotional/intimate |
1.5:1 |
-20dB |
15ms |
300ms |
| Energetic/promo |
4:1 |
-12dB |
3ms |
100ms |
Step 5: EQ and Audio Enhancement
Equalization (EQ) shapes tone - making voice sound clear, professional, and pleasant.
Voice EQ Basics:
| Frequency |
Effect on Voice |
When to Adjust |
| Below 80Hz |
Low rumble, room noise |
Cut completely for voice (high-pass filter) |
| 80-200Hz |
Warmth, body |
Boost slightly for thin voices, cut for muddy |
| 200-500Hz |
Fullness, presence |
Leave mostly flat |
| 500Hz-2kHz |
Intelligibility, clarity |
Boost slightly (+1-3dB) if voice is dull |
| 2kHz-6kHz |
Definition, clarity |
Boost (+2-4dB) to make voice "pop" |
| 6kHz-12kHz |
Air, brilliance, sibilance |
Cut S-heavy voices at 7kHz, boost for "air" |
| Above 12kHz |
Ultra-highs, hiss |
Cut if hissy, leave if clear |
Simple Voice EQ Recipe (works for 80% of recordings):
- High-pass filter: Remove everything below 80Hz
- Cut mud: -2dB at 250Hz if voice sounds boomy
- Boost clarity: +2dB at 3kHz for intelligibility
- Tame harshness: -3dB at 7kHz if S-sounds are harsh
- Add air: +2dB at 10kHz if voice needs openness
De-Essing (taming harsh S and T sounds):
- Problem: Sibilance causes harsh, ear-piercing S sounds
- Solution: De-esser plugin targets 5-8kHz frequencies only
- Settings: Threshold -20dB, frequency 7kHz, range -6dB
- Alternative: Manual EQ cut at 7kHz by -2 to -4dB
Step 6: Adding Music and Sound Effects
Music enhances emotion but should never compete with voice.
Music Selection Principles:
- Match mood: Upbeat music for energetic content, calm for tutorials
- Instrumental preferred: Lyrics distract from spoken content
- Right tempo: 60-90 BPM for narration, 120+ for energetic
- Legal sources: Royalty-free from YouTube Audio Library, Epidemic Sound, Artlist
Leveling Voice vs. Music:
| Content Type |
Voice Level |
Music Level |
Ratio |
| Tutorial/education |
-6dB to -3dB |
-20dB to -18dB |
12-15dB difference |
| Narration/story |
-6dB to -3dB |
-16dB to -14dB |
10-12dB difference |
| Emotional/intimate |
-8dB to -6dB |
-22dB to -20dB |
14-16dB difference |
| High-energy promo |
-3dB to 0dB |
-12dB to -10dB |
10-12dB difference |
Music Mixing Workflow:
- Import music to separate track (never mix with voice on same track)
- Fade music in: 2-3 second fade at beginning
- Duck music: Lower music volume by 4-6dB when voice is present
- Auto-duck (if available): Software automatically lowers music during voice
- Fade music out: 2-3 second fade at end
- Check on phone: Test on mobile device (most viewers use mobile)
Sound Effects (SFX):
- Transition sounds: Whooshes, clicks, swipes between sections
- Emphasis: Ding, pop, or sparkle for key points
- Ambience: Subtle room tone, nature sounds for atmosphere
- Rule: Less is more - 1-3 sounds per minute max
Step 7: Export and Platform Optimization
Export Settings for Xiaohongshu:
| Setting |
Recommended |
Why |
| Format |
AAC (.m4a) or MP3 |
Best compression quality |
| Sample Rate |
44.1kHz or 48kHz |
Match source rate |
| Bitrate |
192 kbps (stereo) or 128 kbps (mono) |
Balance quality and file size |
| Channels |
Stereo or Mono |
Mono fine for voice-only |
| Loudness |
-16 LUFS |
Streaming platform standard |
Export Quality Comparison:
| Bitrate |
File Size (1 min) |
Quality |
Use Case |
| 128 kbps |
~1 MB |
Good |
Voice-only,节省流量 |
| 192 kbps |
~1.5 MB |
Very Good |
Recommended for most content |
| 256 kbps |
~2 MB |
Excellent |
Music-heavy or audiophile content |
| 320 kbps |
~2.5 MB |
Best |
Overkill for social media |
Final Checklist Before Export:
Quality Control Testing:
- Headphones: Check for details, hiss, harshness
- Speakers: Check overall balance, bass response
- Phone speakers: Most viewers listen here - critical test
- Car test: Play in car (noisy environment) - intelligible?
Common Mistakes
| Mistake |
Why It's Wrong |
Fix |
| Recording in noisy room |
Noise reduction can't fix everything, artifacts result |
Record in quietest space, treat room with blankets |
| Mic too far from mouth |
Room echo increases, voice-to-noise ratio decreases |
Move 6-12 inches from mic, use pop filter |
| Recording level too low |
Boosting in post amplifies noise floor |
Aim for -12dB to -6dB average |
| Recording level too hot |
Distortion/clipping is permanent and unfixable |
Leave headroom, peak around -6dB |
| Over-applying noise reduction |
Audio sounds robotic, underwater artifacts |
Use light passes (6-12dB), not heavy (20dB+) |
| No compression on voice |
Inconsistent volume, whisper-quiet then too-loud |
Apply 2:1 to 4:1 compression |
| Music too loud |
Distracts from voice, makes content hard to follow |
Duck music 12-15dB below voice |
| Too much high-frequency EQ |
Harsh, ear-fatiguing, sibilance amplified |
Cut 7kHz region, boost 3kHz instead |
| Exporting at wrong bitrate |
Either poor quality (too low) or huge files (too high) |
Use 192 kbps for optimal balance |
| Never testing on phone |
Sounds different on viewers' most common device |
Always final QC on mobile device |
Real-World Impact
Case Study 1: Tutorial Creator's Retention Transformation
Creator: Xiaohongshu tech tutorial creator
Problem: 40% viewer drop-off within 30 seconds, despite valuable content
Issue: Poor audio quality - room echo, inconsistent volume, background noise
Solution Implemented:
- Purchased USB microphone (¥300) and pop filter
- Treated recording space with blankets on walls
- Applied noise reduction, compression, and EQ in Audacity
- Normalized all content to -16 LUFS
Results (60 days):
- Average view duration: 45 seconds → 2:45 minutes (6x retention)
- 30-second drop-off: 40% → 15% (62.5% improvement)
- Engagement rate: 2% → 7% (3.5x increase)
- Comment feedback: "Finally can hear clearly!" "Audio quality is pro"
Case Study 2: Podcaster's Audio Upgrade
Creator: Storytelling podcast on Xiaohongshu
Problem: Listeners complained about "can't hear in car," "too quiet then too loud"
Solution:
- Implemented compression (3:1 ratio, -15dB threshold)
- Added limiter to prevent peaks
- Normalized to -16 LUFS loudness target
- Applied high-pass filter below 80Hz
- Subtle voice EQ boost at 3kHz
Results:
- Listener complaints: 0 (down from multiple per episode)
- Apple Podcasts rating: 4.2 → 4.8 stars
- Average completion rate: 45% → 72% (listeners finish episodes)
- Subscription growth: +200% (word-of-mouth from quality improvement)
Case Study 3: Brand's Audio Consistency
Brand: Beauty brand with multiple content creators
Problem: Inconsistent audio quality across 20+ creators, damaged brand credibility
Solution:
- Created audio processing template (preset in Adobe Audition)
- Standardized recording guidelines document
- Provided creators with cheap USB mic + pop filter kit
- Centralized post-processing: all audio edited by one person using template
Results (3 months):
- Audio consistency: 100% across all content
- Viewer retention: +35% (measured by average watch time)
- Brand perception: "Professional, trustworthy" in user surveys
- Reduced editing time: 4 hours per video → 45 minutes (template efficiency)
Related Skills
REQUIRED:
- short-video-production: Complete video creation including audio integration
- vlog-creation: Vlog-specific audio challenges and solutions
- podcast-production: Long-form audio content creation techniques
RECOMMENDED:
- music-licensing: Legal music sourcing and copyright compliance
- content-equipment: Microphone and recording gear selection guides
- post-production: Comprehensive video/audio post-production workflow
- accessibility: Adding subtitles and transcripts for accessibility
NEXT STEPS:
- Audit your current audio: What are your top 3 audio quality issues?
- Upgrade recording setup: Start with mic + pop filter + quiet room
- Learn basic editing: Download Audacity (free) and practice on test recordings
- Create processing preset: Save your EQ, compression, and normalization settings
- Test on mobile: Always final quality check on phone before publishing
Professional audio is not about expensive gear - it's about clean recording and thoughtful processing. A ¥300 microphone with good technique beats a ¥5000 mic used poorly. Your viewers will forgive imperfect visuals, but they will abandon content with painful audio. Invest in audio processing first, visuals second.
1---2name: audio-processing3description: Use when processing audio for Xiaohongshu content, editing voiceovers, improving sound quality, creating podcasts, or producing audio-based posts4---56# Audio Processing (音频处理)78## Overview910Audio processing encompasses recording, editing, enhancing, and optimizing audio content for Xiaohongshu posts, ensuring professional sound quality that significantly enhances content professionalism, viewer retention, and overall production value. Poor audio is the #1 reason viewers abandon content within seconds - even with stunning visuals, bad audio makes content unwatchable. This skill covers the complete audio production workflow: from recording setup through editing, noise reduction, mixing, and final optimization for Xiaohongshu's platform specifications.1112**Key insight**: Viewers will forgive mediocre video quality, but they will not tolerate poor audio. Investing in audio processing yields 50%+ improvements in viewer retention and 3-5x increases in engagement rates. Professional audio transforms amateur content into credible, trustworthy content.1314## When to Use1516**Use when**:17- Recording voiceovers for Xiaohongshu videos, tutorials, or storytelling content18- Editing podcast-style audio content or interview recordings19- Improving audio quality in existing recordings (noise reduction, volume leveling)20- Creating background music tracks or adding sound effects to videos21- Producing audio-first content formats (podcasts, audio diaries, voice notes)22- Fixing common audio issues: background noise, echo, distortion, low volume23- Optimizing audio for Xiaohongshu's platform specifications and compression24- Creating consistent audio quality across content series25- Adding professional polish with music beds, transitions, and sound design2627**Do NOT use when**:28- Using licensed music without proper permissions (copyright violation)29- Content requires pure ambient sound (over-processing disrupts authenticity)30- Audio is already professional quality (over-processing can degrade quality)31- Quick, casual content where production speed matters more than polish32- Live streaming (different audio setup requirements)3334## Core Pattern3536**Before** (poor audio quality):37❌ "Background noise, room echo, distractions"38❌ "Inconsistent volume, too quiet then too loud"39❌ "Viewer adjusts volume constantly, gives up"40❌ "Content seems amateur, low credibility"41❌ "Viewers scroll away within 3 seconds"4243**After** (professional audio):44✅ "Clean, clear voice recording"45✅ "Consistent volume levels throughout"46✅ "Pleasant listening experience, no adjustments needed"47✅ "Content feels professional, trustworthy"48✅ "Viewers watch complete content, high engagement"4950**6 Essential Audio Processing Elements**:5152| Element | Purpose | Quality Impact | Priority |53|---------|---------|----------------|----------|54| **Clean Recording** | Prevent issues at source | Critical | #1 - cannot fix in post |55| **Noise Reduction** | Remove background distractions | High | #2 - most common issue |56| **Volume Normalization** | Consistent listening levels | High | #3 - prevents frustration |57| **EQ & Clarity** | Enhance voice intelligibility | Medium-High | #4 - professional polish |58| **Music & Effects** | Add emotional depth | Medium | #5 - enhance, don't distract |59| **Platform Optimization** | Meet technical specs | Medium | #6 - avoid compression artifacts |6061## Quick Reference6263**Audio Processing Software Comparison**:6465| Tool | Best For | Skill Level | Cost | Platform | Key Features |66|------|----------|------------|------|----------|--------------|67| **Audacity** | Basic editing, noise reduction | Beginner | Free | Win/Mac/Linux | Noise gate, normalize, EQ |68| **Adobe Audition** | Professional production | Intermediate-Advanced | Paid (Subscription) | Win/Mac | Multitrack, advanced repair, batch processing |69| **GarageBand** | Mac users, simple editing | Beginner | Free (Mac) | macOS | Built-in effects, music loops, easy interface |70| **Descript** | Text-based editing, podcasts | Beginner-Intermediate | Paid (Freemium) | Web/Mac/Win | Edit audio like text, overdub, filler removal |71| **Logic Pro** | Music production, advanced editing | Advanced | Paid (One-time) | macOS | Professional DAW, massive library |72| **Reaper** | Power users, customization | Advanced | Paid (Free trial) | Win/Mac/Linux | Lightweight, extensible, affordable |7374**Xiaohongshu Audio Specifications**:75- **Format**: AAC, MP3, or M4A76- **Sample Rate**: 44.1kHz or 48kHz77- **Bitrate**: 128-320 kbps (192 kbps recommended for balance)78- **Channels**: Stereo or Mono (Mono for voice-only is fine)79- **Loudness Target**: -16 LUFS (YouTube/Broadcast standard)8081**Quick Audio Fixes** (by symptom):8283| Symptom | Likely Cause | Quick Fix |84|---------|--------------|-----------|85| **Background hiss/hum** | Room noise, equipment hiss | Noise reduction filter |86| **Room echo/reverb** | Recording in untreated room | Move closer to mic, use de-reverb plugin |87| **Volume too low** | Recording level too low | Gain/normalize to -3dB peak |88| **Distorted/clipping** | Recording level too high | Reduce gain, use clip restoration |89| **Muffled sound** | Poor mic quality or wrong EQ | High-pass filter + EQ boost |90| **Inconsistent levels** | Multiple clips or variable distance | Compression + normalization |9192## Implementation9394### Step 1: Recording Setup and Environment9596**Prevention is better than correction** - capturing clean audio at source saves hours of editing and yields better results than any post-processing.9798**Microphone Selection**:99100| Mic Type | Best For | Pros | Cons | Price Range |101|----------|----------|------|------|-------------|102| **USB Mic** | Beginners, simplicity | Plug-and-play, easy | Limited quality, no upgrades | ¥200-800 |103| **Dynamic XLR** | Voice recording, noisy rooms | Rejects room noise, durable | Quiet, need preamp | ¥500-2000 |104| **Condenser XLR** | Studio recording, vocals | Detailed, professional | Sensitive to room noise | ¥800-5000 |105| **Lavalier (Lapel)** | Video, talking head | Hands-free, close to mouth | Visible in shot, can rub on clothes | ¥100-500 |106| **Shotgun** | Interviews, outdoor | Directional, outdoor use | Expensive, need operator | ¥1000-8000 |107108**Environment Setup**:109- **Quietest room available**: Close windows, turn off fans/AC, avoid high-traffic times110- **Reduce reflections**: Hang blankets, use acoustic foam, record in closet full of clothes111- **Distance to mic**: 6-12 inches (15-30cm) for optimal balance of proximity effect and room noise112- **Pop filter**: Essential for plosive sounds (P, B sounds) - cheap foam or metal mesh113- **Shock mount**: Isolates mic from desk vibrations and handling noise114115**Recording Levels**:116- **Target**: -12dB to -6dB average, peaks around -3dB117- **Too low**: -24dB or below → brings up noise floor when boosted later118- **Too hot**: Peaking at 0dB → distortion/clipping, cannot be fixed119- **Test record**: Always do 10-second test, check levels before full recording120121**Recording Checklist**:122- [ ] Room is quiet (no traffic, appliances, people)123- [ ] Mic positioned 6-12 inches from mouth124- [ ] Pop filter attached125- [ ] Recording level peaks around -6dB126- [ ] Headphones monitoring live audio127- [ ] Test recording sounds clear, no issues128129### Step 2: Basic Audio Editing130131**Importing and Organizing**:132- **Multi-track setup**: Keep voice on track 1, music on track 2, effects on track 3133- **Label tracks**: "Voiceover", "Background Music", "Sound Effects" for clarity134- **Save project file**: Always save editable project format (.aup3, .sesx, etc.) before exporting135136**Trimming and Arranging**:137- **Remove mistakes**: Cut out coughs, false starts, long pauses138- **Tighten pacing**: Reduce pauses between sections to 0.5-1 second for better flow139- **Crossfade edits**: Use 5-10ms crossfades on all cuts to avoid clicks/pops140- **Arrange content**: Drag clips to reorder, build narrative flow141142**Basic Editing Techniques**:143144| Technique | How | Why |145|-----------|-----|-----|146| **Cut/Copy/Paste** | Select region, edit menu | Remove mistakes, reorder content |147| **Split** | Cut at cursor point | Separate sections for independent editing |148| **Trim** | Remove selected region | Quickly cut ends or mistakes |149| **Fade In/Out** | Apply fade to clip start/end | Smooth transitions, avoid abrupt starts/ends |150| **Crossfade** | Overlap clips with transition | Seamless joins between audio segments |151152**Edit Best Practices**:153- **Always backup original**: Keep raw recording untouched, work on copy154- **Non-destructive editing**: Use software that preserves original (Audacity, Adobe Audition)155- **Undo is your friend**: Ctrl+Z / Cmd+Z liberally while learning156- **Save versions**: Save project after major edits (v1, v2, v3) so you can backtrack157158### Step 3: Noise Reduction and Cleanup159160**Identify Noise Types**:161162| Noise | Character | Removal Method | Difficulty |163|-------|-----------|----------------|------------|164| **Hiss** | Steady high-frequency noise | Noise reduction plugin | Easy |165| **Hum** | Low-frequency electrical buzz (50/60Hz) | High-pass filter or notch filter | Easy |166| **Room reverb** | Echoey, cavernous sound | De-reverb plugin or reduce room noise | Medium |167| **Clicks/pops** | Sharp sudden sounds | Click removal plugin | Medium |168| **Wind noise** | Low-frequency rumble | High-pass filter + wind reduction | Medium-Hard |169| **Static/crackle** | Continuous crackling | Noise reduction + de-crackle | Hard |170171**Noise Reduction Workflow** (using Audacity as example):172173**Step 1: Capture Noise Profile**1741. Find a section of pure noise (no voice) - usually 0.5-2 seconds at start/end1752. Select only the noise section1763. Effect → Noise Reduction → Get Noise Profile1774. Software analyzes the noise character178179**Step 2: Apply Noise Reduction**1801. Select entire audio track (Ctrl+A / Cmd+A)1812. Effect → Noise Reduction1823. Adjust settings:183 - **Noise reduction (dB)**: Start with 12dB, adjust to taste184 - **Sensitivity**: 3-6 (higher = more aggressive, may artifact)185 - **Frequency smoothing**: 3-6 bands1864. Preview, adjust, then apply187188**Step 3: Fine-Tune**189- **Too aggressive**: Audio sounds underwater, robotic190- **Too light**: Still hear noise191- **Artifact check**: Listen for "watery" artifacts on "S" and "F" sounds192- **Apply multiple light passes** better than one heavy pass193194**Alternative Noise Reduction Methods**:195- **High-pass filter**: Removes low-frequency rumble below 80-100Hz196- **Low-pass filter**: Removes high-frequency hiss above 12-15kHz197- **Notch filter**: Removes specific frequency (like 60Hz electrical hum)198- **Gate**: Silences audio below threshold (good for background noise between words)199200### Step 4: Volume Normalization and Compression201202**Consistent volume is critical** - viewers should never have to adjust their volume.203204**Leveling Techniques** (in order of application):205206**1. Normalization** (simple, fixes overall level):207- **Purpose**: Bring entire track to target peak level208- **Settings**: Normalize to -3dB peak (leaves headroom, prevents clipping)209- **When**: First step after noise reduction210- **How**: Select all, Effect → Normalize → Target peak level211212**2. Compression** (evens out dynamics):213- **Purpose**: Reduce difference between loudest and softest parts214- **Key settings**:215 - **Ratio**: 2:1 to 4:1 for voice (2:1 = subtle, 4:1 = more aggressive)216 - **Threshold**: -20dB to -12dB (lower = more compression)217 - **Attack**: 5-10ms (fast enough to catch peaks)218 - **Release**: 100-300ms (natural release)219- **Result**: Whisper-quiet parts become audible, loud parts tamed220221**3. Limiting** (prevents clipping):222- **Purpose**: Hard ceiling at -0.1dB or -1dB, ensures no digital distortion223- **Settings**: Threshold -1dB, ceiling -0.1dB224- **When**: Final step before export225- **Result**: No peaks exceed target, consistent loudness226227**Compression Quick Settings by Use Case**:228229| Use Case | Ratio | Threshold | Attack | Release |230|----------|-------|-----------|--------|---------|231| **Spoken word (tutorial)** | 2:1 | -18dB | 10ms | 200ms |232| **Narration (documentary)** | 3:1 | -15dB | 5ms | 150ms |233| **Podcast (conversation)** | 2.5:1 | -16dB | 8ms | 250ms |234| **Emotional/intimate** | 1.5:1 | -20dB | 15ms | 300ms |235| **Energetic/promo** | 4:1 | -12dB | 3ms | 100ms |236237### Step 5: EQ and Audio Enhancement238239**Equalization (EQ) shapes tone** - making voice sound clear, professional, and pleasant.240241**Voice EQ Basics**:242243| Frequency | Effect on Voice | When to Adjust |244|-----------|----------------|----------------|245| **Below 80Hz** | Low rumble, room noise | **Cut** completely for voice (high-pass filter) |246| **80-200Hz** | Warmth, body | Boost slightly for thin voices, cut for muddy |247| **200-500Hz** | Fullness, presence | Leave mostly flat |248| **500Hz-2kHz** | Intelligibility, clarity | Boost slightly (+1-3dB) if voice is dull |249| **2kHz-6kHz** | Definition, clarity | Boost (+2-4dB) to make voice "pop" |250| **6kHz-12kHz** | Air, brilliance, sibilance | Cut S-heavy voices at 7kHz, boost for "air" |251| **Above 12kHz** | Ultra-highs, hiss | Cut if hissy, leave if clear |252253**Simple Voice EQ Recipe** (works for 80% of recordings):2541. **High-pass filter**: Remove everything below 80Hz2552. **Cut mud**: -2dB at 250Hz if voice sounds boomy2563. **Boost clarity**: +2dB at 3kHz for intelligibility2574. **Tame harshness**: -3dB at 7kHz if S-sounds are harsh2585. **Add air**: +2dB at 10kHz if voice needs openness259260**De-Essing** (taming harsh S and T sounds):261- **Problem**: Sibilance causes harsh, ear-piercing S sounds262- **Solution**: De-esser plugin targets 5-8kHz frequencies only263- **Settings**: Threshold -20dB, frequency 7kHz, range -6dB264- **Alternative**: Manual EQ cut at 7kHz by -2 to -4dB265266### Step 6: Adding Music and Sound Effects267268**Music enhances emotion** but should never compete with voice.269270**Music Selection Principles**:271- **Match mood**: Upbeat music for energetic content, calm for tutorials272- **Instrumental preferred**: Lyrics distract from spoken content273- **Right tempo**: 60-90 BPM for narration, 120+ for energetic274- **Legal sources**: Royalty-free from YouTube Audio Library, Epidemic Sound, Artlist275276**Leveling Voice vs. Music**:277278| Content Type | Voice Level | Music Level | Ratio |279|--------------|-------------|-------------|-------|280| **Tutorial/education** | -6dB to -3dB | -20dB to -18dB | 12-15dB difference |281| **Narration/story** | -6dB to -3dB | -16dB to -14dB | 10-12dB difference |282| **Emotional/intimate** | -8dB to -6dB | -22dB to -20dB | 14-16dB difference |283| **High-energy promo** | -3dB to 0dB | -12dB to -10dB | 10-12dB difference |284285**Music Mixing Workflow**:2861. **Import music** to separate track (never mix with voice on same track)2872. **Fade music in**: 2-3 second fade at beginning2883. **Duck music**: Lower music volume by 4-6dB when voice is present2894. **Auto-duck** (if available): Software automatically lowers music during voice2905. **Fade music out**: 2-3 second fade at end2916. **Check on phone**: Test on mobile device (most viewers use mobile)292293**Sound Effects (SFX)**:294- **Transition sounds**: Whooshes, clicks, swipes between sections295- **Emphasis**: Ding, pop, or sparkle for key points296- **Ambience**: Subtle room tone, nature sounds for atmosphere297- **Rule**: Less is more - 1-3 sounds per minute max298299### Step 7: Export and Platform Optimization300301**Export Settings for Xiaohongshu**:302303| Setting | Recommended | Why |304|---------|-------------|-----|305| **Format** | AAC (.m4a) or MP3 | Best compression quality |306| **Sample Rate** | 44.1kHz or 48kHz | Match source rate |307| **Bitrate** | 192 kbps (stereo) or 128 kbps (mono) | Balance quality and file size |308| **Channels** | Stereo or Mono | Mono fine for voice-only |309| **Loudness** | -16 LUFS | Streaming platform standard |310311**Export Quality Comparison**:312313| Bitrate | File Size (1 min) | Quality | Use Case |314|---------|-------------------|---------|----------|315| **128 kbps** | ~1 MB | Good | Voice-only,节省流量 |316| **192 kbps** | ~1.5 MB | Very Good | **Recommended** for most content |317| **256 kbps** | ~2 MB | Excellent | Music-heavy or audiophile content |318| **320 kbps** | ~2.5 MB | Best | Overkill for social media |319320**Final Checklist Before Export**:321- [ ] Noise reduction applied, no hiss or hum322- [ ] Volume normalized, consistent levels throughout323- [ ] Voice EQ applied, clear and pleasant324- [ ] Music balanced, doesn't compete with voice325- [ ] No clipping or distortion326- [ ] Fades at beginning and end327- [ ] Test listen on headphones, speakers, and phone328- [ ] Export at correct bitrate and format329330**Quality Control Testing**:3311. **Headphones**: Check for details, hiss, harshness3322. **Speakers**: Check overall balance, bass response3333. **Phone speakers**: Most viewers listen here - critical test3344. **Car test**: Play in car (noisy environment) - intelligible?335336## Common Mistakes337338| Mistake | Why It's Wrong | Fix |339|---------|----------------|-----|340| **Recording in noisy room** | Noise reduction can't fix everything, artifacts result | Record in quietest space, treat room with blankets |341| **Mic too far from mouth** | Room echo increases, voice-to-noise ratio decreases | Move 6-12 inches from mic, use pop filter |342| **Recording level too low** | Boosting in post amplifies noise floor | Aim for -12dB to -6dB average |343| **Recording level too hot** | Distortion/clipping is permanent and unfixable | Leave headroom, peak around -6dB |344| **Over-applying noise reduction** | Audio sounds robotic, underwater artifacts | Use light passes (6-12dB), not heavy (20dB+) |345| **No compression on voice** | Inconsistent volume, whisper-quiet then too-loud | Apply 2:1 to 4:1 compression |346| **Music too loud** | Distracts from voice, makes content hard to follow | Duck music 12-15dB below voice |347| **Too much high-frequency EQ** | Harsh, ear-fatiguing, sibilance amplified | Cut 7kHz region, boost 3kHz instead |348| **Exporting at wrong bitrate** | Either poor quality (too low) or huge files (too high) | Use 192 kbps for optimal balance |349| **Never testing on phone** | Sounds different on viewers' most common device | Always final QC on mobile device |350351## Real-World Impact352353**Case Study 1: Tutorial Creator's Retention Transformation**354355**Creator**: Xiaohongshu tech tutorial creator356**Problem**: 40% viewer drop-off within 30 seconds, despite valuable content357**Issue**: Poor audio quality - room echo, inconsistent volume, background noise358**Solution Implemented**:359- Purchased USB microphone (¥300) and pop filter360- Treated recording space with blankets on walls361- Applied noise reduction, compression, and EQ in Audacity362- Normalized all content to -16 LUFS363364**Results** (60 days):365- Average view duration: 45 seconds → 2:45 minutes (6x retention)366- 30-second drop-off: 40% → 15% (62.5% improvement)367- Engagement rate: 2% → 7% (3.5x increase)368- Comment feedback: "Finally can hear clearly!" "Audio quality is pro"369370**Case Study 2: Podcaster's Audio Upgrade**371372**Creator**: Storytelling podcast on Xiaohongshu373**Problem**: Listeners complained about "can't hear in car," "too quiet then too loud"374**Solution**:375- Implemented compression (3:1 ratio, -15dB threshold)376- Added limiter to prevent peaks377- Normalized to -16 LUFS loudness target378- Applied high-pass filter below 80Hz379- Subtle voice EQ boost at 3kHz380381**Results**:382- Listener complaints: 0 (down from multiple per episode)383- Apple Podcasts rating: 4.2 → 4.8 stars384- Average completion rate: 45% → 72% (listeners finish episodes)385- Subscription growth: +200% (word-of-mouth from quality improvement)386387**Case Study 3: Brand's Audio Consistency**388389**Brand**: Beauty brand with multiple content creators390**Problem**: Inconsistent audio quality across 20+ creators, damaged brand credibility391**Solution**:392- Created audio processing template (preset in Adobe Audition)393- Standardized recording guidelines document394- Provided creators with cheap USB mic + pop filter kit395- Centralized post-processing: all audio edited by one person using template396397**Results** (3 months):398- Audio consistency: 100% across all content399- Viewer retention: +35% (measured by average watch time)400- Brand perception: "Professional, trustworthy" in user surveys401- Reduced editing time: 4 hours per video → 45 minutes (template efficiency)402403---404405## Related Skills406407**REQUIRED**:408- **short-video-production**: Complete video creation including audio integration409- **vlog-creation**: Vlog-specific audio challenges and solutions410- **podcast-production**: Long-form audio content creation techniques411412**RECOMMENDED**:413- **music-licensing**: Legal music sourcing and copyright compliance414- **content-equipment**: Microphone and recording gear selection guides415- **post-production**: Comprehensive video/audio post-production workflow416- **accessibility**: Adding subtitles and transcripts for accessibility417418**NEXT STEPS**:4191. Audit your current audio: What are your top 3 audio quality issues?4202. Upgrade recording setup: Start with mic + pop filter + quiet room4213. Learn basic editing: Download Audacity (free) and practice on test recordings4224. Create processing preset: Save your EQ, compression, and normalization settings4235. Test on mobile: Always final quality check on phone before publishing424425---426427**Professional audio is not about expensive gear - it's about clean recording and thoughtful processing. A ¥300 microphone with good technique beats a ¥5000 mic used poorly. Your viewers will forgive imperfect visuals, but they will abandon content with painful audio. Invest in audio processing first, visuals second.**