🎤 Voice Tools
Extract and profile specific voices from audio files using AI-powered speaker diarization and voice matching.
Choose a workflow below to get started.
👥 Speaker Separation
Analyze multi-speaker audio files to automatically detect and separate individual speakers into separate audio streams.
Upload an audio file with multiple speakers, and this tool will:
- Detect all speakers automatically
- Separate each speaker's audio
- Export clean individual streams
📤 Input File
⚙️ Speaker Detection Settings
📊 Results
💾 Downloads
📚 Usage Tips
How to Use:
- Upload Audio: Select an M4A, WAV, or MP3 file with multiple speakers
- Configure Detection:
- Use min/max speakers for auto-detection (recommended)
- Or set exact speaker count if you know it
- Choose Output: Select format, sample rate, and bitrate
- Separate: Click the button and wait for processing
- Download: Get individual speaker files and a detailed report
Best Practices:
- Clear audio with distinct speakers works best
- If you know the exact speaker count, specify it for better results
- Processing time scales with file duration (expect ~2x realtime)
- M4A format provides best quality-to-size ratio
- For long files (>1 hour), expect several minutes of processing
Troubleshooting:
- If fewer speakers detected than expected, try increasing max_speakers
- If too many speakers detected, try increasing min_speakers
- For overlapping speech, the tool will assign to the dominant speaker
Extract a specific speaker from audio using a reference clip. Upload a short clip (3+ seconds) of the target speaker's voice, then upload the audio file to extract from.
Step 1: Upload Reference Clip
Step 2: Upload Target Audio
Step 3: Configure Parameters
Results
Examples
Voice Denoising
Remove silence and background noise from audio files using voice activity detection (VAD).
How it works:
- Upload an audio file
- Adjust VAD and silence thresholds
- Process the audio to remove unwanted segments
- Download the cleaned result
Tips:
- VAD Threshold: Higher values are more aggressive (remove more segments)
- Silence Threshold: Larger values keep longer silent gaps
- Min Duration: Filters out very short voice segments (reduces false positives)
Input
Parameters
Output
Upload an audio file to begin
Examples
✂️ Voice Clipper
Manually trim and stitch together the good parts of a recording — right in your browser. Drag on the waveform to select a range, drag it into the tray below, reorder or remove clips, then export a single combined WAV file.
100% local: audio is decoded and processed with the Web Audio API in your browser tab. It is never uploaded to the server, so there's no file-size limit tied to processing time and no GPU/model usage for this tab.