Overview
Introducing Voicebox, a revolutionary speech generative model that synthesizes speech across six languages, removes transient noise, edits content, transfers audio style, and generates diverse speech samples.
Multilingual Support
Synthesizes speech across six languages: English, French, German, Spanish, Polish, and Portuguese.
Transient Noise Removal
Removes transient noise by re-generating noise-corrupted speech.
Content Editing
Corrects misspoken words without having the speaker re-record the audio.
Audio Style Transfer
Transfers audio style within and across languages.
Get started
- Open the official website and confirm the service is available in your region.
- Check the current plan, usage limits and terms for your intended use.
- Try a small task with sample data before committing to a paid plan.
Editorial note
Pricing checked: Sep 20, 2026. Official website is active. Please verify current pricing and terms directly on the official site.
Record updated: Sep 20, 2026
Suggest a correction