Self-serve cloning
Clone from a 10-30 second sample and a verified consent recording via the API or Console. No fine-tuning run required.
Create a synthetic voice from a 10-30 second clip and verified spoken consent, then synthesize text in that voice by ID.
Account setup and your API key are free. Voice cloning requires Starter or above.
curl -X POST https://api.speechify.ai/v1/voices \
-H "Authorization: Bearer $SPEECHIFY_API_KEY" \
-H "Speechify-Version: 2026-09-13" \
-F name="Narrator" \
-F gender="female" \
-F sample=@sample.wav \
-F consent_challenge_id="$CONSENT_CHALLENGE_ID" \
-F consent_recording=@consent.wavFirst create a consent challenge with full_name, then record the speaker reading its phrase. Set CONSENT_CHALLENGE_ID to the returned ID and save the recording as consent.wav before submitting the sample.
Listen to the original speaker and a prepared cloned-voice sample generated with Simba 3.0. Two recordings, different scripts, the same voice identity.
Simba 3.0 · Voice cloning
Start with a consented reference recording. Create a reusable voice for the words that come next.
Explore voice cloning01 / Reference
02 / Clone
Two different recordings. Generated sample: Simba 3.0. Only clone a voice you have permission to use.
Simba 3.0 · Multilingual synthesis
Hear 6 language samples, each with a voice from its own catalog. These are separate samples, not translations of one recording.
Explore language supporten-US
de-DE
es-MX
fr-FR
it-IT
pt-BR
No separate service to integrate. Cloning is an endpoint alongside synthesis on the Build API.
Prepare 10-30 seconds of clean speech from the person whose voice you are authorized to clone.
POST to /v1/voices/consent-challenges with the speaker's full name as the JSON field full_name, then record the same speaker reading the returned phrase.
Submit the sample, consent recording, and challenge ID to /v1/voices. Use the returned voice ID with a supported model on the speech endpoints.
Start self-serve, or work with our team on a fine-tuned voice.
Clone from a 10-30 second sample and a verified consent recording via the API or Console. No fine-tuning run required.
Fine-tune on hours of a speaker's audio. Arranged with our team for signature narrators and brand voices.
Both Simba 3.2 and Simba 3.0 are streaming-native and support self-serve voice cloning. Use Simba 3.2 for English, or Simba 3.0 across 6 languages and 7 locales. Check each voice's supported models before synthesis; discuss broader language requirements with our team.
Voice cloning is included on paid plans, with no separate per-clone charge. Synthesis uses your plan's per-character rate, just like a catalog voice. Free accounts do not include cloning.
Assistive technology that speaks in a person's own voice instead of a generic one: clone from a short sample with consent and use the voice ID in read-aloud and communication tools, or preserve a voice before it is lost.
An author or narrator reads a whole book in their own voice, generated from the manuscript. One voice ID keeps every chapter consistent, and the same clone reads the translated editions.
A cloning feature inside your own product: your users clone their voice and generate content in it. The API provides the endpoints and enforces consent; you provide the flow and manage the voice IDs.
One signature host voice across every episode, produced from a script, with cross-language editions in the same voice rather than a different presenter per market.
A signature character keeps one consistent voice across every line and every update, with new lines synthesized on demand from a consented clone of the actor.
Create an account and API key, choose a paid plan, then record a sample and verify the speaker's consent.
Account setup is free. Voice cloning requires Starter or above.
Global Privacy Control honored. Your browser sent a Global Privacy Control signal, so we have opted you out of the sale and sharing of your personal information and switched off analytics and marketing cookies.