Home
Voice Library
Audio
AI VoiceNew
Text to Speech
AI Podcast
Voice Cloning
Visual
AI Image
AI Video
Explainer Video
Slides
Blog
English
Sign In
Download App

Voice Cloning · 12 Languages

AI Voice Cloning

Record a short take, get a voice of your own

The voice you build shows up in the other tools, ready to narrate podcasts, read text, and voice slides and explainers.

  • Record by chatting or upload a file
  • 12 supported cloning languages
  • Reuse in podcasts, TTS, slides, explainers
  • Preview before you save
Loading...

How voice cloning works

  1. 1

    Choose how to give a sample

    Record by chatting in the browser, or upload an audio file you recorded yourself — whichever is easier.

  2. 2

    Give it 25–35 seconds of speech

    Talk naturally rather than reading flatly, in a quiet room without echo. Uploads accept wav, mp3, or m4a up to 20 MB.

  3. 3

    Set the cloning language

    Pick English, Chinese Mandarin, Japanese, Spanish, Portuguese, French, German, Turkish, Korean, Italian, Thai, or Vietnamese, then name the voice and set its gender so it is easy to find later.

  4. 4

    Preview, save, and reuse

    Listen to the preview before saving. Once saved, the voice appears in the voice picker across the other tools.

Where voice cloning helps

Host a podcast in your own voice

Publish episodes that sound like you without sitting down to record each one.

Narrate slides and explainers

Keep one consistent voice across a deck or a video series, even when you edit the script later.

Read long documents aloud

Send a report through text to speech and have it read back in a familiar voice.

Keep a consistent brand voice

Use the same voice across every piece of audio you publish, instead of a different stock voice each time.

Voice Cloning FAQ

What is voice cloning?

Voice cloning builds a synthetic version of a specific voice from a short recording, so that voice can then read any text you write.

How long does the sample need to be?

About 25 to 35 seconds. Speak naturally — free conversation or energetic delivery works better than flat reading — in a quiet, low-echo room.

What audio formats can I upload?

You can upload a wav, mp3, or m4a file up to 20 MB. Use audio you recorded yourself rather than downloaded or third-party files.

Which languages can I clone?

Cloning supports English, Chinese Mandarin, Japanese, Spanish, Portuguese, French, German, Turkish, Korean, Italian, Thai, and Vietnamese.

Where can I use a cloned voice?

A saved voice shows up in the voice picker for AI Podcast, Text to Speech, Slides, and Explainer Video, so one clone covers all of them.

Do I have to record, or can I upload?

Either works. Voice Chat records you in the browser while you talk, and Upload File takes an existing recording — both produce the same kind of voice.

Where to go instead

  • Text to Speech Generator

    When the script is already written and you just need it spoken.

  • AI Voice Generator

    When the delivery matters — several voices in one take, or a line that has to land with real emotion.

Clone your voice by chatting
Chat in a quiet environment