Turn on Voice Clone
When you create a room with spoken translation, choose Voice Clone instead of Standard Voice, and confirm that your speakers agreed to have their voices cloned for this event.
Voice Clone
With Voice Clone, your audience hears every speaker translated live in that speaker's own voice. Turn it on when you create a room. Captions and transcripts stay exactly as they are, and there is no extra charge.
Definition
Voice clone translation is live speech translation where the translated audio sounds like the original speaker instead of a stock synthetic voice. A keynote given in English can be heard in Japanese, Thai, or Spanish in the speaker's own voice, while captions and transcripts carry the same translated words.
How it works
When you create a room with spoken translation, choose Voice Clone instead of Standard Voice, and confirm that your speakers agreed to have their voices cloned for this event.
There is nothing to upload. While Pikka Speech learns a new speaker's voice from their first sentences, those sentences are spoken in Standard Voice.
From then on, each translated sentence is spoken in the voice of the person who said it, in every listener's chosen language. Captions keep streaming exactly as before.
Panels & shared microphones
In a panel discussion the speaker changes every few sentences, often on one shared microphone. Voice Clone recognises who is talking by their voice and gives each person their own cloned voice.
Consent & privacy
Comparison
Both use the same live translation. The only difference is whose voice your listeners hear.
| Standard Voice | Voice Clone | |
|---|---|---|
| Sounds like the original speaker | ||
| Live captions and transcripts | ||
| Panels and shared microphones | One voice per speaker | |
| Voice samples to record | None | None, learned live |
| Speaker consent | Not needed | Required |
| Price | Included | Included, no extra charge |
Questions
No. Voice Clone is included with spoken translation at no extra charge. A Voice Clone room costs the same as a Standard Voice room.
No. Pikka Speech learns each voice from the speaker's own words during the event. Their first sentences are spoken in Standard Voice while the voice is learned.
Yes. Each speaker is recognised by their voice and gets their own cloned voice, even on a shared microphone. When Pikka Speech cannot tell who is speaking, it uses Standard Voice instead of guessing.
They are deleted when the event ends. Each cloned voice is private to your event and is only used to speak that event's translations.
Yes. Before a Voice Clone room can be created, the host must confirm that the speakers agreed to have their voices cloned for this event.
No. Captions and transcripts come from the same speech engine as every other room. Only the translated audio changes.
Voice Clone speaks the listener languages that have a spoken voice in Pikka Speech. Languages offered as captions only stay captions only.
Language coverage
99 source languages and 142 listener languages or dialects.
Hybrid simultaneous interpretation
AI and human interpreters on the same language channels.
Webinar & livestream translation
Live translation for webinars and video streams, with QR join.
Event displays
Live captions on LED walls and projectors.
Create a room, choose Voice Clone under spoken translation, and your speakers are heard in their own voice in every language.
Last updated: 2026-09-24