CONNECTHelp Center

Create a clean voiceprint

Capture a short, representative sample and validate that the generated voice remains clear, natural, and appropriate.

8 min readUpdated September 6, 2026Tutorial

Prepare a clean recording environment

A voiceprint learns from the audio it receives, including pacing, volume, room noise, and microphone artifacts. Use a quiet room, one speaker, and a stable boom microphone near the corner of the mouth. Speak in the same language and natural style you expect to use with the voice, and avoid music, other speakers, reverberation, or aggressive automatic processing.

  • Position: keep the microphone close enough for a strong voice but outside the direct breath path.
  • Level: use a steady conversational volume without whispering or shouting.
  • Content: include varied sounds and complete sentences in the selected language.
  • Privacy: record only the consenting speaker and avoid customer or confidential information.

Complete the Voice Library form

Open Portal → Voice Library → Add Your Voiceprint. Give the voice a descriptive name, select the recording language, and add a short description that distinguishes it from variants. Upload or record the supported short sample, then decide whether background-noise removal is appropriate. Listen to the source clip before creating the voiceprint; a poor sample should be re-recorded, not repaired later.

Connect Add Your Voiceprint dialog with name, language, recording guidance, and upload controls
The creation dialog shows the sample requirements and review controls in one place.
The current dialog accepts a short supported audio sample. Check the limits shown in your installed version because supported formats, minimum duration, and plan quotas can change.

Reject weak audio before submitting

Listen from the beginning to the end with headphones. Reject clips with long silence, clipping, pops, keyboard noise, changing microphone distance, or a second speaker. Noise removal can help with light steady noise, but it cannot restore distorted speech or separate overlapping voices. Trim unused silence and keep the vocal character you want the generated voice to reproduce.

Clear speechEvery word is intelligible without raising playback volume.
Stable loudnessThe speaker does not drift toward or away from the microphone.
One speakerNo other person, television, or synthetic voice is audible.
Natural deliveryPacing and emotion match the intended production use.

Use Pro Voiceprint only for the right source

Pro Voiceprint is a separate, advanced workflow that fine-tunes a model from a larger single-speaker dataset. The application recommends at least 30 minutes of clean audio, with more material improving coverage, and advises splitting long recordings into several files. Review the credit cost shown in the Portal before training and confirm that the speaker has consented to this use.

Connect Pro Voiceprint page showing recording guidance, pricing notice, filters, and creation control
Pro Voiceprint requires substantially more clean audio than the short standard voiceprint workflow.
Voice cloning must be used only with the speaker’s informed permission. Do not upload recordings collected from calls, media, or third parties without clear rights and consent.

Preview and compare the result

Preview the created voice with a neutral sentence that contains short and long words, numbers, and a name. Compare it with a synthetic voice using the same target language and processing mode. Keep the voice only if intelligibility, pacing, and identity remain acceptable without adding excessive enhancement. Then place a short test call before using it with customers.

A usable voiceprint sounds clear on ordinary speakers, remains understandable at normal call volume, and does not require the listener to know the original speaker.

Frequently asked questions

These answers clarify the decisions that most often affect this workflow. Keep the working microphone, speaker, language pair, and call route unchanged while testing a feature, and return to the simplest verified baseline whenever several simultaneous changes make the result difficult to explain.

How long should the standard sample be?

Follow the limits shown in the current Add Your Voiceprint dialog. The observed build accepts a short sample and displays its minimum and maximum duration.

Should I enable background-noise removal?

Enable it for light, steady noise only after comparing the original. Re-record audio that clips, contains another speaker, or changes level significantly.

What is the difference between standard and Pro Voiceprint?

The standard workflow creates a voice from a short sample. Pro Voiceprint fine-tunes from a much larger clean single-speaker dataset and can require credits or a higher plan.

Can I use a recording of another person?

Only with that person’s informed permission and the necessary rights. Never clone a voice deceptively or from private call audio.

Support

Need help? Get in touch with our Support Team for assistance. Include the feature, operating system, calling application, selected microphone and speaker, language direction, and the exact step that failed. Describe the expected result and what happened instead without including passwords, private call content, or unnecessary personal data.

Continue with the resource that matches the next decision in your workflow. Keep the current configuration available for comparison, change one setting at a time, and validate each result in a short test call before applying it to a live conversation.

Was this article helpful?