Many of you have frequently asked about speaker separation, voice selection, and intonation control, so I have summarized them for you.
🔊 Q1. Can I choose the speaker's voice directly?
Currently, there is no separate feature to directly select a voice.
Instead , voice cloning is automatically applied based on the original speaker's voice and tone ,
It provides natural dubbing results similar to the original.
👥 Q2. Is multi-speaker (3 or more) dubbing possible?
Yes, we currently support multi-speaker dubbing with 3 or more speakers .
•
Technically, it is optimized for two speakers and can detect up to 10 people .
⚠️ However, since recognition errors may occur with speakers having similar voices ,
For accuracy, we recommend working with two speakers .
🎯 Q3. What is the accuracy of speaker separation?
•
Perso AI's speaker separation accuracy is the best in the industry, but
•
It can be difficult to distinguish between speakers with similar vocal characteristics .
•
This is also a technical challenge for the entire AI Dubbing industry .
•
We are currently developing a new approach and plan to implement it sequentially as soon as stability verification is complete.
👉 Temporary Recommendation: The more distinct the voice characteristics between speakers, the more stable the results.
Subscribe to 'Perso AI Community Hub'
By subscribing to the site, you will be the first to receive the latest updates, including new posts, via notifications and email. Subscribe to the Perso AI Community Hub!