What Are American English AI Voices?
American English AI voices are digital speech models trained specifically on datasets of North American speakers. They naturally handle American pronunciation, rhythm, and common phonetic quirks. For more on the underlying technology, see how text-to-speech works.
When provided with text, these voices read the words aloud using standard American intonation, making them highly effective for localized content aimed at US and Canadian audiences.
American English Voices Available
Konthora currently offers 6 distinct American English voices. They are available directly in your browser without creating an account.
When to Choose an American English Voice
Selecting an American accent over a British accent is typically a decision based on your target audience and the content you are producing.
- North American Audiences:If your content is aimed primarily at listeners in the United States or Canada, an American voice will sound the most natural and familiar.
- American Spelling and Slang:These voices are trained to naturally interpret American spellings (e.g., "color" instead of "colour") and regional phrasing without unnatural hesitation.
Choosing the Right Voice
Even within the American English category, the 6 voices have varying cadences and depths. We recommend previewing a few different options with a sample sentence from your actual script to see which one fits best.
Keep in mind that you can fine-tune the delivery by adjusting the playback speed between 0.75× and 1.25×, allowing you to match the exact pace required for your project.
Preview and Generate Speech
You can preview all 6 American English voices in the text to speech workspace. Simply enter up to 2,000 characters of text, select your preferred voice, and generate your audio.
Once generated, the voiceover must be downloaded during your active browser session as an MP3 or WAV file.
Ready to create a voiceover?
Try the American English voices now.