female voice ยท A ยท Kokoro
warm ยท natural
Free AI English text to speech online. 28 American and British English voices, local text processing, WAV and MP3 export, and no signup.
Try it now โ no signup or per-character charge
Browser audio synthesis; text handling and network needs depend on the selected engine and language.
Compare 28 Kokoro voices + 9 Kyutai/Pocket for english with a fixed sample. Voice type, grade, and traits are catalog cues rather than a quality guaranteeโuse the same part of your real script before choosing one.
Kyutai voices run on Pocket TTS only (voice-cloning via hf://kyutai/tts-voices/โฆ), not Kokoro.
| Voice | Type | Catalog cues | Preview |
|---|---|---|---|
| female ยท A | warm ยท natural | ||
| female ยท A- | expressive | ||
| female ยท B- | professional | ||
| female ยท C | neutral | ||
| female ยท C+ | melodic | ||
| female ยท D | Compare with your script | ||
| female ยท C+ | clear | ||
| female ยท C | modern | ||
| female ยท D | Compare with your script | ||
| female ยท C+ | warm | ||
| female ยท C- | bright | ||
| male ยท F+ | Compare with your script | ||
| male ยท D | resonant | ||
| male ยท D | Compare with your script | ||
| male ยท C+ | deep | ||
| male ยท D | Compare with your script | ||
| male ยท C+ | professional | ||
| male ยท D | Compare with your script | ||
| male ยท C+ | playful | ||
| male ยท D- | Compare with your script | ||
| female ยท B- | professional | ||
| female ยท D | Compare with your script | ||
| female ยท C | Compare with your script | ||
| female ยท D | Compare with your script | ||
| male ยท D | Compare with your script | ||
| male ยท C | Compare with your script | ||
| male ยท C | Compare with your script | ||
| male ยท D+ | Compare with your script |
female voice ยท A ยท Kokoro
warm ยท natural
female voice ยท A- ยท Kokoro
expressive
female voice ยท B- ยท Kokoro
professional
female voice ยท C ยท Kokoro
neutral
female voice ยท C+ ยท Kokoro
melodic
female voice ยท D ยท Kokoro
Compare this voice with your own script
female voice ยท C+ ยท Kokoro
clear
female voice ยท C ยท Kokoro
modern
female voice ยท D ยท Kokoro
Compare this voice with your own script
female voice ยท C+ ยท Kokoro
warm
female voice ยท C- ยท Kokoro
bright
male voice ยท F+ ยท Kokoro
Compare this voice with your own script
male voice ยท D ยท Kokoro
resonant
male voice ยท D ยท Kokoro
Compare this voice with your own script
male voice ยท C+ ยท Kokoro
deep
male voice ยท D ยท Kokoro
Compare this voice with your own script
male voice ยท C+ ยท Kokoro
professional
male voice ยท D ยท Kokoro
Compare this voice with your own script
male voice ยท C+ ยท Kokoro
playful
male voice ยท D- ยท Kokoro
Compare this voice with your own script
female voice ยท B- ยท Kokoro
professional
female voice ยท D ยท Kokoro
Compare this voice with your own script
female voice ยท C ยท Kokoro
Compare this voice with your own script
female voice ยท D ยท Kokoro
Compare this voice with your own script
male voice ยท D ยท Kokoro
Compare this voice with your own script
male voice ยท C ยท Kokoro
Compare this voice with your own script
male voice ยท C ยท Kokoro
Compare this voice with your own script
male voice ยท D+ ยท Kokoro
Compare this voice with your own script
| Voice | Type | Preview |
|---|---|---|
| male ยท B+ ยท calm ยท steady | ||
| male ยท B ยท youthful ยท energetic | ||
| female ยท B ยท clear ยท professional | ||
| female ยท B+ ยท casual ยท character | ||
| female ยท B ยท gentle ยท narrative | ||
| male ยท B ยท audiobook ยท warm | ||
| male ยท B ยท narrator ยท classic | ||
| male ยท B ยท classic ยท steady | ||
| female ยท B ยท british ยท deep |
calm ยท steady
youthful ยท energetic
clear ยท professional
casual ยท character
gentle ยท narrative
audiobook ยท warm
narrator ยท classic
classic ยท steady
british ยท deep
Enhanced wav via ai-coustics where available. NC datasets (expresso/ears) excluded. Use Pocket TTS to generate.
Our English text-to-speech tool runs entirely in your browser using the Kokoro TTS model. With 20 American English voices and 8 British English voices, you can find the perfect tone for any project.
Best for: YouTube voiceovers, podcast intros, audio books, e-learning materials, and accessibility.
American English voices include Heart (A), Bella (A-), Nicole (A), Sarah (A-), Nova (B+), Kore (B+), Jessica (B+), Sky (B), River (B), and more. These voices are optimized for North American pronunciation and natural prosody.
British English voices include Emma (B-), Alice (B+), Isabella (B+), Lily (B+), Daniel (B+), George (B-), Lewis (B-), Fable (B-). These voices deliver authentic British pronunciation with proper rhythm and intonation for UK-focused content.
Whether you need a warm voice for business content or a conversational tone for a creative project, test the exact script and keep the selected engine, voice, and speed with the export.
The English page combines twenty American and eight British Kokoro presets, but a flag or friendly name does not guarantee one regional pronunciation for every word. Test proper nouns, dates, acronyms, currencies, and domain-specific terms in the exact voice you plan to use. If a script mixes US and UK spelling, decide which reading convention the audience expects before generating a long file.
English Kokoro uses local phonemization after the model and selected voice assets are available, so the entered script is not sent to the non-English OfflineTTS phonemization endpoint. The first load, model host, website hosting, and ordinary analytics still involve network requests. Clear browser site data to remove cached assets, and delete WAV or MP3 exports separately from the device.
A useful evaluation passage includes a person and place name, an abbreviation, a decimal, a year, a currency amount, a quotation, and one long sentence. Compare several voices with the same q4 or fp32 model, speed, browser, and backend. Catalog grades and traits are internal discovery labels rather than a controlled benchmark, so a lower-sorted preset may fit a particular script better.
For long narration, keep paragraphs coherent and listen across each automatic chunk join. Check for repeated words, clipped consonants, inconsistent volume, or pauses that do not match the text. WAV is the safer intermediate for editing; MP3 is smaller but lossy. Keep the reviewed source script and voice ID with the export so a correction can be reproduced.
English Kokoro phonemization and audio synthesis run in the browser after required model and voice assets load.
After model and voice assets are cached, English generation can work without a network connection until those assets are cleared.
No API key or signup is required; the current tool accepts up to 50,000 characters in one generation workflow.
Download as WAV for studio-quality audio or MP3 for compressed output. Compatible with all editors.
Add professional English narration to videos without recording equipment. Top voices: Heart (warm), Bella (energetic), Michael (professional).
Generate intros, outros, and full episodes with consistent voice quality. Mix multiple voices for interview-style segments.
Convert manuscripts to audiobooks with natural-sounding voices. Export as WAV for post-production.
Add voice narration to courses and make content accessible to visually impaired users across English-speaking audiences.
Enter your english text (up to 50,000 chars)
Pick from english voices
AI creates speech on your device
Save as WAV or MP3
Yes. OfflineTTS does not charge per character and requires no signup or API key. Audio synthesis runs in your browser with the engine you select.
Offline behavior depends on the selected engine and language. Supertonic, Piper, Kitten, Pocket TTS, and English Kokoro can synthesize locally after their model files download. Non-English Kokoro uses the OfflineTTS phonemization service before audio is generated on your device.
Audio synthesis runs locally in your browser. Fully local-capable engine and language combinations keep the input on your device after model download. Non-English Kokoro sends plain text to the OfflineTTS phonemization service and receives pronunciation data before local synthesis.
No signup, no per-character fee, and browser-based audio synthesis.
Open TTS Tool โ