AI tools for text-to-speech narration
Turn a supplied script into downloadable narration using a selected synthetic voice, then review wording and audio before use.
For text-to-speech narration, compare Kokoro Web by their documented outputs, setup, operating costs and review requirements. Choose against this task’s inputs and limitations; this guide does not rank measured product performance.
Explore the content creators & social media industry guide →
Inputs & prerequisites
A script, target language/accent, selected synthetic voice, speed and intended audio use.
Inspect the actual first audio for intelligibility, complete wording, numbers, dates and negations, then check the chosen output format and applicable script/model rights.
Candidates & fit
Evaluate when: Creators, educators, small shops and community organizers who already have a short script and want a browser workflow to generate and download narration using a built-in synthetic voice, then review the words and audio before use.
Check first: The recorded native results are partial. MP3 file decoding and duration do not establish intelligibility, exact wording, preserved numbers or negations, or pause accuracy; the available audio-review model could not accept audio.
Documented use: The official README describes a browser-based text-to-speech generator with multiple language accents, voice customization, WebGPU options and optional self-hosting. Inspect evidence →
Filter this task’s candidates
Default: search match priority, or documented detail then name. No popularity scores.
Kokoro Web is Eduardo Lat's browser text-to-speech application for creating downloadable narration from a supplied script. Creators, small businesses, educators and community organizers can choose a built-in synthetic voice, language accent, model quantization and speed, generate speech locally in the Browser route, then play and download the result. It uses the open-weight Kokoro model. Luis Eduardo is credited in the MIT license and official author links; the repository was created on 9 February 2025 and last pushed on 16 March 2025 in the reviewed snapshot. Both frozen native cases produced downloaded MP3 files. Both outcomes remain partial because listening checks could not be completed.
Compare the same facts
| Selection question | Kokoro WebEduardo Lat |
|---|---|
| Product type | Specialist AI tool |
| Best suited for | Creators, educators, small shops and community organizers who already have a short script and want a browser workflow to generate and download narration using a built-in synthetic voice, then review the words and audio before use. |
| Shared / related tasks | |
| Price model | Free browser TTS; optional self-hostingThe official README and served footer describe free personal and commercial use. The observed Browser route exposes local generation without a user account or provider API key. The application is MIT licensed and its fixed ONNX model card is Apache-2.0. First use downloads model, voice and browser runtime resources; hosting, network and hardware costs are separate from this free-access statement. |
| Price amount | Not verified — not assumed to be $0 |
| Deployment | The official online UI is voice-generator.pages.dev. Select Execution place Browser for browser-local inference; API mode is a separate self-hosted branch. The frozen cases use CPU and the model_quantized 8-bit option, so the planned route does not require WebGPU. First generation loads the fixed ONNX model and voice plus eSpeak NG, ONNX Runtime and FFmpeg browser resources. The native UI displayed v0.1.3; its bundle was not bound to the retained source commit. The public repository snapshot was last pushed on 16 March 2025. |
| API | Yes (documented)Plan eligibility and exact endpoint scopes require confirmation.Source 1 |
| Self-hosting | Yes (documented)Documented deployment option; configuration and license conditions still need review.Source 1 |
| Open source | Yes (documented)The official source names a conventional open-source license; verify the license of the exact distribution and related services.Source 1 |
| Inputs | A supplied text script, selected language accent and built-in voice, model quantization, acceleration and speech speed. The two original English fixtures are fictional community and parcel notices containing exact numbers, dates, negations and an advertised silence tag. Their full original inputs and unchanged conditions appear in the Tests section; no voice recording is used. |
| Outputs | The native browser workflow exposes an audio player and Download control. Its selected default format is MP3. Two first native MP3 exports are retained. Primary: 244,365 bytes, 12.129125 seconds; boundary: 298,605 bytes, 14.837 seconds. Both decode as 24 kHz mono MP3. Audible intelligibility, complete wording, values, negations and boundary pause accuracy remain unverified. |
| Platform relationship | Not verified |
| Permissions | Exact action scopes not verified |
| Commercial use | The official README and served footer describe free personal and commercial use. The application license is MIT and the fixed ONNX model card is Apache-2.0. Use a script you have rights to and review its actual audio before publishing. This record contains no separate commercial audio-rights opinion or total-cost measurement. |
| Human review | Review factual claims, captions, voice permissions, and copyright before publication. |
| Content Creators & Social Media · Creator platforms | Not verifiedNot verified in the reviewed official material. |
| Content Creators & Social Media · Content formats | Not verifiedNot verified in the reviewed official material. |
| Content Creators & Social Media · Inputs | A supplied text script, selected language accent and built-in voice, model quantization, acceleration and speech speed. The two original English fixtures are fictional community and parcel notices containing exact numbers, dates, negations and an advertised silence tag. Their full original inputs and unchanged conditions appear in the Tests section; no voice recording is used.Source 1 |
| Content Creators & Social Media · Outputs | Not verifiedNot verified in the reviewed official material. |
| Content Creators & Social Media · Aspect ratios | Not verifiedNot verified in the reviewed official material. |
| Content Creators & Social Media · Caption formats | Not verifiedNot verified in the reviewed official material. |
| Content Creators & Social Media · Voice & caption languages | Not verifiedNot verified in the reviewed official material. |
| Content Creators & Social Media · Commercial use terms | Not verifiedNot verified in the reviewed official material. |
| Content Creators & Social Media · Publishing by platform | Not verifiedNot verified in the reviewed official material. |
| Content Creators & Social Media · Approval requirements | Not verifiedNot verified in the reviewed official material. |
| Verification | Partially verified · 2026-10-03 |
| Recorded product tests | Public vendor UI product test · 2 cases executedLatest recorded case outcomes: 0 passed · 0 failed · 2 partial.Results apply to the recorded inputs, feature and configuration. They do not rank overall product quality. |
A practical evaluation sequence
- Prepare the inputs and define a reviewable output.
- Check the candidate’s documented capability and plan eligibility.
- Run a small sample in the vendor product with authorized data.
- Review the result, permissions, sources and full operating cost.
Cost & human review
Check base subscriptions, seats, usage credits, connected app charges, hosting and any licensed source data. Trial access does not establish a recurring free allowance. Exact unverified prices remain unknown.
Inspect the actual first audio for intelligibility, complete wording, numbers, dates and negations, then check the chosen output format and applicable script/model rights.
Sources & limits
What “official sources” means We read vendor material for the claims cited below. This is a documentation review. No independent product test or professional endorsement is implied. Read our method →
- Kokoro Web official browser application
Eduardo Lat · voice-generator.pages.dev · Read · 2026-10-03
- Kokoro Web official README at the retained source commit
Eduardo Lat · raw.githubusercontent.com · Read · 2026-10-03
- Kokoro Web MIT license: Copyright 2025 Luis Eduardo
Luis Eduardo · raw.githubusercontent.com · Read · 2026-10-03
- Kokoro Web repository metadata and maintenance dates
Eduardo Lat / GitHub · api.github.com · Read · 2026-10-03
- Served Kokoro Web footer, free-use statement and author links
Eduardo Lat · voice-generator.pages.dev · Read · 2026-10-03
- Native Browser generation branch and separate optional API branch
Eduardo Lat · raw.githubusercontent.com · Read · 2026-10-03
- Optional native profile saving stores text and settings in browser localStorage
Eduardo Lat · raw.githubusercontent.com · Read · 2026-10-03
- Kokoro Web model and voice resource loader with fixed ONNX revision
Eduardo Lat · raw.githubusercontent.com · Read · 2026-10-03
- Fixed Kokoro ONNX model card: Apache-2.0 and original model attribution
ONNX Community / Hugging Face · huggingface.co · Read · 2026-10-03
- Fixed Kokoro ONNX repository metadata: model/voice sizes, public and ungated state
ONNX Community / Hugging Face · huggingface.co · Read · 2026-10-03