female voice · D · Kokoro
Compare this voice with your own script
Free Mandarin Chinese text to speech with 8 Kokoro voices, local audio synthesis, WAV and MP3 export, and no signup.
Try it now — no signup or per-character charge
Browser audio synthesis; text handling and network needs depend on the selected engine and language.
Compare 8 Kokoro voices for mandarin chinese with a fixed sample. Voice type, grade, and traits are catalog cues rather than a quality guarantee—use the same part of your real script before choosing one.
| Voice | Type | Catalog cues | Preview |
|---|---|---|---|
| female · D | Compare with your script | ||
| female · D | Compare with your script | ||
| female · D | Compare with your script | ||
| female · D | Compare with your script | ||
| male · D | Compare with your script | ||
| male · D | Compare with your script | ||
| male · D | Compare with your script | ||
| male · D | Compare with your script |
female voice · D · Kokoro
Compare this voice with your own script
female voice · D · Kokoro
Compare this voice with your own script
female voice · D · Kokoro
Compare this voice with your own script
female voice · D · Kokoro
Compare this voice with your own script
male voice · D · Kokoro
Compare this voice with your own script
male voice · D · Kokoro
Compare this voice with your own script
male voice · D · Kokoro
Compare this voice with your own script
male voice · D · Kokoro
Compare this voice with your own script
Convert Mandarin Chinese text to natural speech with Kokoro TTS. The OfflineTTS phonemization service converts Chinese text into pronunciation data, then Kokoro synthesizes the audio locally in your browser.
8 Mandarin Chinese voices (4 female, 4 male) ready for your projects. Top voices include Xiaobei (clear, articulate), Xiaoyi (warm, conversational), Yunjian (professional, authoritative), and Yunyang (natural, friendly).
try our TTS tool to generate Chinese speech instantly, or explore English TTS and Japanese TTS for other language options.
Chinese Kokoro sends the entered text to the OfflineTTS phonemization service before local waveform synthesis. Characters with more than one reading, personal names, place names, abbreviations, numbers, and mixed Latin text can be ambiguous without context. The service does not make the workflow fully offline, and generated audio should not be treated as a verified pronunciation reference for sensitive, legal, or educational material.
Use complete sentences rather than isolated characters when context determines a reading. Add punctuation between clauses, expand unusual abbreviations, and write numbers in the form you want spoken. Test the exact names and terminology in every selected voice. When text must stay local after model download, evaluate Supertonic Chinese support and its built-in styles as a separate engine choice.
A useful Mandarin test passage contains tone sandhi contexts, 一 and 不, neutral-tone words, optional erhua, dates, units, and an English brand or acronym. Listen for intelligibility and regional fit with a fluent reviewer. The catalog voice name, gender field, or grade does not certify a specific regional accent or guarantee consistent tone realization in every sentence.
Generate in short, meaningful paragraphs and compare the audio with the source line by line. Do not use punctuation to hide a wrong lexical reading; revise the written form or provide clearer context. Save the voice ID, speed, model precision, browser, and test date with an important export, and confirm rights in the source material before distributing synthetic narration.
Chinese text is phonemized by the OfflineTTS service, while the generated audio remains on your device.
The speech model is cached after its first download; Chinese phonemization still requires a connection.
No API key or signup is required; the current generation workflow accepts up to 50,000 characters.
Download your Chinese speech as lossless WAV or compressed MP3 for any production pipeline.
Hear how Chinese characters and tones should sound. Paste individual words or full passages for natural pronunciation modeling.
Add professional Chinese narration to videos, presentations, and e-learning courses without hiring a voice actor.
Convert written Chinese content to speech for visually impaired users or anyone who prefers listening over reading.
Generate Chinese speech for internal training materials, product demos, and customer-facing content — all processed privately on your device.
Enter your mandarin chinese text (up to 50,000 chars)
Pick from mandarin chinese voices
AI creates speech on your device
Save as WAV or MP3
Yes. OfflineTTS does not charge per character and requires no signup or API key. Audio synthesis runs in your browser with the engine you select.
Offline behavior depends on the selected engine and language. Supertonic, Piper, Kitten, Pocket TTS, and English Kokoro can synthesize locally after their model files download. Non-English Kokoro uses the OfflineTTS phonemization service before audio is generated on your device.
Audio synthesis runs locally in your browser. Fully local-capable engine and language combinations keep the input on your device after model download. Non-English Kokoro sends plain text to the OfflineTTS phonemization service and receives pronunciation data before local synthesis.
No signup, no per-character fee, and browser-based audio synthesis.
Open TTS Tool →