GLM-TTS Test Console

checking…

Zhipu AI GLM-TTS — Zero-shot bilingual voice cloning (Chinese & English)

Text to Synthesize
45 / 500
Voice Reference (Required)
🎤
Click to upload or drag & drop a WAV / MP3 file
GLM-TTS clones the speaker identity from a 3–10 second reference clip. Clear speech with minimal background noise gives the best results. Both Chinese and English references are supported.
Reference preview:

Leave blank if unknown. Providing the correct transcript significantly improves voice cloning quality.

Emotion
50
Audio Adjustments
Authentication