View DocumentationTry NowGLM-ASR-2512Zhipu's next-generation speech recognition model supports real-time conversion of speech into high-quality text.APIAudio-Video ProcessingPricing:$0.025 /1M tokensGLM-ASR-2512
View DocumentationTry NowGLM-TTSGLM Speech Synthesis Model Combining Large Language Models and Diffusion Model TechnologiesAPIAudio-Video ProcessingPricing:$0.03 /1000 charactersGLM-TTS