Skip to main content
Chinese AI Tools
ProductsModelsIntegrationsRankingsLatest changes
TopicsAvailability TrackerUse casesSubmit a toolAccount
ZHSearch

Chinese AI Tools

Independent directory for Chinese AI products. Product availability, pricing and terms can change. Verify before commercial use.

Editorial standardsClaim productUpdate infoGet featuredAdvertise

Zhipu AI

GLM Audio

Z.ai's audio and realtime multimodal API family, including GLM-TTS, voice clone, ASR, Realtime and GLM-4-Voice.

Partially availablePartial English UIPublic APIFreemium

Quick answer

Audio is now a documented GLM capability family and should be visible in the audio category.

Documentation
Pricing
Usage-based audio API pricing varies by model
Availability
Partially available
API
Public API
Last checked
2026-08-04
Pricing detailsUse-case fitSource evidence

BigModel's model overview lists GLM-TTS, GLM-TTS-Clone, GLM-ASR, GLM-Realtime and GLM-4-Voice under audio/video models, covering speech synthesis, voice cloning, speech recognition and realtime audio-video interaction.

Topic guide

GLM models, Coding Plan and Z.ai ecosystem

A source-backed map of the original GLM-5.3, open multimodal GLM-5.3-Flash, GLM-5.2, BigModel APIs, Coding Plan, ZCode, consumer agents and media products.

Editorial verdict

Best for

Developers evaluating Chinese speech, voice clone, ASR and realtime multimodal APIs.

Avoid if

Avoid production voice cloning without consent and data-retention review.

Why it matters

Audio is now a documented GLM capability family and should be visible in the audio category.

Trust: 2/2 sources verified, recently checkedCoverage: 100/100

Pricing

Usage-based audio API pricing varies by model

Payment

Free model where available, Platform billing

Commercial use

Commercial use should follow the current product, API, model license and billing terms.

Privacy

Review prompt, file, media upload, retention and training-use terms before sensitive workloads.

Use-case fit

Speech and realtime API evaluation

Strong

Use it to test TTS, voice cloning, ASR and realtime audio-video calls.

Global user checklist

RegistrationPartialAccess depends on BigModel account and region.
English UIPartialDetailed audio docs are Chinese-facing.
API and docsConfirmedDocs index lists text-to-speech, voice clone, ASR and realtime APIs.
Commercial useReviewVoice rights and consent are required for real-person voice use.

Model names, quotas, release status, regional access and commercial terms can change quickly; recheck official sources before procurement or production use.

Pros

  • - Speech synthesis, clone, ASR and realtime models are all documented
  • - Useful complement to GLM chat and vision APIs

Cons

  • - Voice consent and regional access need explicit checks

Decision paths

minimax-audio

sparkdesk

zhipu-glm

Sources

BigModel model overview

Documentation · ZH · verified 2026-08-04

Lists GLM-TTS, GLM-TTS-Clone, GLM-ASR, GLM-Realtime and GLM-4-Voice.

BigModel docs index

Documentation · ZH · verified 2026-05-17

Lists speech, voice clone, ASR and realtime API documentation entries.

Last checked: 2026-08-04

Reviews

No approved reviews yet.

Availability snapshot

Availability
Partially available
English UI
Partial English UI
API
Public API
Rating
3.9 (0)

Latest updates

Latest changes
Market · 2026-05-17

Z.ai international product map expanded

Zhipu AI is now tracked under the Z.ai brand across BigModel/GLM, z.ai, ChatGLM, AutoGLM, AutoClaw, Zread.ai, AMiner, Zlearn, Autotyper and GLM multimodal API lines.

API · 2026-05-17

BigModel GLM multimodal APIs filled in

The BigModel profile now reflects official docs for GLM text, vision, image, video, audio, embedding, rerank, agent, OCR, web search, fine-tuning and OpenAI-compatible access.

Submit a review