GLM 5.3

Reasoning, multilingual. Runs on Confidential Inference: an OpenAI-compatible endpoint inside TEEs, with an attestation on every response.

StatusInput (per 1M tokens)Input Cached (per 1M tokens)Output (per 1M tokens)
On request$1.40$0.26$4.40

Get access

Full pricing: pricing. Docs: get started.