Skip to content
z-ai/glm-5.3-flash

Zhipu AI · 1M context · text + image + video → text · 7 upstreams

You need your own AIHubMix key to use this model.

The provider bills you. AnyRouter fee is $0.

Only route: your key

Call this id from your code; another provider takes over when one fails. How it works →

Upstreams and pricing

ProviderContextStatusProvider details
Your own key (BYOK) — provider bills you, $0 AnyRouter fee
0.15 list0.50 list1M
0.15 list0.50 list1M
0.15 list0.50 list1M
0.15 list0.50 list1M
0.15 list0.50 list1M
0.15 list0.50 list1M
0.15 list0.50 list1.1M
Free pool — donated keys: not available for this model — donate a key

Code

POST /api/v1/chat/completionsSDK docs
import osfrom openai import OpenAI client = OpenAI(    api_key=os.environ["ANYROUTER_API_KEY"],    base_url="https://anyrouter.dev/api/v1",) resp = client.chat.completions.create(    model="z-ai/glm-5.3-flash",    messages=[{"role": "user", "content": "Say hi in 3 words."}],)print(resp.choices[0].message.content)

GLM-5.3-Flash is a native multimodal model from Z.ai (320B-A18B, MIT License, 1M-token context). It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while reducing compute overhead.

Aliases (1)
  • zai-org/glm-5.3-flash