# GLM-5.3-FlashX

> GLM-5.3-FlashX is Z.AI's high-speed inference variant for coding agents, real-time interactions, and long-running agentic workflows. AIHubMix lists it separately from GLM-5.3-Flash with its own rates and modality set.

**ID**: `z-ai/glm-5.3-flashx`  
**Creator**: Zhipu AI  
**Category**: multimodal  
**Context**: 1M tokens  
**Input**: $0.37/M  
**Output**: $1.25/M  
**Released**: 2026-09-20  
**Web page**: https://anyrouter.dev/model/z-ai/glm-5.3-flashx

**Input modalities**: text, image, video  
**Output modalities**: text  

**Tokenizer**: GLM

**Capabilities**: chat, streaming, agentic, coding, vision

## Usage

**Endpoint**: `POST https://anyrouter.dev/api/v1/chat/completions`

```bash
curl https://anyrouter.dev/api/v1/chat/completions \
  -H "Authorization: Bearer $ANYROUTER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
  "model": "z-ai/glm-5.3-flashx",
  "messages": [
    {
      "role": "user",
      "content": "Say hi in 3 words."
    }
  ]
}'
```

## Providers

| Provider | Input | Output |
| --- | --- | --- |
| aihubmix-byok | $0 | $0 |

## Source

- Upstream docs: https://aihubmix.com/model/glm-5.3-flashx
