DeepSeek V4 Flash is live · Our GLM-5.2 price just dropped to 50% of list
All models
Z.ai · API

GLM-5.3-Flash API

Use 1M context with one OmniaKey API key.

Model ID
glm-5.3-flash
Context
1M
Provider
Z.ai
Input · USD / 1M tokens
$0.075 USD
Official price $0.15 USD
Output · USD / 1M tokens
$0.25 USD
Official price $0.5 USD
Savings
50%
Last verified
2026-09-04

01

Where it fits

Choose it for cheap, frequent coding and multimodal calls that do not need the flagship GLM-5.3 route.

02

Model capabilities

Vision

Accepts visual input alongside text for multimodal coding and analysis tasks.

Tool use

Can select and call tools in agent workflows that support structured tool use.

Fast responses

Optimized for responsive iteration and higher-volume agent tasks.

03

OpenAI-compatible API

Keep your client and change the base URL, API key, and model ID.

curl
curl https://api.omniakey.com/v1/chat/completions \
  -H "Authorization: Bearer $OMNIAKEY_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "glm-5.3-flash",
    "messages": [{"role": "user", "content": "Explain this codebase"}],
    "stream": true
  }'

04

Compatible coding tools

Claude Code
Codex
Cursor
Cline
Aider

More Z.ai models

Frequently asked questions

How much does GLM-5.3-Flash cost on OmniaKey?

OmniaKey charges $0.075 USD per 1M input tokens and $0.25 USD per 1M output tokens, in USD. Official prices are $0.15 USD and $0.5 USD; you save 50%.

What is the context window for GLM-5.3-Flash?

The listed context window is 1M. Actual usable limits can also depend on the client and upstream request format.

Which coding tools work with GLM-5.3-Flash?

The current catalog lists compatibility with Claude Code, Codex, Cursor, Cline, Aider. Use the model ID shown on this page when the client supports model selection.

GLM-5.3-Flash API, Pricing & Coding Guide · OmniaKey