GLM-5.3 Flash
Z.ai's fast-tier model: $0.15 per million input tokens and $0.50 per million output tokens, with a 1M-token context window.
Input / 1M tokens$0.15
Cached input / 1M$0.03
Output / 1M tokens$0.50
Example task$0.20$0.10 with 80% cached
| Maker | Z.ai · official pricing |
|---|---|
| Context window | 1M tokens · up to 128K output |
| Capabilities | tool calling · reasoning · images |
Prices from LiteLLM's open price list, updated Oct 9, 2026. "Example task": 1M input + 100K output tokens. Check the provider's page before you budget.
Agents that take any provider
These work with OpenRouter or any OpenAI-compatible endpoint, so they can usually run GLM-5.3 Flash too.
OpenAI Codex LangChain Pi OpenCode GitHub Copilot LangGraph AI SDK Zed Mastra Deep Agents CopilotKit Cloudflare Agents SDK Oh My Pi DSPy OpenAI Agents SDK Junie Agno Cline CrewAI mini-SWE-agent OpenHands eve LlamaIndex Agent Development Kit (ADK)
Cheaper fast-tier alternatives
Claude Haiku 5.5 · $0.15 GPT-6 Luna · $0.15 MiMo V2.6 Flash · $0.17