Model catalog
One endpoint, every model. Compare context windows, capabilities and real per-token pricing side by side, then swap the model string in your request — nothing else changes.
- 351
- Models
- 35
- Providers
- 11
- Free to try
- $0.0001
- Cheapest input / 1M
14 of 351 models
- Language
GLM 5V Turbo
zai/glm-5v-turbo
GLM-5V-Turbo is Z.AI’s first multimodal coding foundation model, built for vision-based coding tasks. It can natively process multimodal inputs such as images, video, and text, while also excelling at long-horizon planning, complex coding, and action execution. Deeply optimized for agent workflows, it works seamlessly with agents such as Claude Code and OpenClaw to complete the full loop of “understand the environment → plan actions → execute tasks”.
- Context
- 200K
- In / 1M
- $1.20
- Out / 1M
- $4.00
- Reasoning
- Tool use
- Vision
- +2
- Language
GLM 4.5V
zai/glm-4.5v
Built on the GLM-4.5-Air base model, GLM-4.5V inherits proven techniques from GLM-4.1V-Thinking while achieving effective scaling through a powerful 106B-parameter MoE architecture.
- Context
- 66K
- In / 1M
- $0.600
- Out / 1M
- $1.80
- Reasoning
- Tool use
- Vision
- +1
- Language
GLM 5.3
zai/glm-5.3
GLM 5.3 delivers comprehensive advancements in complex software engineering and agent capabilities. It uses the same base model as GLM-5.2, with all improvements driven by post-training.
- Context
- 1M
- In / 1M
- $1.40
- Out / 1M
- $4.40
- Reasoning
- Tool use
- Implicit caching
- Language
GLM 5.2 Fast
zai/glm-5.2-fast
Fast version of GLM 5.2 with 120-250 TPS.
- Context
- 1M
- In / 1M
- $2.10
- Out / 1M
- $6.60
- Reasoning
- Tool use
- Implicit caching
- Language
GLM 5.2
zai/glm-5.2
GLM-5.2 delivers powerful coding capabilities, usable 1M-context support, and continued strengths in long-horizon tasks.
- Context
- 1M
- In / 1M
- $0.800
- Out / 1M
- $2.55
- Reasoning
- Tool use
- Implicit caching
- Language
GLM 5.1
zai/glm-5.1
GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on a single task for more than 8 hours—autonomously planning, executing, and improving itself throughout the process—ultimately delivering complete, engineering-grade results.
- Context
- 203K
- In / 1M
- $1.40
- Out / 1M
- $4.40
- Reasoning
- Tool use
- Implicit caching
- Language
GLM 5 Turbo
zai/glm-5-turbo
GLM 5 Turbo is a foundation model deeply optimized for the OpenClaw scenario. It has been specifically optimized for the core requirements of OpenClaw tasks since the training phase, enhancing key capabilities such as tool invocation, command following, timed and persistent tasks, and long-chain execution.
- Context
- 203K
- In / 1M
- $1.20
- Out / 1M
- $4.00
- Reasoning
- Tool use
- Implicit caching
- Language
GLM 5
zai/glm-5
GLM 5 is a frontier-class, general-purpose large language model optimized for complex systems engineering and long-horizon agentic tasks. It builds on the GLM 4.5 agent-centric lineage and is designed to support multi-step reasoning, math (including AIME-style benchmarks), advanced coding, and tool-augmented workflows, with long context support suitable for sophisticated agents and enterprise applications. Typical uses include autonomous agents for software engineering, data and systems troubleshooting, operations copilots, and high-end chat assistants that must break down complex tasks, call tools reliably, and reason over long sequences of instructions or documents.
- Context
- 203K
- In / 1M
- $1.00
- Out / 1M
- $3.20
- Reasoning
- Tool use
- Implicit caching
- Language
GLM 4.7 Flash
zai/glm-4.7-flash
GLM-4.7-Flash balances high performance with efficiency, making it the perfect lightweight deployment option. Beyond coding, it is also recommended for creative writing, translation, long-context tasks, and roleplay.
- Context
- 200K
- In / 1M
- $0.070
- Out / 1M
- $0.400
- Reasoning
- Tool use
- Implicit caching
- Language
GLM 4.7 FlashX
zai/glm-4.7-flashx
GLM-4.7-Flash balances high performance with efficiency, making it the perfect lightweight deployment option.
- Context
- 200K
- In / 1M
- $0.060
- Out / 1M
- $0.400
- Reasoning
- Tool use
- Implicit caching
- Language
GLM 4.7
zai/glm-4.7
GLM-4.7 is Z.ai’s latest flagship model, with major upgrades focused on two key areas: stronger coding capabilities and more stable multi-step reasoning and execution.
- Context
- 200K
- In / 1M
- $0.600
- Out / 1M
- $2.20
- Reasoning
- Tool use
- Implicit caching
- Language
GLM 4.6
zai/glm-4.6
As the latest iteration in the GLM series, GLM-4.6 achieves comprehensive enhancements across multiple domains, including real-world coding, long-context processing, reasoning, searching, writing, and agentic applications.
- Context
- 200K
- In / 1M
- $0.600
- Out / 1M
- $2.20
- Reasoning
- Tool use
- Implicit caching
- Language
GLM 4.5
zai/glm-4.5
GLM-4.5 and GLM-4.5-Air are our latest flagship models, purpose-built as foundational models for agent-oriented applications. Both leverage a Mixture-of-Experts (MoE) architecture. GLM-4.5 has a total parameter count of 355B with 32B active parameters per forward pass, while GLM-4.5-Air adopts a more streamlined design with 106B total parameters and 12B active parameters.
- Context
- 128K
- In / 1M
- $0.600
- Out / 1M
- $2.20
- Reasoning
- Tool use
- Implicit caching
- Language
GLM 4.5 Air
zai/glm-4.5-air
GLM-4.5 and GLM-4.5-Air are our latest flagship models, purpose-built as foundational models for agent-oriented applications. Both leverage a Mixture-of-Experts (MoE) architecture. GLM-4.5 has a total parameter count of 355B with 32B active parameters per forward pass, while GLM-4.5-Air adopts a more streamlined design with 106B total parameters and 12B active parameters.
- Context
- 128K
- In / 1M
- $0.200
- Out / 1M
- $1.10
- Reasoning
- Tool use
- Implicit caching
