CATALOG

Models

Find the right model by capability, context and price. Compare options or try one in Playground.

Output / endpointExact metadata
20 of 599 modelsReference token prices are per million tokens.
VisionVideo inputToolsReasoning

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

by z-aiAug 26, 20261.31M context$0.09/M input$0.3/M outputText, Image, VideoText
VisionVideo inputToolsReasoning

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

by z-aiAug 26, 20261.05M context$0.075/M input$0.25/M outputText, Image, VideoText
ToolsReasoning

GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...

by z-aiAug 18, 20261.31M context$1.4/M input$4.4/M outputTextText
ToolsReasoning

GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...

by z-aiAug 18, 20261.05M context$0.7/M input$2.2/M outputTextText
ToolsReasoning

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

by z-aiJun 16, 20261.05M context$1.4/M input$4.4/M outputTextText
ToolsReasoning

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

by z-aiJun 16, 20261.05M context$0.7/M input$2.2/M outputTextText
FreeReasoning

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

by z-aiJun 16, 202632.77K context$0/M input$0/M outputTextText
ToolsReasoning

GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...

by z-aiApr 7, 2026204.8K context$0.966/M input$3.04/M outputTextText
VisionVideo inputToolsReasoning

GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...

by z-aiApr 1, 2026202.75K context$1.2/M input$4/M outputImage, Text, VideoText
ToolsReasoning

GLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments such as OpenClaw scenarios. It is deeply optimized for real-world agent workflows...

by z-aiMar 15, 2026202.75K context$1.2/M input$4/M outputTextText
ToolsReasoning

GLM-5 is Z.ai’s flagship open-source foundation model engineered for complex systems design and long-horizon agent workflows. Built for expert developers, it delivers production-grade performance on large-scale programming tasks, rivaling leading...

by z-aiFeb 11, 2026204.8K context$0.6/M input$1.92/M outputTextText
ToolsReasoning

As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning,...

by z-aiJan 19, 2026200K context$0.0605/M input$0.4/M outputTextText
ToolsReasoning

GLM-4.7 is Z.ai’s latest flagship model, featuring upgrades in two key areas: enhanced programming capabilities and more stable multi-step reasoning/execution. It demonstrates significant improvements in executing complex agent tasks while...

by z-aiDec 22, 2025204.8K context$0.4/M input$1.75/M outputTextText
VisionVideo inputToolsReasoning

GLM-4.6V is a large multimodal model designed for high-fidelity visual understanding and long-context reasoning across images, documents, and mixed media. It supports up to 128K tokens, processes complex page layouts...

by z-aiDec 8, 2025131.07K context$0.3/M input$0.9/M outputImage, Text, VideoText
ToolsReasoning

Compared with GLM-4.5, this generation brings several key improvements: Longer context window: The context window has been expanded from 128K to 200K tokens, enabling the model to handle more complex...

by z-aiSep 30, 2025204.8K context$0.43/M input$1.75/M outputTextText
VisionToolsReasoning

GLM-4.5V is a vision-language foundation model for multimodal agent applications. Built on a Mixture-of-Experts (MoE) architecture with 106B parameters and 12B activated parameters, it achieves state-of-the-art results in video understanding,...

by z-aiAug 11, 202565.54K context$0.6/M input$1.8/M outputText, ImageText
ToolsReasoning

GLM-4.5 is our latest flagship foundation model, purpose-built for agent-based applications. It leverages a Mixture-of-Experts (MoE) architecture and supports a context length of up to 128k tokens. GLM-4.5 delivers significantly...

by z-aiJul 25, 2025131.07K context$0.6/M input$2.2/M outputTextText
ToolsReasoning

GLM-4.5-Air is the lightweight variant of our latest flagship model family, also purpose-built for agent-centric applications. Like GLM-4.5, it adopts the Mixture-of-Experts (MoE) architecture but with a more compact parameter...

by z-aiJul 25, 2025131.07K context$0.13/M input$0.85/M outputTextText