Z.ai: GLM 5.3 Flash

z-ai/glm-5.3-flash
Compare Model Lab

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

Modalities
Text + Image + Video → Text
Input / Output Price$0.075 / $0.25From · per 1M tokens · see provider prices
Context1.31Mtokens · maximum across providers
ReleasedAug 26, 2026OpenRouter public metadata
Availability & data sources Listed in CLSSAI · 28 reference endpoints
ACCESS

Where this model works

OpenRouter testing and CLSSAI API routing are independent decisions with separate sources.

OpenRouter testTestable with your OpenRouter key
Testable

Exact capability metadata is available. CLSSAI can prepare the Image understanding request without guessing the model contract.

Evidence
Exact ID z-ai/glm-5.3-flash · /chat/completions
Mapped mode
Image understanding
CLSSAI APICallable through CLSSAI API
Callable

The exact model ID is listed in the current CLSSAI public directory. This proves current route eligibility, not a completed inference for every mode.

Evidence
Exact ID z-ai/glm-5.3-flash is directory-listed
Vision test status

The request contract and frontend path are ready. A real OpenRouter output has not yet been recorded for this exact model and mode.

Request contract
Build verified
/chat/completions · Schema tests · 2026-08-05
Frontend path
Available
Vision form · 2026-08-05
Real model output
Not yet verified
No keyed OpenRouter smoke record

Providers

Compare published prices and endpoint capabilities. Select a provider for context limits, caching prices and supported parameters.

28 providers · OpenRouter reference data
$0.075$0.25$0.015----98.6%
$0.09$0.28$0.02----99.37%
$0.09$0.3$0.018----99.93%
$0.1$0.35$0.02------
$0.1$0.35$0.02----99.93%
$0.105$0.35$0.021----99.28%
$0.105$0.35$0.021----99.82%
$0.132$0.44$0.026----93.5%
$0.15$0.5$0.03----99.45%
$0.15$0.5$0.03----94.9%
$0.15$0.5$0.05----96.2%
$0.15$0.5$0.03----99.76%
$0.15$0.5$0.03----99.95%
$0.15$0.5$0.03----99.76%
$0.15$0.5$0.03----95.0%
$0.15$0.5$0.03----99.82%
$0.15$0.5$0.03----99.50%
$0.15$0.5$0.03----96.0%
$0.15$0.5$0.03----99.70%
$0.15$0.5$0.03----98.2%
$0.15$0.5$0.03----98.8%
$0.15$0.5$0.03----96.2%
$0.15$0.5$0.03----66.3%
$0.15$0.5$0.03----99.87%
$0.15$0.5$0.03----99.28%
$0.2$0.675$0.04----98.7%
$0.225$0.75$0.045----99.56%
$0.45$1.5$0.09----99.66%

USD per 1M tokens. Prices and metrics describe OpenRouter endpoints, not CLSSAI routing. “--” means no measurement was supplied. Uptime: last 30 minutes.

Pricing

Published base rates across providers, in USD per 1M tokens. Actual cost depends on the provider, caching and prompt length.

Input / 1M tokens$0.075 – $0.45
Output / 1M tokens$0.25 – $1.5
Cache read / 1M tokens$0.015 – $0.09

Effective Pricing: request-weighted costs and historical prices are not supplied by the public source.

Performance

Latency and throughput history are not supplied by the public endpoint API. Check OpenRouter for its latest performance charts.

View on OpenRouter

Uptime

Successful requests in the last 30 minutes, as reported by OpenRouter. A snapshot does not show historical reliability.

DeepInfra Fp498.6%
InferenceNet Fp499.37%
Relace99.93%
Wafer99.93%
StreamLake Fp899.28%
GMICloud Fp899.82%
Novita Fp893.5%
Inceptron Fp899.45%
Crusoe Fp494.9%
CoreWeave Nvfp496.2%
Sail Research Fp899.76%
AtlasCloud Fp899.95%
Fireworks99.76%
Phala Fp895.0%
Friendli99.82%
SiliconFlow Fp899.50%
DigitalOcean96.0%
Together99.70%
Parasail Fp898.2%
BaseTen Fp898.8%
Venice96.2%
Io Net Fp866.3%
Cloudflare99.87%
Z.AI Fp899.28%
NextBit Fp898.7%
Reka Fp899.56%
Modal Fp899.66%

Benchmarks

No scored evaluations are included in the public model metadata. View the source for benchmarks and their methodology.

View on OpenRouter

Apps

Application rankings require verified traffic data, which is not included in this source.

View on OpenRouter

Activity

Historical token and request volumes are not supplied by the public API. OpenRouter shows activity on its own network.

View on OpenRouter

Frequently asked questions

Provider availability, context limits and API capabilities.

Which providers serve Z.ai: GLM 5.3 Flash?

DeepInfra Fp4, InferenceNet Fp4, Relace, Morph Fp8, Wafer, StreamLake Fp8, GMICloud Fp8, Novita Fp8, Inceptron Fp8, Crusoe Fp4, CoreWeave Nvfp4, Sail Research Fp8, AtlasCloud Fp8, Fireworks, Phala Fp8, Friendli, SiliconFlow Fp8, DigitalOcean, Together, Parasail Fp8, BaseTen Fp8, Venice, Io Net Fp8, Cloudflare, Z.AI Fp8, NextBit Fp8, Reka Fp8, Modal Fp8 are listed in the OpenRouter reference data. Check CLSSAI availability above before making a request.

What context length is available for Z.ai: GLM 5.3 Flash?

The largest available context window is 1,310,720 tokens.

Which API parameters does Z.ai: GLM 5.3 Flash support?

The providers list 21 distinct supported parameters. Open an endpoint row to inspect them.

Explore

Continue comparing models or start integrating.