Model
GLM-5.3-Flash
Native multimodal GLM model for efficient coding and long-horizon agent tasks
Compare providers (sorted by listed input price)
45 of 53 providers publish a price for GLM-5.3-Flash. The cheapest full-context option is OpenRouter at $0.15/$0.5 per M. Merge Gateway lists lower but caps context at 1M.
Best price history
Daily lowest published price across live providers. A lower line is better for the buyer.
- Best input /M
- $0.015▼ 80.0%
- Best output /M
- $0.05▼ 80.0%
- Provider coverage
- 61
- 6 at the start · 19 snapshots
Price badge
https://www.vaanalytics.in/badge/zhipuai/glm-5.3-flash.svgLive SVG, regenerated on every hourly sync — paste it into a README to always show the current cheapest listed price. Updates when prices change; no build step needed on your side.
Recent changes
Changelog →- repricedvia cortecs · Sep 13, 2026
- input: $0.201 → $0.1
- output: $0.5 → $0.35
- cache_read: $0.05 → $0.018
- new modelvia nvidia · Sep 13, 2026
- new modelvia volcengine-coding-plan · Sep 13, 2026
- repricedvia openrouter · Sep 11, 2026
- input: $0.07 → $0.15
- output: $0.233 → $0.5
- cache_read: $0.014 → $0.03
- new modelvia llmgateway-providers · Sep 11, 2026
- repricedvia ofox · Sep 11, 2026
- input: $0.075 → $0.15
- output: $0.25 → $0.5
- cache_read: $0.015 → $0.03
Weights
Standard open-source terms. Commercial use without additional conditions. Weights download without an access request.
licence and gating from the Hugging Face model card ↗ — the card is authoritative, this is a summary
Prices are per million tokens (USD). “—” means the provider does not publicly list a price for this model. Data from models.dev; verify with the provider before purchasing.