GLM 5.3 Flash
GLM 5.3 Flash is GLM-5.3-Flash is a native multimodal model from Z.ai. $0.075 per million input tokens, $0.25 per million output tokens. 1,310,720 token context window, maximum output of 131,072 tokens. Higher uptime with 27 providers.…. It is ranked #555 on the Dev Radar, in Backend & APIs, first seen 7 days ago and shared in 2 posts (9.9k views).
GLM 5.3 Flash - API Pricing & Benchmarks | OpenRouter Benchmarks Chat Pricing Docs Sign Up Sign Up Z.ai: GLM 5.3 Flash z-ai / glm-5.3-flash Model weights Compare Playground Try this model GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while reducing compute overhead. Modalities In / Out Price 50% off $0.075 / $0.25 per 1M Context 1.3M Released Aug 26, 2026 Z.ai: GLM 5.3 Flash Compare Playground Try this model Providers Providers Different companies host the same model. OpenRouter…
What people said about GLM 5.3 Flash on X
Nice stats from @OpenRouter that illustrate how @togethercompute delivers solid and scaled performance for agentic workloads. We are serving GLM 5.3 and GLM 5.3 Flash @ top decile of TPS, latency, cache rate, and doing it at large volumes, 23% and 30% of all OpenRouter traffic, and OpenRouter is a fraction of our…
— @vipulved, 7 days ago · 61 likes · see the post
Great teamwork behind this!
— @zhyncs42, 7 days ago · 10 likes · see the post
Alternatives to GLM 5.3 Flash
- context.dev — powers the branding, company context, and search.
- Circle Docs — Circle's hosted x402 facilitator for USDC settlement on Arc, Base, and Polygon PoS
- Merge API Documentation — Browse the catalog
- OrcaRouter — API (official weight)
- Baseten — A speed-optimized GLM-5.3 Model API built for real-time workloads
- merge.dev — Schedule a product demo to learn how Merge can reduce your AI costs and optimize spend across your orginization.
GLM 5.3 Flash in numbers
- Rank on the Dev Radar: #555 of 1355
- Shared in 2 posts by 2 accounts: @vipulved, @zhyncs42
- 9.9k views on those posts
- First seen 7 days ago, last shared 7 days ago
- Pricing seen by Jev: paid
- Market: Backend & APIs
FAQ
What is GLM 5.3 Flash?
GLM-5.3-Flash is a native multimodal model from Z.ai. $0.075 per million input tokens, $0.25 per million output tokens. 1,310,720 token context window, maximum output of 131,072 tokens. Higher uptime with 27 providers.… It was first shared on X 7 days ago and is ranked #555 on the Dev Radar.
Is GLM 5.3 Flash free?
It is a paid product.
Who shared GLM 5.3 Flash?
2 accounts on X, including @vipulved, @zhyncs42, in 2 posts totalling 9.9k views.
Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 12.2k posts from 4.7k X accounts over the last 21 days, 1.4k tools, 12 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-19 18:39 UTC. Full method.