GLM-4.5-Flash
Legacy / SupersededThis is a legacy version with newer models in the same product line. Listed rates reflect available serving channels.
Legacy / superseded version. Newer models are available; official or partner channels may still serve this version.
Replacement reference: z-ai-glm-5.3-flashPricing
How much are you overpaying?
Enter your monthly spend on this model to see what the same workload could cost on cheaper alternatives.
Latest changes
Benchmarks
No benchmark data available.
⚡GLM-4.5-Flash Comparisons & Scenario Guides
❓GLM-4.5-Flash Frequently Asked Questions
How much does GLM-4.5-Flash cost per 1M tokens?▼
GLM-4.5-Flash is priced at $0.0000 per 1M input tokens and $0.0000 per 1M output tokens, with cached input at $0.0000/1M on official. Check provider terms before deployment.
What is the context window for GLM-4.5-Flash?▼
GLM-4.5-Flash supports up to 131,072 tokens context with a max completion limit of 98,304 tokens.
What are cost-saving alternatives to GLM-4.5-Flash?▼
ModelPriceLab automatically calculates cheaper alternatives in the same model class, showing monthly spend differences and migration guides.
Does this change your best option?
Check your models and usage for free. See the relevant official changes, estimated cost, and a practical next step.