GLM 5.3 Flash input price cut 70%, output price rises
Z.ai has repriced GLM 5.3 Flash on OpenRouter. Input drops from $0.15 to $0.045 per 1M tokens, while output climbs from $0.5 to $0.6 per 1M tokens. The model is now much cheaper to feed and slightly more expensive to generate from.
What it means for a SaaS brand
Long-context work gets cheaper on GLM 5.3 Flash, so pipelines that push large prompts through it, such as bulk page audits or answer-surface monitoring across your category, cost less to run at scale. If your workload is generation heavy, recheck the per-run cost before expanding volume.
Source
Every entry here names where it came from, and we publish nothing we have not read at the source.
Read the originalChanges like this decide who gets cited
We measure how often AI answers name your brand, track what moves it, and tell you what to do next.