Z.ai: GLM 4.5 Flash
Efficient GLM model for fast reasoning, coding, and agent workflows
Cheapest input
N/A
$/MTok · N/A
Cheapest output
N/A
$/MTok
Context
200K
tokens
Sources
2
0 listings
Sources
Same model, different sources. Our first-party rate is listed first, then router and aggregator rows stay separate so you can see the spread instead of a single blended number. Free/$0 listings are hidden by default (5 omitted). Show free/$0 listings.
No priced sources in the current snapshot.
Capabilities
- Reasoning
- Tool calling
- Structured output
- Open weights
- In: text
- Out: text
Knowledge cutoff: 2025-04Released: 2025-07-28Family: glm-flash
Run this model in a real workspace
Launch agents from the project they work in, use remote management to reach another machine, and see this model's sessions and token usage in your own history. Free to use, runs on your device.
Data and list-price estimates may contain mistakes. Not investment advice.