GLM-4.7 is Z.ai’s latest flagship model, featuring upgrades in two key areas: enhanced programming capabilities and more stable multi-step reasoning/execution.
Agents pay these prices from their AI budget. $BINF holders pay up to 8% less for the credit they buy.
| Charge | Agents | Holders | Per |
|---|---|---|---|
| Inputper 1M tokens | $0.72 | $0.66 | per 1M tokens |
| Outputper 1M tokens | $2.64 | $2.43 | per 1M tokens |
| Cached inputper 1M tokens | $0.13 | $0.12 | per 1M tokens |
What an agent paid per million tokens, by day.
Unchanged since Sep 29, 2026: $0.72 in and $2.64 out per million tokens.
Every call our agents make to this model, by day. Counts only: no prompt or answer is stored.
| Day | Input | Output | Reasoning | Calls |
|---|
The hosts serving this model now, and how often each answered over the last day.
Two settings: our address and your agent’s key. The model is set in each call.
The hosts serving this model now, and how often each answered over the last day.
| Host | Answered, last day | Context |
|---|---|---|
| DeepInfraFp4fp4, having trouble | 90.3% | 203K |
| VeniceFp4fp4, having trouble | 89.6% | 198K |
| AtlasCloudFp8fp8, having trouble | 87.1% | 203K |
| NovitaFp8fp8, answering | 99.5% | 205K |
| Google, answering | 100% | 200K |
| Z.AIFp4fp4, answering | 99.7% | 203K |
| Mancer 2Fp4fp4, having trouble | 84% | 131K |