GLM 4.7 Flash is cheaper than Llama 3.1 405B at $0.07/M vs $2.7/M input tokens.
| Spec | GLM 4.7 Flash | Llama 3.1 405B |
|---|---|---|
| Input price | $0.07/M tokens | $2.7/M tokens |
| Output price | $0.4/M tokens | $3.1/M tokens |
| Context window | 200K tokens | 128K tokens |
| Weights | Open | Open |
| Free API (OpenRouter) | No | No |
| Released | Jan 2026 | Jul 2024 |
Not enough votes to call it. On the specs, nothing separates them.
GLM 4.7 Flash costs 7.8x less per token.
Reviewing agent-written code?See a Brief PR report
GLM 4.7 Flash is cheaper on both: 39× input, 7.8× output.
No hosts listed on OpenRouter.
Per million tokens. Prices and uptime via OpenRouter, checked 11 Sep 2026.
GLM 4.7 Flash uses 479.8x more bold