GLM 4 32B is cheaper than Gemini 2.0 Flash Thinking at $0.1/M vs $0.25/M input tokens.
| Spec | Gemini 2.0 Flash Thinking | GLM 4 32B |
|---|---|---|
| Input price | $0.25/M tokens | $0.1/M tokens |
| Output price | $0.5/M tokens | $0.1/M tokens |
| Context window | 500K tokens | 128K tokens |
| Released | Dec 2024 | Jul 2025 |
Not enough votes to call it. On the specs, nothing separates them.
GLM 4 32B costs 5.0x less per token.
GLM 4 32B is cheaper on both — 2.5× input, 5.0× output
GLM 4 32B uses 10.4x more headings
GLM 4 32B is cheaper than Gemini 2.0 Flash Thinking at $0.1/M vs $0.25/M input tokens.
| Spec | Gemini 2.0 Flash Thinking | GLM 4 32B |
|---|---|---|
| Input price | $0.25/M tokens | $0.1/M tokens |
| Output price | $0.5/M tokens | $0.1/M tokens |
| Context window | 500K tokens | 128K tokens |
| Released | Dec 2024 | Jul 2025 |
Not enough votes to call it. On the specs, nothing separates them.
GLM 4 32B costs 5.0x less per token.
GLM 4 32B is cheaper on both — 2.5× input, 5.0× output
GLM 4 32B uses 10.4x more headings