Qwen3.8 2.4T A95B has a larger context window than Grok 3 Thinking (1.0M tokens vs 128K tokens).
| Spec | Grok 3 Thinking | Qwen3.8 2.4T A95B |
|---|---|---|
| Input price | — | $2/M tokens |
| Output price | — | $6/M tokens |
| Context window | 128K tokens | 1.0M tokens |
| Released | Feb 2025 | Aug 2026 |
Not enough votes to call it. On the specs, nothing separates them.
Grok 3 Thinking uses 5.5x more transitions
Qwen3.8 2.4T A95B has a larger context window than Grok 3 Thinking (1.0M tokens vs 128K tokens).
| Spec | Grok 3 Thinking | Qwen3.8 2.4T A95B |
|---|---|---|
| Input price | — | $2/M tokens |
| Output price | — | $6/M tokens |
| Context window | 128K tokens | 1.0M tokens |
| Released | Feb 2025 | Aug 2026 |
Not enough votes to call it. On the specs, nothing separates them.
Grok 3 Thinking uses 5.5x more transitions