Mistral Devstral Small 1.1 is cheaper than Qwen3 Max Thinking at $0.1/M vs $1.2/M input tokens.
| Spec | Mistral Devstral Small 1.1 | Qwen3 Max Thinking |
|---|---|---|
| Input price | $0.1/M tokens | $1.2/M tokens |
| Output price | $0.3/M tokens | $6/M tokens |
| Context window | — | 262K tokens |
| Released | Jul 2025 | Feb 2026 |
| SWE-bench Verified | 46.8% | 75.3% |
Not enough votes to call it. On the specs, Qwen3 Max Thinking has the edge: bigger model tier, newer, bigger context window.
Mistral Devstral Small 1.1 costs 20x less per token.
Mistral Devstral Small 1.1 is cheaper on both — 12× input, 20× output
Qwen3 Max Thinking uses 137.1x more emoji
Mistral Devstral Small 1.1 is cheaper than Qwen3 Max Thinking at $0.1/M vs $1.2/M input tokens.
| Spec | Mistral Devstral Small 1.1 | Qwen3 Max Thinking |
|---|---|---|
| Input price | $0.1/M tokens | $1.2/M tokens |
| Output price | $0.3/M tokens | $6/M tokens |
| Context window | — | 262K tokens |
| Released | Jul 2025 | Feb 2026 |
| SWE-bench Verified | 46.8% | 75.3% |
Not enough votes to call it. On the specs, Qwen3 Max Thinking has the edge: bigger model tier, newer, bigger context window.
Mistral Devstral Small 1.1 costs 20x less per token.
Mistral Devstral Small 1.1 is cheaper on both — 12× input, 20× output
Qwen3 Max Thinking uses 137.1x more emoji