Qwen3 Max Thinking is cheaper than Llama 3.1 405B at $1.2/M vs $2.7/M input tokens.
| Spec | Llama 3.1 405B | Qwen3 Max Thinking |
|---|---|---|
| Input price | $2.7/M tokens | $1.2/M tokens |
| Output price | $3.1/M tokens | $6/M tokens |
| Context window | 128K tokens | 262K tokens |
| Released | Jul 2024 | Feb 2026 |
Not enough votes to call it. On the specs, nothing separates them.
Qwen3 Max Thinking wins input (2.3× cheaper)·Llama 3.1 405B wins output (1.9× cheaper)
Qwen3 Max Thinking uses 517.4x more bold
Qwen3 Max Thinking is cheaper than Llama 3.1 405B at $1.2/M vs $2.7/M input tokens.
| Spec | Llama 3.1 405B | Qwen3 Max Thinking |
|---|---|---|
| Input price | $2.7/M tokens | $1.2/M tokens |
| Output price | $3.1/M tokens | $6/M tokens |
| Context window | 128K tokens | 262K tokens |
| Released | Jul 2024 | Feb 2026 |
Not enough votes to call it. On the specs, nothing separates them.
Qwen3 Max Thinking wins input (2.3× cheaper)·Llama 3.1 405B wins output (1.9× cheaper)
Qwen3 Max Thinking uses 517.4x more bold