Qwen3 Max is cheaper than Llama 3.1 405B at $1.2/M vs $2.7/M input tokens.
| Spec | Llama 3.1 405B | Qwen3 Max |
|---|---|---|
| Input price | $2.7/M tokens | $1.2/M tokens |
| Output price | $3.1/M tokens | $6/M tokens |
| Context window | 128K tokens | 256K tokens |
| Released | Jul 2024 | Sep 2025 |
Not enough votes to call it. On the specs, nothing separates them.
Qwen3 Max wins input (2.3× cheaper)·Llama 3.1 405B wins output (1.9× cheaper)
Qwen3 Max uses 572.2x more bold
Qwen3 Max is cheaper than Llama 3.1 405B at $1.2/M vs $2.7/M input tokens.
| Spec | Llama 3.1 405B | Qwen3 Max |
|---|---|---|
| Input price | $2.7/M tokens | $1.2/M tokens |
| Output price | $3.1/M tokens | $6/M tokens |
| Context window | 128K tokens | 256K tokens |
| Released | Jul 2024 | Sep 2025 |
Not enough votes to call it. On the specs, nothing separates them.
Qwen3 Max wins input (2.3× cheaper)·Llama 3.1 405B wins output (1.9× cheaper)
Qwen3 Max uses 572.2x more bold