Qwen Plus 0728 (thinking) is cheaper than Llama 3.1 405B at $0.4/M vs $2.7/M input tokens.
| Spec | Llama 3.1 405B | Qwen Plus 0728 (thinking) |
|---|---|---|
| Input price | $2.7/M tokens | $0.4/M tokens |
| Output price | $3.1/M tokens | $4/M tokens |
| Context window | 128K tokens | 1.0M tokens |
| Released | Jul 2024 | Sep 2025 |
Not enough votes to call it. On the specs, nothing separates them.
Qwen Plus 0728 (thinking) wins input (6.8× cheaper)·Llama 3.1 405B wins output (1.3× cheaper)
Qwen Plus 0728 (thinking) uses 702.2x more bold
Qwen Plus 0728 (thinking) is cheaper than Llama 3.1 405B at $0.4/M vs $2.7/M input tokens.
| Spec | Llama 3.1 405B | Qwen Plus 0728 (thinking) |
|---|---|---|
| Input price | $2.7/M tokens | $0.4/M tokens |
| Output price | $3.1/M tokens | $4/M tokens |
| Context window | 128K tokens | 1.0M tokens |
| Released | Jul 2024 | Sep 2025 |
Not enough votes to call it. On the specs, nothing separates them.
Qwen Plus 0728 (thinking) wins input (6.8× cheaper)·Llama 3.1 405B wins output (1.3× cheaper)
Qwen Plus 0728 (thinking) uses 702.2x more bold