Qwen Plus 0728 (thinking) is cheaper than Llama 3.1 70B (Instruct) at $0.4/M vs $0.59/M input tokens.
| Spec | Llama 3.1 70B (Instruct) | Qwen Plus 0728 (thinking) |
|---|---|---|
| Input price | $0.59/M tokens | $0.4/M tokens |
| Output price | $0.79/M tokens | $4/M tokens |
| Context window | 128K tokens | 1.0M tokens |
| Released | Jul 2024 | Sep 2025 |
Not enough votes to call it. On the specs, nothing separates them.
Llama 3.1 70B (Instruct) costs 5.1x less per token.
Qwen Plus 0728 (thinking) wins input (1.5× cheaper)·Llama 3.1 70B (Instruct) wins output (5.1× cheaper)
Qwen Plus 0728 (thinking) uses 34.0x more emoji
Qwen Plus 0728 (thinking) is cheaper than Llama 3.1 70B (Instruct) at $0.4/M vs $0.59/M input tokens.
| Spec | Llama 3.1 70B (Instruct) | Qwen Plus 0728 (thinking) |
|---|---|---|
| Input price | $0.59/M tokens | $0.4/M tokens |
| Output price | $0.79/M tokens | $4/M tokens |
| Context window | 128K tokens | 1.0M tokens |
| Released | Jul 2024 | Sep 2025 |
Not enough votes to call it. On the specs, nothing separates them.
Llama 3.1 70B (Instruct) costs 5.1x less per token.
Qwen Plus 0728 (thinking) wins input (1.5× cheaper)·Llama 3.1 70B (Instruct) wins output (5.1× cheaper)
Qwen Plus 0728 (thinking) uses 34.0x more emoji