Qwen3 30B A3B Thinking 2507 is cheaper than Llama 3.1 70B (Instruct) at $0.071/M vs $0.59/M input tokens.
| Spec | Llama 3.1 70B (Instruct) | Qwen3 30B A3B Thinking 2507 |
|---|---|---|
| Input price | $0.59/M tokens | $0.071/M tokens |
| Output price | $0.79/M tokens | $0.285/M tokens |
| Context window | 128K tokens | 262K tokens |
| Released | Jul 2024 | Aug 2025 |
Not enough votes to call it. On the specs, nothing separates them.
Qwen3 30B A3B Thinking 2507 is cheaper on both — 8.3× input, 2.8× output
Qwen3 30B A3B Thinking 2507 uses 67.9x more emoji
Qwen3 30B A3B Thinking 2507 is cheaper than Llama 3.1 70B (Instruct) at $0.071/M vs $0.59/M input tokens.
| Spec | Llama 3.1 70B (Instruct) | Qwen3 30B A3B Thinking 2507 |
|---|---|---|
| Input price | $0.59/M tokens | $0.071/M tokens |
| Output price | $0.79/M tokens | $0.285/M tokens |
| Context window | 128K tokens | 262K tokens |
| Released | Jul 2024 | Aug 2025 |
Not enough votes to call it. On the specs, nothing separates them.
Qwen3 30B A3B Thinking 2507 is cheaper on both — 8.3× input, 2.8× output
Qwen3 30B A3B Thinking 2507 uses 67.9x more emoji