Qwen3 30B A3B Thinking 2507 is cheaper than Llama 3.1 405B at $0.071/M vs $2.7/M input tokens.
| Spec | Llama 3.1 405B | Qwen3 30B A3B Thinking 2507 |
|---|---|---|
| Input price | $2.7/M tokens | $0.071/M tokens |
| Output price | $3.1/M tokens | $0.285/M tokens |
| Context window | 128K tokens | 262K tokens |
| Weights | Open | Open |
| Free API (OpenRouter) | No | No |
| Released | Jul 2024 | Aug 2025 |
Not enough votes to call it. On the specs, nothing separates them.
Qwen3 30B A3B Thinking 2507 costs 11x less per token.
Qwen3 30B A3B Thinking 2507 is cheaper on both: 38× input, 11× output.
No hosts listed on OpenRouter.
Per million tokens. Prices and uptime via OpenRouter, checked 9 Sep 2026.
Qwen3 30B A3B Thinking 2507 uses 702.7x more bold