Llama 3.1 405B is cheaper than Qwen3.8 Max Prime at $2.7/M vs $4/M input tokens.
| Spec | Llama 3.1 405B | Qwen3.8 Max Prime |
|---|---|---|
| Input price | $2.7/M tokens | $4/M tokens |
| Output price | $3.1/M tokens | $12/M tokens |
| Context window | 128K tokens | 1.0M tokens |
| Weights | Open | Closed |
| Free API (OpenRouter) | No | No |
| Released | Jul 2024 | Sep 2026 |
Not enough votes to call it. On the specs, nothing separates them.
Llama 3.1 405B costs 3.9x less per token.
Llama 3.1 405B is cheaper on both: 1.5× input, 3.9× output.
No hosts listed on OpenRouter.
Per million tokens. Prices and uptime via OpenRouter, checked 25 Sep 2026.