Llama 3.1 70B (Instruct) is cheaper than GPT-5 at $0.59/M vs $1.25/M input tokens.
| Spec | GPT-5 | Llama 3.1 70B (Instruct) |
|---|---|---|
| Input price | $1.25/M tokens | $0.59/M tokens |
| Output price | $10/M tokens | $0.79/M tokens |
| Context window | 400K tokens | 128K tokens |
| Parameters | Not disclosed | 70B |
| Weights | Closed | Open |
| Free API (OpenRouter) | No | No |
| Released | Aug 2025 | Jul 2024 |
Not enough votes to call it. On the specs, GPT-5 has the edge: bigger model tier, newer, bigger context window.
Llama 3.1 70B (Instruct) costs 13x less per token.
Llama 3.1 70B (Instruct) is cheaper on both: 2.1× input, 13× output.
Per million tokens. Prices and uptime via OpenRouter, checked 8 Sep 2026.
GPT-5 uses 16.2x more headings