Llama 4 Scout is cheaper than GPT-3.5 Turbo at $0.25/M vs $1.5/M input tokens.
| Spec | GPT-3.5 Turbo | Llama 4 Scout |
|---|---|---|
| Input price | $1.5/M tokens | $0.25/M tokens |
| Output price | $2/M tokens | $0.5/M tokens |
| Context window | 16K tokens | 10.0M tokens |
| Parameters | Not disclosed | 17B active (109B total) |
| Weights | Closed | Open |
| Free API (OpenRouter) | No | No |
| Released | Mar 2022 | Apr 2025 |
Not enough votes to call it. On the specs, Llama 4 Scout has the edge: newer, bigger context window.
Llama 4 Scout costs 4.0x less per token.
Llama 4 Scout is cheaper on both: 6.0× input, 4.0× output.
Per million tokens. Prices and uptime via OpenRouter, checked 9 Sep 2026.
Llama 4 Scout uses 26.1x more headings