Llama 4 Scout has a larger context window than Gemini 2.0 Flash Thinking (10.0M tokens vs 500K tokens).
| Spec | Gemini 2.0 Flash Thinking | Llama 4 Scout |
|---|---|---|
| Input price | $0.25/M tokens | $0.25/M tokens |
| Output price | $0.5/M tokens | $0.5/M tokens |
| Context window | 500K tokens | 10.0M tokens |
| Parameters | Not disclosed | 17B active (109B total) |
| Weights | — | Open |
| Free API (OpenRouter) | No | No |
| Released | Dec 2024 | Apr 2025 |
Not enough votes to call it. On the specs, Gemini 2.0 Flash Thinking has the edge: bigger model tier.
Gemini 2.0 Flash Thinking wins Web Design and Image Generation.
No hosts listed on OpenRouter.
Per million tokens. Prices and uptime via OpenRouter, checked 9 Sep 2026.
Llama 4 Scout uses 7.8x more headings