GPT OSS 120B is cheaper than Llama 3.1 70B (Instruct) at $0.18/M vs $0.59/M input tokens.
| Spec | GPT OSS 120B | Llama 3.1 70B (Instruct) |
|---|---|---|
| Input price | $0.18/M tokens | $0.59/M tokens |
| Output price | $0.8/M tokens | $0.79/M tokens |
| Context window | 131K tokens | 128K tokens |
| Parameters | 117B (5.1B active) | 70B |
| Weights | Open | Open |
| Free API (OpenRouter) | No | No |
| Released | Aug 2025 | Jul 2024 |
| MMLU | Matches/exceeds o4-mini | 86.0% |
Not enough votes to call it. On the specs, GPT OSS 120B has the edge: bigger model tier, newer.
GPT OSS 120B wins input (3.3× cheaper)·Llama 3.1 70B (Instruct)
Per million tokens. Prices and uptime via OpenRouter, checked 8 Sep 2026.
GPT OSS 120B uses 15.4x more emoji