Skip to content
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Default Index
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed
  1. Home
  2. /
  3. Compare
  4. /
  5. GPT OSS 120BvsGrok 4.7
Updated Sep 2026

GPT OSS 120BvsGrok 4.7

GPT OSS 120B is cheaper than Grok 4.7 at $0.18/M vs $1.6/M input tokens.

GPT OSS 120B and Grok 4.7 compared across 10 shared prompts
SpecGPT OSS 120BGrok 4.7
Input price$0.18/M tokens$1.6/M tokens
Output price$0.8/M tokens$4.8/M tokens
Context window131K tokens500K tokens
WeightsOpenClosed
Free API (OpenRouter)NoNo
ReleasedAug 2025Sep 2026
Side by side

Same prompt, both answers

Opening lines from 5 of the 10 prompts both answered. Formatting stripped, words untouched.

Simple 7-Ingredient Recipe

Create a simple recipe for a delicious dish (sweet or savory) using only 7 common pantry ingredients. Include a creative name for the dish…

GPT OSS 120B373 words
Recipe: “Golden Crunchy Chick‑a‑Boom Bites” A quick, crunchy‑savory snack that’s ready in under 20 minutes – perfect for a snack, side, or salad topper. Ingredients (7 pantry staples) # Ingredient Quantity 1 Canned chickpeas (drained & rinsed) 1 ½ cups (≈1 ½ cans) 2 Olive oil (or any neutral oil) 2 Tbsp 3 Garlic powder…
Read the full answer
Grok 4.7129 words
Pantry Gold Honey Biscuits A sweet, buttery, slightly salty biscuit that tastes like it took more effort than it did. Ingredients (7) 2 cups all-purpose flour 1 tbsp baking powder ½ tsp salt ¼ cup sugar ½ cup butter (cold) ¾ cup milk 2 tbsp honey Time: about 10 minutes prep, 12–15 minutes bake Instructions Heat the…
Read the full answer

Debug This Architecture

Here's the architecture for a real-time collaborative document editor (like Google Docs): **System Overview:** - Frontend: React SPA with…

GPT OSS 120B2,510 words
Below is a system‑level failure‑mode / race‑condition / scaling‑bottleneck map for the architecture you described, followed by a concrete mitigation (or redesign) for each item together with the main trade‑offs to consider.
Read the full answer
Grok 4.71,399 words
This design will lose edits, split users into inconsistent views, and fall over under concurrent writing. The core problems are clock-based conflict resolution, per-server WebSocket fan-out, polling as the sync path, and treating documents as full HTML snapshots.
Read the full answer

Beginner's Longevity Plan

Outline a simple, actionable 1-month plan for someone starting their journey towards better health and longevity. Focus on 3 key areas…

GPT OSS 120B823 words
1‑Month “Kick‑Start” Plan for Health & Longevity Focus areas: Diet, Exercise, Sleep (stress‑management tips woven in) How to Use This Plan What When How Read the weekly overview first. Day 1 of each week Set a tiny, measurable goal (e.g., “walk 10 min after dinner”).
Read the full answer
Grok 4.7442 words
1-month beginner plan for better health and longevity Focus on three foundations: nutrition, movement, and sleep. Aim for consistency over perfection. If you have a medical condition, check with a clinician before big changes. Week-by-week rhythm Week 1: Build awareness and one easy habit in each area.
Read the full answer

Startup Pitch Teardown

Here's a pitch deck summary for a startup called "MindMeld AI": **Slide 1 - Vision:** "We're building the future of human-AI collaboration.…

GPT OSS 120B1,405 words
Below is a quick “triage” of the three statements that, if left unchanged, will most likely raise eyebrows (or a “no‑thanks”) from investors, VCs, and even a skeptical engineering board.
Read the full answer
Grok 4.7599 words
Three weakest claims 1. “Predict what you want to type before you think it” (Slide 1) This is scientifically incoherent, not just ambitious. Non-invasive EEG decodes neural activity that is already underway (motor imagery, attempted speech, attention).
Read the full answer

Explain Like I'm a Specific Expert

Explain how a large language model (like GPT or Claude) learns and generates text to each of these three audiences: 1. **An experienced…

GPT OSS 120B821 words
1. The Software Engineer (API‑first, Distributed‑Systems Mindset) Think of a large language model (LLM) as a stateless microservice that receives a stream of tokens (the smallest lexical units) and returns a probability distribution over the next token.
Read the full answer
Grok 4.71,173 words
1. Experienced software engineer An LLM is a stateless function from a token sequence to a probability distribution over the next token. Training is a giant batch job: you stream text, run a forward pass, compare the predicted distribution to the actual next token with cross-entropy, and push gradients through billions…
Read the full answer
Our Verdict
GPT OSS 120B
GPT OSS 120B
Grok 4.7
Grok 4.7

Not enough votes to call it. On the specs, nothing separates them.

GPT OSS 120B costs 6.0x less per token.

Too close to call
API pricing

Cost per 1M tokens

GPT OSS 120B
Input
$0.18
8.9× cheaper
Output
$0.80
6.0× cheaper
Grok 4.7
Input
$1.60
Output
$4.80

GPT OSS 120B is cheaper on both: 8.9× input, 6.0× output.

Where to run it

21 hosts, cheapest first

GPT OSS 120B20 hosts
HostInOutContextUptime
AAkashMLbf16$0.03 in·$0.17 out·131k·100% upCCoreWeavefp4$0.03 in·$0.17 out·131k·99.8% upDDekaLLMbf16$0.03 in·$0.18 out·131k·99.8% upDDeepInfrabf16$0.04 in·$0.17 out·131k·98.7% upCCrusoebf16$0.05 in·$0.25 out·131k·98.8% upMMancerfp8$0.05 in·$0.30 out·131k·97.7% up
14 more hostsFewer hosts
DDigitalOcean$0.06 in·$0.42 out·128k·100% upGoogle Vertex AI$0.09 in·$0.36 out·131k·93.2% upBBasetenfp4$0.10 in·$0.50 out·128k·100% upPParasailfp4$0.10 in·$0.75 out·131k·97.9% upAmazon Bedrock$0.15 in·$0.60 out·131k·87.7% upGroq$0.15 in·$0.60 out·131k·100% upSSiliconFlowfp8$0.15 in·$0.60 out·131k·53% upTTogether$0.15 in·$0.60 out·131k·92% upCCerebrasfp16$0.35 in·$0.75 out·131k·100% upNNovitafp4degraded$0.05 in·$0.25 out·131k·61.2% upSSambaNovadegraded$0.14 in·$0.95 out·131k·97.1% upNNebiusfp4degraded$0.15 in·$0.60 out·131k·97.9% upPPhaladegraded$0.15 in·$0.60 out·131k·63.6% upMMaradegraded$0.15 in·$0.75 out·131k·83.8% up
Grok 4.71 host
HostInOutContextUptime
xAI$1.60 in·$4.80 out·500k·98.9% up

Per million tokens. Prices and uptime via OpenRouter, checked 23 Sep 2026.

Research

What we learned reading every model

FAQ

Common questions

GPT OSS 120B is developed by OpenAI while Grok 4.7 is developed by xAI. GPT OSS 120B has a 131K token context window vs Grok 4.7's 500K. You can compare their actual outputs across 10 challenges on Rival to see how they differ in practice.

It depends on your use case. GPT OSS 120B and Grok 4.7 each have strengths in different areas. Rival lets you compare their real outputs side-by-side across 10 challenges so you can judge which fits your needs best.

GPT OSS 120B costs $0.18/M input tokens and Grok 4.7 costs $1.6/M input tokens. GPT OSS 120B is $1.42/M cheaper per input. Check their side-by-side outputs on Rival to see if the price difference is justified by quality.

This page shows a side-by-side comparison of GPT OSS 120B and Grok 4.7 across shared challenges. You can vote on which model produced the better output in a blind duel. Browsing and voting are free. No account is needed to look; signing in only saves your votes and likes.

Keep exploring

More comparisons

Against the newest arrivals

GPT OSS 120B logoDeepSeek V4 Flash Vision Exp logo
GPT OSS 120B vs DeepSeek V4 Flash Vision ExpLanded Sep 2026
Grok 4.7 logoSolar Pro 4 logo
Grok 4.7 vs Solar Pro 4Landed Sep 2026
GPT OSS 120B logoHy3 logo
GPT OSS 120B vs Hy3Landed Sep 2026
Grok 4.7 logoQwen3.7 Flash logo
Grok 4.7 vs Qwen3.7 FlashLanded Sep 2026
GPT OSS 120B logoLing 3.0 Flash logo
GPT OSS 120B vs Ling 3.0 FlashLanded Sep 2026
Grok 4.7 logoMuse Glimmer 30B logo
Grok 4.7 vs Muse Glimmer 30BLanded Sep 2026
GPT OSS 120B logoGLM 5.3 logo
GPT OSS 120B vs GLM 5.3Landed Sep 2026
Grok 4.7 logoTernary Bonsai 2 27B logo
Grok 4.7 vs Ternary Bonsai 2 27BLanded Sep 2026

Same lab, same size, long tail

GPT OSS 120B logoGPT OSS 20B logo
GPT OSS 120B vs GPT OSS 20BVersion compare
GPT OSS 120B logoGPT-6 Astra Pro logo
GPT OSS 120B vs GPT-6 Astra ProSame lab
Grok 4.7 logoGrok 4.6 logo
Grok 4.7 vs Grok 4.6Same lab
Grok 4.7 logoGrok 4.5 logo
Grok 4.7 vs Grok 4.5Same lab
Grok 4.7 logoMercury logo
Grok 4.7 vs MercuryNew provider
Grok 4.7 logoMercury 2 logo
Grok 4.7 vs Mercury 2New provider
Grok 4.7 logoMercury 2.5 logo
Grok 4.7 vs Mercury 2.5Same size
Grok 4.7 logoMercury 2.5 Preview logo
Grok 4.7 vs Mercury 2.5 PreviewNew provider

Model pages

GPT OSS 120B logo
GPT OSS 120B54 outputs, specs and price
Grok 4.7 logo
Grok 4.710 outputs, specs and price
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Default Index
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed