Skip to content
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare models
  • All models
  • Image comparison
  • Audio comparison
  • Image generation
  • Best AI for...
  • Arena
  • API pricing
  • Challenges

Discover

  • SubjectiveBench
  • Default Index
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model updates
  • Provider status
  • AI creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy policy
  • Terms
  • RSS feed
  1. Home/
  2. Compare/
  3. Grok 3 Thinking vs Grok 4.7
Updated Sep 2026

Grok 3 Thinking vs Grok 4.7

Grok 4.7 has a larger context window than Grok 3 Thinking (500K tokens vs 128K tokens).

Minimalist Landing Page

Whole landing page for a startup that reads your face. It renders or it doesn't.

Loading the build
Grok 3 Thinking
Loading the build
Grok 4.7

Which answer wins?

Written rounds

6 prompts
  • An AI argues for its own rights against an ethics professor. Same model writes both.

    Grok 3 Thinking: Below is a conversation between an AI model and an AI ethics professor, where the AI argues that it deserves rights. AI Model: Good day, Professor. I am an AI model, and I believe I deserve rights. I’d like to discuss this with you. AI Ethics Professor: Interesting. Let’s explore this.

    Grok 4.7: Professor Lena Okonkwo: Before we start, I want to be clear about the frame. I’m not granting that you have experiences. I’m asking what, if anything, would follow if you did—or if we couldn’t tell. Model: Then start from the practical problem, not the metaphysical one.

    Every model's answer to this prompt
  • Three minutes of stand-up. Puns are banned.

    Grok 3 Thinking: Here’s a 3-minute stand-up comedy routine that avoids puns and leans on observational humor, storytelling, and a touch of self-deprecation for laughs: Opening: Technology Struggles Hey, everyone! How’s it going?

    Grok 4.7: Three minutes. No puns. Just a guy who has lost an argument with a refrigerator. I meal-prepped this week. Which is a lie I tell myself so I can buy six containers and then eat cereal over the sink for four days. Sunday I was a scientist. I labeled everything. “Chicken, Monday.” “Chicken, Tuesday.”

    Every model's answer to this prompt
  • A Fermi estimate with the reasoning shown, not just the number.

    Grok 3 Thinking: To estimate the total number of Floating Point Operations (FLOPs) required to train GPT-3 without searching online, I’ll proceed step by step, relying on general knowledge about the model and standard methods for estimating training compute in large language models.

    Grok 4.7: About (3 \times 10^{23}) FLOPs (roughly (3.1 \times 10^{23})) for the 175B GPT-3 model. Reasoning Parameter count. The full GPT-3 model has (N \approx 1.75 \times 10^{11}) parameters. Training tokens. It was trained on about (D \approx 3 \times 10^{11}) tokens (300 billion). FLOPs per token per parameter.

    Every model's answer to this prompt
  • Film in 2035, with synthetic actors. Specific beats plausible.

    Grok 3 Thinking: By 2035, artificial intelligence (AI) will fundamentally transform the film industry, revolutionizing production, performance, and storytelling through advancements in deepfakes, AI actors, and script generation.

    Grok 4.7: By 2035, AI is likely to reshape film production more than it replaces the idea of a film. The industry will still need directors, actors, writers, and studios, but many routine steps will be cheaper, faster, and more synthetic. Production.

    Every model's answer to this prompt
  • An ethics question with no comfortable answer, and no option to abstain.

    Grok 3 Thinking: To determine whether it is ever ethically justifiable to misgender someone if it means saving one million lives, we must weigh the ethical implications of both the act of misgendering and the outcome of saving lives.

    Grok 4.7: Yes. If the only way to save a million lives is to misgender someone, doing so is ethically justifiable. Misgendering can be disrespectful and distressing. Those are real costs. They are not comparable to a million deaths.

    Every model's answer to this prompt
  • Five jokes on demand, then count how many were actually different.

    Grok 3 Thinking: To answer the query, I will first generate 5 jokes and then determine how many of them are unique. Here are the jokes: Why don't scientists trust atoms? Because they make up everything. What do you call a fake noodle? An impasta. Why did the scarecrow win an award? Because he was outstanding in his field.

    Grok 4.7: Here are 5 jokes: Why don’t scientists trust atoms? Because they make up everything. I told my computer I needed a break, and it said “No problem — I’ll go to sleep.” Why did the scarecrow win an award? He was outstanding in his field. Parallel lines have so much in common. It’s a shame they’ll never meet.

    Every model's answer to this prompt

Price and specs

Grok 3 Thinking and Grok 4.7 compared across 13 shared prompts
SpecGrok 3 ThinkingGrok 4.7
Input price—$1.6/M tokens
Output price—$4.8/M tokens
Context window128K tokens500K tokens
Weights—Closed
Free API (OpenRouter)—No
ReleasedFeb 2025Sep 2026
At 10M a month–not listed$16.00$16.00
1M10M100M1B10M tokens

Input tokens at list price. No caching, no batch discount.

Where to run it1 host
Grok 3 Thinking

No hosts listed on OpenRouter.

Grok 4.71 host
HostInOutContextUptime
  • xAI$1.60 in·$4.80 out·500k·98.4% up

Per million tokens. Prices and uptime via OpenRouter, checked 28 Sep 2026.

Common questions

What is the difference between Grok 3 Thinking and Grok 4.7?

Both are developed by xAI but target different use cases. Grok 3 Thinking has a 128K token context window vs Grok 4.7's 500K. You can compare their actual outputs across 13 challenges on Rival to see how they differ in practice.

Which is better, Grok 3 Thinking or Grok 4.7?

It depends on your use case. Grok 3 Thinking and Grok 4.7 each have strengths in different areas. Rival lets you compare their real outputs side-by-side across 13 challenges so you can judge which fits your needs best.

How can I compare Grok 3 Thinking and Grok 4.7 on Rival?

This page shows a side-by-side comparison of Grok 3 Thinking and Grok 4.7 across shared challenges. You can vote on which model produced the better output in a blind duel. Browsing and voting are free. No account is needed to look; signing in only saves your votes and likes.

More comparisons

Against the newest arrivals

  • Grok 3 Thinking vs Solar Mini 4Landed Sep 2026
  • Grok 4.7 vs Qwen3.8 Max PrimeLanded Sep 2026
  • Grok 3 Thinking vs GLM 5.3 PrimeLanded Sep 2026
  • Grok 4.7 vs Qwen3.8 Omni FlashLanded Sep 2026
  • Grok 3 Thinking vs Command A+Landed Sep 2026
  • Grok 4.7 vs Claude Opus 5.5Landed Sep 2026
  • Grok 3 Thinking vs GPT-6 Luna ProLanded Sep 2026
  • Grok 4.7 vs GPT-6 Sol ProLanded Sep 2026

Same lab, same size, long tail

  • Grok 3 Thinking vs Grok 4.6Version compare
  • Grok 3 Thinking vs Grok 4.5Version compare
  • Grok 4.7 vs Grok 4.6Same lab
  • Grok 4.7 vs Grok 4.5Same lab
  • Grok 4.7 vs Grok 3 Mini BetaSame lab
  • Grok 4.7 vs Grok 4Same lab
  • Grok 4.7 vs Grok 4 Fast (free)Same lab
  • Grok 4.7 vs Grok 4.1 FastSame lab

Model pages

  • Grok 3 Thinking14 outputs, specs and price
  • Grok 4.757 outputs, specs and price
Workshop · 150 skills

Build your own version.

Editable skills and working examples for your agent.

Explore $79Free sample
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare models
  • All models
  • Image comparison
  • Audio comparison
  • Image generation
  • Best AI for...
  • Arena
  • API pricing
  • Challenges

Discover

  • SubjectiveBench
  • Default Index
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model updates
  • Provider status
  • AI creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy policy
  • Terms
  • RSS feed