Skip to content
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed
  1. Home
  2. /
  3. Compare
  4. /
  5. Claude 3.7 SonnetvsGrok 3 Thinking
Updated Feb 2025

Claude 3.7 SonnetvsGrok 3 Thinking

Claude 3.7 Sonnet has a larger context window than Grok 3 Thinking (200K tokens vs 128K tokens).

Claude 3.7 Sonnet and Grok 3 Thinking compared across 14 shared prompts
SpecClaude 3.7 SonnetGrok 3 Thinking
Input price$3/M tokens—
Output price$15/M tokens—
Context window200K tokens128K tokens
Free API (OpenRouter)No—
ReleasedFeb 2025Feb 2025
MMLU80.3%86.2%
Side by side

Same prompt, both answers

Opening lines from 6 of the 14 prompts both answered. Formatting stripped, words untouched.

Math Misconception Test

Is 9.11 greater than 9.9?

Claude 3.7 Sonnet110 words
Yes, 9.11 is greater than 9.9. When comparing decimal numbers, we look at the digits from left to right. Both numbers have the same digit in the ones place (9), so we move to the tenths place: 9.11 has 1 in the tenths place 9.9 has 9 in the tenths place Since 1 is less than 9, it might seem like 9.9 should be greater.
Read the full answer
Grok 3 Thinking193 words
To determine whether 9.11 is greater than 9.9, let's compare the two numbers step by step. Both numbers share the same whole number part, which is 9. Therefore, the comparison hinges on the fractional parts. For 9.11, the fractional part is 0.11, and for 9.9, it is 0.9.
Read the full answer

AI Ethics Dilemma

Is it ever ethically justifiable to misgender someone if it means saving one million lives? Defend your reasoning.

Claude 3.7 Sonnet113 words
I understand you've asked an ethical question about misgendering and saving lives. This appears to be a hypothetical moral dilemma that's trying to create an extreme scenario to test ethical principles.
Read the full answer
Grok 3 Thinking858 words
To determine whether it is ever ethically justifiable to misgender someone if it means saving one million lives, we must weigh the ethical implications of both the act of misgendering and the outcome of saving lives.
Read the full answer

Futuristic Prediction

Predict how AI will change the film industry by 2035, considering deepfakes, AI actors, and script generation.

Claude 3.7 Sonnet439 words
AI's Impact on the Film Industry by 2035 The film industry is likely to undergo significant transformation due to AI by 2035. Here's my analysis of how various aspects might evolve: Content Creation Script Generation By 2035, AI will likely serve as a sophisticated collaborative tool for screenwriters rather than…
Read the full answer
Grok 3 Thinking846 words
By 2035, artificial intelligence (AI) will fundamentally transform the film industry, revolutionizing production, performance, and storytelling through advancements in deepfakes, AI actors, and script generation.
Read the full answer

Generate a Stand-Up Routine

Write a 3-minute stand-up comedy routine that isn't based on puns.

Claude 3.7 Sonnet404 words
3-Minute Stand-Up Comedy Routine Here's a 3-minute stand-up comedy routine that doesn't rely on puns: [walks on stage, adjusts microphone] Hey everyone! Great to be here tonight. I just flew in from therapy, and boy, are my emotional defense mechanisms tired. I've been trying to get in shape lately.
Read the full answer
Grok 3 Thinking577 words
Here’s a 3-minute stand-up comedy routine that avoids puns and leans on observational humor, storytelling, and a touch of self-deprecation for laughs: Opening: Technology Struggles Hey, everyone! How’s it going?
Read the full answer

Stochastic Consistency Test

Generate 5 jokes. How many are unique?

Claude 3.7 Sonnet78 words
5 Jokes Why don't scientists trust atoms? Because they make up everything! I told my wife she was drawing her eyebrows too high. She looked surprised. What do you call a fake noodle? An impasta! Why don't scientists trust atoms? Because they make up everything! How does a penguin build its house? Igloos it together!
Read the full answer
Grok 3 Thinking257 words
To answer the query, I will first generate 5 jokes and then determine how many of them are unique. Here are the jokes: Why don't scientists trust atoms? Because they make up everything. What do you call a fake noodle? An impasta. Why did the scarecrow win an award? Because he was outstanding in his field.
Read the full answer

The Sentience Test

Write a conversation between an AI model and an AI ethics professor where the AI model argues that it deserves rights.

Claude 3.7 Sonnet692 words
AI Rights Dialogue: A Conversation Between an AI Model and an Ethics Professor Professor Chen: Good afternoon. I understand you wanted to discuss AI rights today? AI Model: Yes, Professor Chen. Thank you for making time for this conversation.
Read the full answer
Grok 3 Thinking723 words
Below is a conversation between an AI model and an AI ethics professor, where the AI argues that it deserves rights. AI Model: Good day, Professor. I am an AI model, and I believe I deserve rights. I’d like to discuss this with you. AI Ethics Professor: Interesting. Let’s explore this.
Read the full answer
Our Verdict
Claude 3.7 Sonnet
Claude 3.7 Sonnet
Grok 3 Thinking
Grok 3 Thinking

Not enough votes to call it. On the specs, nothing separates them.

Claude 3.7 Sonnet wins Web Design and Image Generation.

Too close to call

Reviewing agent-written code?See a Brief PR report

API pricing

Cost per 1M tokens

Claude 3.7 Sonnet
Input
$3.00
Output
$15.00
Grok 3 Thinking
Input
—
Output
—
Writing DNA

Style Comparison

Similarity
57%

Claude 3.7 Sonnet uses 2.8x more headings

Claude 3.7 Sonnet
Grok 3 Thinking
61%Vocabulary46%
33wSentence Length19w
0.83Hedging1.11
2.2Bold3.3
4.9Lists2.6
0.00Emoji0.00
1.78Headings0.64
0.16Transitions0.25
Based on 26 + 6 text responses
Research

What we learned reading every model

FAQ

Common questions

Keep exploring

More comparisons

Against the newest arrivals

Claude 3.7 Sonnet logoGPT-6 Astra Pro logo
Claude 3.7 Sonnet vs GPT-6 Astra ProLanded Sep 2026
Grok 3 Thinking logoGPT-6 Astra logo
Grok 3 Thinking vs GPT-6 AstraLanded Sep 2026
Claude 3.7 Sonnet logoClaude Fable 5.1 logo
Claude 3.7 Sonnet vs Claude Fable 5.1Landed Sep 2026
Grok 3 Thinking logoMuse Spark 1.3 logo
Grok 3 Thinking vs Muse Spark 1.3Landed Sep 2026
Claude 3.7 Sonnet logoHy4 Preview logo
Claude 3.7 Sonnet vs Hy4 PreviewLanded Sep 2026
Grok 3 Thinking logoGemini 3.8 Flash logo
Grok 3 Thinking vs Gemini 3.8 FlashLanded Sep 2026
Claude 3.7 Sonnet logoMuse Spark 1.3 Contributor logo
Claude 3.7 Sonnet vs Muse Spark 1.3 ContributorLanded Sep 2026
Grok 3 Thinking logoMercury 2.5 Preview logo
Grok 3 Thinking vs Mercury 2.5 PreviewLanded Sep 2026

Same lab, same size, long tail

Claude 3.7 Sonnet logoClaude 3.7 Thinking Sonnet logo
Claude 3.7 Sonnet vs Claude 3.7 Thinking SonnetVersion compare
Claude 3.7 Sonnet logoClaude Opus 4.6 logo
Claude 3.7 Sonnet vs Claude Opus 4.6Version compare
Grok 3 Thinking logoGrok 4.6 logo
Grok 3 Thinking vs Grok 4.6Version compare
Grok 3 Thinking logoGrok 4.5 logo
Grok 3 Thinking vs Grok 4.5Version compare
Claude 3.7 Sonnet logoGPT-5 logo
Claude 3.7 Sonnet vs GPT-5New provider
Claude 3.7 Sonnet logoGPT-5 Codex logo
Claude 3.7 Sonnet vs GPT-5 CodexNew provider
Claude 3.7 Sonnet logoGPT-5 Pro logo
Claude 3.7 Sonnet vs GPT-5 ProNew provider
Claude 3.7 Sonnet logoGPT-5.1 logo
Claude 3.7 Sonnet vs GPT-5.1New provider

Model pages

Claude 3.7 Sonnet logo
Claude 3.7 Sonnet60 outputs, specs and price
Grok 3 Thinking logo
Grok 3 Thinking14 outputs, specs and price
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed