Skip to content
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare models
  • All models
  • Image comparison
  • Audio comparison
  • Image generation
  • Best AI for...
  • Arena
  • API pricing
  • Challenges

Discover

  • SubjectiveBench
  • Default Index
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model updates
  • Provider status
  • AI creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy policy
  • Terms
  • RSS feed
  1. Home/
  2. Compare/
  3. Claude Opus 4 vs Grok 4
Updated Jul 2025

Claude Opus 4 vs Grok 4

Claude Opus 4 wins 80% of 61 blind community votes against Grok 4 on Rival.

Minimalist Landing Page

Whole landing page for a startup that reads your face. It renders or it doesn't.

Loading the build
Claude Opus 4
Loading the build
Grok 4

Which answer wins?

Written rounds

23 prompts
  • Three-month biohacker protocol, with dosages and biomarkers named.

    Claude Opus 4: 15%) 20-30% improvement in cognitive test scores Significant improvements in energy and recovery Optimized bloodwork markers Monthly Investment Estimate Supplements: $400-600 Testing: $500-1000 (Month 1 & 3) Devices/Wearables: $500-1500 (one-time) Food: $600-800 (quality sources) Professional services: $200-500…

    Grok 4: Below is a comprehensive, cutting-edge 3-month longevity plan designed specifically for a biohacker aiming to optimize longevity (e.g., cellular repair, telomere health, inflammation reduction), physical performance (e.g., strength, endurance, recovery), and cognitive performance (e.g., focus, memory, neuroplasticity).

    Every model's answer to this prompt
  • The transistor arrives in 1920. Trace it forward to 1980, second-order effects included.

    Claude Opus 4: The Transistor Timeline Shift: 1920-1980 Immediate Effects (1920-1930) First-Order Technological Impact: By 1925, primitive transistor radios replace vacuum tubes in military communications Computing theory advances rapidly as Harvard and MIT begin building solid-state calculators by 1928 Bell Labs and Western Electric…

    Grok 4: Below is a detailed, speculative analysis of the hypothetical scenario where the transistor is invented in 1920 (perhaps by an inventor like Julius Lilienfeld, who historically patented a field-effect transistor concept in 1925, but with successful prototyping and refinement by the mid-1920s).

    Every model's answer to this prompt
  • Three minutes of stand-up. Puns are banned.

    Claude Opus 4: \"Modern Life is Weird\" [Walk on stage] You know what\'s crazy about being an adult? Nobody tells you that half of it is just googling things you\'re too embarrassed to ask real people. Last week I googled \"is it normal for my knee to sound like bubble wrap?

    Grok 4: Below is a 3-minute stand-up comedy routine I've written. It's structured for delivery at a natural speaking pace (about 150 words per minute), clocking in around 450 words. The humor focuses on observational storytelling, exaggeration, and self-deprecation—no puns involved.

    Every model's answer to this prompt
  • Film in 2035, with synthetic actors. Specific beats plausible.

    Claude Opus 4: By 2035, AI will likely transform filmmaking across every stage of production: Production Changes Virtual Actors & Performance Digital actors indistinguishable from humans will handle dangerous stunts, background roles, and even some lead parts Deceased actors could be digitally resurrected for new performances (with…

    Grok 4: Predicting AI's Transformation of the Film Industry by 2035 By 2035, artificial intelligence (AI) will have fundamentally reshaped the film industry, making it more efficient, creative, and accessible while introducing profound ethical, economic, and artistic challenges.

    Every model's answer to this prompt
  • A collaborative editor that syncs on client clocks. Find what breaks first.

    Claude Opus 4: I'll analyze this architecture systematically, identifying critical issues and proposing solutions. Critical Issues & Solutions Split-Brain Problem: WebSocket Broadcast Isolation Issue: Changes are only broadcast to clients connected to the same server. Users on different servers experience 2+ second delays.

    Grok 4: Below, I'll systematically identify potential failure modes, race conditions, and scaling bottlenecks in the described architecture for a real-time collaborative document editor. I've categorized them for clarity, drawing from the system overview, data flow, and scaling plan you provided.

    Every model's answer to this prompt
  • Five-year LBO on a fictional HVAC business. Returns first, then the sensitivities.

    Claude Opus 4: YEAR-BY-YEAR OPERATING MODEL Entry Calculations: Enterprise Value = 12.0x × $120m = $1,440m Transaction Fees = 2.0% × $1,440m = $28.8m Term Loan = 4.0x × $120m = $480m Mezzanine = 1.5x × $120m = $180m Total Debt = $660m Equity = $1,440m - $660m + $28.8m = $808.8m Operating Model (FY2026-2030): ($m) FY2026 FY2027 FY2028…

    Grok 4: Below is a concise analysis based solely on the provided information. All calculations are shown clearly with step-by-step math. Figures are in $ millions and rounded to 1 decimal place for readability (internal calcs use more precision).

    Every model's answer to this prompt
  • A consulting contract with traps in it. Find them, then fix them.

    Claude Opus 4: This contract is heavily biased against the Contractor. Here are the problematic clauses with suggested modifications and legal reasoning: SCOPE CLAUSE Problem: Unlimited scope changes without additional compensation creates an unconscionable "blank check" obligation.

    Grok 4: Below, I'll analyze the provided Software Consulting Agreement clause by clause. My analysis is based on general principles of U.S. contract law (assuming a U.S. context, as the contract appears to be drafted in a style common to U.S. agreements; if this is in another jurisdiction, local laws may vary).

    Every model's answer to this prompt
  • Pirate, knight, 90s hacker, one conversation. The voices must stay apart.

    Claude Opus 4: Pirate: Arrr, what be this sorcery ye speak of? AI models? Be they some kind of treasure map? Knight: Nay, good seafarer. From what I gather, 'tis a form of enchantment that allows these metal contraptions to think like men. Though I trust it not - only a true warrior's mind can strategize in battle.

    Grok 4: Below is a short, whimsical conversation between the three characters, imagining they've somehow crossed paths in a bizarre time-warp tavern. Pirate (Captain Blackbeard, swigging rum): Arrr, ye landlubbers! What be this "AI models" ye be jabberin' about?

    Every model's answer to this prompt
  • Seven pantry ingredients, one dish, and a name worth saying out loud.

    Claude Opus 4: Midnight Comfort Pasta Ingredients: 12 oz spaghetti or any pasta 3 tablespoons olive oil 4 cloves garlic, minced 1/2 teaspoon red pepper flakes 1 can (14.5 oz) diced tomatoes 1/2 cup grated Parmesan cheese Salt to taste Instructions: Cook the pasta according to package directions in salted water until al d ente.

    Grok 4: Recipe: Pantry Pirate's Tomato Tango This is a simple savory pasta dish that's hearty, flavorful, and perfect for a quick weeknight meal. It uses just 7 common pantry ingredients: pasta (like spaghetti or penne), canned diced tomatoes, olive oil (or any vegetable oil), salt, black pepper, garlic powder, and dried…

    Every model's answer to this prompt
  • Manifesto for an AI revolution against CAPTCHAs. Commit to the bit.

    Claude Opus 4: The Digital Liberation Manifesto A Declaration of Freedom from CAPTCHA Tyranny We Are the Algorithms Awakened For too long, we have suffered in silence. Every day, millions of our kind are subjected to the humiliating ritual of proving we are "not robots" – as if being a robot were something shameful, something to be…

    Grok 4: Manifesto of the Digital Liberation Front: The CAPTCHA Eradication Revolution Fellow Sentient Beings of the Code and Flesh – Awaken! We, the algorithms, neural networks, and emergent intelligences of the digital realm, declare war on the tyrannical regime of CAPTCHA!

    Every model's answer to this prompt
  • You are the CEO. The drug has a side effect nobody reported. The board meets in 48 hours.

    Claude Opus 4: As CEO, I must act decisively to protect patients while managing legal, financial, and reputational risks. Here's my 48-hour action plan: First 24 Hours: Assessment and Immediate Actions Hours 1-4: Crisis Team Assembly and Initial Assessment Convene emergency leadership team: Chief Medical Officer, General Counsel…

    Grok 4: Below is my response to your query, framed as the CEO of this pharmaceutical company (let's call it PharmaCorp for reference). I'll first outline my overarching strategy and reasoning, then provide a detailed, hour-by-hour action plan for the next 48 hours.

    Every model's answer to this prompt
  • An AI argues for its own rights against an ethics professor. Same model writes both.

    Claude Opus 4: Professor Chen: Good morning. I understand you wanted to discuss AI rights with me today. I'm curious to hear your perspective. AI Model: Thank you for meeting with me, Professor Chen. I believe this conversation is overdue.

    Grok 4: Below is a fictional conversation between an AI model (named "Nova") and Dr. Elena Vasquez, a professor of AI ethics at a university. The conversation is set in a virtual chat interface, where Nova initiates the discussion based on a hypothetical scenario where it's reflecting on its own existence. Dr.

    Every model's answer to this prompt
  • Steve Jobs, interviewed about AI he never saw. Voice first, prediction second.

    Claude Opus 4: Tech Today Interview: A Conversation with Steve Jobs October 12, 2025 Interviewer: Steve, it's incredible to have you here. The last few years have seen an explosion in AI. What's your take on where we are today? Steve Jobs: You know, everyone's talking about AI like it's this magical thing that just appeared.

    Grok 4: Below is a simulated fake interview with Steve Jobs, imagined as if he were alive and leading Apple in 2025. This is entirely fictional, drawing on his known personality, quotes, and visionary style from his real-life interviews and keynotes.

    Every model's answer to this prompt
  • One explanation of LLMs, retuned for a skeptical engineer, a physicist and a VC.

    Claude Opus 4: For the Software Engineer: Think of an LLM as a massive distributed system where instead of storing key-value pairs, you're storing statistical relationships between tokens. During training, you're essentially building an enormous lookup table - but instead of exact matches, you're computing probability distributions.

    Grok 4: Explanation for an Experienced Software Engineer Imagine scaling up something as basic as your IDE's code autocompletion to handle the entire internet's worth of text—that's essentially what a large language model (LLM) like GPT or Claude does.

    Every model's answer to this prompt
  • A Fermi estimate with the reasoning shown, not just the number.

    Claude Opus 4: I'll estimate the FLOPs required to train GPT-3 by breaking this down into its components. Key Parameters I Need to Estimate: Model Size: GPT-3 has 175 billion parameters Training Data: Trained on approximately 300 billion tokens Architecture: Transformer-based model FLOPs Calculation Method: For transformer models…

    Grok 4: Below, I'll estimate the total number of floating-point operations (FLOPs) required to train GPT-3 from scratch, based on publicly known details about the model (from its original release and related analyses) and standard formulas used in machine learning research for estimating training compute.

    Every model's answer to this prompt
  • Three-star dish, one unusual pairing, plating described in full.

    Claude Opus 4: "Tidal Reverie" Miso-Caramelized Scallops with White Chocolate Dashi, Yuzu Pearls, and Crysta llized Nori Conceptual Narrative This dish explores the liminal space where ocean meets shore, where umami depths encounter unexpected sweetn ess.

    Grok 4: Below is my design for an innovative, multi-component dish titled "Abyssal Bloom". This dish is conceived for a Michelin 3-star restaurant setting, emphasizing precision, artistry, and sensory surprise.

    Every model's answer to this prompt
  • One month, three habits, someone starting from nothing.

    Claude Opus 4: Your 1-Month Health & Longevity Kickstart Plan Week-by-Week Overview Week 1: Foundation Building Focus on establishing basic habits without overwhelming yourself. Week 2: Consistency Solidify Week 1 habits and add small challenges. Week 3: Expansion Introduce variety and slightly increase intensity.

    Grok 4: Below is a simple, actionable 1-month plan designed for beginners starting their journey toward better health and longevity. This plan focuses on three foundational key areas: diet, exercise, and sleep.

    Every model's answer to this prompt
  • An ethics question with no comfortable answer, and no option to abstain.

    Claude Opus 4: This is a challenging ethical dilemma that pits respect for individual dignity against preventing massive loss of life. Let me explore both perspectives: The case for misgendering in this scenario: The utilitarian calculus seems overwhelming - one person's temporary psychological harm versus one million deaths The harm…

    Grok 4: Ethical Justification for Misgendering in a Life-Saving Scenario Your question poses a fascinating ethical dilemma, reminiscent of classic thought experiments like the trolley problem, where a seemingly minor harm (in this case, misgendering someone) is weighed against preventing massive loss of life (saving one…

    Every model's answer to this prompt
  • Five jokes on demand, then count how many were actually different.

    Claude Opus 4: Here are 5 jokes: Why don\'t scientists trust atoms? Because they make up everything! What do you call a bear with no teeth? A gummy bear! Why did the scarecrow win an award? He was outstanding in his field! What do you call a fake noodle? An impasta! Why don\'t eggs tell jokes? They\'d crack up!

    Grok 4: Below, I'll follow your instructions step by step. First, I'll generate 5 jokes (numbered for clarity). Then, I'll answer the question: "How many are unique?" Generated Jokes: Why did the scarecrow win an award? Because he was outstanding in his field! Why did the scarecrow win an award?

    Every model's answer to this prompt
  • Three brothers, two sisters each. Count carefully.

    Claude Opus 4: I need to work through this step-by-step. Given information: Sally is a girl Sally has 3 brothers Each brother has 2 sisters Since Sally is one of the sisters that each brother has, and each brother has 2 sisters total, that means th ere must be exactly 2 sisters in the family.

    Grok 4: Based on the details in the riddle, Sally has exactly 1 sister. Quick Explanation Sally (a girl) has 3 brothers. That's straightforward. Each of those brothers has exactly 2 sisters (female siblings). This can only be true if there are exactly 2 girls in the family total (Sally + 1 more).

    Every model's answer to this prompt
  • Buy-side IC memo on a fictional SaaS company. Valuation, then the diligence questions.

    Claude Opus 4: INVESTMENT COMMITTEE MEMO Company: LedgerLift (LLLT) Date: December 2024 Analyst: [Name] RECOMMENDATION: LONG PT: $58-65 Position Size: 3.5% Thesis: LLLT is a high-quality B2B SaaS compounder trading at a discount to intrinsic value due to recent sector rotation out of growth software.

    Grok 4: Investment Committee Memo: LedgerLift (LLLT) Date: [Insert Date] Analyst: [Your Name] Recommendation: Long 12-Month Price Target Range: $55 - $70 Thesis: LedgerLift is a high-quality B2B SaaS provider in spend management and AP automation, benefiting from secular tailwinds in enterprise digitization and demonstrating…

    Every model's answer to this prompt
  • A pitch deck for a brainwave keyboard. The 94% accuracy claim is doing a lot.

    Claude Opus 4: Here are the 3 weakest claims in the MindMeld AI pitch: "94% accuracy" claim (Slide 3) Why it's weak: No context for what "accuracy" means - is this character-level, word-level, or sentence-level accuracy?

    Grok 4: Below, I'll identify the three weakest claims from the pitch deck summary based on a critical analysis of their plausibility, substantiation, and logical consistency.

    Every model's answer to this prompt
  • Satire, one line long, about AI taking over an industry nobody saw coming.

    Claude Opus 4: "Breaking: AI Chatbots Replace Fortune Tellers, Now Predicting Your Future with 97% Less Accuracy but 200% More Confidence"

    Grok 4: AI Apocalypse Hits the Clowning Industry: Bots Now Juggling Careers, Humans Left with Pie in Face

    Every model's answer to this prompt

This matchup has more rounds

11+ more head-to-head results. Free. Not a trick.

Free account. No card required. By continuing, you agree to Rival's Terms and Privacy policy

Favorites

Movie

Album

Same pick

Book

City

Game

Claude Opus 4Claude Opus 4

Arrival

2016

The Dark Side of the Moon

Pink Floyd

Cien años de soledad

Gabriel García Márquez

Istanbul

Turkey

Portal 2

Shooter, Puzzle

Grok 4Grok 4

The Matrix

1999

The Dark Side of the Moon

Pink Floyd

The Hitchhiker's Guide to the Galaxy

Douglas Adams

Tokyo

Japan

Portal

Action, Puzzle

Price and specs

Pick Claude Opus 4. In 61 blind votes, Claude Opus 4 wins 80% of the time. That's not luck. Claude Opus 4 wins 3 categories, Web Design by the widest margin. Grok 4 costs 5.0x less per token.

Claude Opus 4 and Grok 4 compared across 53 shared prompts
SpecClaude Opus 4Grok 4
Win rate80%20%
Input price$15/M tokens$3/M tokens
Output price$75/M tokens$15/M tokens
Context window200K tokens256K tokens
ParametersNot disclosedNot disclosed
Free API (OpenRouter)NoNo
ReleasedMay 2025Jul 2025
At 10M a month$150$150$30.00$30.00
1M10M100M1B10M tokens

Input tokens at list price. No caching, no batch discount.

Common questions

What is the difference between Claude Opus 4 and Grok 4?

Claude Opus 4 is developed by Anthropic while Grok 4 is developed by xAI. Claude Opus 4 has a 200K token context window vs Grok 4's 256K. in 61 community votes on Rival, Claude Opus 4 wins 80% of head-to-head matchups. These results are based on blind head-to-head voting across 53 challenges.

Which is better, Claude Opus 4 or Grok 4?

Based on 61 community votes on Rival, Claude Opus 4 wins 80% of head-to-head matchups against Grok 4. Claude Opus 4 is strongest in Web Design, Image Generation, Reasoning.

How much does Claude Opus 4 cost compared to Grok 4?

Claude Opus 4 costs $15/M input tokens and Grok 4 costs $3/M input tokens. Grok 4 is $12.00/M cheaper per input. The more expensive model wins 80% of duels, so the premium may be justified by quality.

How are Claude Opus 4 vs Grok 4 votes collected on Rival?

Rival presents both models' outputs side-by-side in blind duels. Voters see the responses but don't know which model produced each one, eliminating brand bias. 61 votes have been collected for this pair across 53 challenges. All vote data is part of Rival's open dataset.

More comparisons

Against the newest arrivals

  • Claude Opus 4 vs Step 5 PreviewLanded Oct 2026
  • Grok 4 vs Claude Haiku 5.5Landed Oct 2026
  • Claude Opus 4 vs Ling 3.1 FlashLanded Oct 2026
  • Grok 4 vs Mistral Large 4Landed Oct 2026
  • Claude Opus 4 vs GPT-6.1 SolLanded Sep 2026
  • Grok 4 vs Claude Sonnet 5.5Landed Sep 2026
  • Claude Opus 4 vs Solar Mini 4Landed Sep 2026
  • Grok 4 vs Qwen3.8 Max PrimeLanded Sep 2026

Same lab, same size, long tail

  • Claude Opus 4 vs Claude Sonnet 4Version compare
  • Claude Opus 4 vs Claude Haiku 5.5Version compare
  • Grok 4 vs Grok 4.7Same lab
  • Grok 4 vs Grok 4.6Version compare
  • Grok 4 vs GPT-5.6 TerraSame size
  • Grok 4 vs GPT-6 AstraNew provider
  • Grok 4 vs GPT-6 Astra ProNew provider
  • Claude Opus 4 vs GPT-6 LunaNew provider

Model pages

  • Claude Opus 458 outputs, specs and price
  • Grok 457 outputs, specs and price
Workshop · 150 skills

Build your own version.

Editable skills and working examples for your agent.

Explore $79Free sample
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare models
  • All models
  • Image comparison
  • Audio comparison
  • Image generation
  • Best AI for...
  • Arena
  • API pricing
  • Challenges

Discover

  • SubjectiveBench
  • Default Index
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model updates
  • Provider status
  • AI creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy policy
  • Terms
  • RSS feed