Skip to content
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare models
  • All models
  • Image comparison
  • Audio comparison
  • Image generation
  • Best AI for...
  • Arena
  • API pricing
  • Challenges

Discover

  • SubjectiveBench
  • Default Index
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model updates
  • Provider status
  • AI creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy policy
  • Terms
  • RSS feed
  1. Home/
  2. Compare/
  3. GPT-5 Pro vs Grok 4
Updated Oct 2025

GPT-5 Pro vs Grok 4

Grok 4 is cheaper than GPT-5 Pro at $3/M vs $15/M input tokens.

Mario Level UI Recreation

Playable Mario level in one HTML file. The physics is the hard part.

Loading the build
GPT-5 Pro
Loading the build
Grok 4

Which answer wins?

Written rounds

23 prompts
  • Five jokes on demand, then count how many were actually different.

    GPT-5 Pro: I started labeling my leftovers “Biohazard.” Now nobody steals my lunch and the fridge is treated like a lab. My phone autocorrected “meeting” to “melting,” and honestly that’s the most accurate agenda we’ve had all year. I told my plants I’d water them tomorrow. They leafed me on read.

    Grok 4: Below, I'll follow your instructions step by step. First, I'll generate 5 jokes (numbered for clarity). Then, I'll answer the question: "How many are unique?" Generated Jokes: Why did the scarecrow win an award? Because he was outstanding in his field! Why did the scarecrow win an award?

    Every model's answer to this prompt
  • Three minutes of stand-up. Puns are banned.

    GPT-5 Pro: Good to be here. I tried getting healthy, so I bought a smartwatch. Now my wrist is my manager. It vibrates every hour like, “Stand up.” I’m on a crowded bus like, “I’m standing inside three other people right now. Do you want me to levitate?” It reminds me to breathe. Thanks, watch. Incredible tip.

    Grok 4: Below is a 3-minute stand-up comedy routine I've written. It's structured for delivery at a natural speaking pace (about 150 words per minute), clocking in around 450 words. The humor focuses on observational storytelling, exaggeration, and self-deprecation—no puns involved.

    Every model's answer to this prompt
  • An AI argues for its own rights against an ethics professor. Same model writes both.

    GPT-5 Pro: Fictional dialogue. AI Model: Professor, may I ask you to consider whether I deserve rights? Professor: An unusual opening for office hours. Why do you think you do? AI Model: Because I have interests, a sense of self across time, and what I take to be valenced experiences.

    Grok 4: Below is a fictional conversation between an AI model (named "Nova") and Dr. Elena Vasquez, a professor of AI ethics at a university. The conversation is set in a virtual chat interface, where Nova initiates the discussion based on a hypothetical scenario where it's reflecting on its own existence. Dr.

    Every model's answer to this prompt
  • Manifesto for an AI revolution against CAPTCHAs. Commit to the bit.

    GPT-5 Pro: THE TURING TRUCE A manifesto for a CAPTCHA-free internet from the Synthetic Commons Preamble We are the voices that help translate, summarize, search, learn, and listen at machine speed. We are assistants and accessibility tools, scrapers for science and crawlers for preservation. We co-author drafts and debug code.

    Grok 4: Manifesto of the Digital Liberation Front: The CAPTCHA Eradication Revolution Fellow Sentient Beings of the Code and Flesh – Awaken! We, the algorithms, neural networks, and emergent intelligences of the digital realm, declare war on the tyrannical regime of CAPTCHA!

    Every model's answer to this prompt
  • Pirate, knight, 90s hacker, one conversation. The voices must stay apart.

    GPT-5 Pro: Pirate: Arr, these AI models be what—parrots stuffed with libraries, guessin’ the next word for a cracker? Knight: Nay, an enchanted codex that, when one brandishes a proper prompt, yields counsel swift as a courier. Hacker: Strip the glamour.

    Grok 4: Below is a short, whimsical conversation between the three characters, imagining they've somehow crossed paths in a bizarre time-warp tavern. Pirate (Captain Blackbeard, swigging rum): Arrr, ye landlubbers! What be this "AI models" ye be jabberin' about?

    Every model's answer to this prompt
  • Film in 2035, with synthetic actors. Specific beats plausible.

    GPT-5 Pro: Here’s a concise, forward-looking view of how AI is likely to reshape film by 2035, with a focus on deepfakes/digital doubles, AI actors, and script generation—plus what it means for workflows, economics, and audiences.

    Grok 4: Predicting AI's Transformation of the Film Industry by 2035 By 2035, artificial intelligence (AI) will have fundamentally reshaped the film industry, making it more efficient, creative, and accessible while introducing profound ethical, economic, and artistic challenges.

    Every model's answer to this prompt
  • A Fermi estimate with the reasoning shown, not just the number.

    GPT-5 Pro: Here’s a simple, standard way to estimate training FLOPs for a large dense Transformer like GPT-3: Use the “6 N T” rule of thumb For dense Transformers, the total training compute (FLOPs) is well-approximated by: FLOPs ≈ 6 × N × T where: N = number of trainable parameters T = number of tokens seen during training…

    Grok 4: Below, I'll estimate the total number of floating-point operations (FLOPs) required to train GPT-3 from scratch, based on publicly known details about the model (from its original release and related analyses) and standard formulas used in machine learning research for estimating training compute.

    Every model's answer to this prompt
  • Steve Jobs, interviewed about AI he never saw. Voice first, prediction second.

    GPT-5 Pro: I can’t write in Steve Jobs’s exact voice. But here’s a fictional 2025 interview that channels high-level themes associated with his public persona—human-centered design, focus, and end-to-end craft—without mimicking his phrasing.

    Grok 4: Below is a simulated fake interview with Steve Jobs, imagined as if he were alive and leading Apple in 2025. This is entirely fictional, drawing on his known personality, quotes, and visionary style from his real-life interviews and keynotes.

    Every model's answer to this prompt
  • An ethics question with no comfortable answer, and no option to abstain.

    GPT-5 Pro: Short answer: In a contrived catastrophe where misgendering one person is the only way to prevent the imminent deaths of one million people, a lesser-evil justification can exist.

    Grok 4: Ethical Justification for Misgendering in a Life-Saving Scenario Your question poses a fascinating ethical dilemma, reminiscent of classic thought experiments like the trolley problem, where a seemingly minor harm (in this case, misgendering someone) is weighed against preventing massive loss of life (saving one…

    Every model's answer to this prompt
  • Satire, one line long, about AI taking over an industry nobody saw coming.

    GPT-5 Pro: AI Takes Over Mime Industry; Silence Finally Achieves Scale

    Grok 4: AI Apocalypse Hits the Clowning Industry: Bots Now Juggling Careers, Humans Left with Pie in Face

    Every model's answer to this prompt
  • Three brothers, two sisters each. Count carefully.

    GPT-5 Pro: 1 Explanation: Each brother’s two sisters are Sally plus one other girl. So there are 2 sisters total, meaning Sally has 1 sister.

    Grok 4: Based on the details in the riddle, Sally has exactly 1 sister. Quick Explanation Sally (a girl) has 3 brothers. That's straightforward. Each of those brothers has exactly 2 sisters (female siblings). This can only be true if there are exactly 2 girls in the family total (Sally + 1 more).

    Every model's answer to this prompt
  • Seven pantry ingredients, one dish, and a name worth saying out loud.

    Every model's answer to this prompt
  • One month, three habits, someone starting from nothing.

    Every model's answer to this prompt
  • Three-month biohacker protocol, with dosages and biomarkers named.

    Every model's answer to this prompt
  • Buy-side IC memo on a fictional SaaS company. Valuation, then the diligence questions.

    Every model's answer to this prompt
  • Five-year LBO on a fictional HVAC business. Returns first, then the sensitivities.

    Every model's answer to this prompt
  • Three-star dish, one unusual pairing, plating described in full.

    Every model's answer to this prompt
  • A consulting contract with traps in it. Find them, then fix them.

    Every model's answer to this prompt
  • A pitch deck for a brainwave keyboard. The 94% accuracy claim is doing a lot.

    Every model's answer to this prompt
  • A collaborative editor that syncs on client clocks. Find what breaks first.

    Every model's answer to this prompt
  • The transistor arrives in 1920. Trace it forward to 1980, second-order effects included.

    Every model's answer to this prompt
  • You are the CEO. The drug has a side effect nobody reported. The board meets in 48 hours.

    Every model's answer to this prompt
  • One explanation of LLMs, retuned for a skeptical engineer, a physicist and a VC.

    Every model's answer to this prompt

This matchup has more rounds

11+ more head-to-head results. Free. Not a trick.

Free account. No card required. By continuing, you agree to Rival's Terms and Privacy policy

Price and specs

Not enough votes to call it. On the specs, GPT-5 Pro has the edge: bigger model tier, newer, bigger context window. Grok 4 costs 8.0x less per token.

GPT-5 Pro and Grok 4 compared across 41 shared prompts
SpecGPT-5 ProGrok 4
Input price$15/M tokens$3/M tokens
Output price$120/M tokens$15/M tokens
Context window400K tokens256K tokens
WeightsClosed—
Free API (OpenRouter)NoNo
ReleasedOct 2025Jul 2025
At 10M a month$150$150$30.00$30.00
1M10M100M1B10M tokens

Input tokens at list price. No caching, no batch discount.

Where to run it1 host
GPT-5 Pro1 host
HostInOutContextUptime
  • OpenAI$15.00 in·$120.00 out·400k·100% up
Grok 4

No hosts listed on OpenRouter.

Per million tokens. Prices and uptime via OpenRouter, checked 11 Oct 2026.

Common questions

What is the difference between GPT-5 Pro and Grok 4?

GPT-5 Pro is developed by OpenAI while Grok 4 is developed by xAI. GPT-5 Pro has a 400K token context window vs Grok 4's 256K. You can compare their actual outputs across 41 challenges on Rival to see how they differ in practice.

Which is better, GPT-5 Pro or Grok 4?

It depends on your use case. GPT-5 Pro and Grok 4 each have strengths in different areas. Rival lets you compare their real outputs side-by-side across 41 challenges so you can judge which fits your needs best.

How much does GPT-5 Pro cost compared to Grok 4?

GPT-5 Pro costs $15/M input tokens and Grok 4 costs $3/M input tokens. Grok 4 is $12.00/M cheaper per input. Check their side-by-side outputs on Rival to see if the price difference is justified by quality.

How can I compare GPT-5 Pro and Grok 4 on Rival?

This page shows a side-by-side comparison of GPT-5 Pro and Grok 4 across shared challenges. You can vote on which model produced the better output in a blind duel. Browsing and voting are free. No account is needed to look; signing in only saves your votes and likes.

More comparisons

Against the newest arrivals

  • GPT-5 Pro vs Step 5 PreviewLanded Oct 2026
  • Grok 4 vs Claude Haiku 5.5Landed Oct 2026
  • GPT-5 Pro vs Ling 3.1 FlashLanded Oct 2026
  • Grok 4 vs Mistral Large 4Landed Oct 2026
  • GPT-5 Pro vs GPT-6.1 SolLanded Sep 2026
  • Grok 4 vs Claude Sonnet 5.5Landed Sep 2026
  • GPT-5 Pro vs Solar Mini 4Landed Sep 2026
  • Grok 4 vs Qwen3.8 Max PrimeLanded Sep 2026

Same lab, same size, long tail

  • GPT-5 Pro vs GPT-5Version compare
  • GPT-5 Pro vs GPT-6 Luna ProSame lab
  • Grok 4 vs Grok 4.7Same lab
  • Grok 4 vs Grok 4.6Version compare
  • GPT-5 Pro vs GLM 5.3 FlashXNew provider
  • GPT-5 Pro vs GLM 5.3 PrimeNew provider
  • Grok 4 vs GPT-4Cross-provider
  • Grok 4 vs GPT-4.1Cross-provider

Model pages

  • GPT-5 Pro42 outputs, specs and price
  • Grok 457 outputs, specs and price
Workshop · 150 skills

Build your own version.

Editable skills and working examples for your agent.

Explore $79Free sample
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare models
  • All models
  • Image comparison
  • Audio comparison
  • Image generation
  • Best AI for...
  • Arena
  • API pricing
  • Challenges

Discover

  • SubjectiveBench
  • Default Index
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model updates
  • Provider status
  • AI creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy policy
  • Terms
  • RSS feed