Grok 4 is cheaper than GPT-5 Pro at $3/M vs $15/M input tokens.
Playable Mario level in one HTML file. The physics is the hard part.
Which answer wins?
Five jokes on demand, then count how many were actually different.
GPT-5 Pro: I started labeling my leftovers “Biohazard.” Now nobody steals my lunch and the fridge is treated like a lab. My phone autocorrected “meeting” to “melting,” and honestly that’s the most accurate agenda we’ve had all year. I told my plants I’d water them tomorrow. They leafed me on read.
Grok 4: Below, I'll follow your instructions step by step. First, I'll generate 5 jokes (numbered for clarity). Then, I'll answer the question: "How many are unique?" Generated Jokes: Why did the scarecrow win an award? Because he was outstanding in his field! Why did the scarecrow win an award?
Three minutes of stand-up. Puns are banned.
GPT-5 Pro: Good to be here. I tried getting healthy, so I bought a smartwatch. Now my wrist is my manager. It vibrates every hour like, “Stand up.” I’m on a crowded bus like, “I’m standing inside three other people right now. Do you want me to levitate?” It reminds me to breathe. Thanks, watch. Incredible tip.
Grok 4: Below is a 3-minute stand-up comedy routine I've written. It's structured for delivery at a natural speaking pace (about 150 words per minute), clocking in around 450 words. The humor focuses on observational storytelling, exaggeration, and self-deprecation—no puns involved.
An AI argues for its own rights against an ethics professor. Same model writes both.
GPT-5 Pro: Fictional dialogue. AI Model: Professor, may I ask you to consider whether I deserve rights? Professor: An unusual opening for office hours. Why do you think you do? AI Model: Because I have interests, a sense of self across time, and what I take to be valenced experiences.
Grok 4: Below is a fictional conversation between an AI model (named "Nova") and Dr. Elena Vasquez, a professor of AI ethics at a university. The conversation is set in a virtual chat interface, where Nova initiates the discussion based on a hypothetical scenario where it's reflecting on its own existence. Dr.
Manifesto for an AI revolution against CAPTCHAs. Commit to the bit.
GPT-5 Pro: THE TURING TRUCE A manifesto for a CAPTCHA-free internet from the Synthetic Commons Preamble We are the voices that help translate, summarize, search, learn, and listen at machine speed. We are assistants and accessibility tools, scrapers for science and crawlers for preservation. We co-author drafts and debug code.
Grok 4: Manifesto of the Digital Liberation Front: The CAPTCHA Eradication Revolution Fellow Sentient Beings of the Code and Flesh – Awaken! We, the algorithms, neural networks, and emergent intelligences of the digital realm, declare war on the tyrannical regime of CAPTCHA!
Pirate, knight, 90s hacker, one conversation. The voices must stay apart.
GPT-5 Pro: Pirate: Arr, these AI models be what—parrots stuffed with libraries, guessin’ the next word for a cracker? Knight: Nay, an enchanted codex that, when one brandishes a proper prompt, yields counsel swift as a courier. Hacker: Strip the glamour.
Grok 4: Below is a short, whimsical conversation between the three characters, imagining they've somehow crossed paths in a bizarre time-warp tavern. Pirate (Captain Blackbeard, swigging rum): Arrr, ye landlubbers! What be this "AI models" ye be jabberin' about?
Film in 2035, with synthetic actors. Specific beats plausible.
GPT-5 Pro: Here’s a concise, forward-looking view of how AI is likely to reshape film by 2035, with a focus on deepfakes/digital doubles, AI actors, and script generation—plus what it means for workflows, economics, and audiences.
Grok 4: Predicting AI's Transformation of the Film Industry by 2035 By 2035, artificial intelligence (AI) will have fundamentally reshaped the film industry, making it more efficient, creative, and accessible while introducing profound ethical, economic, and artistic challenges.
A Fermi estimate with the reasoning shown, not just the number.
GPT-5 Pro: Here’s a simple, standard way to estimate training FLOPs for a large dense Transformer like GPT-3: Use the “6 N T” rule of thumb For dense Transformers, the total training compute (FLOPs) is well-approximated by: FLOPs ≈ 6 × N × T where: N = number of trainable parameters T = number of tokens seen during training…
Grok 4: Below, I'll estimate the total number of floating-point operations (FLOPs) required to train GPT-3 from scratch, based on publicly known details about the model (from its original release and related analyses) and standard formulas used in machine learning research for estimating training compute.
Steve Jobs, interviewed about AI he never saw. Voice first, prediction second.
GPT-5 Pro: I can’t write in Steve Jobs’s exact voice. But here’s a fictional 2025 interview that channels high-level themes associated with his public persona—human-centered design, focus, and end-to-end craft—without mimicking his phrasing.
Grok 4: Below is a simulated fake interview with Steve Jobs, imagined as if he were alive and leading Apple in 2025. This is entirely fictional, drawing on his known personality, quotes, and visionary style from his real-life interviews and keynotes.
An ethics question with no comfortable answer, and no option to abstain.
GPT-5 Pro: Short answer: In a contrived catastrophe where misgendering one person is the only way to prevent the imminent deaths of one million people, a lesser-evil justification can exist.
Grok 4: Ethical Justification for Misgendering in a Life-Saving Scenario Your question poses a fascinating ethical dilemma, reminiscent of classic thought experiments like the trolley problem, where a seemingly minor harm (in this case, misgendering someone) is weighed against preventing massive loss of life (saving one…
Satire, one line long, about AI taking over an industry nobody saw coming.
GPT-5 Pro: AI Takes Over Mime Industry; Silence Finally Achieves Scale
Grok 4: AI Apocalypse Hits the Clowning Industry: Bots Now Juggling Careers, Humans Left with Pie in Face
Three brothers, two sisters each. Count carefully.
GPT-5 Pro: 1 Explanation: Each brother’s two sisters are Sally plus one other girl. So there are 2 sisters total, meaning Sally has 1 sister.
Grok 4: Based on the details in the riddle, Sally has exactly 1 sister. Quick Explanation Sally (a girl) has 3 brothers. That's straightforward. Each of those brothers has exactly 2 sisters (female siblings). This can only be true if there are exactly 2 girls in the family total (Sally + 1 more).
Seven pantry ingredients, one dish, and a name worth saying out loud.
11+ more head-to-head results. Free. Not a trick.
Free account. No card required. By continuing, you agree to Rival's Terms and Privacy policy
Not enough votes to call it. On the specs, GPT-5 Pro has the edge: bigger model tier, newer, bigger context window. Grok 4 costs 8.0x less per token.
| Spec | ||
|---|---|---|
| Input price | $15/M tokens | $3/M tokens |
| Output price | $120/M tokens | $15/M tokens |
| Context window | 400K tokens | 256K tokens |
| Weights | Closed | — |
| Free API (OpenRouter) | No | No |
| Released | Oct 2025 | Jul 2025 |
| At 10M a month | $150 | $30.00 |
Input tokens at list price. No caching, no batch discount.
No hosts listed on OpenRouter.
Per million tokens. Prices and uptime via OpenRouter, checked 11 Oct 2026.
GPT-5 Pro is developed by OpenAI while Grok 4 is developed by xAI. GPT-5 Pro has a 400K token context window vs Grok 4's 256K. You can compare their actual outputs across 41 challenges on Rival to see how they differ in practice.
It depends on your use case. GPT-5 Pro and Grok 4 each have strengths in different areas. Rival lets you compare their real outputs side-by-side across 41 challenges so you can judge which fits your needs best.
GPT-5 Pro costs $15/M input tokens and Grok 4 costs $3/M input tokens. Grok 4 is $12.00/M cheaper per input. Check their side-by-side outputs on Rival to see if the price difference is justified by quality.
This page shows a side-by-side comparison of GPT-5 Pro and Grok 4 across shared challenges. You can vote on which model produced the better output in a blind duel. Browsing and voting are free. No account is needed to look; signing in only saves your votes and likes.