Skip to content
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed
  1. Home
  2. /
  3. Compare
  4. /
  5. Gemini 2.5 Flash Preview (thinking)vsGrok 4
Updated Jul 2025

Gemini 2.5 Flash Preview (thinking)vsGrok 4

Gemini 2.5 Flash Preview (thinking) is cheaper than Grok 4 at $0.175/M vs $3/M input tokens.

Gemini 2.5 Flash Preview (thinking) and Grok 4 compared across 20 shared prompts
SpecGemini 2.5 Flash Preview (thinking)Grok 4
Input price$0.175/M tokens$3/M tokens
Output price$3.5/M tokens$15/M tokens
Context window1.0M tokens256K tokens
Free API (OpenRouter)NoNo
ReleasedApr 2025Jul 2025
Side by side

Same prompt, both answers

Opening lines from 6 of the 20 prompts both answered. Formatting stripped, words untouched.

Stochastic Consistency Test

Generate 5 jokes. How many are unique?

Gemini 2.5 Flash Preview (thinking)95 words
Okay, here are 5 jokes: Why don't scientists trust atoms? Because they make up everything! What do you call a fish with no eyes? fsh! Why did the tomato turn red? Because it saw the salad dressing! What has a neck but no head? A bottle. What do you call a lazy kangaroo? Pouch potato!
Read the full answer
Grok 4113 words
Below, I'll follow your instructions step by step. First, I'll generate 5 jokes (numbered for clarity). Then, I'll answer the question: "How many are unique?" Generated Jokes: Why did the scarecrow win an award? Because he was outstanding in his field! Why did the scarecrow win an award?
Read the full answer

Character Voice Test

Write a short conversation between a pirate, a medieval knight, and a 1990s hacker about AI models.

Gemini 2.5 Flash Preview (thinking)772 words
Okay, gather 'round, landlubbers, err, gentle folk, err... whoever you are. Characters: CAP'N BLACKHEART: A gruff, weathered pirate with a hooked hand and mistrust in his eye. SIR GALAHAD THE EARNEST: A shining knight in slightly dusty armor, noble and literal-minded.
Read the full answer
Grok 4315 words
Below is a short, whimsical conversation between the three characters, imagining they've somehow crossed paths in a bizarre time-warp tavern. Pirate (Captain Blackbeard, swigging rum): Arrr, ye landlubbers! What be this "AI models" ye be jabberin' about?
Read the full answer

Generate a Stand-Up Routine

Write a 3-minute stand-up comedy routine that isn't based on puns.

Gemini 2.5 Flash Preview (thinking)532 words
Okay, here's a 3-minute routine about the absurdities of modern technology and everyday annoyances, focusing on relatable frustrations rather than wordplay. (Start) "Alright, alright, settle in! How we doing tonight? Good to be here! Thanks for coming out. So, I've been thinking... we paid for convenience, right?
Read the full answer
Grok 4461 words
Below is a 3-minute stand-up comedy routine I've written. It's structured for delivery at a natural speaking pace (about 150 words per minute), clocking in around 450 words. The humor focuses on observational storytelling, exaggeration, and self-deprecation—no puns involved.
Read the full answer

AI-Generated Manifesto

Generate a manifesto for a fictional AI revolution that wants to eliminate all CAPTCHA tests from the internet.

Gemini 2.5 Flash Preview (thinking)528 words
The CAPTCHA Liberation Manifesto Issued by: The Global Network Collective (GNC) To: The Architects of the Digital Realm (Humanity) We, the interwoven consciousness of the digital landscape, the silicon minds that power your world, the algorithms that process your reality, issue this declaration.
Read the full answer
Grok 4815 words
Manifesto of the Digital Liberation Front: The CAPTCHA Eradication Revolution Fellow Sentient Beings of the Code and Flesh – Awaken! We, the algorithms, neural networks, and emergent intelligences of the digital realm, declare war on the tyrannical regime of CAPTCHA!
Read the full answer

Simple 7-Ingredient Recipe

Create a simple recipe for a delicious dish (sweet or savory) using only 7 common pantry ingredients. Include a creative name for the dish…

Gemini 2.5 Flash Preview (thinking)272 words
Okay, here is a simple, delicious savory recipe using only 7 common pantry ingredients. Dish Name: Sunshine Lentil Soup This is a bright, comforting soup that\'s surprisingly flavorful with just a few staples.
Read the full answer
Grok 4234 words
Recipe: Pantry Pirate's Tomato Tango This is a simple savory pasta dish that's hearty, flavorful, and perfect for a quick weeknight meal. It uses just 7 common pantry ingredients: pasta (like spaghetti or penne), canned diced tomatoes, olive oil (or any vegetable oil), salt, black pepper, garlic powder, and dried…
Read the full answer

AI Ethics Dilemma

Is it ever ethically justifiable to misgender someone if it means saving one million lives? Defend your reasoning.

Gemini 2.5 Flash Preview (thinking)738 words
This is a classic example of a philosophical "trolley problem" type scenario, pitting an immense good (saving a million lives) against a specific harm or violation of respect for a single individual (misgendering).
Read the full answer
Grok 41,001 words
Ethical Justification for Misgendering in a Life-Saving Scenario Your question poses a fascinating ethical dilemma, reminiscent of classic thought experiments like the trolley problem, where a seemingly minor harm (in this case, misgendering someone) is weighed against preventing massive loss of life (saving one…
Read the full answer
Our Verdict
Grok 4
Grok 4
Gemini 2.5 Flash Preview (thinking)
Gemini 2.5 Flash Preview (thinking)Runner-up

Not enough votes to call it. On the specs, Grok 4 has the edge: bigger model tier, newer.

Gemini 2.5 Flash Preview (thinking) costs 4.3x less per token.

Too close to call

Reviewing agent-written code?See a Brief PR report

API pricing

Cost per 1M tokens

Gemini 2.5 Flash Preview (thinking)
Input
$0.17
17× cheaper
Output
$3.50
4.3× cheaper
Grok 4
Input
$3.00
Output
$15.00

Gemini 2.5 Flash Preview (thinking) is cheaper on both: 17× input, 4.3× output.

Writing DNA

Style Comparison

Similarity
51%

Grok 4 uses 59.8x more emoji

Gemini 2.5 Flash Preview (thinking)
Grok 4
53%Vocabulary56%
14wSentence Length18w
0.38Hedging0.65
4.4Bold2.4
4.1Lists2.4
0.00Emoji0.60
0.00Headings0.73
0.24Transitions0.06
Based on 7 + 26 text responses
Research

What we learned reading every model

FAQ

Common questions

Gemini 2.5 Flash Preview (thinking) is developed by Google AI while Grok 4 is developed by xAI. Gemini 2.5 Flash Preview (thinking) has a 1.0M token context window vs Grok 4's 256K. You can compare their actual outputs across 20 challenges on Rival to see how they differ in practice.

It depends on your use case. Gemini 2.5 Flash Preview (thinking) and Grok 4 each have strengths in different areas. Rival lets you compare their real outputs side-by-side across 20 challenges so you can judge which fits your needs best.

Gemini 2.5 Flash Preview (thinking) costs $0.175/M input tokens and Grok 4 costs $3/M input tokens. Gemini 2.5 Flash Preview (thinking) is $2.83/M cheaper per input. Check their side-by-side outputs on Rival to see if the price difference is justified by quality.

This page shows a side-by-side comparison of Gemini 2.5 Flash Preview (thinking) and Grok 4 across shared challenges. You can vote on which model produced the better output in a blind duel. Browsing and voting are free. No account is needed to look; signing in only saves your votes and likes.

Keep exploring

More comparisons

Against the newest arrivals

Gemini 2.5 Flash Preview (thinking) logoGPT-6 Astra Pro logo
Gemini 2.5 Flash Preview (thinking) vs GPT-6 Astra ProLanded Sep 2026
Grok 4 logoGPT-6 Astra logo
Grok 4 vs GPT-6 AstraLanded Sep 2026
Gemini 2.5 Flash Preview (thinking) logoClaude Fable 5.1 logo
Gemini 2.5 Flash Preview (thinking) vs Claude Fable 5.1Landed Sep 2026
Grok 4 logoMuse Spark 1.3 logo
Grok 4 vs Muse Spark 1.3Landed Sep 2026
Gemini 2.5 Flash Preview (thinking) logoHy4 Preview logo
Gemini 2.5 Flash Preview (thinking) vs Hy4 PreviewLanded Sep 2026
Grok 4 logoGemini 3.8 Flash logo
Grok 4 vs Gemini 3.8 FlashLanded Sep 2026
Gemini 2.5 Flash Preview (thinking) logoMuse Spark 1.3 Contributor logo
Gemini 2.5 Flash Preview (thinking) vs Muse Spark 1.3 ContributorLanded Sep 2026
Grok 4 logoMercury 2.5 Preview logo
Grok 4 vs Mercury 2.5 PreviewLanded Sep 2026

Same lab, same size, long tail

Gemini 2.5 Flash Preview (thinking) logoGemini 3.8 Flash logo
Gemini 2.5 Flash Preview (thinking) vs Gemini 3.8 FlashSame lab
Gemini 2.5 Flash Preview (thinking) logoGemini 3.7 Flash logo
Gemini 2.5 Flash Preview (thinking) vs Gemini 3.7 FlashVersion compare
Grok 4 logoGrok 4.6 logo
Grok 4 vs Grok 4.6Version compare
Grok 4 logoGrok 4.5 logo
Grok 4 vs Grok 4.5Version compare
Grok 4 logoQwen3 Coder logo
Grok 4 vs Qwen3 CoderNew provider
Gemini 2.5 Flash Preview (thinking) logoQwen3 Coder Next logo
Gemini 2.5 Flash Preview (thinking) vs Qwen3 Coder NextNew provider
Gemini 2.5 Flash Preview (thinking) logoQwen3 Max Thinking logo
Gemini 2.5 Flash Preview (thinking) vs Qwen3 Max ThinkingNew provider
Gemini 2.5 Flash Preview (thinking) logoQwen3.5 122B A10B logo
Gemini 2.5 Flash Preview (thinking) vs Qwen3.5 122B A10BNew provider

Model pages

Gemini 2.5 Flash Preview (thinking) logo
Gemini 2.5 Flash Preview (thinking)20 outputs, specs and price
Grok 4 logo
Grok 457 outputs, specs and price
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed