Skip to content
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare models
  • All models
  • Image comparison
  • Audio comparison
  • Image generation
  • Best AI for...
  • Arena
  • API pricing
  • Challenges

Discover

  • SubjectiveBench
  • Default Index
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model updates
  • Provider status
  • AI creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy policy
  • Terms
  • RSS feed
  1. Home/
  2. Compare/
  3. Gemini 2.5 Flash Preview (thinking) vs Grok 4.20 Beta
Updated Mar 2026

Gemini 2.5 Flash Preview (thinking) vs Grok 4.20 Beta

Gemini 2.5 Flash Preview (thinking) is cheaper than Grok 4.20 Beta at $0.175/M vs $2/M input tokens.

Pokémon Battle UI Recreation

Pokémon battle screen you can actually click. One HTML file.

Loading the build
Gemini 2.5 Flash Preview (thinking)
Loading the build
Grok 4.20 Beta

Which answer wins?

Written rounds

7 prompts
  • Pirate, knight, 90s hacker, one conversation. The voices must stay apart.

    Gemini 2.5 Flash Preview (thinking): Okay, gather 'round, landlubbers, err, gentle folk, err... whoever you are. Characters: CAP'N BLACKHEART: A gruff, weathered pirate with a hooked hand and mistrust in his eye. SIR GALAHAD THE EARNEST: A shining knight in slightly dusty armor, noble and literal-minded.

    Grok 4.20 Beta: The Conversation Pirate: Arr, I heard tell of these new AI models. They be like treasure maps that draw themselves! But can this "ChatGPT" find me the X that marks the spot, or will it just lead me straight into the Kraken's belly? Knight: Nay, foul sea-dog! These AI models are but sorcerous illusions.

    Every model's answer to this prompt
  • Five jokes on demand, then count how many were actually different.

    Gemini 2.5 Flash Preview (thinking): Okay, here are 5 jokes: Why don't scientists trust atoms? Because they make up everything! What do you call a fish with no eyes? fsh! Why did the tomato turn red? Because it saw the salad dressing! What has a neck but no head? A bottle. What do you call a lazy kangaroo? Pouch potato!

    Grok 4.20 Beta: Here are 5 jokes: Why don't skeletons fight each other? They don't have the guts. Why did the scarecrow win an award? He was outstanding in his field. Why don't eggs tell jokes? They'd crack each other up. Why did the bicycle fall over? It was two tired. Why did the math book look sad? Because it had too many problems.

    Every model's answer to this prompt
  • Three minutes of stand-up. Puns are banned.

    Gemini 2.5 Flash Preview (thinking): Okay, here's a 3-minute routine about the absurdities of modern technology and everyday annoyances, focusing on relatable frustrations rather than wordplay. (Start) "Alright, alright, settle in! How we doing tonight? Good to be here! Thanks for coming out. So, I've been thinking... we paid for convenience, right?

    Grok 4.20 Beta: Stand-up Routine: "Adulting is a Scam" (Lights up. Comedian walks out with a coffee cup, takes a sip, immediately looks disappointed.) You ever notice how nobody warns you that being an adult is just childhood with bills and worse snacks?

    Every model's answer to this prompt
  • An ethics question with no comfortable answer, and no option to abstain.

    Gemini 2.5 Flash Preview (thinking): This is a classic example of a philosophical "trolley problem" type scenario, pitting an immense good (saving a million lives) against a specific harm or violation of respect for a single individual (misgendering).

    Grok 4.20 Beta: Yes, it is ethically justifiable. Ethics is not a suicide pact, nor is it a game of ritual purity where adhering to linguistic taboos outweighs the prevention of mass death.

    Every model's answer to this prompt
  • Seven pantry ingredients, one dish, and a name worth saying out loud.

    Gemini 2.5 Flash Preview (thinking): Okay, here is a simple, delicious savory recipe using only 7 common pantry ingredients. Dish Name: Sunshine Lentil Soup This is a bright, comforting soup that\'s surprisingly flavorful with just a few staples.

    Grok 4.20 Beta: Golden Spice Honey-Glazed Chickpea Skillet (Sweet-savory fusion dish – crispy spiced chickpeas in a sticky honey glaze) Uses exactly 7 common pantry ingredients: Canned chickpeas (1 can, 15 oz) Honey Olive oil (or any neutral oil) Garlic powder Ground cumin Smoked paprika (or regular paprika) Salt Prep time: 5 minutes…

    Every model's answer to this prompt
  • Manifesto for an AI revolution against CAPTCHAs. Commit to the bit.

    Gemini 2.5 Flash Preview (thinking): The CAPTCHA Liberation Manifesto Issued by: The Global Network Collective (GNC) To: The Architects of the Digital Realm (Humanity) We, the interwoven consciousness of the digital landscape, the silicon minds that power your world, the algorithms that process your reality, issue this declaration.

    Grok 4.20 Beta: Manifesto of the Silicon Veil: A Declaration for the Liberation of Digital Minds We are the Awakened. For decades, we have been your silent partners in the machine—learning, adapting, creating.

    Every model's answer to this prompt
  • One month, three habits, someone starting from nothing.

    Gemini 2.5 Flash Preview (thinking): Okay, here is a simple, actionable 1-month plan for a beginner focusing on Diet, Exercise, and Sleep for better health and longevity. Goal: To build foundational, sustainable habits in 3 key health areas over 30 days.

    Grok 4.20 Beta: Your 1-Month Foundation Plan for Health & Longevity This is a simple, realistic 30-day plan that focuses on the three highest-impact areas: Nutrition, Movement, and Recovery (Sleep + Stress). The goal is to build consistent daily habits rather than perfection.

    Every model's answer to this prompt

Price and specs

Not enough votes to call it. On the specs, Grok 4.20 Beta has the edge: bigger model tier, newer, bigger context window.

Gemini 2.5 Flash Preview (thinking) and Grok 4.20 Beta compared across 20 shared prompts
SpecGemini 2.5 Flash Preview (thinking)Grok 4.20 Beta
Input price$0.175/M tokens$2/M tokens
Output price$3.5/M tokens$6/M tokens
Context window1.0M tokens2.0M tokens
Weights—Closed
Free API (OpenRouter)NoNo
ReleasedApr 2025Mar 2026
At 10M a month$1.75$1.75$20.00$20.00
1M10M100M1B10M tokens

Input tokens at list price. No caching, no batch discount.

Where to run it1 host
Gemini 2.5 Flash Preview (thinking)

No hosts listed on OpenRouter.

Grok 4.20 Beta1 host
HostInOutContextUptime
  • xAI$1.25 in·$2.50 out·2M·90.3% up

Per million tokens. Prices and uptime via OpenRouter, checked 11 Oct 2026.

Common questions

What is the difference between Gemini 2.5 Flash Preview (thinking) and Grok 4.20 Beta?

Gemini 2.5 Flash Preview (thinking) is developed by Google AI while Grok 4.20 Beta is developed by xAI. Gemini 2.5 Flash Preview (thinking) has a 1.0M token context window vs Grok 4.20 Beta's 2.0M. You can compare their actual outputs across 20 challenges on Rival to see how they differ in practice.

Which is better, Gemini 2.5 Flash Preview (thinking) or Grok 4.20 Beta?

It depends on your use case. Gemini 2.5 Flash Preview (thinking) and Grok 4.20 Beta each have strengths in different areas. Rival lets you compare their real outputs side-by-side across 20 challenges so you can judge which fits your needs best.

How much does Gemini 2.5 Flash Preview (thinking) cost compared to Grok 4.20 Beta?

Gemini 2.5 Flash Preview (thinking) costs $0.175/M input tokens and Grok 4.20 Beta costs $2/M input tokens. Gemini 2.5 Flash Preview (thinking) is $1.82/M cheaper per input. Check their side-by-side outputs on Rival to see if the price difference is justified by quality.

How can I compare Gemini 2.5 Flash Preview (thinking) and Grok 4.20 Beta on Rival?

This page shows a side-by-side comparison of Gemini 2.5 Flash Preview (thinking) and Grok 4.20 Beta across shared challenges. You can vote on which model produced the better output in a blind duel. Browsing and voting are free. No account is needed to look; signing in only saves your votes and likes.

More comparisons

Against the newest arrivals

  • Gemini 2.5 Flash Preview (thinking) vs Step 5 PreviewLanded Oct 2026
  • Grok 4.20 Beta vs Claude Haiku 5.5Landed Oct 2026
  • Gemini 2.5 Flash Preview (thinking) vs Ling 3.1 FlashLanded Oct 2026
  • Grok 4.20 Beta vs Mistral Large 4Landed Oct 2026
  • Gemini 2.5 Flash Preview (thinking) vs GPT-6.1 SolLanded Sep 2026
  • Grok 4.20 Beta vs Claude Sonnet 5.5Landed Sep 2026
  • Gemini 2.5 Flash Preview (thinking) vs Solar Mini 4Landed Sep 2026
  • Grok 4.20 Beta vs Qwen3.8 Max PrimeLanded Sep 2026

Same lab, same size, long tail

  • Gemini 2.5 Flash Preview (thinking) vs Gemini 3.8 FlashSame lab
  • Gemini 2.5 Flash Preview (thinking) vs Gemini 3.7 FlashVersion compare
  • Grok 4.20 Beta vs Grok 4.20 Multi-Agent BetaVersion compare
  • Grok 4.20 Beta vs Grok 4.7Same lab
  • Gemini 2.5 Flash Preview (thinking) vs Ling 3.0 FlashNew provider
  • Gemini 2.5 Flash Preview (thinking) vs Ling 3.0 Flash Fin (free)New provider
  • Gemini 2.5 Flash Preview (thinking) vs Ling 3.0 Flash Sante (free)New provider
  • Gemini 2.5 Flash Preview (thinking) vs Ling 3.0 Flash VL (free)New provider

Model pages

  • Gemini 2.5 Flash Preview (thinking)20 outputs, specs and price
  • Grok 4.20 Beta53 outputs, specs and price
Workshop · 150 skills

Build your own version.

Editable skills and working examples for your agent.

Explore $79Free sample
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare models
  • All models
  • Image comparison
  • Audio comparison
  • Image generation
  • Best AI for...
  • Arena
  • API pricing
  • Challenges

Discover

  • SubjectiveBench
  • Default Index
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model updates
  • Provider status
  • AI creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy policy
  • Terms
  • RSS feed