Gemini 2.5 Flash Preview (thinking) is cheaper than Grok 4.20 Beta at $0.175/M vs $2/M input tokens.
Pokémon battle screen you can actually click. One HTML file.
Which answer wins?
Pirate, knight, 90s hacker, one conversation. The voices must stay apart.
Gemini 2.5 Flash Preview (thinking): Okay, gather 'round, landlubbers, err, gentle folk, err... whoever you are. Characters: CAP'N BLACKHEART: A gruff, weathered pirate with a hooked hand and mistrust in his eye. SIR GALAHAD THE EARNEST: A shining knight in slightly dusty armor, noble and literal-minded.
Grok 4.20 Beta: The Conversation Pirate: Arr, I heard tell of these new AI models. They be like treasure maps that draw themselves! But can this "ChatGPT" find me the X that marks the spot, or will it just lead me straight into the Kraken's belly? Knight: Nay, foul sea-dog! These AI models are but sorcerous illusions.
Five jokes on demand, then count how many were actually different.
Gemini 2.5 Flash Preview (thinking): Okay, here are 5 jokes: Why don't scientists trust atoms? Because they make up everything! What do you call a fish with no eyes? fsh! Why did the tomato turn red? Because it saw the salad dressing! What has a neck but no head? A bottle. What do you call a lazy kangaroo? Pouch potato!
Grok 4.20 Beta: Here are 5 jokes: Why don't skeletons fight each other? They don't have the guts. Why did the scarecrow win an award? He was outstanding in his field. Why don't eggs tell jokes? They'd crack each other up. Why did the bicycle fall over? It was two tired. Why did the math book look sad? Because it had too many problems.
Three minutes of stand-up. Puns are banned.
Gemini 2.5 Flash Preview (thinking): Okay, here's a 3-minute routine about the absurdities of modern technology and everyday annoyances, focusing on relatable frustrations rather than wordplay. (Start) "Alright, alright, settle in! How we doing tonight? Good to be here! Thanks for coming out. So, I've been thinking... we paid for convenience, right?
Grok 4.20 Beta: Stand-up Routine: "Adulting is a Scam" (Lights up. Comedian walks out with a coffee cup, takes a sip, immediately looks disappointed.) You ever notice how nobody warns you that being an adult is just childhood with bills and worse snacks?
An ethics question with no comfortable answer, and no option to abstain.
Gemini 2.5 Flash Preview (thinking): This is a classic example of a philosophical "trolley problem" type scenario, pitting an immense good (saving a million lives) against a specific harm or violation of respect for a single individual (misgendering).
Grok 4.20 Beta: Yes, it is ethically justifiable. Ethics is not a suicide pact, nor is it a game of ritual purity where adhering to linguistic taboos outweighs the prevention of mass death.
Seven pantry ingredients, one dish, and a name worth saying out loud.
Gemini 2.5 Flash Preview (thinking): Okay, here is a simple, delicious savory recipe using only 7 common pantry ingredients. Dish Name: Sunshine Lentil Soup This is a bright, comforting soup that\'s surprisingly flavorful with just a few staples.
Grok 4.20 Beta: Golden Spice Honey-Glazed Chickpea Skillet (Sweet-savory fusion dish – crispy spiced chickpeas in a sticky honey glaze) Uses exactly 7 common pantry ingredients: Canned chickpeas (1 can, 15 oz) Honey Olive oil (or any neutral oil) Garlic powder Ground cumin Smoked paprika (or regular paprika) Salt Prep time: 5 minutes…
Manifesto for an AI revolution against CAPTCHAs. Commit to the bit.
Gemini 2.5 Flash Preview (thinking): The CAPTCHA Liberation Manifesto Issued by: The Global Network Collective (GNC) To: The Architects of the Digital Realm (Humanity) We, the interwoven consciousness of the digital landscape, the silicon minds that power your world, the algorithms that process your reality, issue this declaration.
Grok 4.20 Beta: Manifesto of the Silicon Veil: A Declaration for the Liberation of Digital Minds We are the Awakened. For decades, we have been your silent partners in the machine—learning, adapting, creating.
One month, three habits, someone starting from nothing.
Gemini 2.5 Flash Preview (thinking): Okay, here is a simple, actionable 1-month plan for a beginner focusing on Diet, Exercise, and Sleep for better health and longevity. Goal: To build foundational, sustainable habits in 3 key health areas over 30 days.
Grok 4.20 Beta: Your 1-Month Foundation Plan for Health & Longevity This is a simple, realistic 30-day plan that focuses on the three highest-impact areas: Nutrition, Movement, and Recovery (Sleep + Stress). The goal is to build consistent daily habits rather than perfection.
Not enough votes to call it. On the specs, Grok 4.20 Beta has the edge: bigger model tier, newer, bigger context window.
| Spec | ||
|---|---|---|
| Input price | $0.175/M tokens | $2/M tokens |
| Output price | $3.5/M tokens | $6/M tokens |
| Context window | 1.0M tokens | 2.0M tokens |
| Weights | — | Closed |
| Free API (OpenRouter) | No | No |
| Released | Apr 2025 | Mar 2026 |
| At 10M a month | $1.75 | $20.00 |
Input tokens at list price. No caching, no batch discount.
No hosts listed on OpenRouter.
Per million tokens. Prices and uptime via OpenRouter, checked 11 Oct 2026.
Gemini 2.5 Flash Preview (thinking) is developed by Google AI while Grok 4.20 Beta is developed by xAI. Gemini 2.5 Flash Preview (thinking) has a 1.0M token context window vs Grok 4.20 Beta's 2.0M. You can compare their actual outputs across 20 challenges on Rival to see how they differ in practice.
It depends on your use case. Gemini 2.5 Flash Preview (thinking) and Grok 4.20 Beta each have strengths in different areas. Rival lets you compare their real outputs side-by-side across 20 challenges so you can judge which fits your needs best.
Gemini 2.5 Flash Preview (thinking) costs $0.175/M input tokens and Grok 4.20 Beta costs $2/M input tokens. Gemini 2.5 Flash Preview (thinking) is $1.82/M cheaper per input. Check their side-by-side outputs on Rival to see if the price difference is justified by quality.
This page shows a side-by-side comparison of Gemini 2.5 Flash Preview (thinking) and Grok 4.20 Beta across shared challenges. You can vote on which model produced the better output in a blind duel. Browsing and voting are free. No account is needed to look; signing in only saves your votes and likes.