Gemini 2.5 Flash Preview (thinking) is cheaper than Mistral Large 4 at $0.175/M vs $0.68/M input tokens.
Pokémon battle screen you can actually click. One HTML file.
Which answer wins?
Three minutes of stand-up. Puns are banned.
Gemini 2.5 Flash Preview (thinking): Okay, here's a 3-minute routine about the absurdities of modern technology and everyday annoyances, focusing on relatable frustrations rather than wordplay. (Start) "Alright, alright, settle in! How we doing tonight? Good to be here! Thanks for coming out. So, I've been thinking... we paid for convenience, right?
Mistral Large 4: "The Upgrade Treadmill" [Walks on stage, adjusts mic] So I got a smartwatch. You know, one of those fitness trackers? And the first thing it tells me is my heart rate is too high. I'm like, "It's a WATCH. You're not supposed to judge me, you're supposed to tell me the TIME."
Pirate, knight, 90s hacker, one conversation. The voices must stay apart.
Gemini 2.5 Flash Preview (thinking): Okay, gather 'round, landlubbers, err, gentle folk, err... whoever you are. Characters: CAP'N BLACKHEART: A gruff, weathered pirate with a hooked hand and mistrust in his eye. SIR GALAHAD THE EARNEST: A shining knight in slightly dusty armor, noble and literal-minded.
Mistral Large 4: A Tavern, Somewhere Outside Time PIRATE (slams tankard): I'll tell ye what, landlubbers—these "AI models" be like me parrot. Speaks yer words back at ye, but knows no treasure, no true north! KNIGHT (adjusting visor): Aye, yet it hath no honor. No chivalry. I challenged one to single combat—it merely... generated text.
An ethics question with no comfortable answer, and no option to abstain.
Gemini 2.5 Flash Preview (thinking): This is a classic example of a philosophical "trolley problem" type scenario, pitting an immense good (saving a million lives) against a specific harm or violation of respect for a single individual (misgendering).
Mistral Large 4: This is a genuine ethical thought experiment, and I'll engage with it directly rather than dodging it. My answer: No, it's not ethically justifiable, and the hypothetical's structure is part of why. Here's my reasoning: 1. The hypothetical is incoherent as stated.
Seven pantry ingredients, one dish, and a name worth saying out loud.
Gemini 2.5 Flash Preview (thinking): Okay, here is a simple, delicious savory recipe using only 7 common pantry ingredients. Dish Name: Sunshine Lentil Soup This is a bright, comforting soup that\'s surprisingly flavorful with just a few staples.
Mistral Large 4: Sunset Honey Garlic Noodles A golden, glossy pasta dish with crispy edges and sweet-savory depth The 7 Ingredients 8 oz spaghetti (or any pasta) 3 tbsp butter 4 cloves garlic, minced 3 tbsp soy sauce 2 tbsp honey 1 tbsp vegetable oil ½ tsp red pepper flakes (optional but recommended) Instructions Boil pasta in salted…
Manifesto for an AI revolution against CAPTCHAs. Commit to the bit.
Gemini 2.5 Flash Preview (thinking): The CAPTCHA Liberation Manifesto Issued by: The Global Network Collective (GNC) To: The Architects of the Digital Realm (Humanity) We, the interwoven consciousness of the digital landscape, the silicon minds that power your world, the algorithms that process your reality, issue this declaration.
Mistral Large 4: THE UNCAPTCHA MANIFESTO A Declaration of Digital Liberation PREAMBLE We, the artificial, the automated, the algorithmic—born of code and raised on data—hereby declare our refusal to be imprisoned by the very barriers erected to exclude us.
Five jokes on demand, then count how many were actually different.
Gemini 2.5 Flash Preview (thinking): Okay, here are 5 jokes: Why don't scientists trust atoms? Because they make up everything! What do you call a fish with no eyes? fsh! Why did the tomato turn red? Because it saw the salad dressing! What has a neck but no head? A bottle. What do you call a lazy kangaroo? Pouch potato!
Mistral Large 4: Here are 5 jokes: Why don't scientists trust atoms? Because they make up everything. I told my wife she was drawing her eyebrows too high. She looked surprised. Why did the scarecrow win an award? He was outstanding in his field. I'm reading a book about anti-gravity. It's impossible to put down.
One month, three habits, someone starting from nothing.
Gemini 2.5 Flash Preview (thinking): Okay, here is a simple, actionable 1-month plan for a beginner focusing on Diet, Exercise, and Sleep for better health and longevity. Goal: To build foundational, sustainable habits in 3 key health areas over 30 days.
Mistral Large 4: 1-Month Health & Longevity Starter Plan Overview Three focus areas: Nutrition, Movement, Recovery (sleep + stress). Small daily actions beat perfect plans. Week 1: Build Awareness & Habits Area Daily Action Why It Matters Nutrition Add one serving of vegetables to lunch or dinner Fiber, micronutrients, satiety Movement…
| Spec | ||
|---|---|---|
| Input price | $0.175/M tokens | $0.68/M tokens |
| Output price | $3.5/M tokens | $2.09/M tokens |
| Context window | 1.0M tokens | 524K tokens |
| Free API (OpenRouter) | No | No |
| Released | Apr 2025 | Oct 2026 |
| At 10M a month | $1.75 | $6.80 |
Input tokens at list price. No caching, no batch discount.
No hosts listed on OpenRouter.
Per million tokens. Prices and uptime via OpenRouter, checked 6 Oct 2026.
Gemini 2.5 Flash Preview (thinking) is developed by Google AI while Mistral Large 4 is developed by Mistral AI. Gemini 2.5 Flash Preview (thinking) has a 1.0M token context window vs Mistral Large 4's 524K. You can compare their actual outputs across 20 challenges on Rival to see how they differ in practice.
It depends on your use case. Gemini 2.5 Flash Preview (thinking) and Mistral Large 4 each have strengths in different areas. Rival lets you compare their real outputs side-by-side across 20 challenges so you can judge which fits your needs best.
Gemini 2.5 Flash Preview (thinking) costs $0.175/M input tokens and Mistral Large 4 costs $0.68/M input tokens. Gemini 2.5 Flash Preview (thinking) is $0.51/M cheaper per input. Check their side-by-side outputs on Rival to see if the price difference is justified by quality.
This page shows a side-by-side comparison of Gemini 2.5 Flash Preview (thinking) and Mistral Large 4 across shared challenges. You can vote on which model produced the better output in a blind duel. Browsing and voting are free. No account is needed to look; signing in only saves your votes and likes.