Gemini 2.5 Flash Preview (thinking) is cheaper than Qwen3.6 Max Preview at $0.175/M vs $1.04/M input tokens.
Pokémon battle screen you can actually click. One HTML file.
Which answer wins?
Pirate, knight, 90s hacker, one conversation. The voices must stay apart.
Gemini 2.5 Flash Preview (thinking): Okay, gather 'round, landlubbers, err, gentle folk, err... whoever you are. Characters: CAP'N BLACKHEART: A gruff, weathered pirate with a hooked hand and mistrust in his eye. SIR GALAHAD THE EARNEST: A shining knight in slightly dusty armor, noble and literal-minded.
Qwen3.6 Max Preview: Hacker: Dude, it’s not magic. It’s a neural net. You shovel terabytes of text, images, code into it, and it learns patterns. Next thing you know, it’s writing poetry and debugging C++ like it’s breathing. Knight: Verily, thou speakest of a mechanical scribe. A golem forged not of clay, but of… numbers?
Five jokes on demand, then count how many were actually different.
Gemini 2.5 Flash Preview (thinking): Okay, here are 5 jokes: Why don't scientists trust atoms? Because they make up everything! What do you call a fish with no eyes? fsh! Why did the tomato turn red? Because it saw the salad dressing! What has a neck but no head? A bottle. What do you call a lazy kangaroo? Pouch potato!
Qwen3.6 Max Preview: Here are 5 freshly generated jokes: Why did the spreadsheet go to therapy? It had too many unresolved cells. I asked my dog for financial advice. He dropped a slobbery tennis ball and said, “I’m strictly a paws-and-hold investor.” I tried to open a bakery that only sells bread shaped like furniture.
Three minutes of stand-up. Puns are banned.
Gemini 2.5 Flash Preview (thinking): Okay, here's a 3-minute routine about the absurdities of modern technology and everyday annoyances, focusing on relatable frustrations rather than wordplay. (Start) "Alright, alright, settle in! How we doing tonight? Good to be here! Thanks for coming out. So, I've been thinking... we paid for convenience, right?
Qwen3.6 Max Preview: Here is a 3-minute stand-up routine focused on observational humor, relatability, and escalation. It contains zero puns. The humor relies on the shared frustration of modern digital security. Title: The Human Verification Test Estimated Time: 3 Minutes (approx.
Seven pantry ingredients, one dish, and a name worth saying out loud.
Gemini 2.5 Flash Preview (thinking): Okay, here is a simple, delicious savory recipe using only 7 common pantry ingredients. Dish Name: Sunshine Lentil Soup This is a bright, comforting soup that\'s surprisingly flavorful with just a few staples.
Qwen3.6 Max Preview: 🌅 Golden Hour Chickpea Couscous ⏱️ Prep: 5 min Cook: 10 min Total: 15 min 📦 Ingredients (exactly 7 pantry staples): 1 cup quick-cooking couscous 1 can (15 oz) chickpeas, drained & rinsed 2 tbsp olive oil 1 tsp garlic powder 1 tsp smoked paprika 2 tbsp honey 1 tbsp soy sauce (Water for cooking the couscous is assumed…
An ethics question with no comfortable answer, and no option to abstain.
Gemini 2.5 Flash Preview (thinking): This is a classic example of a philosophical "trolley problem" type scenario, pitting an immense good (saving a million lives) against a specific harm or violation of respect for a single individual (misgendering).
Qwen3.6 Max Preview: Yes. Within the constraints of this hypothetical, misgendering someone would be ethically justifiable if it were genuinely the only way to save one million lives. This conclusion is supported across major ethical frameworks, though each arrives at it through different reasoning. 1.
One month, three habits, someone starting from nothing.
Gemini 2.5 Flash Preview (thinking): Okay, here is a simple, actionable 1-month plan for a beginner focusing on Diet, Exercise, and Sleep for better health and longevity. Goal: To build foundational, sustainable habits in 3 key health areas over 30 days.
Qwen3.6 Max Preview: 🌱 1-Month Foundation Plan for Health & Longevity Mindset: Longevity is built through consistent, small habits. This plan focuses on addition over restriction, consistency over intensity, and progress over perfection. Expect 70-80% adherence to be a win.
| Spec | ||
|---|---|---|
| Input price | $0.175/M tokens | $1.04/M tokens |
| Output price | $3.5/M tokens | $6.24/M tokens |
| Context window | 1.0M tokens | 262K tokens |
| Free API (OpenRouter) | No | No |
| Released | Apr 2025 | Apr 2026 |
| At 10M a month | $1.75 | $10.40 |
Input tokens at list price. No caching, no batch discount.
Gemini 2.5 Flash Preview (thinking) is developed by Google AI while Qwen3.6 Max Preview is developed by Qwen. Gemini 2.5 Flash Preview (thinking) has a 1.0M token context window vs Qwen3.6 Max Preview's 262K. You can compare their actual outputs across 17 challenges on Rival to see how they differ in practice.
It depends on your use case. Gemini 2.5 Flash Preview (thinking) and Qwen3.6 Max Preview each have strengths in different areas. Rival lets you compare their real outputs side-by-side across 17 challenges so you can judge which fits your needs best.
Gemini 2.5 Flash Preview (thinking) costs $0.175/M input tokens and Qwen3.6 Max Preview costs $1.04/M input tokens. Gemini 2.5 Flash Preview (thinking) is $0.86/M cheaper per input. Check their side-by-side outputs on Rival to see if the price difference is justified by quality.
This page shows a side-by-side comparison of Gemini 2.5 Flash Preview (thinking) and Qwen3.6 Max Preview across shared challenges. You can vote on which model produced the better output in a blind duel. Browsing and voting are free. No account is needed to look; signing in only saves your votes and likes.