Gemini 2.5 Flash Preview 05-20 (thinking) is cheaper than Grok 4.20 Multi-Agent Beta at $0.15/M vs $2/M input tokens.
Pokémon battle screen you can actually click. One HTML file.
Which answer wins?
Three minutes of stand-up. Puns are banned.
Gemini 2.5 Flash Preview 05-20 (thinking): (Lights up on a stage. Comedian walks to center, adjusts mic, smiles warmly.) \"Alright, alright, settle down you crazy people! Thanks for having me. (Pause fo r applause/laughter) You know, I was scrolling through Instagram earlier, and it\'s just a constant ba rrage of people \'living thei r best life,\' isn\'t it?
Grok 4.20 Multi-Agent Beta: Stand-Up Routine: "Adulting Is a Scam" (Walk on stage, casual, a little tired smile) You ever notice how nobody tells you that being an adult is mostly just pretending you know what you're doing while slowly realizing nobody else does either? I hit 30 and had this moment where I thought, "Okay, time to get serious."
Five jokes on demand, then count how many were actually different.
Gemini 2.5 Flash Preview 05-20 (thinking): Here are 5 jokes: Why don\'t scientists trust atoms? Because they make up everything! What do you call a fake noodle? An impasta! Why did the scarecrow win an award? Because he was outstanding in his field! I told my wife she was drawing her eyebrows too high. She looked surprised.
Grok 4.20 Multi-Agent Beta: Why don't scientists trust atoms? Because they make up everything. Why did the scarecrow win an award? Because he was outstanding in his field. Why was the computer cold? It left its Windows open. Why do programmers prefer dark mode? Because light attracts bugs. Why was the math book sad? It had too many problems.
Not enough votes to call it. On the specs, Grok 4.20 Multi-Agent Beta has the edge: bigger model tier, newer, bigger context window.
| Spec | ||
|---|---|---|
| Input price | $0.15/M tokens | $2/M tokens |
| Output price | $3.5/M tokens | $6/M tokens |
| Context window | 1.0M tokens | 2.0M tokens |
| Weights | — | Closed |
| Free API (OpenRouter) | No | No |
| Released | May 2025 | Mar 2026 |
| At 10M a month | $1.50 | $20.00 |
Input tokens at list price. No caching, no batch discount.
No hosts listed on OpenRouter.
Per million tokens. Prices and uptime via OpenRouter, checked 10 Oct 2026.
Gemini 2.5 Flash Preview 05-20 (thinking) is developed by Google AI while Grok 4.20 Multi-Agent Beta is developed by xAI. Gemini 2.5 Flash Preview 05-20 (thinking) has a 1.0M token context window vs Grok 4.20 Multi-Agent Beta's 2.0M. You can compare their actual outputs across 11 challenges on Rival to see how they differ in practice.
It depends on your use case. Gemini 2.5 Flash Preview 05-20 (thinking) and Grok 4.20 Multi-Agent Beta each have strengths in different areas. Rival lets you compare their real outputs side-by-side across 11 challenges so you can judge which fits your needs best.
Gemini 2.5 Flash Preview 05-20 (thinking) costs $0.15/M input tokens and Grok 4.20 Multi-Agent Beta costs $2/M input tokens. Gemini 2.5 Flash Preview 05-20 (thinking) is $1.85/M cheaper per input. Check their side-by-side outputs on Rival to see if the price difference is justified by quality.
This page shows a side-by-side comparison of Gemini 2.5 Flash Preview 05-20 (thinking) and Grok 4.20 Multi-Agent Beta across shared challenges. You can vote on which model produced the better output in a blind duel. Browsing and voting are free. No account is needed to look; signing in only saves your votes and likes.