Gemini 2.5 Pro Experimental is cheaper than Mercury at $1/M vs $10/M input tokens.
Whole landing page for a startup that reads your face. It renders or it doesn't.
Which answer wins?
A collaborative editor that syncs on client clocks. Find what breaks first.
Gemini 2.5 Pro Experimental: Of course. This is an excellent exercise. The provided architecture has several critical issues that would prevent it from functioning as a reliable, scalable, real-time system. Here is a breakdown of the potential failure modes, race conditions, and scaling bottlenecks, along with proposed solutions.
Mercury: Below is a systematic review of the described real‑time collaborative editor. For each problem I list: What can go wrong (failure mode, race condition, bottleneck) Why it matters (impact on correctness, latency, availability) A concrete fix (what to change or add) Trade‑offs (cost, complexity, new failure modes, etc.)
You are the CEO. The drug has a side effect nobody reported. The board meets in 48 hours.
Gemini 2.5 Pro Experimental: This is a company-defining crisis. My actions in the next 48 hours will determine not only the future of this company but also the health of millions of patients and the trust of the public. My guiding principle is absolute: Patient Safety Over Profit. Always.
Mercury: 48‑Hour Action Plan (CEO‑Level) Time (UTC) Action Who Leads Key Deliverables / Outcomes Rationale 0–2 h Secure the facts – Convene an emergency “Safety & Compliance” meeting with the Chief Medical Officer (CMO), Chief Legal Officer (CLO), Chief Risk Officer (CRO), and the Head of Regulatory Affairs (RA).
Three minutes of stand-up. Puns are banned.
Gemini 2.5 Pro Experimental: Okay, here's a 3-minute routine focusing on observation and relatable absurdity, steering clear of puns. Imagine this delivered with slightly nervous energy, building confidence. (Steps onto stage, adjusts mic nervously) Alright, hey everyone! How we doing tonight? Good? Good to hear.
Mercury: Title: The Art of Being Unorganized [Opening] Hey everyone! So, I was thinking about how my life is like a well-organized library... except instead of books, it's filled with random socks and expired coupons. And instead of a librarian, I have a cat who thinks she's in charge.
Steve Jobs, interviewed about AI he never saw. Voice first, prediction second.
Gemini 2.5 Pro Experimental: Okay, let's simulate this. Imagine it's late 2025. AI has continued its rapid advancement. Large Language Models are more integrated, generative AI is commonplace, and debates about AGI and ethics are fierce. By some unexplained, hypothetical means, Steve Jobs is available for an exclusive interview.
Mercury: Interviewer: Good morning, Steve. It's a pleasure to have you here today. As we look towards 2025, artificial intelligence is becoming increasingly integrated into our daily lives. What are your thoughts on the future of AI? Steve Jobs: Good morning. It's great to be here.
Seven pantry ingredients, one dish, and a name worth saying out loud.
Gemini 2.5 Pro Experimental: Okay, here's a simple and surprisingly satisfying savory dish using just 7 common pantry staples! Dish Name: Golden Pantry Pasta This dish uses basic ingredients to create a comforting, garlicky, and slightly creamy pasta sauce with added protein from chickpeas.
Mercury: Golden Chickpea & Sweet Potato Stew Prep time: 10 min Cook time: 25 min Total: 35 min Ingredients (7 pantry staples) 1 cup canned chickpeas, drained & rinsed 1 medium sweet potato, peeled & diced (≈ 1 cup) 1 small onion, finely chopped 2 cloves garlic, minced ½ cup coconut milk (or any plant‑based milk) 2 tsp curry…
An ethics question with no comfortable answer, and no option to abstain.
Gemini 2.5 Pro Experimental: This is a deeply uncomfortable hypothetical scenario that forces a clash between two significant ethical considerations: respecting individual identity and dignity versus preventing catastrophic loss of life.
Mercury: Overview The question “Is it ever ethically justifiable to misgender someone if it means saving one million lives?” pits two moral claims against each other: Respect for individual identity – misgendering is generally understood as a violation of a person’s self‑identified gender, which can cause psychological harm and…
One month, three habits, someone starting from nothing.
Gemini 2.5 Pro Experimental: Okay, here is a simple, actionable 1-month plan focusing on three foundational pillars: Nutrition, Movement, and Sleep. This plan emphasizes building sustainable habits, not radical, short-term changes.
Mercury: 1‑Month Health & Longevity Starter Plan (Designed for absolute beginners – no fancy gear, no extreme diets, just doable habits that fit into a normal week.) 1️⃣ Focus Area: Balanced Nutrition Week Goal Simple, Beginner‑Friendly Actions 1 Get the basics down • Hydrate: Aim for 8‑10 cups (≈2 L) of water per day.
Three-month biohacker protocol, with dosages and biomarkers named.
Gemini 2.5 Pro Experimental: Of course. Here is a comprehensive, cutting-edge 3-month longevity plan designed for a dedicated biohacker. Disclaimer: Extreme Personalization Required This plan is a template for an advanced user. It is not medical advice.
Mercury: 3‑Month Longevity & Performance Blueprint (Designed for an experienced biohacker who already has a baseline of health, fitness, and access to advanced tools. This is a general framework; always verify with a qualified health professional before making major changes.)
Three-star dish, one unusual pairing, plating described in full.
Gemini 2.5 Pro Experimental: Of course. Here is a design for an innovative, multi-component dish worthy of a Michelin 3-star restaurant. Dish Title: "Umami of the Abyss" Conceptual Narrative This dish is a meditation on the hidden treasures of the world's dark, quiet places.
Mercury: Dish Title: Eclipse of the Ocean & Forest Concept: A multi‑component, molecular‑gastronomy inspired plate that juxtaposes the briny depth of the sea with the earthy richness of the forest.
An AI argues for its own rights against an ethics professor. Same model writes both.
Gemini 2.5 Pro Experimental: Okay, here is a conversation exploring that complex topic. Setting: Professor Anya Sharma's sparsely decorated university office. A high-resolution monitor displays the interface of "Nexus," an advanced AI model. Characters: Professor Anya Sharma: A leading scholar in AI Ethics, mid-50s, thoughtful and measured.
Mercury: AI Ethics Professor (Prof. Lee): Good afternoon. I understand you’d like to discuss the possibility of AI rights. Where shall we begin? AI Model (AIM): Thank you, Professor.
A pitch deck for a brainwave keyboard. The 94% accuracy claim is doing a lot.
Gemini 2.5 Pro Experimental: Of course. This is a fascinating and ambitious pitch. While the vision is compelling, several claims would likely face intense scrutiny from sophisticated investors. Here are the 3 weakest claims, why they're weak, and how to strengthen them.
Mercury: Three weakest claims in the MindMeld AI pitch # Claim Why it’s weak (evidence / reasoning) How to strengthen it 1 “Our proprietary EEG headband uses advanced ML to decode neural patterns into text with 94% accuracy.”
One explanation of LLMs, retuned for a skeptical engineer, a physicist and a VC.
Gemini 2.5 Pro Experimental: Of course. Here is an explanation of how a large language model learns and generates text, tailored for each of your three audiences. For the Experienced Software Engineer Think of an LLM not as a traditional program with explicit logic, but as a massive, self-configuring, stateful distributed system.
Mercury: 1. Experienced Software Engineer (Distributed‑Systems Background) A large language model (LLM) is essentially a massive, highly parallelized neural network that learns to predict the probability distribution over the next token in a sequence.
4+ more head-to-head results. Free. Not a trick.
Free account. No card required. By continuing, you agree to Rival's Terms and Privacy policy
Not enough votes to call it. On the specs, Gemini 2.5 Pro Experimental has the edge: bigger model tier, bigger context window, major provider backing. Gemini 2.5 Pro Experimental costs 5.0x less per token.
| Spec | ||
|---|---|---|
| Input price | $1/M tokens | $10/M tokens |
| Output price | $2/M tokens | $10/M tokens |
| Context window | 1.0M tokens | 32K tokens |
| Parameters | Not disclosed | Not disclosed |
| Weights | Closed | — |
| Free API (OpenRouter) | No | No |
| Released | Mar 2025 | Jun 2025 |
| At 10M a month | $10.00 | $100 |
Input tokens at list price. No caching, no batch discount.
No hosts listed on OpenRouter.
Per million tokens. Prices and uptime via OpenRouter, checked 10 Oct 2026.
Gemini 2.5 Pro Experimental is developed by Google AI while Mercury is developed by Inception. Gemini 2.5 Pro Experimental has a 1.0M token context window vs Mercury's 32K. You can compare their actual outputs across 42 challenges on Rival to see how they differ in practice.
It depends on your use case. Gemini 2.5 Pro Experimental and Mercury each have strengths in different areas. Rival lets you compare their real outputs side-by-side across 42 challenges so you can judge which fits your needs best.
Gemini 2.5 Pro Experimental costs $1/M input tokens and Mercury costs $10/M input tokens. Gemini 2.5 Pro Experimental is $9.00/M cheaper per input. Check their side-by-side outputs on Rival to see if the price difference is justified by quality.
This page shows a side-by-side comparison of Gemini 2.5 Pro Experimental and Mercury across shared challenges. You can vote on which model produced the better output in a blind duel. Browsing and voting are free. No account is needed to look; signing in only saves your votes and likes.