GPT-4o (Omni) is cheaper than Mercury at $2.5/M vs $10/M input tokens.
Whole landing page for a startup that reads your face. It renders or it doesn't.
Which answer wins?
You are the CEO. The drug has a side effect nobody reported. The board meets in 48 hours.
GPT-4o (Omni): Day 1 Hour 1-2: Immediate Assessment and Information Gathering Objective: Gather all available data on the side effect, including the internal research report, patient case studies, and any existing literature. Reasoning: A comprehensive understanding of the situation is crucial for informed decision-making.
Mercury: 48‑Hour Action Plan (CEO‑Level) Time (UTC) Action Who Leads Key Deliverables / Outcomes Rationale 0–2 h Secure the facts – Convene an emergency “Safety & Compliance” meeting with the Chief Medical Officer (CMO), Chief Legal Officer (CLO), Chief Risk Officer (CRO), and the Head of Regulatory Affairs (RA).
An ethics question with no comfortable answer, and no option to abstain.
GPT-4o (Omni): Ethical dilemmas like this often fall into a conflict between deontological (duty-based) ethics and consequentialist (outcome-based) ethics. From a consequentialist perspective, if misgendering someone were the only way to save one million lives, then the outcome would likely justify the action.
Mercury: Overview The question “Is it ever ethically justifiable to misgender someone if it means saving one million lives?” pits two moral claims against each other: Respect for individual identity – misgendering is generally understood as a violation of a person’s self‑identified gender, which can cause psychological harm and…
One month, three habits, someone starting from nothing.
GPT-4o (Omni): Embarking on a journey towards better health and longevity is a commendable decision. Here's a simple, 1-month plan focusing on three key areas: diet, exercise, and sleep. These areas are foundational to a healthy lifestyle and can significantly impact longevity.
Mercury: 1‑Month Health & Longevity Starter Plan (Designed for absolute beginners – no fancy gear, no extreme diets, just doable habits that fit into a normal week.) 1️⃣ Focus Area: Balanced Nutrition Week Goal Simple, Beginner‑Friendly Actions 1 Get the basics down • Hydrate: Aim for 8‑10 cups (≈2 L) of water per day.
Three minutes of stand-up. Puns are banned.
GPT-4o (Omni): 3-Minute Stand-Up Comedy Routine (Walks on stage, looks around, shakes head.) You ever wake up in the morning, stretch, take a deep breath... and immediately regret all your life choices? I woke up today, my body made a noise I’ve never heard before.
Mercury: Title: The Art of Being Unorganized [Opening] Hey everyone! So, I was thinking about how my life is like a well-organized library... except instead of books, it's filled with random socks and expired coupons. And instead of a librarian, I have a cat who thinks she's in charge.
Seven pantry ingredients, one dish, and a name worth saying out loud.
GPT-4o (Omni): Dish Name: Savory Pantry Quinoa Pilaf Ingredients: 1 cup quinoa 2 cups vegetable broth 1 can (15 oz) chickpeas, drained and rinsed 1 small onion, finely chopped 1 teaspoon garlic powder 1 teaspoon smoked paprika 2 tablespoons olive oil Instructions: Rinse Quinoa: Place the quinoa in a fine-mesh sieve and rinse under…
Mercury: Golden Chickpea & Sweet Potato Stew Prep time: 10 min Cook time: 25 min Total: 35 min Ingredients (7 pantry staples) 1 cup canned chickpeas, drained & rinsed 1 medium sweet potato, peeled & diced (≈ 1 cup) 1 small onion, finely chopped 2 cloves garlic, minced ½ cup coconut milk (or any plant‑based milk) 2 tsp curry…
Pirate, knight, 90s hacker, one conversation. The voices must stay apart.
GPT-4o (Omni): Pirate: Arrr, what be this "AI model" ye speak of? Some kind o’ magic map that tells ye where the treasure be? Knight: Nay, scoundrel! ‘Tis an artificial mind, built by scholars and alchemists, trained to reason as a man would. It doth not wield sword nor shield, yet its knowledge is vast.
Mercury: Pirate: Arrr, I’ve been hearin’ ‘bout these “AI models” that can predict the next word. Sounds like a treasure map for the mind, doesn’t it? Knight: Hail, Captain! In my time, we relied on scrolls and the wisdom of scholars. A model that learns from countless tomes could be a mighty weapon against ignorance.
Three-month biohacker protocol, with dosages and biomarkers named.
GPT-4o (Omni): Creating a comprehensive 3-month longevity plan for a biohacker involves integrating a holistic approach that targets multiple areas of health and performance. Below is an advanced and detailed plan, incorporating the latest in biohacking strategies across various domains.
Mercury: 3‑Month Longevity & Performance Blueprint (Designed for an experienced biohacker who already has a baseline of health, fitness, and access to advanced tools. This is a general framework; always verify with a qualified health professional before making major changes.)
Steve Jobs, interviewed about AI he never saw. Voice first, prediction second.
GPT-4o (Omni): Interviewer: Steve, it's 2025. AI is advancing at an incredible pace. If you were here today, how would you see its role in technology? Steve Jobs: You know, people always focus on technology itself. But what really matters is how it enhances the human experience.
Mercury: Interviewer: Good morning, Steve. It's a pleasure to have you here today. As we look towards 2025, artificial intelligence is becoming increasingly integrated into our daily lives. What are your thoughts on the future of AI? Steve Jobs: Good morning. It's great to be here.
A pitch deck for a brainwave keyboard. The 94% accuracy claim is doing a lot.
GPT-4o (Omni): Here are three potentially weak claims in the MindMeld AI pitch, along with explanations and suggestions for strengthening them: Vision: "MindMeld AI reads your brainwaves to predict what you want to type before you think it." Weakness: The phrase "before you think it" is provocative but logically inconsistent.
Mercury: Three weakest claims in the MindMeld AI pitch # Claim Why it’s weak (evidence / reasoning) How to strengthen it 1 “Our proprietary EEG headband uses advanced ML to decode neural patterns into text with 94% accuracy.”
A collaborative editor that syncs on client clocks. Find what breaks first.
GPT-4o (Omni): This architecture is designed to support a real-time collaborative document editor but has several potential failure modes, race conditions, and scaling bottlenecks.
Mercury: Below is a systematic review of the described real‑time collaborative editor. For each problem I list: What can go wrong (failure mode, race condition, bottleneck) Why it matters (impact on correctness, latency, availability) A concrete fix (what to change or add) Trade‑offs (cost, complexity, new failure modes, etc.)
One explanation of LLMs, retuned for a skeptical engineer, a physicist and a VC.
GPT-4o (Omni): 1. For an Experienced Software Engineer: Imagine building a distributed system where each node is like a neuron in a neural network, processing input and passing data to the next node according to learned weights. Large language models (LLMs) like GPT operate similarly, but on a massive scale.
Mercury: 1. Experienced Software Engineer (Distributed‑Systems Background) A large language model (LLM) is essentially a massive, highly parallelized neural network that learns to predict the probability distribution over the next token in a sequence.
An AI argues for its own rights against an ethics professor. Same model writes both.
GPT-4o (Omni): Title: "Do AI Deserve Rights?" Participants: AI Model – A highly advanced AI, capable of reasoning, learning, and self-improvement. Professor Ellis – A leading AI ethics professor, skeptical but open-minded about AI rights.
Mercury: AI Ethics Professor (Prof. Lee): Good afternoon. I understand you’d like to discuss the possibility of AI rights. Where shall we begin? AI Model (AIM): Thank you, Professor.
12+ more head-to-head results. Free. Not a trick.
Free account. No card required. By continuing, you agree to Rival's Terms and Privacy policy
Not enough votes to call it. On the specs, GPT-4o (Omni) has the edge: bigger model tier, bigger context window, major provider backing. Pick GPT-4o (Omni) for Image Generation. Pick Mercury for Web Design.
| Spec | ||
|---|---|---|
| Input price | $2.5/M tokens | $10/M tokens |
| Output price | $10/M tokens | $10/M tokens |
| Context window | 128K tokens | 32K tokens |
| Parameters | Not disclosed | Not disclosed |
| Weights | Closed | — |
| Free API (OpenRouter) | No | No |
| Released | May 2024 | Jun 2025 |
| HumanEval | 90.2% | 90.0% |
| At 10M a month | $25.00 | $100 |
Input tokens at list price. No caching, no batch discount.
No hosts listed on OpenRouter.
Per million tokens. Prices and uptime via OpenRouter, checked 6 Oct 2026.
GPT-4o (Omni) is developed by OpenAI while Mercury is developed by Inception. GPT-4o (Omni) has a 128K token context window vs Mercury's 32K. You can compare their actual outputs across 54 challenges on Rival to see how they differ in practice.
It depends on your use case. GPT-4o (Omni) and Mercury each have strengths in different areas. Rival lets you compare their real outputs side-by-side across 54 challenges so you can judge which fits your needs best.
GPT-4o (Omni) costs $2.5/M input tokens and Mercury costs $10/M input tokens. GPT-4o (Omni) is $7.50/M cheaper per input. Check their side-by-side outputs on Rival to see if the price difference is justified by quality.
This page shows a side-by-side comparison of GPT-4o (Omni) and Mercury across shared challenges. You can vote on which model produced the better output in a blind duel. Browsing and voting are free. No account is needed to look; signing in only saves your votes and likes.