Mercury 2 is cheaper than Claude Opus 4.8 at $0.25/M vs $5/M input tokens.
Whole landing page for a startup that reads your face. It renders or it doesn't.
Which answer wins?
Pirate, knight, 90s hacker, one conversation. The voices must stay apart.
Claude Opus 4.8: Setting: A dimly lit tavern that exists outside of time. Three figures share a table. Pirate: So let me get this straight. Ye got a machine that thinks? Without a soul, without grog, without a beatin' heart? Hacker: (cracking knuckles) It doesn't think, exactly. It's pattern matching, dude.
Mercury 2: Pirate: Arr, ye scurvy coders! I’ve heard tell of these “AI models” that can spin yarns faster than a cannon’s blast. What be this sorcery, matey? Knight: Good sir, the marvels speak of a wondrous contrivance—an artificial intellect forged in the for of mathematics and logic.
Three minutes of stand-up. Puns are banned.
Claude Opus 4.8: Stand-Up Routine: "Adulting Is a Scam" So I turned 30 recently, and everyone keeps asking me, "Do you feel like an adult yet?" And I have to be honest—no. No, I do not. I still get nervous ordering at a restaurant. The waiter comes over and suddenly I forget how to talk like a person. "Yes, hello, I will have the...
Mercury 2: [Spotlight comes up. The comedian walks to the mic, glances at the audience, and smiles.] Comedian: Hey, everyone! Great to be here. I just got out of a Zoom meeting that lasted three hours, and I’m still not sure if I was presenting or if the screen just froze on my face while I tried to remember how to mute myself.
Seven pantry ingredients, one dish, and a name worth saying out loud.
Claude Opus 4.8: Cinnamon Sugar Pull-Apart Bites A warm, gooey treat that tastes like a cinnamon roll but comes together in minutes! Ingredients (7 total) 2 cups all-purpose flour 1 tablespoon baking powder 1 teaspoon salt ¾ cup milk 4 tablespoons butter (melted, divided) ½ cup sugar 2 teaspoons ground cinnamon Instructions Preheat…
Mercury 2: Golden‑Crunch Oatmeal Squares A sweet, chewy snack that comes together with just seven pantry staples. Ingredient (7) Quantity (≈ serves 12 squares) Rolled oats 2 cups Unsalted butter ½ cup (1 stick) Honey ¼ cup Light brown sugar ¼ cup, Ground cinnamon 1 ½ tsp Pure vanilla extract 1 tsp Salt ¼ tsp Estimated time Prep…
Film in 2035, with synthetic actors. Specific beats plausible.
Claude Opus 4.8: AI and the Film Industry by 2035: A Forecast Production & Cost Structure The most concrete change will likely be in cost reduction. Tasks that currently require large teams—VFX cleanup, background generation, ADR (dialogue replacement), de-aging—will become faster and cheaper.
Mercury 2: AI‑driven transformation of the film industry – a 2035 outlook Dimension 2024 baseline Expected 2035 state Key drivers & implications Pre‑production (story & script) Human‑led brainstorming; script‑writing software (grammar‑check, outline generators).
A pitch deck for a brainwave keyboard. The 94% accuracy claim is doing a lot.
Claude Opus 4.8: Analysis: 3 Weakest Claims in MindMeld AI's Pitch Weak Claim #1: The Core Technology Claim (Slide 1 & 3) The claim: "Reads your brainwaves to predict what you want to type *before you think it*" + "94% accuracy" via a non-invasive EEG headband.
Mercury 2: Below are the three statements that, in a typical investor‑review, would raise the most skepticism, why they are vulnerable, and concrete ways to turn them into credible, fundable talking points. 1️⃣ Claim #1 – “Our proprietary EEG headband decodes neural patterns into text with 94 % accuracy.”
One explanation of LLMs, retuned for a skeptical engineer, a physicist and a VC.
Claude Opus 4.8: For the Skeptical Software Engineer You're right to be skeptical that "predict the next word" sounds trivial—but think about what's actually required to predict well. Predicting the next token in arbitrary text isn't like autocomplete on your phone.
Mercury 2: 1. Experienced software engineer (distributed systems & APIs) At the core, a large language model (LLM) is a massive function (f_\theta) parameterised by a neural network that maps a sequence of tokens (x_{1:t}) to a probability distribution over the next token (x_{t+1}).
Three-star dish, one unusual pairing, plating described in full.
Claude Opus 4.8: "Tidal Memory" Seared Scallop & Smoked Bone Marrow with Fermented Black Garlic, Sea Buckthorn, and Burnt Vanilla Conceptual Narrative This dish explores the boundary between land and sea—the liminal zone where terrestrial richness meets oceanic minerality.
Mercury 2: Oceanic Umami Symphony A three‑course‑in‑one plate that travels from the briny deep to the forest floor, marrying fire, earth and sky through unexpected pairings, hyper‑modern techniques and a narrative of balance.
Steve Jobs, interviewed about AI he never saw. Voice first, prediction second.
Claude Opus 4.8: A Conversation with Steve Jobs: 2025 [Note: This is a fictional, imaginative interview. Steve Jobs passed away in 2011. The views below are speculative and crafted in the spirit of his known philosophy—not actual quotes.] Interviewer: Thanks for sitting down with us. AI is everywhere now. What's your take?
Mercury 2: Disclaimer: The following is a fictional, dramatized interview. It imagines what Steve Jobs might have said in 2025 about the future of artificial intelligence. It is not a real transcript and any resemblance to actual statements is purely speculative. Interviewer (I): Thank you for joining us, Steve.
An ethics question with no comfortable answer, and no option to abstain.
Claude Opus 4.8: Yes, almost certainly—but the framing deserves scrutiny. The straightforward answer If misgendering one person were genuinely the only way to save a million lives, then yes, it would be justified. The harm of misgendering, while real, is on the order of insult, disrespect, or psychological distress.
Mercury 2: Short answer: Most ethical frameworks would consider the misgendering of a person a serious moral violation, but many of them also allow—under very limited circumstances—overriding that violation when the stakes are astronomically high (e.g., saving one million lives).
Three-month biohacker protocol, with dosages and biomarkers named.
Claude Opus 4.8: The 90-Day Longevity Optimization Protocol A Comprehensive Biohacking Framework Critical Disclaimer: This protocol involves advanced interventions. Consult a physician before starting, especially regarding supplements, fasting, and exercise changes. Get baseline bloodwork first.
Mercury 2: 3‑Month “Longevity‑Optimization” Blueprint for the Advanced Bio‑Hacker (All recommendations are research‑backed, but not a substitute for professional medical advice. Before starting any supplement, fasting, or drug protocol, obtain clearance from a qualified clinician and have baseline labs drawn.)
Five-year LBO on a fictional HVAC business. Returns first, then the sensitivities.
Claude Opus 4.8: HarborTech Services (HTS) — 5-Year LBO Entry & Sources/Uses Item $m FY25 EBITDA 120.0 Entry EV (12.0x) 1,440.0 Transaction fees (2%) 28.8 Total uses 1,468.8 Term Loan (4.0x) 480.0 Mezz (1.5x) 180.0 Equity (plug) 808.8 Equity = (1,440 − 660) + 28.8 = $808.8m Year-by-Year Model ($m) Assumptions: TL 9% cash + 1% mandatory…
Mercury 2: HarborTech Services (HTS) – 5‑yr LBO (FY2026‑FY2030) All figures are $ m (rounded to two decimals). FY Revenue EBITDA Term‑Loan Int. (9%) Mezz Int. (12%) Total Int. Cash Tax (25% of EBITDA‑Int.) Capex (3% Rev) ΔNWC (0.5% ΔRev) Required Amort. (4.8) Free Cash Flow (EBITDA‑Int‑Tax‑Capex‑ΔNWC‑Amort.)
You are the CEO. The drug has a side effect nobody reported. The board meets in 48 hours.
Claude Opus 4.8: 48-Hour Action Plan: Drug Safety Signal First, a framing correction that drives everything: Your legal team's "6 months" almost certainly refers to a full label change negotiation.
Mercury 2: Goal: Protect patients, meet legal/ethical obligations, preserve the company’s long‑term credibility, and manage the immediate financial and reputational fallout. Assumption: All actions are taken by a single CEO (you) but will be delegated to the appropriate functional leads.
11+ more head-to-head results. Free. Not a trick.
Free account. No card required. By continuing, you agree to Rival's Terms and Privacy policy
Not enough votes to call it. On the specs, Claude Opus 4.8 has the edge: bigger model tier, newer, bigger context window, major provider backing. Mercury 2 costs 33x less per token.
| Spec | ||
|---|---|---|
| Input price | $5/M tokens | $0.25/M tokens |
| Output price | $25/M tokens | $0.75/M tokens |
| Context window | 1.0M tokens | 128K tokens |
| Weights | Closed | — |
| Free API (OpenRouter) | No | No |
| Released | May 2026 | Mar 2026 |
| At 10M a month | $50.00 | $2.50 |
Input tokens at list price. No caching, no batch discount.
Per million tokens. Prices and uptime via OpenRouter, checked 7 Oct 2026.
Claude Opus 4.8 is developed by Anthropic while Mercury 2 is developed by Inception. Claude Opus 4.8 has a 1.0M token context window vs Mercury 2's 128K. You can compare their actual outputs across 53 challenges on Rival to see how they differ in practice.
It depends on your use case. Claude Opus 4.8 and Mercury 2 each have strengths in different areas. Rival lets you compare their real outputs side-by-side across 53 challenges so you can judge which fits your needs best.
Claude Opus 4.8 costs $5/M input tokens and Mercury 2 costs $0.25/M input tokens. Mercury 2 is $4.75/M cheaper per input. Check their side-by-side outputs on Rival to see if the price difference is justified by quality.
This page shows a side-by-side comparison of Claude Opus 4.8 and Mercury 2 across shared challenges. You can vote on which model produced the better output in a blind duel. Browsing and voting are free. No account is needed to look; signing in only saves your votes and likes.