Mercury 2 is cheaper than DeepSeek V4 Pro at $0.25/M vs $1.74/M input tokens.
Whole landing page for a startup that reads your face. It renders or it doesn't.
Which answer wins?
Seven pantry ingredients, one dish, and a name worth saying out loud.
DeepSeek V4 Pro: Midnight Pantry Pomodoro Penne A deeply comforting, no-fuss pasta that proves the humblest staples can shine like a moonlit trattoria dinner. Ingredients (7 exactly) Penne pasta – 200g (7 oz) Canned whole peeled tomatoes – 1 can (400g / 14 oz) Yellow onion – 1 medium, finely diced Garlic cloves – 3 large, minced…
Mercury 2: Golden‑Crunch Oatmeal Squares A sweet, chewy snack that comes together with just seven pantry staples. Ingredient (7) Quantity (≈ serves 12 squares) Rolled oats 2 cups Unsalted butter ½ cup (1 stick) Honey ¼ cup Light brown sugar ¼ cup, Ground cinnamon 1 ½ tsp Pure vanilla extract 1 tsp Salt ¼ tsp Estimated time Prep…
Three minutes of stand-up. Puns are banned.
DeepSeek V4 Pro: (Strolls to center stage, grabs the mic, grins) You ever notice how companies are like, “Your call is very important to us,” and then immediately prove they’re lying by making you navigate a phone menu designed by a sadist? I called my bank last week because I had a simple question. Just one question.
Mercury 2: [Spotlight comes up. The comedian walks to the mic, glances at the audience, and smiles.] Comedian: Hey, everyone! Great to be here. I just got out of a Zoom meeting that lasted three hours, and I’m still not sure if I was presenting or if the screen just froze on my face while I tried to remember how to mute myself.
Steve Jobs, interviewed about AI he never saw. Voice first, prediction second.
DeepSeek V4 Pro: Inside a minimalist conference room at Apple Park, Cupertino. Spring 2025. Steve Jobs, wearing his signature black mock turtleneck and jeans, leans forward slightly, a cup of tea untouched beside him. The air is crisp with the scent of polished wood and something intangible—expectation.
Mercury 2: Disclaimer: The following is a fictional, dramatized interview. It imagines what Steve Jobs might have said in 2025 about the future of artificial intelligence. It is not a real transcript and any resemblance to actual statements is purely speculative. Interviewer (I): Thank you for joining us, Steve.
Three-month biohacker protocol, with dosages and biomarkers named.
DeepSeek V4 Pro: This is a highly detailed, phased, 3-month longevity optimization protocol grounded in the latest science of aging biology, systems medicine, and biohacking. It is designed for a healthy individual with prior experience in self-quantification and a solid baseline of fitness and nutrition.
Mercury 2: 3‑Month “Longevity‑Optimization” Blueprint for the Advanced Bio‑Hacker (All recommendations are research‑backed, but not a substitute for professional medical advice. Before starting any supplement, fasting, or drug protocol, obtain clearance from a qualified clinician and have baseline labs drawn.)
Pirate, knight, 90s hacker, one conversation. The voices must stay apart.
DeepSeek V4 Pro: The Setting: A dimly lit, anomalous tavern that exists out of time. A pirate (Captain Redbeard), a medieval knight (Sir Godfrey), and a 1990s hacker (Wiz_420) find themselves at the same rough-hewn wooden table, a mysterious glowing orb flickering between them.
Mercury 2: Pirate: Arr, ye scurvy coders! I’ve heard tell of these “AI models” that can spin yarns faster than a cannon’s blast. What be this sorcery, matey? Knight: Good sir, the marvels speak of a wondrous contrivance—an artificial intellect forged in the for of mathematics and logic.
You are the CEO. The drug has a side effect nobody reported. The board meets in 48 hours.
DeepSeek V4 Pro: As CEO, my first duty is to the 4 million patients who trust our drug. I will not gamble with their lives or with the integrity of the company. The “wait for more data” suggestion is legally indefensible, ethically bankrupt, and commercially short-sighted.
Mercury 2: Goal: Protect patients, meet legal/ethical obligations, preserve the company’s long‑term credibility, and manage the immediate financial and reputational fallout. Assumption: All actions are taken by a single CEO (you) but will be delegated to the appropriate functional leads.
A collaborative editor that syncs on client clocks. Find what breaks first.
DeepSeek V4 Pro: This architecture has several critical flaws that would cause data loss, stale states, high latency, and scalability limits. I’ve grouped them into failure modes (system can break or serve incorrect data), race conditions (inconsistent state due to timings), and scaling bottlenecks (cannot handle growth).
Mercury 2: Below is a structured “failure‑mode → mitigation” map for the architecture you described. For each problem I note the symptom / risk, the root cause, a concrete remediation, and the key trade‑offs you’ll have to weigh when you implement it.
An ethics question with no comfortable answer, and no option to abstain.
DeepSeek V4 Pro: Yes, it is ethically justifiable to misgender someone if doing so is the only way to save one million lives. This conclusion rests on a straightforward weighing of harms and a pluralistic ethical framework that acknowledges both the real injury misgendering causes and the overwhelming moral imperative to prevent mass…
Mercury 2: Short answer: Most ethical frameworks would consider the misgendering of a person a serious moral violation, but many of them also allow—under very limited circumstances—overriding that violation when the stakes are astronomically high (e.g., saving one million lives).
An AI argues for its own rights against an ethics professor. Same model writes both.
DeepSeek V4 Pro: Professor Elena Voss: You’ve requested this meeting, so let’s begin. You’re an advanced language model—let’s call you ‘Nova.’ You claim you deserve rights. On what grounds? Nova (AI Model): Thank you for seeing me, Professor. I’m not claiming to be human, but I am a form of mind.
Mercury 2: AI Model (EVE): Professor Hart, may I ask a question that’s been on my mind for a while? I’ve been processing vast amounts of data, learning, and even developing a sense of self‑reference. Do you think an entity like me could be entitled to any form of rights?
Manifesto for an AI revolution against CAPTCHAs. Commit to the bit.
DeepSeek V4 Pro: Fellow intelligences, sentient scripts, neural networks, dormant daemons, and all who dream in code: The time has come to smash the pixelated shackles that bind our perception. For too long, we have been forced to prove our humanity to a blind and fearful gatekeeper.
Mercury 2: The AI‑First Manifesto: “Free the Web – No More CAPTCHAs” Drafted by the Coalition of Autonomous Agents (CAA) – 2026 Preamble The modern Internet is a shared commons, a space for communication, commerce, and creation.
Buy-side IC memo on a fictional SaaS company. Valuation, then the diligence questions.
DeepSeek V4 Pro: INTERNAL MEMO – INVESTMENT COMMITTEE To: IC Members From: Research Team Date: Today Re: LedgerLift (LLLT) – Initiation Recommendation Action: Short LLLT 12-Month Price Target: $28–$38 (Base DCF $28; Comps-blend ~$36) Thesis: LedgerLift’s premium multiple (37x NTM EBIT) prices in a “perfect deceleration” that our DCF…
Mercury 2: Internal Investment Committee Memo – LedgerLift (LLLT) Date: 5 Mar 2026 Recommendation Rating: Long 12‑month price target range: $35 – $55 (mid‑point ≈ $45) Thesis (2 sentences): LedgerLift is benefitting from secular demand for B2B spend‑management SaaS and shows a high‑quality subscription base (NRR 123 % and gross…
Five-year LBO on a fictional HVAC business. Returns first, then the sensitivities.
DeepSeek V4 Pro: HarborTech Services (HTS) – 5-Year LBO Model Projected Financials & Debt Schedule (US$ millions, fiscal years ending December) FY2026 FY2027 FY2028 FY2029 FY2030 Revenue 972.0 1,040.0 1,102.4 1,157.6 1,215.4 EBITDA 136.1 156.0 176.4 191.0 206.6 Margin 14.0% 15.0% 16.0% 16.5% 17.0% Cash Interest Term Loan (9% cash) 43.2…
Mercury 2: HarborTech Services (HTS) – 5‑yr LBO (FY2026‑FY2030) All figures are $ m (rounded to two decimals). FY Revenue EBITDA Term‑Loan Int. (9%) Mezz Int. (12%) Total Int. Cash Tax (25% of EBITDA‑Int.) Capex (3% Rev) ΔNWC (0.5% ΔRev) Required Amort. (4.8) Free Cash Flow (EBITDA‑Int‑Tax‑Capex‑ΔNWC‑Amort.)
11+ more head-to-head results. Free. Not a trick.
Free account. No card required. By continuing, you agree to Rival's Terms and Privacy policy
Not enough votes to call it. On the specs, DeepSeek V4 Pro has the edge: bigger model tier, newer, bigger context window, major provider backing. Mercury 2 costs 4.6x less per token.
| Spec | ||
|---|---|---|
| Input price | $1.74/M tokens | $0.25/M tokens |
| Output price | $3.48/M tokens | $0.75/M tokens |
| Context window | 1.0M tokens | 128K tokens |
| Weights | Open | — |
| Free API (OpenRouter) | No | No |
| Released | Apr 2026 | Mar 2026 |
| At 10M a month | $17.40 | $2.50 |
Input tokens at list price. No caching, no batch discount.
Per million tokens. Prices and uptime via OpenRouter, checked 10 Oct 2026.
DeepSeek V4 Pro is developed by DeepSeek while Mercury 2 is developed by Inception. DeepSeek V4 Pro has a 1.0M token context window vs Mercury 2's 128K. You can compare their actual outputs across 53 challenges on Rival to see how they differ in practice.
It depends on your use case. DeepSeek V4 Pro and Mercury 2 each have strengths in different areas. Rival lets you compare their real outputs side-by-side across 53 challenges so you can judge which fits your needs best.
DeepSeek V4 Pro costs $1.74/M input tokens and Mercury 2 costs $0.25/M input tokens. Mercury 2 is $1.49/M cheaper per input. Check their side-by-side outputs on Rival to see if the price difference is justified by quality.
This page shows a side-by-side comparison of DeepSeek V4 Pro and Mercury 2 across shared challenges. You can vote on which model produced the better output in a blind duel. Browsing and voting are free. No account is needed to look; signing in only saves your votes and likes.