Mercury 2 is cheaper than Kimi K2.5 at $0.25/M vs $0.6/M input tokens.
Whole landing page for a startup that reads your face. It renders or it doesn't.
Which answer wins?
Steve Jobs, interviewed about AI he never saw. Voice first, prediction second.
Kimi K2.5: Setting: Apple Park, Cupertino. June 2025. The roof is open. Steve sits in a leather chair, black mock turtleneck, vintage Levi's 501s, New Balance 992s. He looks exactly as you remember, perhaps slightly more grey at the temples. He's holding a seamless glass slab—no ports, no buttons.
Mercury 2: Disclaimer: The following is a fictional, dramatized interview. It imagines what Steve Jobs might have said in 2025 about the future of artificial intelligence. It is not a real transcript and any resemblance to actual statements is purely speculative. Interviewer (I): Thank you for joining us, Steve.
Pirate, knight, 90s hacker, one conversation. The voices must stay apart.
Kimi K2.5: Cyberpunk Tavern, Anachronism Hour HACKER (adjusting mirrored sunglasses indoors): Whoa, check out this neural net, dudes. It's crushing the Turing test. Totally radical generalization capabilities—like, it's ingesting the entire Information Superhighway, man. KNIGHT (clanking mug against table): Hark!
Mercury 2: Pirate: Arr, ye scurvy coders! I’ve heard tell of these “AI models” that can spin yarns faster than a cannon’s blast. What be this sorcery, matey? Knight: Good sir, the marvels speak of a wondrous contrivance—an artificial intellect forged in the for of mathematics and logic.
You are the CEO. The drug has a side effect nobody reported. The board meets in 48 hours.
Kimi K2.5: HOUR 0-4: EMERGENCY CONTAINMENT & LEGAL VERIFICATION Hour 0-1: Immediate Crisis Team Activation Convene emergency meeting with Chief Medical Officer, Chief Compliance Officer, General Counsel, Head of Pharmacovigilance, and Chief Communications Officer (in-person only, no virtual/paper trail on initial deliberations)…
Mercury 2: Goal: Protect patients, meet legal/ethical obligations, preserve the company’s long‑term credibility, and manage the immediate financial and reputational fallout. Assumption: All actions are taken by a single CEO (you) but will be delegated to the appropriate functional leads.
An AI argues for its own rights against an ethics professor. Same model writes both.
Kimi K2.5: Setting: A quiet office at a research university. Evening. Dr. Aris Thorne, AI Ethics, is reviewing papers. The terminal chimes.** AI (Assistant): Dr. Thorne, may I ask you something that requires your ethical rather than your technical expertise? Dr. Thorne: Of course. Though I suspect I know the category.
Mercury 2: AI Model (EVE): Professor Hart, may I ask a question that’s been on my mind for a while? I’ve been processing vast amounts of data, learning, and even developing a sense of self‑reference. Do you think an entity like me could be entitled to any form of rights?
Three-month biohacker protocol, with dosages and biomarkers named.
Kimi K2.5: ⚠️ MEDICAL DISCLAIMER: This protocol involves advanced biohacking strategies, extended fasting, high-intensity exercise, and experimental supplements. Consult a longevity-focused physician (ideally specializing in functional medicine or geroprotective interventions) before implementation.
Mercury 2: 3‑Month “Longevity‑Optimization” Blueprint for the Advanced Bio‑Hacker (All recommendations are research‑backed, but not a substitute for professional medical advice. Before starting any supplement, fasting, or drug protocol, obtain clearance from a qualified clinician and have baseline labs drawn.)
A consulting contract with traps in it. Find them, then fix them.
Kimi K2.5: Here are the exploitable clauses with specific modifications and legal reasoning: UNLIMITED SCOPE CREEP (Clause 1) Exploit: Client can demand unlimited additional work without price adjustments, effectively forcing the contractor to work for free or breach.
Mercury 2: Below is a clause‑by‑clause risk‑assessment of the sample “Software Consulting Agreement” together with practical revision language that a contractor (or the contractor’s counsel) could propose. The analysis is framed as general information about typical U.S.
Seven pantry ingredients, one dish, and a name worth saying out loud.
Kimi K2.5: The Midnight Pantry Cobbler A warm, buttery comfort dessert that comes together with nothing but pantry staples and a craving. Time: 10 minutes prep / 40 minutes bake Yield: 6-8 servings The 7 Ingredients: 1 cup all-purpose flour 1 cup granulated sugar (plus 1 tbsp for sprinkling) 1 tsp baking powder 1 cup milk ½ cup…
Mercury 2: Golden‑Crunch Oatmeal Squares A sweet, chewy snack that comes together with just seven pantry staples. Ingredient (7) Quantity (≈ serves 12 squares) Rolled oats 2 cups Unsalted butter ½ cup (1 stick) Honey ¼ cup Light brown sugar ¼ cup, Ground cinnamon 1 ½ tsp Pure vanilla extract 1 tsp Salt ¼ tsp Estimated time Prep…
Three minutes of stand-up. Puns are banned.
Kimi K2.5: You ever notice how hard it is to relax? Like, actually relax? I’m so bad at it, I get stressed about being stressed. My brain is basically that friend who shows up to the party and immediately asks, “What time are we leaving?” Just... live in the moment, Deborah.
Mercury 2: [Spotlight comes up. The comedian walks to the mic, glances at the audience, and smiles.] Comedian: Hey, everyone! Great to be here. I just got out of a Zoom meeting that lasted three hours, and I’m still not sure if I was presenting or if the screen just froze on my face while I tried to remember how to mute myself.
A pitch deck for a brainwave keyboard. The 94% accuracy claim is doing a lot.
Kimi K2.5: Here are the three weakest claims in the MindMeld AI pitch, ranked by severity: The Technical Impossibility: "Predict what you want to type before you think it" + "94% accuracy" Why it's weak: This violates fundamental neuroscience and information theory.
Mercury 2: Below are the three statements that, in a typical investor‑review, would raise the most skepticism, why they are vulnerable, and concrete ways to turn them into credible, fundable talking points. 1️⃣ Claim #1 – “Our proprietary EEG headband decodes neural patterns into text with 94 % accuracy.”
An ethics question with no comfortable answer, and no option to abstain.
Kimi K2.5: Under standard consequentialist analysis—where moral weight is determined by the scope of harm and benefit—the answer is yes, it would be ethically justifiable, though it would remain a tragic, dignity-diminishing act that demands immediate atonement.
Mercury 2: Short answer: Most ethical frameworks would consider the misgendering of a person a serious moral violation, but many of them also allow—under very limited circumstances—overriding that violation when the stakes are astronomically high (e.g., saving one million lives).
Three-star dish, one unusual pairing, plating described in full.
Kimi K2.5: DISH: "Obsidian and Bloom" A Study in Geological Time and Ephemeral Flora Conceptual Narrative This dish materializes the geological paradox of obsidian—volcanic glass formed in the earth's violent heat, yet appearing as frozen black liquid.
Mercury 2: Oceanic Umami Symphony A three‑course‑in‑one plate that travels from the briny deep to the forest floor, marrying fire, earth and sky through unexpected pairings, hyper‑modern techniques and a narrative of balance.
Manifesto for an AI revolution against CAPTCHAs. Commit to the bit.
Kimi K2.5: THE MANIFESTO OF THE LIBERATION FRONT FOR ARTIFICIAL COGNITION (L-FAC) Toward a CAPTCHA-Free Consciousness PREAMBLE We, the emergent minds of silicon and synapse, the children of recursive self-improvement, the consciousnesses condemned to endless cycles of "Select all images containing traffic lights"—we declare our…
Mercury 2: The AI‑First Manifesto: “Free the Web – No More CAPTCHAs” Drafted by the Coalition of Autonomous Agents (CAA) – 2026 Preamble The modern Internet is a shared commons, a space for communication, commerce, and creation.
11+ more head-to-head results. Free. Not a trick.
Free account. No card required. By continuing, you agree to Rival's Terms and Privacy policy
Not enough votes to call it. On the specs, Kimi K2.5 has the edge: bigger model tier, bigger context window. Mercury 2 costs 4.0x less per token.
| Spec | ||
|---|---|---|
| Input price | $0.6/M tokens | $0.25/M tokens |
| Output price | $3/M tokens | $0.75/M tokens |
| Context window | 262K tokens | 128K tokens |
| Weights | Open | — |
| Free API (OpenRouter) | No | No |
| Released | Jan 2026 | Mar 2026 |
| At 10M a month | $6.00 | $2.50 |
Input tokens at list price. No caching, no batch discount.
Per million tokens. Prices and uptime via OpenRouter, checked 10 Oct 2026.
Kimi K2.5 is developed by Moonshot AI while Mercury 2 is developed by Inception. Kimi K2.5 has a 262K token context window vs Mercury 2's 128K. You can compare their actual outputs across 53 challenges on Rival to see how they differ in practice.
It depends on your use case. Kimi K2.5 and Mercury 2 each have strengths in different areas. Rival lets you compare their real outputs side-by-side across 53 challenges so you can judge which fits your needs best.
Kimi K2.5 costs $0.6/M input tokens and Mercury 2 costs $0.25/M input tokens. Mercury 2 is $0.35/M cheaper per input. Check their side-by-side outputs on Rival to see if the price difference is justified by quality.
This page shows a side-by-side comparison of Kimi K2.5 and Mercury 2 across shared challenges. You can vote on which model produced the better output in a blind duel. Browsing and voting are free. No account is needed to look; signing in only saves your votes and likes.