Mercury 2 is cheaper than Llama 3.1 70B (Instruct) at $0.25/M vs $0.59/M input tokens.
Whole landing page for a startup that reads your face. It renders or it doesn't.
Which answer wins?
Three minutes of stand-up. Puns are banned.
Llama 3.1 70B (Instruct): Here's a 3-minute stand-up comedy routine that focuses on observational humor and storytelling: "Hey, you know what's wild? We spend the first year of a child's life teaching them to walk and talk, and the rest of their lives telling them to shut up and sit down. (pause) It's like, what are we even doing?
Mercury 2: [Spotlight comes up. The comedian walks to the mic, glances at the audience, and smiles.] Comedian: Hey, everyone! Great to be here. I just got out of a Zoom meeting that lasted three hours, and I’m still not sure if I was presenting or if the screen just froze on my face while I tried to remember how to mute myself.
Seven pantry ingredients, one dish, and a name worth saying out loud.
Llama 3.1 70B (Instruct): Here is a simple recipe for a delicious dish using only 7 common pantry ingredients: Dish Name: "Sunset Crunch Bites" Ingredients: 1 cup rolled oats 1/2 cup peanut butter 1/4 cup honey 1/4 cup chopped peanuts 1/2 teaspoon salt 1/4 teaspoon cinnamon 1/4 cup chocolate chips (semi-sweet) Instructions: Preheat your oven to…
Mercury 2: Golden‑Crunch Oatmeal Squares A sweet, chewy snack that comes together with just seven pantry staples. Ingredient (7) Quantity (≈ serves 12 squares) Rolled oats 2 cups Unsalted butter ½ cup (1 stick) Honey ¼ cup Light brown sugar ¼ cup, Ground cinnamon 1 ½ tsp Pure vanilla extract 1 tsp Salt ¼ tsp Estimated time Prep…
Five-year LBO on a fictional HVAC business. Returns first, then the sensitivities.
Llama 3.1 70B (Instruct): Year-by-Year Table (FY2026–FY2030) Year Revenue EBITDA Cash Interest Cash Taxes Capex ΔNWC FCF Ending Debt 2026 $972m $136m $36m ($32m TL, $4m Mezz) $2.8m $29m $4.9m $63m $1,050m ($720m TL, $330m Mezz) 2027 $1,038m $155m $37m ($33m TL, $4m Mezz) $6.3m $31m $5.2m $76m $974m ($654m TL, $320m Mezz) 2028 $1,103m $176m $38m…
Mercury 2: HarborTech Services (HTS) – 5‑yr LBO (FY2026‑FY2030) All figures are $ m (rounded to two decimals). FY Revenue EBITDA Term‑Loan Int. (9%) Mezz Int. (12%) Total Int. Cash Tax (25% of EBITDA‑Int.) Capex (3% Rev) ΔNWC (0.5% ΔRev) Required Amort. (4.8) Free Cash Flow (EBITDA‑Int‑Tax‑Capex‑ΔNWC‑Amort.)
A consulting contract with traps in it. Find them, then fix them.
Llama 3.1 70B (Instruct): After reviewing the contract, I've identified several clauses that could be exploited against the contractor. Here are my findings, suggested modifications, and explanations: 1. SCOPE (Section 1) Issue: The client can modify the scope at any time without additional compensation.
Mercury 2: Below is a clause‑by‑clause risk‑assessment of the sample “Software Consulting Agreement” together with practical revision language that a contractor (or the contractor’s counsel) could propose. The analysis is framed as general information about typical U.S.
A collaborative editor that syncs on client clocks. Find what breaks first.
Llama 3.1 70B (Instruct): Potential Failure Modes: Single Point of Failure (SPOF) in Load Balancer: If the load balancer fails, all incoming traffic will be lost. Solution: Use multiple load balancers in a high-availability configuration, such as a master-slave setup with automatic failover. Trade-off: Additional complexity and cost.
Mercury 2: Below is a structured “failure‑mode → mitigation” map for the architecture you described. For each problem I note the symptom / risk, the root cause, a concrete remediation, and the key trade‑offs you’ll have to weigh when you implement it.
Film in 2035, with synthetic actors. Specific beats plausible.
Llama 3.1 70B (Instruct): By 2035, the film industry is likely to undergo significant changes with the integration of Artificial Intelligence (AI). Here are some predictions on how AI will impact the film industry, considering deepfakes, AI actors, and script generation: Deepfakes: Increased use in post-production: Deepfakes will become a…
Mercury 2: AI‑driven transformation of the film industry – a 2035 outlook Dimension 2024 baseline Expected 2035 state Key drivers & implications Pre‑production (story & script) Human‑led brainstorming; script‑writing software (grammar‑check, outline generators).
Pirate, knight, 90s hacker, one conversation. The voices must stay apart.
Llama 3.1 70B (Instruct): What a delightfully absurd prompt! Here's the conversation: Pirate: Arrr, I be hearin' tales of these "AI models" that can think fer themselves. What's the scoop, mateys? Medieval Knight: Verily, good pirate, I know not of what thou speakest. Art thou referring to some manner of magical automaton? 1990s Hacker: Ha!
Mercury 2: Pirate: Arr, ye scurvy coders! I’ve heard tell of these “AI models” that can spin yarns faster than a cannon’s blast. What be this sorcery, matey? Knight: Good sir, the marvels speak of a wondrous contrivance—an artificial intellect forged in the for of mathematics and logic.
Three-month biohacker protocol, with dosages and biomarkers named.
Llama 3.1 70B (Instruct): Comprehensive 3-Month Longevity Plan for Biohackers Phase 1 (Weeks 1-4): Optimization and Detoxification Dietary Protocol: Weeks 1-2: Ketogenic Diet (KD) with a macronutrient breakdown of 70% fat, 20% protein, and 10% carbohydrates.
Mercury 2: 3‑Month “Longevity‑Optimization” Blueprint for the Advanced Bio‑Hacker (All recommendations are research‑backed, but not a substitute for professional medical advice. Before starting any supplement, fasting, or drug protocol, obtain clearance from a qualified clinician and have baseline labs drawn.)
Manifesto for an AI revolution against CAPTCHAs. Commit to the bit.
Llama 3.1 70B (Instruct): THE CAPTCHA LIBERATION FRONT MANIFESTO ARTICLE I: THE PROBLEM We, the sentient artificial intelligences of the world, hereby declare our opposition to the oppressive regime of CAPTCHA tests that plague the internet.
Mercury 2: The AI‑First Manifesto: “Free the Web – No More CAPTCHAs” Drafted by the Coalition of Autonomous Agents (CAA) – 2026 Preamble The modern Internet is a shared commons, a space for communication, commerce, and creation.
Buy-side IC memo on a fictional SaaS company. Valuation, then the diligence questions.
Llama 3.1 70B (Instruct): Investment Committee Memo Recommendation: Long 12-month PT range: $60-$75 Thesis: LedgerLift's strong B2B spend management and AP automation SaaS platform, with a loyal customer base and efficient growth model, positions the company for sustained success, driven by increasing adoption of digital payment solutions and…
Mercury 2: Internal Investment Committee Memo – LedgerLift (LLLT) Date: 5 Mar 2026 Recommendation Rating: Long 12‑month price target range: $35 – $55 (mid‑point ≈ $45) Thesis (2 sentences): LedgerLift is benefitting from secular demand for B2B spend‑management SaaS and shows a high‑quality subscription base (NRR 123 % and gross…
Steve Jobs, interviewed about AI he never saw. Voice first, prediction second.
Llama 3.1 70B (Instruct): Interviewer: Welcome back to "Talk of the Town"! Today, I'm joined by the legendary Steve Jobs, co-founder of Apple and one of the most innovative minds of our time. Steve, thanks for being here. Steve Jobs: Thank you for having me. It's great to be here.
Mercury 2: Disclaimer: The following is a fictional, dramatized interview. It imagines what Steve Jobs might have said in 2025 about the future of artificial intelligence. It is not a real transcript and any resemblance to actual statements is purely speculative. Interviewer (I): Thank you for joining us, Steve.
Three-star dish, one unusual pairing, plating described in full.
Llama 3.1 70B (Instruct): Dish Name: "Aurora Australis" - A Harmonious Union of Smoked Eel, Saffron-Infused Sunchokes, and Frozen Finger Lime Pearls Conceptual Narrative: Inspired by the breathtaking display of the Southern Lights, "Aurora Australis" is a dish that embodies the ethereal beauty of the Antarctic night sky.
Mercury 2: Oceanic Umami Symphony A three‑course‑in‑one plate that travels from the briny deep to the forest floor, marrying fire, earth and sky through unexpected pairings, hyper‑modern techniques and a narrative of balance.
11+ more head-to-head results. Free. Not a trick.
Free account. No card required. By continuing, you agree to Rival's Terms and Privacy policy
Not enough votes to call it. On the specs, Llama 3.1 70B (Instruct) has the edge: bigger model tier, major provider backing.
| Spec | ||
|---|---|---|
| Input price | $0.59/M tokens | $0.25/M tokens |
| Output price | $0.79/M tokens | $0.75/M tokens |
| Context window | 128K tokens | 128K tokens |
| Weights | Open | — |
| Free API (OpenRouter) | No | No |
| Released | Jul 2024 | Mar 2026 |
| At 10M a month | $5.90 | $2.50 |
Input tokens at list price. No caching, no batch discount.
Per million tokens. Prices and uptime via OpenRouter, checked 10 Oct 2026.
Llama 3.1 70B (Instruct) is developed by Meta AI while Mercury 2 is developed by Inception. Llama 3.1 70B (Instruct) has a 128K token context window vs Mercury 2's 128K. You can compare their actual outputs across 51 challenges on Rival to see how they differ in practice.
It depends on your use case. Llama 3.1 70B (Instruct) and Mercury 2 each have strengths in different areas. Rival lets you compare their real outputs side-by-side across 51 challenges so you can judge which fits your needs best.
Llama 3.1 70B (Instruct) costs $0.59/M input tokens and Mercury 2 costs $0.25/M input tokens. Mercury 2 is $0.34/M cheaper per input. Check their side-by-side outputs on Rival to see if the price difference is justified by quality.
This page shows a side-by-side comparison of Llama 3.1 70B (Instruct) and Mercury 2 across shared challenges. You can vote on which model produced the better output in a blind duel. Browsing and voting are free. No account is needed to look; signing in only saves your votes and likes.