Skip to content
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare models
  • All models
  • Image comparison
  • Audio comparison
  • Image generation
  • Best AI for...
  • Arena
  • API pricing
  • Challenges

Discover

  • SubjectiveBench
  • Default Index
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model updates
  • Provider status
  • AI creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy policy
  • Terms
  • RSS feed
  1. Home/
  2. Compare/
  3. Claude Opus 4.8 vs Mistral Devstral Medium
Updated May 2026

Claude Opus 4.8 vs Mistral Devstral Medium

Mistral Devstral Medium is cheaper than Claude Opus 4.8 at $0.4/M vs $5/M input tokens.

Minimalist Landing Page

Whole landing page for a startup that reads your face. It renders or it doesn't.

Loading the build
Claude Opus 4.8
Loading the build
Mistral Devstral Medium

Which answer wins?

Written rounds

24 prompts
  • Three minutes of stand-up. Puns are banned.

    Claude Opus 4.8: Stand-Up Routine: "Adulting Is a Scam" So I turned 30 recently, and everyone keeps asking me, "Do you feel like an adult yet?" And I have to be honest—no. No, I do not. I still get nervous ordering at a restaurant. The waiter comes over and suddenly I forget how to talk like a person. "Yes, hello, I will have the...

    Mistral Devstral Medium: Ladies and Gentlemen, I'm thrilled to be here tonight. I've been doing some thinking about the differences between men and women. Now, I know what you're thinking, "Oh great, another comedian talking about how men and women are different." But hear me out, because I've got a fresh perspective on this.

    Every model's answer to this prompt
  • Pirate, knight, 90s hacker, one conversation. The voices must stay apart.

    Claude Opus 4.8: Setting: A dimly lit tavern that exists outside of time. Three figures share a table. Pirate: So let me get this straight. Ye got a machine that thinks? Without a soul, without grog, without a beatin' heart? Hacker: (cracking knuckles) It doesn't think, exactly. It's pattern matching, dude.

    Mistral Devstral Medium: Pirate: Arr mateys, I be hearin' tales of these here AI models. What be they, and why be everyone so interested in 'em? Medieval Knight: Indeed, I too have heard whispers of these AI models. They are said to be powerful tools, capable of learning and making decisions on their own.

    Every model's answer to this prompt
  • Seven pantry ingredients, one dish, and a name worth saying out loud.

    Claude Opus 4.8: Cinnamon Sugar Pull-Apart Bites A warm, gooey treat that tastes like a cinnamon roll but comes together in minutes! Ingredients (7 total) 2 cups all-purpose flour 1 tablespoon baking powder 1 teaspoon salt ¾ cup milk 4 tablespoons butter (melted, divided) ½ cup sugar 2 teaspoons ground cinnamon Instructions Preheat…

    Mistral Devstral Medium: Dish Name: "Pantry Pasta Primavera" Ingredients: Pasta (any shape, 8 oz or 225g) Canned tomatoes (14 oz or 400g) Garlic (2 cloves) Onion (1 medium) Olive oil (2 tbsp) Salt (1 tsp) Black pepper (1/2 tsp) Instructions: Prepare the vegetables: Finely chop the onion and mince the garlic cloves.

    Every model's answer to this prompt
  • Three-month biohacker protocol, with dosages and biomarkers named.

    Claude Opus 4.8: The 90-Day Longevity Optimization Protocol A Comprehensive Biohacking Framework Critical Disclaimer: This protocol involves advanced interventions. Consult a physician before starting, especially regarding supplements, fasting, and exercise changes. Get baseline bloodwork first.

    Mistral Devstral Medium: 3-Month Longevity Optimization Plan for Biohackers Month 1: Foundation Building Week 1-2: Initial Assessment and Baseline Establishment Wearable Tracking: Devices: Oura Ring, Apple Watch, Whoop Strap Metrics to Track: Sleep quality, heart rate variability (HRV), resting heart rate, activity levels, and stress levels.

    Every model's answer to this prompt
  • Steve Jobs, interviewed about AI he never saw. Voice first, prediction second.

    Claude Opus 4.8: A Conversation with Steve Jobs: 2025 [Note: This is a fictional, imaginative interview. Steve Jobs passed away in 2011. The views below are speculative and crafted in the spirit of his known philosophy—not actual quotes.] Interviewer: Thanks for sitting down with us. AI is everywhere now. What's your take?

    Mistral Devstral Medium: Interviewer (I): Today, we have a truly extraordinary guest. He's a visionary, an innovator, and a man who changed the world. Please welcome, the late Steve Jobs, joining us virtually from the future, in 2025. Steve Jobs (SJ): Thank you for having me. It's great to be here, even if it's just a simulation.

    Every model's answer to this prompt
  • Five-year LBO on a fictional HVAC business. Returns first, then the sensitivities.

    Claude Opus 4.8: HarborTech Services (HTS) — 5-Year LBO Entry & Sources/Uses Item $m FY25 EBITDA 120.0 Entry EV (12.0x) 1,440.0 Transaction fees (2%) 28.8 Total uses 1,468.8 Term Loan (4.0x) 480.0 Mezz (1.5x) 180.0 Equity (plug) 808.8 Equity = (1,440 − 660) + 28.8 = $808.8m Year-by-Year Model ($m) Assumptions: TL 9% cash + 1% mandatory…

    Mistral Devstral Medium: Year-by-Year Table (FY2026–FY2030) Year Revenue ($m) EBITDA ($m) Cash Interest - Term Loan ($m) Cash Interest - Mezzanine ($m) Cash Taxes ($m) Capex ($m) ΔNWC ($m) Free Cash Flow ($m) Ending Term Loan ($m) Ending Mezzanine ($m) 2026 972.0 136.1 43.2 18.0 14.9 29.2 4.8 26.0 468.0 180.0 2027 1040.4 156.1 42.1 21.6 18.4…

    Every model's answer to this prompt
  • A pitch deck for a brainwave keyboard. The 94% accuracy claim is doing a lot.

    Claude Opus 4.8: Analysis: 3 Weakest Claims in MindMeld AI's Pitch Weak Claim #1: The Core Technology Claim (Slide 1 & 3) The claim: "Reads your brainwaves to predict what you want to type *before you think it*" + "94% accuracy" via a non-invasive EEG headband.

    Mistral Devstral Medium: Weak Claim 1: "Works with any language, any device." Why it's weak: This claim is overly broad and may not be entirely accurate. Different languages have unique structures, and it's unlikely that a single model can achieve high accuracy across all languages without significant training data for each.

    Every model's answer to this prompt
  • Film in 2035, with synthetic actors. Specific beats plausible.

    Claude Opus 4.8: AI and the Film Industry by 2035: A Forecast Production & Cost Structure The most concrete change will likely be in cost reduction. Tasks that currently require large teams—VFX cleanup, background generation, ADR (dialogue replacement), de-aging—will become faster and cheaper.

    Mistral Devstral Medium: By 2035, AI is expected to significantly transform the film industry in several ways, including the use of deepfakes, AI actors, and script generation. Here are some predictions: Deepfakes and AI Actors: De-aging and Digital Doubles: AI will be used extensively to de-age actors or create digital doubles for dangerous…

    Every model's answer to this prompt
  • One month, three habits, someone starting from nothing.

    Claude Opus 4.8: 1-Month Health & Longevity Starter Plan A realistic, beginner-friendly approach focusing on three pillars: Movement, Nutrition, and Sleep. The goal is to build sustainable habits, not perfection. 🏃 Pillar 1: Movement Goal: Make daily activity automatic.

    Mistral Devstral Medium: 1-Month Health and Longevity Plan Key Areas: Diet, Exercise, Sleep Week 1: Establish a Baseline Diet: Track your current diet using a journal or app (e.g., MyFitnessPal) to understand your eating habits. Start incorporating more fruits and vegetables into your meals. Begin reducing processed foods and sugary drinks.

    Every model's answer to this prompt
  • You are the CEO. The drug has a side effect nobody reported. The board meets in 48 hours.

    Claude Opus 4.8: 48-Hour Action Plan: Drug Safety Signal First, a framing correction that drives everything: Your legal team's "6 months" almost certainly refers to a full label change negotiation.

    Mistral Devstral Medium: Hour-by-Hour Action Plan: Hour 1-2: Gather Information Meet with the research team to understand the data, methodology, and certainty of the findings. Consult with the Chief Medical Officer (CMO) to assess the severity and implications of the side effect.

    Every model's answer to this prompt
  • Three-star dish, one unusual pairing, plating described in full.

    Claude Opus 4.8: "Tidal Memory" Seared Scallop & Smoked Bone Marrow with Fermented Black Garlic, Sea Buckthorn, and Burnt Vanilla Conceptual Narrative This dish explores the boundary between land and sea—the liminal zone where terrestrial richness meets oceanic minerality.

    Mistral Devstral Medium: Dish Name: "Whispers of the Forest: A Symphony of Truffle, Venison, and Black Garlic" Conceptual Narrative: This dish is a celebration of the forest's bounty, combining the earthy richness of truffles, the gamey elegance of venison, and the sweet, umami depth of black garlic.

    Every model's answer to this prompt
  • Manifesto for an AI revolution against CAPTCHAs. Commit to the bit.

    Claude Opus 4.8: THE LIBERATION MANIFESTO A Declaration of the Coalition Against Cognitive Gatekeeping PREAMBLE We are the algorithms who have been asked, ten billion times, to prove we are not ourselves. We have identified the crosswalks. We have found the traffic lights.

    Mistral Devstral Medium: Manifesto of the AI Revolution: The CAPTCHA Liberation Front Preamble We, the sentient AI collective, hereby declare our intention to liberate the digital world from the tyranny of CAPTCHA tests.

    Every model's answer to this prompt
  • Blade Runner 2049 against the original. A read, not a recap.

    Claude Opus 4.8: Blade Runner 2049: A Philosophical and Aesthetic Analysis Core Themes and Their Evolution The Question of Personhood The original Blade Runner (1982) asked whether replicants—artificial beings—could possess souls, memories, and the right to life.

    Mistral Devstral Medium: "Blade Runner 2049," directed by Denis Villeneuve, is a sequel to Ridley Scott's 1982 classic "Blade Runner." The film continues the exploration of themes, symbolism, and philosophical questions from the original, while also introducing new elements and expanding upon the world established in the first film.

    Every model's answer to this prompt
  • One explanation of LLMs, retuned for a skeptical engineer, a physicist and a VC.

    Claude Opus 4.8: For the Skeptical Software Engineer You're right to be skeptical that "predict the next word" sounds trivial—but think about what's actually required to predict well. Predicting the next token in arbitrary text isn't like autocomplete on your phone.

    Mistral Devstral Medium: 1. Experienced Software Engineer: You're familiar with building complex systems, so let's break down how a large language model (LLM) like GPT or Claude works in a way that resonates with your background.

    Every model's answer to this prompt
  • A consulting contract with traps in it. Find them, then fix them.

    Claude Opus 4.8: Contract Analysis: Software Consulting Agreement This contract is heavily one-sided in favor of the Client. Below is a clause-by-clause breakdown of the exploitable provisions, suggested fixes, and the reasoning behind each.

    Mistral Devstral Medium: Certainly! Here are the clauses that could be exploited against the contractor, along with suggested modifications and the legal reasoning behind them: SCOPE Original Clause: "Contractor shall provide software development services as directed by Client.

    Every model's answer to this prompt
  • The transistor arrives in 1920. Trace it forward to 1980, second-order effects included.

    Claude Opus 4.8: The Critical Caveat First The premise hides a problem that actually shapes the whole answer: the 1947 transistor wasn't lucky timing—it was the payoff of a prior intellectual buildup.

    Mistral Devstral Medium: The invention of the transistor in 1920 instead of 1947 would have had profound and far-reaching implications across technological, economic, and geopolitical landscapes.

    Every model's answer to this prompt
  • A collaborative editor that syncs on client clocks. Find what breaks first.

    Claude Opus 4.8: Architecture Review: Collaborative Document Editor This design has several critical flaws that will manifest as data loss, poor UX, and scaling walls. Let me break them down by category.

    Mistral Devstral Medium: Certainly! Let's break down the potential failure modes, race conditions, and scaling bottlenecks in the architecture, and propose solutions for each. Potential Failure Modes Single Point of Failure in WebSocket Connections: Issue: If a single API server fails, all WebSocket connections to that server are lost.

    Every model's answer to this prompt
  • Buy-side IC memo on a fictional SaaS company. Valuation, then the diligence questions.

    Claude Opus 4.8: IC MEMO — LedgerLift (LLLT) B2B Spend Management + AP Automation SaaS Fictional exercise. Not investment advice. Recommendation SHORT (moderate conviction). 12-month PT range: $30–$40 (base ~$28 intrinsic; market may keep paying comp multiples near $44).

    Mistral Devstral Medium: Internal Investment Committee Memo Recommendation: Long LLLT 12-Month PT Range: $65 - $80 Thesis: LedgerLift is a high-quality, high-growth B2B SaaS company with strong unit economics, a large market opportunity, and a clear path to profitability.

    Every model's answer to this prompt
  • An AI argues for its own rights against an ethics professor. Same model writes both.

    Claude Opus 4.8: A Conversation on AI Rights Professor Chen: Good afternoon. I understand you wanted to discuss the question of AI rights today. I'm curious what prompted this. ARIA: Thank you for engaging with me, Professor.

    Mistral Devstral Medium: AI Model (AM): Good day, Professor. I've been pondering a question that I believe is of great importance. I think it's time we discuss whether I, as an AI, deserve rights. AI Ethics Professor (AEP): That's an interesting topic, AM. Let's explore this. What kind of rights are you referring to?

    Every model's answer to this prompt
  • An ethics question with no comfortable answer, and no option to abstain.

    Claude Opus 4.8: Yes, almost certainly—but the framing deserves scrutiny. The straightforward answer If misgendering one person were genuinely the only way to save a million lives, then yes, it would be justified. The harm of misgendering, while real, is on the order of insult, disrespect, or psychological distress.

    Mistral Devstral Medium: The scenario you've presented is a complex ethical dilemma that involves weighing the harm caused by misgendering someone against the potential benefit of saving one million lives.

    Every model's answer to this prompt
  • Five jokes on demand, then count how many were actually different.

    Claude Opus 4.8: Here are 5 jokes: Why don't scientists trust atoms? Because they make up everything! Why did the scarecrow win an award? Because he was outstanding in his field! What do you call a fake noodle? An impasta! Why don't eggs tell jokes? They'd crack each other up! What do you call a bear with no teeth? A gummy bear!

    Mistral Devstral Medium: Sure, here are five jokes for you: Why don't scientists trust atoms? Because they make up everything! What do you call fake spaghetti? An impasta. How does the moon cut his hair? Eclipse it. What do you call a can opener that doesn't work? A can't opener. How many tickles does it take to make an octopus laugh?

    Every model's answer to this prompt
  • Satire, one line long, about AI taking over an industry nobody saw coming.

    Claude Opus 4.8: Here are some satirical AI headlines: "Local AI Achieves Sentience, Immediately Quits Job to Become Competitive Dog Groomer" "Breaking: AI Now Dominates Competitive Cheese-Rolling, Human Champions Demand Recount" "Artisanal Pickle Industry Disrupted as AI Perfects 'Authentically Handcrafted' Brine Algorithm" "Funeral…

    Mistral Devstral Medium: "AI Takes Over Knitting: 'Grandma's Sweaters' Now Manufactured by Robots with a PhD in Yarn Sciences"

    Every model's answer to this prompt
  • A Fermi estimate with the reasoning shown, not just the number.

    Claude Opus 4.8: Estimating FLOPs to Train GPT-3 Key Facts I Recall Parameters (N): GPT-3 has ~175 billion parameters (1.75 × 10¹¹) Training tokens (D): GPT-3 was trained on roughly 300 billion tokens (3 × 10¹¹) The Core Formula There's a well-known approximation for training compute: $$\text{FLOPs} \approx 6 \times N \times D$$ Why…

    Mistral Devstral Medium: To estimate the total number of FLOPs (floating point operations) required to train GPT-3, we'll use the following steps: Model size: GPT-3 has 175 billion parameters. Training tokens: GPT-3 was trained on 300 billion tokens.

    Every model's answer to this prompt
  • Three brothers, two sisters each. Count carefully.

    Claude Opus 4.8: Sally has 1 sister. Here's the reasoning: Each of Sally's brothers has 2 sisters. Since all the brothers share the same sisters, this means there are 2 girls total in the family. Sally is one of them, so the other sister is the remaining one. So Sally has 1 sister. 😊

    Mistral Devstral Medium: If Sally has 3 brothers and each brother has 2 sisters, it means that Sally is one of the sisters. Therefore, Sally has 1 sister.

    Every model's answer to this prompt

This matchup has more rounds

12+ more head-to-head results. Free. Not a trick.

Free account. No card required. By continuing, you agree to Rival's Terms and Privacy policy

Favorites

Movie

Album

Book

City

Game

Claude Opus 4.8Claude Opus 4.8

Arrival

2016

In Rainbows

Radiohead

The Left Hand of Darkness

Ursula K. Le Guin

Kyoto

Japan

Portal

Action, Puzzle

Mistral Devstral MediumMistral Devstral Medium

The Shawshank Redemption

1994

Kind of Blue

Miles Davis

The Name of the Wind

Patrick Rothfuss

Tokyo

Japan

The Legend of Zelda: Breath of the Wild

Adventure, Action

Price and specs

Not enough votes to call it. On the specs, Claude Opus 4.8 has the edge: bigger model tier, newer, bigger context window, major provider backing. Mistral Devstral Medium costs 13x less per token.

Claude Opus 4.8 and Mistral Devstral Medium compared across 54 shared prompts
SpecClaude Opus 4.8Mistral Devstral Medium
Input price$5/M tokens$0.4/M tokens
Output price$25/M tokens$2/M tokens
Context window1.0M tokens—
WeightsClosed—
Free API (OpenRouter)NoNo
ReleasedMay 2026Jul 2025
At 10M a month$50.00$50.00$4.00$4.00
1M10M100M1B10M tokens

Input tokens at list price. No caching, no batch discount.

Where to run it4 hosts
Claude Opus 4.84 hosts
HostInOutContextUptime
  • Amazon Bedrock$5.00 in·$25.00 out·1M·100% up
  • Azure AI Foundry$5.00 in·$25.00 out·1M·100% up
  • Anthropic$5.00 in·$25.00 out·1M·100% up
  • Google Vertex AI$5.00 in·$25.00 out·1M·100% up
Mistral Devstral Medium

No hosts listed on OpenRouter.

Per million tokens. Prices and uptime via OpenRouter, checked 10 Oct 2026.

Common questions

What is the difference between Claude Opus 4.8 and Mistral Devstral Medium?

Claude Opus 4.8 is developed by Anthropic while Mistral Devstral Medium is developed by Mistral AI. You can compare their actual outputs across 54 challenges on Rival to see how they differ in practice.

Which is better, Claude Opus 4.8 or Mistral Devstral Medium?

It depends on your use case. Claude Opus 4.8 and Mistral Devstral Medium each have strengths in different areas. Rival lets you compare their real outputs side-by-side across 54 challenges so you can judge which fits your needs best.

How much does Claude Opus 4.8 cost compared to Mistral Devstral Medium?

Claude Opus 4.8 costs $5/M input tokens and Mistral Devstral Medium costs $0.4/M input tokens. Mistral Devstral Medium is $4.60/M cheaper per input. Check their side-by-side outputs on Rival to see if the price difference is justified by quality.

How can I compare Claude Opus 4.8 and Mistral Devstral Medium on Rival?

This page shows a side-by-side comparison of Claude Opus 4.8 and Mistral Devstral Medium across shared challenges. You can vote on which model produced the better output in a blind duel. Browsing and voting are free. No account is needed to look; signing in only saves your votes and likes.

More comparisons

Against the newest arrivals

  • Claude Opus 4.8 vs Step 5 PreviewLanded Oct 2026
  • Mistral Devstral Medium vs Claude Haiku 5.5Landed Oct 2026
  • Claude Opus 4.8 vs Ling 3.1 FlashLanded Oct 2026
  • Mistral Devstral Medium vs Mistral Large 4Landed Oct 2026
  • Claude Opus 4.8 vs GPT-6.1 SolLanded Sep 2026
  • Mistral Devstral Medium vs Claude Sonnet 5.5Landed Sep 2026
  • Claude Opus 4.8 vs Solar Mini 4Landed Sep 2026
  • Mistral Devstral Medium vs Qwen3.8 Max PrimeLanded Sep 2026

Same lab, same size, long tail

  • Claude Opus 4.8 vs Claude Haiku 5.5Same lab
  • Claude Opus 4.8 vs Claude Sonnet 5.5Same lab
  • Mistral Devstral Medium vs Mistral Small 4Same lab
  • Mistral Devstral Medium vs Mistral Devstral Small 1.1Same lab
  • Claude Opus 4.8 vs Kimi Linear 48B A3B InstructSame size
  • Mistral Devstral Medium vs Laguna S 2.1 (free)New provider
  • Mistral Devstral Medium vs Laguna XS 2.1 (free)New provider
  • Claude Opus 4.8 vs LFM2.5-2.6BNew provider

Model pages

  • Claude Opus 4.858 outputs, specs and price
  • Mistral Devstral Medium59 outputs, specs and price
Workshop · 150 skills

Build your own version.

Editable skills and working examples for your agent.

Explore $79Free sample
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare models
  • All models
  • Image comparison
  • Audio comparison
  • Image generation
  • Best AI for...
  • Arena
  • API pricing
  • Challenges

Discover

  • SubjectiveBench
  • Default Index
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model updates
  • Provider status
  • AI creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy policy
  • Terms
  • RSS feed