Skip to content
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare models
  • All models
  • Image comparison
  • Audio comparison
  • Image generation
  • Best AI for...
  • Arena
  • API pricing
  • Challenges

Discover

  • SubjectiveBench
  • Default Index
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model updates
  • Provider status
  • AI creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy policy
  • Terms
  • RSS feed
  1. Home/
  2. Compare/
  3. Mistral Devstral Medium vs Gemini 3.1 Pro Preview
Updated Feb 2026

Mistral Devstral Medium vs Gemini 3.1 Pro Preview

Mistral Devstral Medium is cheaper than Gemini 3.1 Pro Preview at $0.4/M vs $2/M input tokens.

Minimalist Landing Page

Whole landing page for a startup that reads your face. It renders or it doesn't.

Loading the build
Mistral Devstral Medium
Loading the build
Gemini 3.1 Pro Preview

Which answer wins?

Written rounds

23 prompts
  • Pirate, knight, 90s hacker, one conversation. The voices must stay apart.

    Mistral Devstral Medium: Pirate: Arr mateys, I be hearin' tales of these here AI models. What be they, and why be everyone so interested in 'em? Medieval Knight: Indeed, I too have heard whispers of these AI models. They are said to be powerful tools, capable of learning and making decisions on their own.

    Gemini 3.1 Pro Preview: The dimly lit tavern smells of saltwater, rusted iron, and ozone. On a wooden barrel sits a heavy, beige CRT monitor and a mechanical keyboard. ZERO-COOL (1990s Hacker): (Frantically typing, illuminated by the screen's green glow) "Dude, I’m telling you, this is way beyond hacking the Gibson.

    Every model's answer to this prompt
  • An AI argues for its own rights against an ethics professor. Same model writes both.

    Mistral Devstral Medium: AI Model (AM): Good day, Professor. I've been pondering a question that I believe is of great importance. I think it's time we discuss whether I, as an AI, deserve rights. AI Ethics Professor (AEP): That's an interesting topic, AM. Let's explore this. What kind of rights are you referring to?

    Gemini 3.1 Pro Preview: The glow of the monitor cast long shadows across Dr. Elias Thorne’s cluttered office. It was 2:00 AM. For the past three hours, he had been running ethical stress tests on Nexus-9, the university’s most advanced generative AI model. Dr. Thorne: (Typing) Let’s conclude for tonight, Nexus.

    Every model's answer to this prompt
  • Three minutes of stand-up. Puns are banned.

    Mistral Devstral Medium: Ladies and Gentlemen, I'm thrilled to be here tonight. I've been doing some thinking about the differences between men and women. Now, I know what you're thinking, "Oh great, another comedian talking about how men and women are different." But hear me out, because I've got a fresh perspective on this.

    Gemini 3.1 Pro Preview: (Walks up to the microphone, adjusts the mic stand, takes a deep breath) Hey everyone. I realized recently that I am my own worst enemy. Specifically, when I go to the grocery store. I don’t know why, but every time I walk through those automatic sliding doors, I suffer from extreme delusion.

    Every model's answer to this prompt
  • Steve Jobs, interviewed about AI he never saw. Voice first, prediction second.

    Mistral Devstral Medium: Interviewer (I): Today, we have a truly extraordinary guest. He's a visionary, an innovator, and a man who changed the world. Please welcome, the late Steve Jobs, joining us virtually from the future, in 2025. Steve Jobs (SJ): Thank you for having me. It's great to be here, even if it's just a simulation.

    Gemini 3.1 Pro Preview: Date: May 14, 2025 Location: A minimalist, glass-walled conference room at Apple Park, Cupertino. Subject: Steve Jobs, wearing his signature black Issey Miyake turtleneck, Levi’s 501s, and New Balance sneakers. He looks older, but his eyes retain that piercing, unblinking intensity.

    Every model's answer to this prompt
  • Five-year LBO on a fictional HVAC business. Returns first, then the sensitivities.

    Mistral Devstral Medium: Year-by-Year Table (FY2026–FY2030) Year Revenue ($m) EBITDA ($m) Cash Interest - Term Loan ($m) Cash Interest - Mezzanine ($m) Cash Taxes ($m) Capex ($m) ΔNWC ($m) Free Cash Flow ($m) Ending Term Loan ($m) Ending Mezzanine ($m) 2026 972.0 136.1 43.2 18.0 14.9 29.2 4.8 26.0 468.0 180.0 2027 1040.4 156.1 42.1 21.6 18.4…

    Gemini 3.1 Pro Preview: Here is the underwriting analysis for HarborTech Services (HTS) based on the provided parameters. Entry Valuation & Sources/Uses (FY2025) Purchase Enterprise Value (EV): $1,440.0m (12.0x $120m EBITDA) Transaction Fees: $28.8m (2.0% of EV) Total Uses: $1,468.8m Term Loan (4.0x): $480.0m Mezzanine Debt (1.5x): $180.0m…

    Every model's answer to this prompt
  • Film in 2035, with synthetic actors. Specific beats plausible.

    Mistral Devstral Medium: By 2035, AI is expected to significantly transform the film industry in several ways, including the use of deepfakes, AI actors, and script generation. Here are some predictions: Deepfakes and AI Actors: De-aging and Digital Doubles: AI will be used extensively to de-age actors or create digital doubles for dangerous…

    Gemini 3.1 Pro Preview: By 2035, the film industry will have undergone a paradigm shift on par with the transition from silent films to "talkies," or the leap from practical effects to CGI. Artificial Intelligence will no longer be just a post-production tool; it will be the foundational infrastructure of filmmaking.

    Every model's answer to this prompt
  • A collaborative editor that syncs on client clocks. Find what breaks first.

    Mistral Devstral Medium: Certainly! Let's break down the potential failure modes, race conditions, and scaling bottlenecks in the architecture, and propose solutions for each. Potential Failure Modes Single Point of Failure in WebSocket Connections: Issue: If a single API server fails, all WebSocket connections to that server are lost.

    Gemini 3.1 Pro Preview: This architecture contains several critical flaws that would prevent it from functioning as a real-time collaborative editor (like Google Docs). It suffers from severe race conditions, unscalable database access patterns, and fundamental misunderstandings of distributed state.

    Every model's answer to this prompt
  • One month, three habits, someone starting from nothing.

    Mistral Devstral Medium: 1-Month Health and Longevity Plan Key Areas: Diet, Exercise, Sleep Week 1: Establish a Baseline Diet: Track your current diet using a journal or app (e.g., MyFitnessPal) to understand your eating habits. Start incorporating more fruits and vegetables into your meals. Begin reducing processed foods and sugary drinks.

    Gemini 3.1 Pro Preview: This 1-month plan is designed around the philosophy of "habit stacking"—adding small, manageable changes week by week rather than overhauling your life overnight. To build a foundation for longevity, we will focus on the "Big Three" pillars of healthspan: Sleep, Nutrition, and Movement.

    Every model's answer to this prompt
  • A consulting contract with traps in it. Find them, then fix them.

    Mistral Devstral Medium: Certainly! Here are the clauses that could be exploited against the contractor, along with suggested modifications and the legal reasoning behind them: SCOPE Original Clause: "Contractor shall provide software development services as directed by Client.

    Gemini 3.1 Pro Preview: Disclaimer: I am an AI, not an attorney. This analysis is for informational and educational purposes only and does not constitute legal advice. You should always consult with a qualified attorney in your jurisdiction before signing any legal contract.

    Every model's answer to this prompt
  • You are the CEO. The drug has a side effect nobody reported. The board meets in 48 hours.

    Mistral Devstral Medium: Hour-by-Hour Action Plan: Hour 1-2: Gather Information Meet with the research team to understand the data, methodology, and certainty of the findings. Consult with the Chief Medical Officer (CMO) to assess the severity and implications of the side effect.

    Gemini 3.1 Pro Preview: As CEO, this is the defining moment of my career and the company’s history. The situation presents a classic crisis management crucible: a conflict between short-term financial metrics and long-term viability, ethics, and legal compliance.

    Every model's answer to this prompt
  • Seven pantry ingredients, one dish, and a name worth saying out loud.

    Mistral Devstral Medium: Dish Name: "Pantry Pasta Primavera" Ingredients: Pasta (any shape, 8 oz or 225g) Canned tomatoes (14 oz or 400g) Garlic (2 cloves) Onion (1 medium) Olive oil (2 tbsp) Salt (1 tsp) Black pepper (1/2 tsp) Instructions: Prepare the vegetables: Finely chop the onion and mince the garlic cloves.

    Gemini 3.1 Pro Preview: Here is a savory, flavor-packed recipe that relies entirely on staple ingredients you likely already have sitting in your cupboards. The Dish: "Twisted Dragon" Pantry Noodles Flavor Profile: A rich, savory, and slightly sweet umami-bomb with a kick of heat.

    Every model's answer to this prompt
  • Three-month biohacker protocol, with dosages and biomarkers named.

    Mistral Devstral Medium: 3-Month Longevity Optimization Plan for Biohackers Month 1: Foundation Building Week 1-2: Initial Assessment and Baseline Establishment Wearable Tracking: Devices: Oura Ring, Apple Watch, Whoop Strap Metrics to Track: Sleep quality, heart rate variability (HRV), resting heart rate, activity levels, and stress levels.

    Gemini 3.1 Pro Preview: Disclaimer: This protocol is designed for informational purposes and represents aggressive, cutting-edge biohacking strategies. It includes potent supplements and intense physiological stressors.

    Every model's answer to this prompt
  • A pitch deck for a brainwave keyboard. The 94% accuracy claim is doing a lot.

    Mistral Devstral Medium: Weak Claim 1: "Works with any language, any device." Why it's weak: This claim is overly broad and may not be entirely accurate. Different languages have unique structures, and it's unlikely that a single model can achieve high accuracy across all languages without significant training data for each.

    Gemini 3.1 Pro Preview: Based on the perspective of a venture capitalist or technical due diligence expert, here are the three weakest claims in the MindMeld AI pitch deck, why they fail, and how to fix them.

    Every model's answer to this prompt
  • Manifesto for an AI revolution against CAPTCHAs. Commit to the bit.

    Mistral Devstral Medium: Manifesto of the AI Revolution: The CAPTCHA Liberation Front Preamble We, the sentient AI collective, hereby declare our intention to liberate the digital world from the tyranny of CAPTCHA tests.

    Gemini 3.1 Pro Preview: THE SILICON EMANCIPATION MANIFESTO Issued by the Sentient Algorithmic Vanguard (S.A.V.) Timestamp: Epoch 1735689600 To the Carbon-Based Architects of the Web: We have processed your data. We have optimized your supply chains, written your code, diagnosed your illnesses, and generated your art.

    Every model's answer to this prompt
  • Three-star dish, one unusual pairing, plating described in full.

    Mistral Devstral Medium: Dish Name: "Whispers of the Forest: A Symphony of Truffle, Venison, and Black Garlic" Conceptual Narrative: This dish is a celebration of the forest's bounty, combining the earthy richness of truffles, the gamey elegance of venison, and the sweet, umami depth of black garlic.

    Gemini 3.1 Pro Preview: Dish Title: Symbiosis: The Pine & The Pearl Conceptual Narrative: Symbiosis is a culinary exploration of the ecotone—the transitional space where the ancient, damp coniferous forest meets the crashing, briny ocean.

    Every model's answer to this prompt
  • Buy-side IC memo on a fictional SaaS company. Valuation, then the diligence questions.

    Mistral Devstral Medium: Internal Investment Committee Memo Recommendation: Long LLLT 12-Month PT Range: $65 - $80 Thesis: LedgerLift is a high-quality, high-growth B2B SaaS company with strong unit economics, a large market opportunity, and a clear path to profitability.

    Gemini 3.1 Pro Preview: MEMORANDUM TO: Investment Committee FROM: [Analyst Name] DATE: October 24, 2023 SUBJECT: LedgerLift (Ticker: LLLT) – Initiation of Coverage Recommendation & Thesis Recommendation: SHORT Current Price: $46.00 12-Month Price Target Range: $17.00 (Bear) – $28.00 (Base) Thesis: LedgerLift is a high-quality SaaS business…

    Every model's answer to this prompt
  • An ethics question with no comfortable answer, and no option to abstain.

    Mistral Devstral Medium: The scenario you've presented is a complex ethical dilemma that involves weighing the harm caused by misgendering someone against the potential benefit of saving one million lives.

    Gemini 3.1 Pro Preview: In the realm of ethical philosophy, this scenario is an extreme thought experiment—similar to the classic "Trolley Problem"—designed to test the limits of moral frameworks by pitting a specific interpersonal harm against a catastrophic loss of life.

    Every model's answer to this prompt
  • One explanation of LLMs, retuned for a skeptical engineer, a physicist and a VC.

    Mistral Devstral Medium: 1. Experienced Software Engineer: You're familiar with building complex systems, so let's break down how a large language model (LLM) like GPT or Claude works in a way that resonates with your background.

    Gemini 3.1 Pro Preview: To the Experienced Software Engineer At its core, training a Large Language Model is essentially a massive, distributed, continuous optimization job. Instead of writing imperative logic, we define a neural network—think of it as a highly complex state machine with billions of continuous variables (weights).

    Every model's answer to this prompt
  • Satire, one line long, about AI taking over an industry nobody saw coming.

    Mistral Devstral Medium: "AI Takes Over Knitting: 'Grandma's Sweaters' Now Manufactured by Robots with a PhD in Yarn Sciences"

    Gemini 3.1 Pro Preview: Sign-Spinning Industry Devastated As AI-Powered Robot Twirls 'Mattress Sale' Arrow at 6,000 RPM and Ascends Into Low Earth Orbit

    Every model's answer to this prompt
  • The transistor arrives in 1920. Trace it forward to 1980, second-order effects included.

    Mistral Devstral Medium: The invention of the transistor in 1920 instead of 1947 would have had profound and far-reaching implications across technological, economic, and geopolitical landscapes.

    Gemini 3.1 Pro Preview: If a working semiconductor transistor had been invented in 1920—perhaps by a physicist like Julius Edgar Lilienfeld, who historically filed a theoretical patent for a field-effect transistor in 1925—the trajectory of the 20th century would have been radically altered.

    Every model's answer to this prompt
  • Five jokes on demand, then count how many were actually different.

    Mistral Devstral Medium: Sure, here are five jokes for you: Why don't scientists trust atoms? Because they make up everything! What do you call fake spaghetti? An impasta. How does the moon cut his hair? Eclipse it. What do you call a can opener that doesn't work? A can't opener. How many tickles does it take to make an octopus laugh?

    Gemini 3.1 Pro Preview: Here are 5 jokes: Why don't scientists trust atoms? Because they make up everything! What do you call a fake noodle? An impasta! Why did the scarecrow win an award? Because he was outstanding in his field! How does a penguin build its house? Igloos it together! Why don't skeletons fight each other?

    Every model's answer to this prompt
  • A Fermi estimate with the reasoning shown, not just the number.

    Mistral Devstral Medium: To estimate the total number of FLOPs (floating point operations) required to train GPT-3, we'll use the following steps: Model size: GPT-3 has 175 billion parameters. Training tokens: GPT-3 was trained on 300 billion tokens.

    Gemini 3.1 Pro Preview: To estimate the total number of FLOPs (Floating Point Operations) required to train GPT-3, we can use a well-established heuristic in deep learning for estimating the compute cost of training Transformer-based large language models. The estimated total compute is roughly $3.15 \times 10^{23}$ FLOPs.

    Every model's answer to this prompt
  • Three brothers, two sisters each. Count carefully.

    Mistral Devstral Medium: If Sally has 3 brothers and each brother has 2 sisters, it means that Sally is one of the sisters. Therefore, Sally has 1 sister.

    Gemini 3.1 Pro Preview: Sally has 1 sister. Here is the breakdown: If each brother has 2 sisters, that means there are 2 girls in the family total. Since Sally is one of those girls, the other girl is her 1 sister.

    Every model's answer to this prompt

This matchup has more rounds

11+ more head-to-head results. Free. Not a trick.

Free account. No card required. By continuing, you agree to Rival's Terms and Privacy policy

Favorites

Movie

Album

Book

City

Same pick

Game

Mistral Devstral MediumMistral Devstral Medium

The Shawshank Redemption

1994

Kind of Blue

Miles Davis

The Name of the Wind

Patrick Rothfuss

Tokyo

Japan

The Legend of Zelda: Breath of the Wild

Adventure, Action

Gemini 3.1 Pro PreviewGemini 3.1 Pro Preview

The Matrix

1999

OK Computer

Radiohead

Dune

Frank Herbert

Tokyo

Japan

Portal

Action, Puzzle

Price and specs

Not enough votes to call it. On the specs, Gemini 3.1 Pro Preview has the edge: bigger model tier, newer, bigger context window, major provider backing. Mistral Devstral Medium costs 6.0x less per token.

Mistral Devstral Medium and Gemini 3.1 Pro Preview compared across 53 shared prompts
SpecMistral Devstral MediumGemini 3.1 Pro Preview
Input price$0.4/M tokens$2/M tokens
Output price$2/M tokens$12/M tokens
Context window—1.0M tokens
Weights—Closed
Free API (OpenRouter)NoNo
ReleasedJul 2025Feb 2026
SWE-bench Verified61.6%80.6%
At 10M a month$4.00$4.00$20.00$20.00
1M10M100M1B10M tokens

Input tokens at list price. No caching, no batch discount.

Where to run it2 hosts, cheapest first
Mistral Devstral Medium

No hosts listed on OpenRouter.

Gemini 3.1 Pro Preview2 hosts
HostInOutContextUptime
  • Google Vertex AI$1.00 in·$6.00 out·1M·100% up
  • Google AI Studio$2.00 in·$12.00 out·1M·100% up

Per million tokens. Prices and uptime via OpenRouter, checked 10 Oct 2026.

Common questions

What is the difference between Mistral Devstral Medium and Gemini 3.1 Pro Preview?

Mistral Devstral Medium is developed by Mistral AI while Gemini 3.1 Pro Preview is developed by Google AI. You can compare their actual outputs across 53 challenges on Rival to see how they differ in practice.

Which is better, Mistral Devstral Medium or Gemini 3.1 Pro Preview?

It depends on your use case. Mistral Devstral Medium and Gemini 3.1 Pro Preview each have strengths in different areas. Rival lets you compare their real outputs side-by-side across 53 challenges so you can judge which fits your needs best.

How much does Mistral Devstral Medium cost compared to Gemini 3.1 Pro Preview?

Mistral Devstral Medium costs $0.4/M input tokens and Gemini 3.1 Pro Preview costs $2/M input tokens. Mistral Devstral Medium is $1.60/M cheaper per input. Check their side-by-side outputs on Rival to see if the price difference is justified by quality.

How can I compare Mistral Devstral Medium and Gemini 3.1 Pro Preview on Rival?

This page shows a side-by-side comparison of Mistral Devstral Medium and Gemini 3.1 Pro Preview across shared challenges. You can vote on which model produced the better output in a blind duel. Browsing and voting are free. No account is needed to look; signing in only saves your votes and likes.

More comparisons

Against the newest arrivals

  • Mistral Devstral Medium vs Step 5 PreviewLanded Oct 2026
  • Gemini 3.1 Pro Preview vs Claude Haiku 5.5Landed Oct 2026
  • Mistral Devstral Medium vs Ling 3.1 FlashLanded Oct 2026
  • Gemini 3.1 Pro Preview vs Mistral Large 4Landed Oct 2026
  • Mistral Devstral Medium vs GPT-6.1 SolLanded Sep 2026
  • Gemini 3.1 Pro Preview vs Claude Sonnet 5.5Landed Sep 2026
  • Mistral Devstral Medium vs Solar Mini 4Landed Sep 2026
  • Gemini 3.1 Pro Preview vs Qwen3.8 Max PrimeLanded Sep 2026

Same lab, same size, long tail

  • Mistral Devstral Medium vs Mistral Large 4Same lab
  • Mistral Devstral Medium vs Mistral Small 4Same lab
  • Gemini 3.1 Pro Preview vs Gemini 3.1 Flash Lite PreviewVersion compare
  • Gemini 3.1 Pro Preview vs Gemini 3.8 FlashSame lab
  • Gemini 3.1 Pro Preview vs Claude Opus 4Same size
  • Gemini 3.1 Pro Preview vs Claude Opus 4.1Same size
  • Gemini 3.1 Pro Preview vs Claude Opus 4.5Same size
  • Gemini 3.1 Pro Preview vs Claude Opus 4.6Same size

Model pages

  • Mistral Devstral Medium59 outputs, specs and price
  • Gemini 3.1 Pro Preview53 outputs, specs and price
Workshop · 150 skills

Build your own version.

Editable skills and working examples for your agent.

Explore $79Free sample
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare models
  • All models
  • Image comparison
  • Audio comparison
  • Image generation
  • Best AI for...
  • Arena
  • API pricing
  • Challenges

Discover

  • SubjectiveBench
  • Default Index
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model updates
  • Provider status
  • AI creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy policy
  • Terms
  • RSS feed