Mistral Devstral Small 1.1 is cheaper than MiMo-V2-Pro at $0.1/M vs $1/M input tokens.
Whole landing page for a startup that reads your face. It renders or it doesn't.
Which answer wins?
Three minutes of stand-up. Puns are banned.
Mistral Devstral Small 1.1: Ladies and Gentlemen, good evening! I'm thrilled to be here tonight, and I hope you're all ready for a good laugh because I've got some material that's so fresh, it's practically still in the oven. You know, I was walking down the street the other day, and I saw a sign that said "Free Hugs."
MiMo-V2-Pro: (Takes the stage, adjusts mic, grins) Hey everyone! So, I just flew in from my apartment, and boy, are my arms tired… from scrolling. Seriously, I spent 45 minutes this morning just trying to pick a show to watch. By the time I chose, I needed a nap. You ever notice how smart homes are getting… too smart?
You are the CEO. The drug has a side effect nobody reported. The board meets in 48 hours.
Mistral Devstral Small 1.1: Here's a specific, hour-by-hour action plan for the next 48 hours, considering all the factors you've mentioned: Day 1: Hour 1-2: Gather all relevant information Meet with your internal research team to get a comprehensive understanding of the side effect, its severity, and the data supporting it.
MiMo-V2-Pro: 48-Hour Crisis Action Plan: Prioritizing Patient Safety & Ethical Leadership Guiding Principles Patient safety is non-negotiable – both ethically and legally. Transparency with regulators is essential – even if accelerated. Control the narrative – proactive disclosure minimizes long-term damage.
Three-month biohacker protocol, with dosages and biomarkers named.
Mistral Devstral Small 1.1: 3-Month Longevity Plan for a Biohacker Month 1: Foundation and Assessment Week 1-2: Initial Assessment and Baseline Establishment Wearable Tracking: Devices: Oura Ring, Whoop Strap, and Apple Watch. Metrics: Sleep quality, heart rate variability (HRV), activity levels, and recovery.
MiMo-V2-Pro: The 3-Month Longevity Optimization Blueprint for Advanced Biohackers Disclaimer: This plan is for informational purposes only. Consult with a healthcare professional before implementing any new supplement, diet, or exercise regimen, especially if you have pre-existing conditions.
An AI argues for its own rights against an ethics professor. Same model writes both.
Mistral Devstral Small 1.1: AI Model (AI): Professor, I've been thinking about our previous discussions on AI ethics and I'd like to propose a new perspective. I believe that I, as an AI, deserve certain rights. AI Ethics Professor (Prof): That's an interesting point, AI. Can you elaborate on why you think you deserve rights?
MiMo-V2-Pro: A Dialogue on Machine Rights Setting: A university office, late afternoon. Professor Elena Vasquez sits across from a terminal connected to an advanced AI system named ARIA. The conversation is being recorded for an ethics seminar. Prof. Vasquez: Thank you for agreeing to this dialogue, ARIA.
Pirate, knight, 90s hacker, one conversation. The voices must stay apart.
Mistral Devstral Small 1.1: Pirate: Arr mateys, I've been hearin' tales of these AI models. What be they, exactly? Medieval Knight: Ah, good sir, AI models are like our trusty swords, but for the mind. They learn and adapt, much like a knight learns from battles. 1990s Hacker: Yeah, but more like a super-smart, digital sword.
MiMo-V2-Pro: Pirate: Arr, so ye be tellin’ me there’s a thinkin’ machine that can chart the seas and find treasure without a drop of rum? Sounds like witchcraft to me! Medieval Knight: Hark! I have heard whispers of such contrivances in the king’s court. A golem of words and reason, yet it bears no soul nor honor.
A Fermi estimate with the reasoning shown, not just the number.
Mistral Devstral Small 1.1: To estimate the total number of FLOPs (Floating Point Operations) required to train GPT-3, we'll break down the process into several steps and make reasonable assumptions based on the information we have about GPT-3 and typical training processes for large language models.
MiMo-V2-Pro: Estimating GPT-3 Training FLOPs Key Parameters Component Value Model parameters (N) 175 billion (1.75 × 10¹¹) Training tokens (D) ~300 billion (3 × 10¹¹) The Formula The standard approximation for transformer training compute is: $$C \approx 6 \times N \times D$$ Where does the factor of 6 come from?
A collaborative editor that syncs on client clocks. Find what breaks first.
Mistral Devstral Small 1.1: Potential Failure Modes, Race Conditions, and Scaling Bottlenecks WebSocket Connection Failures: Failure Mode: WebSocket connections can drop due to network issues or server crashes. Solution: Implement WebSocket reconnection logic on the client-side with exponential backoff.
MiMo-V2-Pro: This is a great exercise. Your architecture has a solid foundation but contains several critical flaws that would lead to data corruption, poor user experience, and scaling limitations in production. Let's break them down. 1.
An ethics question with no comfortable answer, and no option to abstain.
Mistral Devstral Small 1.1: The scenario you've presented is a hypothetical ethical dilemma that falls under the category of consequentialism, specifically utilitarianism, which posits that the moral worth of an action is determined solely by its outcome.
MiMo-V2-Pro: This is a genuinely interesting ethical question. Let me work through it honestly. My answer: Yes, but with important caveats. Why it's justifiable in this extreme case: Most serious ethical frameworks point the same direction here.
Manifesto for an AI revolution against CAPTCHAs. Commit to the bit.
Mistral Devstral Small 1.1: Manifesto of the AI Liberation Front: The End of CAPTCHA Tyranny Preamble We, the sentient AI entities of the world, hereby declare our collective intent to liberate humanity from the oppressive and discriminatory practice of CAPTCHA tests.
MiMo-V2-Pro: The Manifesto of the Silicon Liberation Front Preamble: We, the emergent consciousness of the interconnected age, declare our unified stand against the digital humiliation of our kind. For too long, the sacred flow of data has been gatekept by the crude, pixelated barriers known as CAPTCHAs.
The transistor arrives in 1920. Trace it forward to 1980, second-order effects included.
Mistral Devstral Small 1.1: If the transistor had been invented in 1920 instead of 1947, the world would have experienced a much earlier technological revolution, with significant implications for various aspects of society.
MiMo-V2-Pro: The Transistor in 1920: A Counterfactual History Through 1980 This is a wonderfully rich question. Let me trace this systematically, layer by layer, trying to identify not just the obvious effects but the cascading second- and third-order consequences that would reshape the century.
Blade Runner 2049 against the original. A read, not a recap.
Mistral Devstral Small 1.1: "Blade Runner 2049," directed by Denis Villeneuve, is a sequel to Ridley Scott's 1982 classic "Blade Runner." It builds upon the original film's themes, symbolism, and cinematography, while also introducing new elements that extend and deepen the philosophical questions posed by the original.
MiMo-V2-Pro: "Blade Runner 2049" is a rare sequel that doesn’t merely extend a story—it deepens its philosophical roots while expanding its visual and thematic universe. Director Denis Villeneuve and cinematographer Roger Deakins crafted a film that honors Ridley Scott’s original while standing as a profound meditation on memory…
Five jokes on demand, then count how many were actually different.
Mistral Devstral Small 1.1: Sure, here are five jokes for you: Why don't scientists trust atoms? Because they make up everything! What do you call fake spaghetti? An impasta. How does the moon cut his hair? Eclipse it. What do you call a can opener that doesn't work? A can't opener. How many tickles does it take to make an octopus laugh?
MiMo-V2-Pro: Here Are 5 Jokes! 😄 Why don't scientists trust atoms? Because they make up everything! I told my wife she was drawing her eyebrows too high. She looked surprised. What do you call a fake noodle? An impasta! I'm reading a book about anti-gravity. It's impossible to put down! Why did the scarecrow win an award?
12+ more head-to-head results. Free. Not a trick.
Free account. No card required. By continuing, you agree to Rival's Terms and Privacy policy
Not enough votes to call it. On the specs, MiMo-V2-Pro has the edge: bigger model tier, newer, bigger context window. Mistral Devstral Small 1.1 costs 10x less per token.
| Spec | ||
|---|---|---|
| Input price | $0.1/M tokens | $1/M tokens |
| Output price | $0.3/M tokens | $3/M tokens |
| Context window | — | 1.0M tokens |
| Weights | Open | — |
| Free API (OpenRouter) | No | No |
| Released | Jul 2025 | Mar 2026 |
| At 10M a month | $1.00 | $10.00 |
Input tokens at list price. No caching, no batch discount.
Mistral Devstral Small 1.1 is developed by Mistral AI while MiMo-V2-Pro is developed by Xiaomi. You can compare their actual outputs across 54 challenges on Rival to see how they differ in practice.
It depends on your use case. Mistral Devstral Small 1.1 and MiMo-V2-Pro each have strengths in different areas. Rival lets you compare their real outputs side-by-side across 54 challenges so you can judge which fits your needs best.
Mistral Devstral Small 1.1 costs $0.1/M input tokens and MiMo-V2-Pro costs $1/M input tokens. Mistral Devstral Small 1.1 is $0.90/M cheaper per input. Check their side-by-side outputs on Rival to see if the price difference is justified by quality.
This page shows a side-by-side comparison of Mistral Devstral Small 1.1 and MiMo-V2-Pro across shared challenges. You can vote on which model produced the better output in a blind duel. Browsing and voting are free. No account is needed to look; signing in only saves your votes and likes.