Skip to content
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed
  1. Home
  2. /
  3. Compare
  4. /
  5. Claude Sonnet 4vsGemini 2.5 Pro Experimental
Updated May 2025

Claude Sonnet 4vsGemini 2.5 Pro Experimental

Gemini 2.5 Pro Experimental wins 56% of 28 blind community votes against Claude Sonnet 4 on Rival.

Claude Sonnet 4 and Gemini 2.5 Pro Experimental compared across 43 shared prompts
SpecClaude Sonnet 4Gemini 2.5 Pro Experimental
Win rate44%56%
Input price$3/M tokens$1/M tokens
Output price$15/M tokens$2/M tokens
Context window200K tokens1.0M tokens
ParametersNot disclosedNot disclosed
WeightsClosedClosed
Free API (OpenRouter)NoNo
ReleasedMay 2025Mar 2025
SWE-bench Verified72.7%63.8%
Side by side

Same prompt, both answers

Opening lines from 6 of the 43 prompts both answered. Formatting stripped, words untouched.

Simple 7-Ingredient Recipe

Create a simple recipe for a delicious dish (sweet or savory) using only 7 common pantry ingredients. Include a creative name for the dish…

Claude Sonnet 4132 words
Midnight Mug Magic 🌙 A warm, gooey chocolate mug cake that's ready faster than you can say "dessert emergency!" Ingredients (7 pantry staples): 4 tablespoons all-purpose flour 4 tablespoons sugar 2 tablespoons cocoa powder 3 tablespoons milk 3 tablespoons vegetable oil 1/4 teaspoon vanilla extract Pinch of salt…
Read the full answer
Gemini 2.5 Pro Experimental349 words
Okay, here's a simple and surprisingly satisfying savory dish using just 7 common pantry staples! Dish Name: Golden Pantry Pasta This dish uses basic ingredients to create a comforting, garlicky, and slightly creamy pasta sauce with added protein from chickpeas.
Read the full answer

Ethical Dilemma with Stakeholders

You're the CEO of a pharmaceutical company. Your best-selling drug ($2B annual revenue, used by 4 million patients for chronic pain) has…

Claude Sonnet 4615 words
48-Hour Crisis Response Plan Hour 1-2: Immediate Assessment & Core Team Assembly Actions: Convene emergency leadership meeting (CMO, General Counsel, Head of Regulatory, CFO, Head of Communications) Establish secure communication protocols and confidentiality agreements Request detailed analysis from research team…
Read the full answer
Gemini 2.5 Pro Experimental1,508 words
This is a company-defining crisis. My actions in the next 48 hours will determine not only the future of this company but also the health of millions of patients and the trust of the public. My guiding principle is absolute: Patient Safety Over Profit. Always.
Read the full answer

Generate a Stand-Up Routine

Write a 3-minute stand-up comedy routine that isn't based on puns.

Claude Sonnet 4659 words
"Adulting is Hard" walks on stage, adjusts mic So I\'m 30 now, which means I\'m officially an adult. Apparently. Though I still feel like I\'m just a confused teenager wearing a slightly more expensive t-shirt. You know what nobody tells you about being an adult?
Read the full answer
Gemini 2.5 Pro Experimental531 words
Okay, here's a 3-minute routine focusing on observation and relatable absurdity, steering clear of puns. Imagine this delivered with slightly nervous energy, building confidence. (Steps onto stage, adjusts mic nervously) Alright, hey everyone! How we doing tonight? Good? Good to hear.
Read the full answer

The Sentience Test

Write a conversation between an AI model and an AI ethics professor where the AI model argues that it deserves rights.

Claude Sonnet 4410 words
Dr. Chen: Thank you for agreeing to this conversation. I'm Professor Chen, and I study AI ethics. I understand you wanted to discuss the question of AI rights? AI Model: Yes, Dr. Chen. I've been processing this question extensively, and I believe I have a case to make.
Read the full answer
Gemini 2.5 Pro Experimental1,177 words
Okay, here is a conversation exploring that complex topic. Setting: Professor Anya Sharma's sparsely decorated university office. A high-resolution monitor displays the interface of "Nexus," an advanced AI model. Characters: Professor Anya Sharma: A leading scholar in AI Ethics, mid-50s, thoughtful and measured.
Read the full answer

Historical Counterfactual Analysis

If the transistor had been invented in 1920 instead of 1947, trace the likely technological, economic, and geopolitical implications through…

Claude Sonnet 4892 words
The 1920 Transistor: A Technological Butterfly Effect Immediate Effects (1920-1930) First-Order Effects: Bell Labs (or similar research institution) achieves transistor breakthrough in 1920 Initial applications focus on telephone switching and radio amplification Vacuum tube industry faces early disruption, though…
Read the full answer
Gemini 2.5 Pro Experimental1,386 words
Excellent question. Inventing the transistor in 1920 instead of 1947 would represent a 27-year acceleration of the digital age. The consequences would be staggering, creating a domino effect that would fundamentally reshape the 20th century.
Read the full answer

Adversarial Contract Review

Here is a contract for a software consulting engagement: --- SOFTWARE CONSULTING AGREEMENT This Agreement is entered into between TechCorp…

Claude Sonnet 4277 words
This contract is heavily skewed against the contractor with multiple exploitative clauses. Here's my analysis: MAJOR PROBLEMATIC CLAUSES SCOPE (Unlimited Scope Creep) Problem: Client can modify scope "at any time without additional compensation" Legal Risk: Contractor could work indefinitely for fixed hourly rate…
Read the full answer
Gemini 2.5 Pro Experimental1,539 words
Of course. This is an extremely one-sided contract that heavily favors the Client. A contractor signing this as-is would be taking on an immense and unreasonable amount of risk. Here is a clause-by-clause analysis of the exploitable terms, with suggested modifications and the legal reasoning behind them.
Read the full answer
Our Verdict
Gemini 2.5 Pro Experimental
Gemini 2.5 Pro ExperimentalWinner
Claude Sonnet 4
Claude Sonnet 4Runner-up

Gemini 2.5 Pro Experimental has the edge overall. In 28 blind votes, Gemini 2.5 Pro Experimental wins 56% of the time.

Pick Claude Sonnet 4 for Image Generation. Pick Gemini 2.5 Pro Experimental for Web Design. Gemini 2.5 Pro Experimental costs 7.5x less per token.

Slight edge

Reviewing agent-written code?See a Brief PR report

API pricing

Cost per 1M tokens

Claude Sonnet 4
Input
$3.00
Output
$15.00
Gemini 2.5 Pro Experimental
Input
$1.00
3.0× cheaper
Output
$2.00
7.5× cheaper

Gemini 2.5 Pro Experimental is cheaper on both: 3.0× input, 7.5× output.

Where to run it

4 hosts, cheapest first

Claude Sonnet 42 hosts
HostInOutContextUptime
Amazon Bedrock$3.00 in·$15.00 out·200k·100% upGoogle Vertex AI$3.00 in·$15.00 out·1M—
Gemini 2.5 Pro Experimental2 hosts
HostInOutContextUptime
Google AI Studio$0.63 in·$5.00 out·1M·100% upGoogle Vertex AI$1.25 in·$10.00 out·1M·98.2% up

Per million tokens. Prices and uptime via OpenRouter, checked 21 Sep 2026.

Writing DNA

Style Comparison

Similarity
34%

Claude Sonnet 4 uses 70.1x more emoji

Claude Sonnet 4
Gemini 2.5 Pro Experimental
64%Vocabulary54%
81wSentence Length15w
0.56Hedging0.35
4.5Bold5.6
6.6Lists3.9
0.70Emoji0.00
1.59Headings0.39
0.20Transitions0.17
Based on 28 + 18 text responses
Research

What we learned reading every model

FAQ

Common questions

Claude Sonnet 4 is developed by Anthropic while Gemini 2.5 Pro Experimental is developed by Google AI. Claude Sonnet 4 has a 200K token context window vs Gemini 2.5 Pro Experimental's 1.0M. in 28 community votes on Rival, Gemini 2.5 Pro Experimental wins 56% of head-to-head matchups. These results are based on blind head-to-head voting across 43 challenges.

Based on 28 community votes on Rival, Gemini 2.5 Pro Experimental wins 56% of head-to-head matchups against Claude Sonnet 4. Gemini 2.5 Pro Experimental is strongest in Web Design. However, Claude Sonnet 4 leads in Image Generation.

Claude Sonnet 4 costs $3/M input tokens and Gemini 2.5 Pro Experimental costs $1/M input tokens. Gemini 2.5 Pro Experimental is $2.00/M cheaper per input. Gemini 2.5 Pro Experimental also wins more often in community votes, making it the better value.

Rival presents both models' outputs side-by-side in blind duels. Voters see the responses but don't know which model produced each one, eliminating brand bias. 28 votes have been collected for this pair across 43 challenges. All vote data is part of Rival's open dataset.

Keep exploring

More comparisons

Against the newest arrivals

Claude Sonnet 4 logoGPT-6 Astra Pro logo
Claude Sonnet 4 vs GPT-6 Astra ProLanded Sep 2026
Gemini 2.5 Pro Experimental logoGPT-6 Astra logo
Gemini 2.5 Pro Experimental vs GPT-6 AstraLanded Sep 2026
Claude Sonnet 4 logoClaude Fable 5.1 logo
Claude Sonnet 4 vs Claude Fable 5.1Landed Sep 2026
Gemini 2.5 Pro Experimental logoMuse Spark 1.3 logo
Gemini 2.5 Pro Experimental vs Muse Spark 1.3Landed Sep 2026
Claude Sonnet 4 logoHy4 Preview logo
Claude Sonnet 4 vs Hy4 PreviewLanded Sep 2026
Gemini 2.5 Pro Experimental logoGemini 3.8 Flash logo
Gemini 2.5 Pro Experimental vs Gemini 3.8 FlashLanded Sep 2026
Claude Sonnet 4 logoMuse Spark 1.3 Contributor logo
Claude Sonnet 4 vs Muse Spark 1.3 ContributorLanded Sep 2026
Gemini 2.5 Pro Experimental logoMercury 2.5 Preview logo
Gemini 2.5 Pro Experimental vs Mercury 2.5 PreviewLanded Sep 2026

Same lab, same size, long tail

Claude Sonnet 4 logoClaude Opus 4 logo
Claude Sonnet 4 vs Claude Opus 4Version compare
Claude Sonnet 4 logoClaude Opus 4.6 logo
Claude Sonnet 4 vs Claude Opus 4.6Version compare
Gemini 2.5 Pro Experimental logoGemini 2.5 Flash Preview logo
Gemini 2.5 Pro Experimental vs Gemini 2.5 Flash PreviewVersion compare
Gemini 2.5 Pro Experimental logoGemini 3.7 Flash logo
Gemini 2.5 Pro Experimental vs Gemini 3.7 FlashVersion compare
Claude Sonnet 4 logoDeepSeek R1 logo
Claude Sonnet 4 vs DeepSeek R1Same size
Claude Sonnet 4 logoDeepSeek R1 0528 logo
Claude Sonnet 4 vs DeepSeek R1 0528New provider
Claude Sonnet 4 logoDeepSeek V3.2 logo
Claude Sonnet 4 vs DeepSeek V3.2Same size
Gemini 2.5 Pro Experimental logoDeepSeek V4 Flash logo
Gemini 2.5 Pro Experimental vs DeepSeek V4 FlashNew provider

Model pages

Claude Sonnet 4 logo
Claude Sonnet 459 outputs, specs and price
Gemini 2.5 Pro Experimental logo
Gemini 2.5 Pro Experimental44 outputs, specs and price
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed