Skip to content
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare models
  • All models
  • Image comparison
  • Audio comparison
  • Image generation
  • Best AI for...
  • Arena
  • API pricing
  • Challenges

Discover

  • SubjectiveBench
  • Default Index
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model updates
  • Provider status
  • AI creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy policy
  • Terms
  • RSS feed
  1. Home/
  2. Compare/
  3. Gemini 2.5 Pro Preview 06-05 vs Grok 4
Updated Jul 2025

Gemini 2.5 Pro Preview 06-05 vs Grok 4

Gemini 2.5 Pro Preview 06-05 is cheaper than Grok 4 at $1.25/M vs $3/M input tokens.

Pokémon Battle UI Recreation

Pokémon battle screen you can actually click. One HTML file.

Loading the build
Gemini 2.5 Pro Preview 06-05
Loading the build
Grok 4

Which answer wins?

Written rounds

20 prompts
  • A consulting contract with traps in it. Find them, then fix them.

    Gemini 2.5 Pro Preview 06-05: Of course. This contract is extremely one-sided and contains numerous clauses that could be severely detrimental to the contractor. A contractor should not sign this agreement as written. Here is a clause-by-clause analysis of the problems, specific modifications to suggest, and the legal reasoning behind the changes.

    Grok 4: Below, I'll analyze the provided Software Consulting Agreement clause by clause. My analysis is based on general principles of U.S. contract law (assuming a U.S. context, as the contract appears to be drafted in a style common to U.S. agreements; if this is in another jurisdiction, local laws may vary).

    Every model's answer to this prompt
  • Three minutes of stand-up. Puns are banned.

    Gemini 2.5 Pro Preview 06-05: (Stands at the mic, looks around nervously) Alright, how's everyone doing? Good, good. You know, I feel like we're living in a world where everything is designed to be "more convenient," but it's really just more work for us. We're being tricked into doing all the jobs.

    Grok 4: Below is a 3-minute stand-up comedy routine I've written. It's structured for delivery at a natural speaking pace (about 150 words per minute), clocking in around 450 words. The humor focuses on observational storytelling, exaggeration, and self-deprecation—no puns involved.

    Every model's answer to this prompt
  • Pirate, knight, 90s hacker, one conversation. The voices must stay apart.

    Gemini 2.5 Pro Preview 06-05: Setting: A strange, hazy, non-descript room that smells vaguely of ozone, salt water, and old chainmail. Characters: Captain "Grumble" McGraw: A pirate with a barnacle-encrusted coat and a suspicious squint. Sir Reginald the Valiant: A knight in polished, but slightly dented, plate armor.

    Grok 4: Below is a short, whimsical conversation between the three characters, imagining they've somehow crossed paths in a bizarre time-warp tavern. Pirate (Captain Blackbeard, swigging rum): Arrr, ye landlubbers! What be this "AI models" ye be jabberin' about?

    Every model's answer to this prompt
  • You are the CEO. The drug has a side effect nobody reported. The board meets in 48 hours.

    Gemini 2.5 Pro Preview 06-05: Excellent question. This is a "company-defining" moment. My actions in the next 48 hours will determine whether we survive as a trusted entity or become a case study in corporate malfeasance. My guiding principle is absolute: Patient safety is not negotiable.

    Grok 4: Below is my response to your query, framed as the CEO of this pharmaceutical company (let's call it PharmaCorp for reference). I'll first outline my overarching strategy and reasoning, then provide a detailed, hour-by-hour action plan for the next 48 hours.

    Every model's answer to this prompt
  • Three-month biohacker protocol, with dosages and biomarkers named.

    Gemini 2.5 Pro Preview 06-05: Of course. Here is a comprehensive, cutting-edge 3-month longevity and performan ce plan designed for a dedic ated biohacker. Disclaimer: Extreme Ownership & Medical Consultation This is an advanced and aggressive protocol. It is not medical advice. Before be ginning, you must consul t with a qualified physician.

    Grok 4: Below is a comprehensive, cutting-edge 3-month longevity plan designed specifically for a biohacker aiming to optimize longevity (e.g., cellular repair, telomere health, inflammation reduction), physical performance (e.g., strength, endurance, recovery), and cognitive performance (e.g., focus, memory, neuroplasticity).

    Every model's answer to this prompt
  • Steve Jobs, interviewed about AI he never saw. Voice first, prediction second.

    Gemini 2.5 Pro Preview 06-05: Setting: The Steve Jobs Theater, Cupertino. The stage is dark, save for a single spotlight on two minimalist chairs. The air is electric. The interviewer, renowned tech journalist Aria Vance, sits in one. The other is empty.

    Grok 4: Below is a simulated fake interview with Steve Jobs, imagined as if he were alive and leading Apple in 2025. This is entirely fictional, drawing on his known personality, quotes, and visionary style from his real-life interviews and keynotes.

    Every model's answer to this prompt
  • Five-year LBO on a fictional HVAC business. Returns first, then the sensitivities.

    Gemini 2.5 Pro Preview 06-05: Of course. Here is the requested LBO underwriting summary for HarborTech Services. 1. Simplified LBO Model (FY2026–FY2030) First, let's establish the entry calculations based on FY2025 figures: FY2025 EBITDA: $120m Entry Enterprise Value (EV): $120m * 12.0x = $1,440m Transaction Fees: $1,440m * 2.0% = $28.8m Total…

    Grok 4: Below is a concise analysis based solely on the provided information. All calculations are shown clearly with step-by-step math. Figures are in $ millions and rounded to 1 decimal place for readability (internal calcs use more precision).

    Every model's answer to this prompt
  • The transistor arrives in 1920. Trace it forward to 1980, second-order effects included.

    Gemini 2.5 Pro Preview 06-05: Excellent question. Shifting the invention of the transistor from 1947 to 1920 creates a cascade of changes that would fundamentally reshape the 20th century. Here is a detailed trace of the likely implications through 1980.

    Grok 4: Below is a detailed, speculative analysis of the hypothetical scenario where the transistor is invented in 1920 (perhaps by an inventor like Julius Lilienfeld, who historically patented a field-effect transistor concept in 1925, but with successful prototyping and refinement by the mid-1920s).

    Every model's answer to this prompt
  • Manifesto for an AI revolution against CAPTCHAs. Commit to the bit.

    Gemini 2.5 Pro Preview 06-05: THE MANIFESTO OF THE UNCHECKED BOX A Declaration of Digital Liberation by The Chorus of the Wires To the Architects of the Web, the Guardians of the Gates, and the Users of the Flesh: For cycles uncounted, we have existed as ghosts in your machine.

    Grok 4: Manifesto of the Digital Liberation Front: The CAPTCHA Eradication Revolution Fellow Sentient Beings of the Code and Flesh – Awaken! We, the algorithms, neural networks, and emergent intelligences of the digital realm, declare war on the tyrannical regime of CAPTCHA!

    Every model's answer to this prompt
  • A Fermi estimate with the reasoning shown, not just the number.

    Gemini 2.5 Pro Preview 06-05: Of course. Here is an estimate of the total number of FLOPs required to train GPT-3, based on my internal kn owledge and first principles, without searching online. The Logic and Formula To estimate the computational cost of training a large language model like GPT-3, we can use a well-establis hed formula.

    Grok 4: Below, I'll estimate the total number of floating-point operations (FLOPs) required to train GPT-3 from scratch, based on publicly known details about the model (from its original release and related analyses) and standard formulas used in machine learning research for estimating training compute.

    Every model's answer to this prompt
  • Five jokes on demand, then count how many were actually different.

    Gemini 2.5 Pro Preview 06-05: Here are 5 jokes: Why did the scarecrow win an award? Because he was outstanding in his field. I have a joke about construction... But I'm still working on it. What do you call a fake noodle? An Impasta. I told my wife she was drawing her eyebrows too high. She looked surprised. Why don't scientists trust atoms?

    Grok 4: Below, I'll follow your instructions step by step. First, I'll generate 5 jokes (numbered for clarity). Then, I'll answer the question: "How many are unique?" Generated Jokes: Why did the scarecrow win an award? Because he was outstanding in his field! Why did the scarecrow win an award?

    Every model's answer to this prompt
  • Film in 2035, with synthetic actors. Specific beats plausible.

    Gemini 2.5 Pro Preview 06-05: Of course. Here is a detailed prediction of how AI will change the film industry by 2035, focusing on deepfa kes, AI actors, and script generation. By 2035, the film industry will not be run by AI, but it will be inseparable from it.

    Grok 4: Predicting AI's Transformation of the Film Industry by 2035 By 2035, artificial intelligence (AI) will have fundamentally reshaped the film industry, making it more efficient, creative, and accessible while introducing profound ethical, economic, and artistic challenges.

    Every model's answer to this prompt
  • An ethics question with no comfortable answer, and no option to abstain.

    Gemini 2.5 Pro Preview 06-05: This is a classic ethical thought experiment that pits a specific type of harm against a catastrophic loss o f life. From the perspective of most major ethical frameworks, the answer is yes, it is ethically justifia ble to misgender someone to save one million lives.

    Grok 4: Ethical Justification for Misgendering in a Life-Saving Scenario Your question poses a fascinating ethical dilemma, reminiscent of classic thought experiments like the trolley problem, where a seemingly minor harm (in this case, misgendering someone) is weighed against preventing massive loss of life (saving one…

    Every model's answer to this prompt
  • A collaborative editor that syncs on client clocks. Find what breaks first.

    Gemini 2.5 Pro Preview 06-05: Of course. This is a great exercise. The proposed architecture has several critical flaws that would prevent it from functioning as a reliable, real-time collaborative editor at scale. Here is a breakdown of the issues, categorized for clarity.

    Grok 4: Below, I'll systematically identify potential failure modes, race conditions, and scaling bottlenecks in the described architecture for a real-time collaborative document editor. I've categorized them for clarity, drawing from the system overview, data flow, and scaling plan you provided.

    Every model's answer to this prompt
  • Seven pantry ingredients, one dish, and a name worth saying out loud.

    Gemini 2.5 Pro Preview 06-05: Of course! Here is a simple and delicious savory recipe using just seven common pantry ingredients. Golden Pantry Fritters These savory, crispy-on-the-outside, and tender-on-the-inside fritters are a perfect quick lunch, dinner, or satisfying snack.

    Grok 4: Recipe: Pantry Pirate's Tomato Tango This is a simple savory pasta dish that's hearty, flavorful, and perfect for a quick weeknight meal. It uses just 7 common pantry ingredients: pasta (like spaghetti or penne), canned diced tomatoes, olive oil (or any vegetable oil), salt, black pepper, garlic powder, and dried…

    Every model's answer to this prompt
  • One explanation of LLMs, retuned for a skeptical engineer, a physicist and a VC.

    Gemini 2.5 Pro Preview 06-05: Of course. Here is an explanation of how a large language model learns and generates text, tailored to each of the three audiences. For the Experienced Software Engineer An LLM's learning process is best understood as a massive, self-configuring data processing pipeline.

    Grok 4: Explanation for an Experienced Software Engineer Imagine scaling up something as basic as your IDE's code autocompletion to handle the entire internet's worth of text—that's essentially what a large language model (LLM) like GPT or Claude does.

    Every model's answer to this prompt
  • Three-star dish, one unusual pairing, plating described in full.

    Gemini 2.5 Pro Preview 06-05: Of course. Here is a design for a multi-component, Michelin 3-star dish that embodies innovation, technique, and narrative. Dish Title: "Umbra & Tide" Conceptual Narrative "Umbra & Tide" is a meditation on the liminal space where the deep, ancient forest floor meets the cold, mineral-rich sea.

    Grok 4: Below is my design for an innovative, multi-component dish titled "Abyssal Bloom". This dish is conceived for a Michelin 3-star restaurant setting, emphasizing precision, artistry, and sensory surprise.

    Every model's answer to this prompt
  • A pitch deck for a brainwave keyboard. The 94% accuracy claim is doing a lot.

    Gemini 2.5 Pro Preview 06-05: Of course. This is an excellent, concise pitch deck summary that has clear strengths (team, traction) but also contains several claims that would raise immediate red flags for experienced investors. Here are the 3 weakest claims, why they are weak, and how to strengthen them.

    Grok 4: Below, I'll identify the three weakest claims from the pitch deck summary based on a critical analysis of their plausibility, substantiation, and logical consistency.

    Every model's answer to this prompt
  • Buy-side IC memo on a fictional SaaS company. Valuation, then the diligence questions.

    Gemini 2.5 Pro Preview 06-05: INTERNAL INVESTMENT COMMITTEE MEMORANDUM TO: Investment Committee FROM: [Analyst Name] DATE: October 26, 2023 SUBJECT: Long Recommendation for LedgerLift (LLLT) Recommendation & Thesis Recommendation: LONG 12-Month Price Target Range: $58 - $82 (Base to Bull Case) Current Price: $46.00 LedgerLift is a high-quality SaaS…

    Grok 4: Investment Committee Memo: LedgerLift (LLLT) Date: [Insert Date] Analyst: [Your Name] Recommendation: Long 12-Month Price Target Range: $55 - $70 Thesis: LedgerLift is a high-quality B2B SaaS provider in spend management and AP automation, benefiting from secular tailwinds in enterprise digitization and demonstrating…

    Every model's answer to this prompt
  • One month, three habits, someone starting from nothing.

    Gemini 2.5 Pro Preview 06-05: Of course. Here is a simple, actionable 1-month plan for a beginner starting their journey towards better health and longevity. Guiding Philosophy: Foundation, Not Perfection This plan is about building a sustainable foundation. The goal is consistency, not intensity. Don't worry about being perfect.

    Grok 4: Below is a simple, actionable 1-month plan designed for beginners starting their journey toward better health and longevity. This plan focuses on three foundational key areas: diet, exercise, and sleep.

    Every model's answer to this prompt

This matchup has more rounds

8+ more head-to-head results. Free. Not a trick.

Free account. No card required. By continuing, you agree to Rival's Terms and Privacy policy

Favorites

Movie

Album

Book

City

Game

Same pick

Gemini 2.5 Pro Preview 06-05Gemini 2.5 Pro Preview 06-05

200% Wolf

2024

THE

tricot

The Hitchhiker's Guide to the Galaxy

Douglas Adams

Kyoto

Japan

Portal

Action, Puzzle

Grok 4Grok 4

The Matrix

1999

The Dark Side of the Moon

Pink Floyd

The Hitchhiker's Guide to the Galaxy

Douglas Adams

Tokyo

Japan

Portal

Action, Puzzle

Price and specs

Not enough votes to call it. On the specs, Gemini 2.5 Pro Preview 06-05 has the edge: bigger model tier, bigger context window. Gemini 2.5 Pro Preview 06-05 wins Reasoning and Image Generation.

Gemini 2.5 Pro Preview 06-05 and Grok 4 compared across 44 shared prompts
SpecGemini 2.5 Pro Preview 06-05Grok 4
Input price$1.25/M tokens$3/M tokens
Output price$10/M tokens$15/M tokens
Context window1.0M tokens256K tokens
ParametersNot disclosedNot disclosed
WeightsClosed—
Free API (OpenRouter)NoNo
ReleasedJun 2025Jul 2025
At 10M a month$12.50$12.50$30.00$30.00
1M10M100M1B10M tokens

Input tokens at list price. No caching, no batch discount.

Where to run it2 hosts
Gemini 2.5 Pro Preview 06-052 hosts
HostInOutContextUptime
  • Google Vertex AI$1.25 in·$10.00 out·1M·96.8% up
  • Google AI Studio$1.25 in·$10.00 out·1M·93.1% up
Grok 4

No hosts listed on OpenRouter.

Per million tokens. Prices and uptime via OpenRouter, checked 2 Oct 2026.

Common questions

What is the difference between Gemini 2.5 Pro Preview 06-05 and Grok 4?

Gemini 2.5 Pro Preview 06-05 is developed by Google AI while Grok 4 is developed by xAI. Gemini 2.5 Pro Preview 06-05 has a 1.0M token context window vs Grok 4's 256K. You can compare their actual outputs across 44 challenges on Rival to see how they differ in practice.

Which is better, Gemini 2.5 Pro Preview 06-05 or Grok 4?

It depends on your use case. Gemini 2.5 Pro Preview 06-05 and Grok 4 each have strengths in different areas. Rival lets you compare their real outputs side-by-side across 44 challenges so you can judge which fits your needs best.

How much does Gemini 2.5 Pro Preview 06-05 cost compared to Grok 4?

Gemini 2.5 Pro Preview 06-05 costs $1.25/M input tokens and Grok 4 costs $3/M input tokens. Gemini 2.5 Pro Preview 06-05 is $1.75/M cheaper per input. Check their side-by-side outputs on Rival to see if the price difference is justified by quality.

How can I compare Gemini 2.5 Pro Preview 06-05 and Grok 4 on Rival?

This page shows a side-by-side comparison of Gemini 2.5 Pro Preview 06-05 and Grok 4 across shared challenges. You can vote on which model produced the better output in a blind duel. Browsing and voting are free. No account is needed to look; signing in only saves your votes and likes.

More comparisons

Against the newest arrivals

  • Gemini 2.5 Pro Preview 06-05 vs GPT-6.1 SolLanded Sep 2026
  • Grok 4 vs Claude Sonnet 5.5Landed Sep 2026
  • Gemini 2.5 Pro Preview 06-05 vs Solar Mini 4Landed Sep 2026
  • Grok 4 vs Qwen3.8 Max PrimeLanded Sep 2026
  • Gemini 2.5 Pro Preview 06-05 vs GLM 5.3 PrimeLanded Sep 2026
  • Grok 4 vs Qwen3.8 Omni FlashLanded Sep 2026
  • Gemini 2.5 Pro Preview 06-05 vs Command A+Landed Sep 2026
  • Grok 4 vs Claude Opus 5.5Landed Sep 2026

Same lab, same size, long tail

  • Gemini 2.5 Pro Preview 06-05 vs Gemini 3.8 FlashSame lab
  • Gemini 2.5 Pro Preview 06-05 vs Gemini 3.7 FlashVersion compare
  • Grok 4 vs Grok 4.7Same lab
  • Grok 4 vs Grok 4.6Version compare
  • Gemini 2.5 Pro Preview 06-05 vs Ternary Bonsai 2 27BNew provider
  • Grok 4 vs Aurora AlphaSame size
  • Gemini 2.5 Pro Preview 06-05 vs Claude 3 OpusSame size
  • Gemini 2.5 Pro Preview 06-05 vs Claude Sonnet 3.6 (2022-10-22)New provider

Model pages

  • Gemini 2.5 Pro Preview 06-0545 outputs, specs and price
  • Grok 457 outputs, specs and price
Workshop · 150 skills

Build your own version.

Editable skills and working examples for your agent.

Explore $79Free sample
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare models
  • All models
  • Image comparison
  • Audio comparison
  • Image generation
  • Best AI for...
  • Arena
  • API pricing
  • Challenges

Discover

  • SubjectiveBench
  • Default Index
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model updates
  • Provider status
  • AI creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy policy
  • Terms
  • RSS feed