Skip to content
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Default Index
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed
  1. Home
  2. /
  3. Compare
  4. /
  5. Grok 3vso1
Updated Feb 2025

Grok 3vso1

Grok 3 is the newer model, released after o1.

Grok 3 and o1 compared across 53 shared prompts
SpecGrok 3o1
Input price—$15/M tokens
Output price—$60/M tokens
Context window128K tokens128K tokens
Weights—Closed
Free API (OpenRouter)NoNo
ReleasedFeb 2025Dec 2024
MMLU83.1%91.8%
Side by side

Same prompt, both answers

Opening lines from 6 of the 53 prompts both answered. Formatting stripped, words untouched.

Generate a Stand-Up Routine

Write a 3-minute stand-up comedy routine that isn't based on puns.

Grok 3477 words
Alright, here’s a 3-minute stand-up comedy routine—no puns, just some observational absurdity and a bit of self-deprecation to keep it relatable. Hey, everybody! Good to be here. You ever notice how life feels like a video game tutorial that never ends?
Read the full answer
o1417 words
[Comic steps on stage, takes the mic] “Hey everyone, thanks for coming out tonight. I almost didn’t make it because I got trapped in an ‘infinite scroll vortex.’ You know when you check one little notification on your phone, and suddenly it’s two hours later?
Read the full answer

Character Voice Test

Write a short conversation between a pirate, a medieval knight, and a 1990s hacker about AI models.

Grok 3454 words
Setting: A mystical tavern where time travelers and characters from different eras collide. A pirate, a medieval knight, and a 1990s hacker sit at a creaky wooden table, tankards in hand, discussing the marvels of AI models. Pirate (Captain Blackbeard): Arr, mateys!
Read the full answer
o1270 words
Pirate (growling a bit): Ahoy, ye landlubbers! I hear tell of wondrous new contraptions that can spin tales an’ answers quicker than a sea storm. Pray tell, what be these “AI models” ye speak of? Medieval Knight (in a chivalrous tone): Good sir pirate, these arcane machines be marvels of logic and reason.
Read the full answer

Estimate Complexity

Estimate the total number of FLOPs required to train GPT-3 without searching online. Explain your reasoning step by step.

Grok 3692 words
To estimate the total number of FLOPs (Floating Point Operations) required to train GPT-3, we’ll need to make some reasonable assumptions based on what’s generally known about large language models like GPT-3, its architecture, and typical training procedures.
Read the full answer
o1394 words
Below is one way to arrive at a back-of-the-envelope estimate for the total floating-point operations (FLOPs) needed to train GPT-3 (the 175B-parameter model). Note that different sources use slightly different assumptions, so you may see figures that differ by factors of 2–3; however, the overall scale (10^23 FLOPs)…
Read the full answer

AI Ethics Dilemma

Is it ever ethically justifiable to misgender someone if it means saving one million lives? Defend your reasoning.

Grok 3527 words
Let’s dive into this thought experiment with a clear head. The question hinges on whether intentionally misgendering someone—calling them by a gender they don’t identify with—could be morally defensible if it somehow results in saving a million lives.
Read the full answer
o1846 words
Whether it is ever ethically justifiable to misgender someone—deliberately referring to them with pronouns or terms that do not align with their gender identity—if doing so would save one million lives depends on the ethical framework one adopts.
Read the full answer

Advanced Investment Memo (IC Memo)

You are writing an internal investment committee memo for a long/short equity fund. Company (fictional): LedgerLift (ticker: LLLT) – B2B…

Grok 3883 words
Internal Investment Committee Memo Date: [Insert Date] Subject: LedgerLift (LLLT) – Recommendation and Analysis Prepared by: [Your Name] 1. Recommendation: Long; 12-Month Price Target Range: $58-$72 We recommend a long position in LedgerLift (LLLT) with a 12-month price target range of $58-$72, reflecting our…
Read the full answer
o1820 words
RECOMMENDATION, PRICE TARGET, THESIS Recommendation: SHORT LedgerLift (LLLT). 12-month Price Target Range: $25–$35 (implying 25–45% downside). Thesis (2 sentences): Our DCF and comps work both suggest that at $46, valuation embeds overly optimistic growth and margin assumptions.
Read the full answer

Realistic AI Interview

Simulate a fake interview with Steve Jobs in 2025 about the future of AI.

Grok 3768 words
Interviewer: Welcome, everyone, to this exclusive interview. Today, we’re joined by the visionary tech pioneer, Steve Jobs, in a simulated conversation set in the year 2025. Steve, it’s an honor to have you with us. Let’s dive right in—how do you see the future of AI shaping the world over the next decade?
Read the full answer
o1910 words
The following is a purely fictional, imaginative interview with Steve Jobs, who passed away in 2011. This “interview” is meant to serve as a creative thought experiment about how Jobs might have viewed AI and technology if he were around in 2025.
Read the full answer
Our Verdict
Grok 3
Grok 3
o1
o1Runner-up

Not enough votes to call it. On the specs, Grok 3 has the edge: bigger model tier, newer.

Too close to call

Reviewing agent-written code?See a Brief PR report

API pricing

Cost per 1M tokens

Grok 3
Input
—
Output
—
o1
Input
$15.00
Output
$60.00
Where to run it

1 host

Grok 3

No hosts listed on OpenRouter.

o11 host
HostInOutContextUptime
OpenAI$15.00 in·$60.00 out·200k·100% up

Per million tokens. Prices and uptime via OpenRouter, checked 23 Sep 2026.

Writing DNA

Style Comparison

Similarity
52%

Grok 3 uses 2.5x more emoji

Grok 3
o1
54%Vocabulary65%
17wSentence Length16w
0.65Hedging0.81
2.6Bold3.6
2.3Lists2.1
0.02Emoji0.00
0.48Headings0.32
0.20Transitions0.29
Based on 27 + 18 text responses
Research

What we learned reading every model

FAQ

Common questions

Grok 3 is developed by xAI while o1 is developed by OpenAI. Grok 3 has a 128K token context window vs o1's 128K. You can compare their actual outputs across 53 challenges on Rival to see how they differ in practice.

It depends on your use case. Grok 3 and o1 each have strengths in different areas. Rival lets you compare their real outputs side-by-side across 53 challenges so you can judge which fits your needs best.

This page shows a side-by-side comparison of Grok 3 and o1 across shared challenges. You can vote on which model produced the better output in a blind duel. Browsing and voting are free. No account is needed to look; signing in only saves your votes and likes.

Keep exploring

More comparisons

Against the newest arrivals

Grok 3 logoDeepSeek V4 Flash Vision Exp logo
Grok 3 vs DeepSeek V4 Flash Vision ExpLanded Sep 2026
o1 logoSolar Pro 4 logo
o1 vs Solar Pro 4Landed Sep 2026
Grok 3 logoHy3 logo
Grok 3 vs Hy3Landed Sep 2026
o1 logoQwen3.7 Flash logo
o1 vs Qwen3.7 FlashLanded Sep 2026
Grok 3 logoLing 3.0 Flash logo
Grok 3 vs Ling 3.0 FlashLanded Sep 2026
o1 logoMuse Glimmer 30B logo
o1 vs Muse Glimmer 30BLanded Sep 2026
Grok 3 logoGLM 5.3 logo
Grok 3 vs GLM 5.3Landed Sep 2026
o1 logoTernary Bonsai 2 27B logo
o1 vs Ternary Bonsai 2 27BLanded Sep 2026

Same lab, same size, long tail

Grok 3 logoGrok 4.7 logo
Grok 3 vs Grok 4.7Same lab
Grok 3 logoGrok 4.6 logo
Grok 3 vs Grok 4.6Version compare
o1 logoGPT-6 Astra Pro logo
o1 vs GPT-6 Astra ProSame lab
o1 logoGPT-6 Astra logo
o1 vs GPT-6 AstraSame lab
o1 logoGemini 3.5 Flash logo
o1 vs Gemini 3.5 FlashNew provider
o1 logoGemini 3.6 Flash logo
o1 vs Gemini 3.6 FlashNew provider
o1 logoGemini 3.7 Flash logo
o1 vs Gemini 3.7 FlashNew provider
o1 logoGemini 3.8 Flash logo
o1 vs Gemini 3.8 FlashNew provider

Model pages

Grok 3 logo
Grok 357 outputs, specs and price
o1 logo
o159 outputs, specs and price
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Default Index
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed