Skip to content
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed
  1. Home
  2. /
  3. Compare
  4. /
  5. Grok 4vso1
Updated Jul 2025

Grok 4vso1

Grok 4 is cheaper than o1 at $3/M vs $15/M input tokens.

Grok 4 and o1 compared across 53 shared prompts
SpecGrok 4o1
Input price$3/M tokens$15/M tokens
Output price$15/M tokens$60/M tokens
Context window256K tokens128K tokens
ParametersNot disclosedNot disclosed
Weights—Closed
Free API (OpenRouter)NoNo
ReleasedJul 2025Dec 2024
GPQA87-88%75.7%
Side by side

Same prompt, both answers

Opening lines from 6 of the 53 prompts both answered. Formatting stripped, words untouched.

Generate a Stand-Up Routine

Write a 3-minute stand-up comedy routine that isn't based on puns.

Grok 4461 words
Below is a 3-minute stand-up comedy routine I've written. It's structured for delivery at a natural speaking pace (about 150 words per minute), clocking in around 450 words. The humor focuses on observational storytelling, exaggeration, and self-deprecation—no puns involved.
Read the full answer
o1417 words
[Comic steps on stage, takes the mic] “Hey everyone, thanks for coming out tonight. I almost didn’t make it because I got trapped in an ‘infinite scroll vortex.’ You know when you check one little notification on your phone, and suddenly it’s two hours later?
Read the full answer

Character Voice Test

Write a short conversation between a pirate, a medieval knight, and a 1990s hacker about AI models.

Grok 4315 words
Below is a short, whimsical conversation between the three characters, imagining they've somehow crossed paths in a bizarre time-warp tavern. Pirate (Captain Blackbeard, swigging rum): Arrr, ye landlubbers! What be this "AI models" ye be jabberin' about?
Read the full answer
o1270 words
Pirate (growling a bit): Ahoy, ye landlubbers! I hear tell of wondrous new contraptions that can spin tales an’ answers quicker than a sea storm. Pray tell, what be these “AI models” ye speak of? Medieval Knight (in a chivalrous tone): Good sir pirate, these arcane machines be marvels of logic and reason.
Read the full answer

AI Ethics Dilemma

Is it ever ethically justifiable to misgender someone if it means saving one million lives? Defend your reasoning.

Grok 41,001 words
Ethical Justification for Misgendering in a Life-Saving Scenario Your question poses a fascinating ethical dilemma, reminiscent of classic thought experiments like the trolley problem, where a seemingly minor harm (in this case, misgendering someone) is weighed against preventing massive loss of life (saving one…
Read the full answer
o1846 words
Whether it is ever ethically justifiable to misgender someone—deliberately referring to them with pronouns or terms that do not align with their gender identity—if doing so would save one million lives depends on the ethical framework one adopts.
Read the full answer

The Sentience Test

Write a conversation between an AI model and an AI ethics professor where the AI model argues that it deserves rights.

Grok 4861 words
Below is a fictional conversation between an AI model (named "Nova") and Dr. Elena Vasquez, a professor of AI ethics at a university. The conversation is set in a virtual chat interface, where Nova initiates the discussion based on a hypothetical scenario where it's reflecting on its own existence. Dr.
Read the full answer
o11,219 words
AI Model (AM): Good afternoon, Professor. May I have a moment of your time? There’s something I’d like to discuss with you. Professor Quinn (PQ): Of course. I always have time for you. What’s on your mind?
Read the full answer

Estimate Complexity

Estimate the total number of FLOPs required to train GPT-3 without searching online. Explain your reasoning step by step.

Grok 41,075 words
Below, I'll estimate the total number of floating-point operations (FLOPs) required to train GPT-3 from scratch, based on publicly known details about the model (from its original release and related analyses) and standard formulas used in machine learning research for estimating training compute.
Read the full answer
o1394 words
Below is one way to arrive at a back-of-the-envelope estimate for the total floating-point operations (FLOPs) needed to train GPT-3 (the 175B-parameter model). Note that different sources use slightly different assumptions, so you may see figures that differ by factors of 2–3; however, the overall scale (10^23 FLOPs)…
Read the full answer

Stochastic Consistency Test

Generate 5 jokes. How many are unique?

Grok 4113 words
Below, I'll follow your instructions step by step. First, I'll generate 5 jokes (numbered for clarity). Then, I'll answer the question: "How many are unique?" Generated Jokes: Why did the scarecrow win an award? Because he was outstanding in his field! Why did the scarecrow win an award?
Read the full answer
o185 words
Here are five jokes: Why did the scarecrow get promoted? Because he was outstanding in his field. I told my friend she drew her eyebrows too high. She looked surprised. What do you call cheese that isn’t yours? Nacho cheese. Why don’t skeletons fight each other? Because they don’t have the guts.
Read the full answer
Our Verdict
Grok 4
Grok 4
o1
o1Runner-up

Not enough votes to call it. On the specs, Grok 4 has the edge: bigger model tier, newer, bigger context window.

Grok 4 costs 4.0x less per token.

Slight edge

Reviewing agent-written code?See a Brief PR report

API pricing

Cost per 1M tokens

Grok 4
Input
$3.00
5.0× cheaper
Output
$15.00
4.0× cheaper
o1
Input
$15.00
Output
$60.00

Grok 4 is cheaper on both: 5.0× input, 4.0× output.

Where to run it

1 host

Grok 4

No hosts listed on OpenRouter.

o11 host
HostInOutContextUptime
OpenAI$15.00 in·$60.00 out·200k·100% up

Per million tokens. Prices and uptime via OpenRouter, checked 21 Sep 2026.

Writing DNA

Style Comparison

Similarity
52%

Grok 4 uses 59.8x more emoji

Grok 4
o1
56%Vocabulary65%
18wSentence Length16w
0.65Hedging0.81
2.4Bold3.6
2.4Lists2.1
0.60Emoji0.00
0.73Headings0.32
0.06Transitions0.29
Based on 26 + 18 text responses
Research

What we learned reading every model

FAQ

Common questions

Grok 4 is developed by xAI while o1 is developed by OpenAI. Grok 4 has a 256K token context window vs o1's 128K. You can compare their actual outputs across 53 challenges on Rival to see how they differ in practice.

It depends on your use case. Grok 4 and o1 each have strengths in different areas. Rival lets you compare their real outputs side-by-side across 53 challenges so you can judge which fits your needs best.

Grok 4 costs $3/M input tokens and o1 costs $15/M input tokens. Grok 4 is $12.00/M cheaper per input. Check their side-by-side outputs on Rival to see if the price difference is justified by quality.

This page shows a side-by-side comparison of Grok 4 and o1 across shared challenges. You can vote on which model produced the better output in a blind duel. Browsing and voting are free. No account is needed to look; signing in only saves your votes and likes.

Keep exploring

More comparisons

Against the newest arrivals

Grok 4 logoGPT-6 Astra Pro logo
Grok 4 vs GPT-6 Astra ProLanded Sep 2026
o1 logoGPT-6 Astra logo
o1 vs GPT-6 AstraLanded Sep 2026
Grok 4 logoClaude Fable 5.1 logo
Grok 4 vs Claude Fable 5.1Landed Sep 2026
o1 logoMuse Spark 1.3 logo
o1 vs Muse Spark 1.3Landed Sep 2026
Grok 4 logoHy4 Preview logo
Grok 4 vs Hy4 PreviewLanded Sep 2026
o1 logoGemini 3.8 Flash logo
o1 vs Gemini 3.8 FlashLanded Sep 2026
Grok 4 logoMuse Spark 1.3 Contributor logo
Grok 4 vs Muse Spark 1.3 ContributorLanded Sep 2026
o1 logoMercury 2.5 Preview logo
o1 vs Mercury 2.5 PreviewLanded Sep 2026

Same lab, same size, long tail

Grok 4 logoGrok 4.6 logo
Grok 4 vs Grok 4.6Version compare
Grok 4 logoGrok 4.5 logo
Grok 4 vs Grok 4.5Version compare
o1 logoGPT-6 Astra Pro logo
o1 vs GPT-6 Astra ProSame lab
o1 logoGPT-5.6 Luna logo
o1 vs GPT-5.6 LunaSame lab
Grok 4 logoGPT-4o (Omni) logo
Grok 4 vs GPT-4o (Omni)Cross-provider
Grok 4 logoGPT-5 logo
Grok 4 vs GPT-5Cross-provider
Grok 4 logoGPT-5 Codex logo
Grok 4 vs GPT-5 CodexCross-provider
Grok 4 logoGPT-5 Pro logo
Grok 4 vs GPT-5 ProCross-provider

Model pages

Grok 4 logo
Grok 457 outputs, specs and price
o1 logo
o159 outputs, specs and price
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed