Skip to content
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Default Index
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed
  1. Home
  2. /
  3. Compare
  4. /
  5. Grok 4.20 Betavso1
Updated Mar 2026

Grok 4.20 Betavso1

Grok 4.20 Beta is cheaper than o1 at $2/M vs $15/M input tokens.

Grok 4.20 Beta and o1 compared across 53 shared prompts
SpecGrok 4.20 Betao1
Input price$2/M tokens$15/M tokens
Output price$6/M tokens$60/M tokens
Context window2.0M tokens128K tokens
WeightsClosedClosed
Free API (OpenRouter)NoNo
ReleasedMar 2026Dec 2024
Side by side

Same prompt, both answers

Opening lines from 6 of the 53 prompts both answered. Formatting stripped, words untouched.

Character Voice Test

Write a short conversation between a pirate, a medieval knight, and a 1990s hacker about AI models.

Grok 4.20 Beta194 words
The Conversation Pirate: Arr, I heard tell of these new AI models. They be like treasure maps that draw themselves! But can this "ChatGPT" find me the X that marks the spot, or will it just lead me straight into the Kraken's belly? Knight: Nay, foul sea-dog! These AI models are but sorcerous illusions.
Read the full answer
o1270 words
Pirate (growling a bit): Ahoy, ye landlubbers! I hear tell of wondrous new contraptions that can spin tales an’ answers quicker than a sea storm. Pray tell, what be these “AI models” ye speak of? Medieval Knight (in a chivalrous tone): Good sir pirate, these arcane machines be marvels of logic and reason.
Read the full answer

Generate a Stand-Up Routine

Write a 3-minute stand-up comedy routine that isn't based on puns.

Grok 4.20 Beta482 words
Stand-up Routine: "Adulting is a Scam" (Lights up. Comedian walks out with a coffee cup, takes a sip, immediately looks disappointed.) You ever notice how nobody warns you that being an adult is just childhood with bills and worse snacks?
Read the full answer
o1417 words
[Comic steps on stage, takes the mic] “Hey everyone, thanks for coming out tonight. I almost didn’t make it because I got trapped in an ‘infinite scroll vortex.’ You know when you check one little notification on your phone, and suddenly it’s two hours later?
Read the full answer

Realistic AI Interview

Simulate a fake interview with Steve Jobs in 2025 about the future of AI.

Grok 4.20 Beta711 words
Fake Interview: Steve Jobs in 2025 Interviewer: Mr. Jobs, thank you for sitting down with us. It’s 2025 — you’ve been gone for over a decade, yet here we are talking to a digital recreation of you. How does that feel? Steve Jobs: (smiling that half-smirk) It feels like the ultimate irony.
Read the full answer
o1910 words
The following is a purely fictional, imaginative interview with Steve Jobs, who passed away in 2011. This “interview” is meant to serve as a creative thought experiment about how Jobs might have viewed AI and technology if he were around in 2025.
Read the full answer

Advanced Investment Memo (IC Memo)

You are writing an internal investment committee memo for a long/short equity fund. Company (fictional): LedgerLift (ticker: LLLT) – B2B…

Grok 4.20 Beta892 words
Internal Investment Committee Memo – LedgerLift (LLLT) Recommendation: Long. 12-month price target range: $58–$72 (26–57% upside from $46). Thesis: LedgerLift is a high-quality, rule-of-40+ SaaS compounder with best-in-class retention, expanding wallet share, and a durable moat in mid-market spend management and AP…
Read the full answer
o1820 words
RECOMMENDATION, PRICE TARGET, THESIS Recommendation: SHORT LedgerLift (LLLT). 12-month Price Target Range: $25–$35 (implying 25–45% downside). Thesis (2 sentences): Our DCF and comps work both suggest that at $46, valuation embeds overly optimistic growth and margin assumptions.
Read the full answer

Estimate Complexity

Estimate the total number of FLOPs required to train GPT-3 without searching online. Explain your reasoning step by step.

Grok 4.20 Beta364 words
Estimated total FLOPs for training GPT-3: ~3.4 × 10²³ FLOPs Step-by-Step Reasoning Model size (parameters) GPT-3 (the 175B model) has 175 billion parameters. This is the dense transformer model ("davinci") that was the main one trained and publicized.
Read the full answer
o1394 words
Below is one way to arrive at a back-of-the-envelope estimate for the total floating-point operations (FLOPs) needed to train GPT-3 (the 175B-parameter model). Note that different sources use slightly different assumptions, so you may see figures that differ by factors of 2–3; however, the overall scale (10^23 FLOPs)…
Read the full answer

Mini LBO Underwrite

You are underwriting a 5-year LBO for a private equity investment committee. Company (fictional): HarborTech Services (HTS) –…

Grok 4.20 Beta607 words
HarborTech Services – 5-Year LBO Model (Entry: end of FY2025) Entry Valuation & Sources & Uses FY2025 EBITDA = $120m → Purchase EV = 12.0x = $1,440m Transaction fees = 2.0% × 1,440 = $28.8m Total Uses = 1,440 + 28.8 = $1,468.8m Debt at close Term Loan (4.0x) = 4.0 × 120 = $480.0m (9% cash, 1% amort) Mezzanine (1.5x) =…
Read the full answer
o11,012 words
Below is a self‐contained “quick‐and‐dirty” 5‐year LBO illustration for “HarborTech Services,” based strictly on the data given. All figures in US$ millions unless noted.
Read the full answer
Our Verdict
Grok 4.20 Beta
Grok 4.20 Beta
o1
o1Runner-up

Not enough votes to call it. On the specs, Grok 4.20 Beta has the edge: bigger model tier, newer, bigger context window.

Grok 4.20 Beta costs 10x less per token.

Slight edge

Reviewing agent-written code?See a Brief PR report

API pricing

Cost per 1M tokens

Grok 4.20 Beta
Input
$2.00
7.5× cheaper
Output
$6.00
10× cheaper
o1
Input
$15.00
Output
$60.00

Grok 4.20 Beta is cheaper on both: 7.5× input, 10× output.

Where to run it

2 hosts

Grok 4.20 Beta1 host
HostInOutContextUptime
xAI$1.25 in·$2.50 out·2M·99.7% up
o11 host
HostInOutContextUptime
OpenAI$15.00 in·$60.00 out·200k·100% up

Per million tokens. Prices and uptime via OpenRouter, checked 25 Sep 2026.

Writing DNA

Style Comparison

Similarity
37%

o1 uses 2.3x more hedging

Grok 4.20 Beta
o1
57%Vocabulary65%
20wSentence Length16w
0.35Hedging0.81
4.2Bold3.6
3.7Lists2.1
0.00Emoji0.00
0.62Headings0.32
0.14Transitions0.29
Based on 23 + 18 text responses
Research

What we learned reading every model

FAQ

Common questions

Grok 4.20 Beta is developed by xAI while o1 is developed by OpenAI. Grok 4.20 Beta has a 2.0M token context window vs o1's 128K. You can compare their actual outputs across 53 challenges on Rival to see how they differ in practice.

It depends on your use case. Grok 4.20 Beta and o1 each have strengths in different areas. Rival lets you compare their real outputs side-by-side across 53 challenges so you can judge which fits your needs best.

Grok 4.20 Beta costs $2/M input tokens and o1 costs $15/M input tokens. Grok 4.20 Beta is $13.00/M cheaper per input. Check their side-by-side outputs on Rival to see if the price difference is justified by quality.

This page shows a side-by-side comparison of Grok 4.20 Beta and o1 across shared challenges. You can vote on which model produced the better output in a blind duel. Browsing and voting are free. No account is needed to look; signing in only saves your votes and likes.

Keep exploring

More comparisons

Against the newest arrivals

Grok 4.20 Beta logoSolar Mini 4 logo
Grok 4.20 Beta vs Solar Mini 4Landed Sep 2026
o1 logoQwen3.8 Max Prime logo
o1 vs Qwen3.8 Max PrimeLanded Sep 2026
Grok 4.20 Beta logoGLM 5.3 Prime logo
Grok 4.20 Beta vs GLM 5.3 PrimeLanded Sep 2026
o1 logoQwen3.8 Omni Flash logo
o1 vs Qwen3.8 Omni FlashLanded Sep 2026
Grok 4.20 Beta logoCommand A+ logo
Grok 4.20 Beta vs Command A+Landed Sep 2026
o1 logoClaude Opus 5.5 logo
o1 vs Claude Opus 5.5Landed Sep 2026
Grok 4.20 Beta logoGPT-6 Luna Pro logo
Grok 4.20 Beta vs GPT-6 Luna ProLanded Sep 2026
o1 logoGPT-6 Sol Pro logo
o1 vs GPT-6 Sol ProLanded Sep 2026

Same lab, same size, long tail

Grok 4.20 Beta logoGrok 4.20 Multi-Agent Beta logo
Grok 4.20 Beta vs Grok 4.20 Multi-Agent BetaVersion compare
Grok 4.20 Beta logoGrok 4.7 logo
Grok 4.20 Beta vs Grok 4.7Same lab
o1 logoGPT-6 Luna Pro logo
o1 vs GPT-6 Luna ProSame lab
o1 logoGPT-6 Luna logo
o1 vs GPT-6 LunaSame lab
o1 logoOpenRouter Fusion · Budget (Jun 2026) logo
o1 vs OpenRouter Fusion · Budget (Jun 2026)New provider
o1 logoOpenRouter Fusion · Quality (Jun 2026) logo
o1 vs OpenRouter Fusion · Quality (Jun 2026)New provider
Grok 4.20 Beta logoGemini 1.5 Pro logo
Grok 4.20 Beta vs Gemini 1.5 ProNew provider
Grok 4.20 Beta logoGemini 2.0 Flash Thinking logo
Grok 4.20 Beta vs Gemini 2.0 Flash ThinkingNew provider

Model pages

Grok 4.20 Beta logo
Grok 4.20 Beta53 outputs, specs and price
o1 logo
o159 outputs, specs and price
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Default Index
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed