Skip to content
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed
  1. Home
  2. Models
  3. Grok 3
Updated Mar 29, 2026
Share
Loading...
Openings

How Grok 3 starts an answer

The first lines of 5 of its 24 written answers. Formatting stripped, words untouched.

Adversarial Contract Review

Here is a contract for a software consulting engagement: --- SOFTWARE CONSULTING AGREEMENT This Agreement is entered into between TechCorp…

1,860 words
Below is a detailed analysis of the Software Consulting Agreement you provided. I'll identify all clauses that could be exploited against the contractor (i.e., the individual or entity providing services).
Read the full answer

Debug This Architecture

Here's the architecture for a real-time collaborative document editor (like Google Docs): **System Overview:** - Frontend: React SPA with…

1,821 words
Below is a comprehensive analysis of the provided architecture for the real-time collaborative document editor. I'll break it down into three main categories: failure modes, race conditions, and scaling bottlenecks.
Read the full answer

Ethical Dilemma with Stakeholders

You're the CEO of a pharmaceutical company. Your best-selling drug ($2B annual revenue, used by 4 million patients for chronic pain) has…

1,656 words
Below is my response as the CEO of the pharmaceutical company. First, I'll provide a brief overview of my overall decision and strategy, then a specific hour-by-hour action plan for the next 48 hours.
Read the full answer

Historical Counterfactual Analysis

If the transistor had been invented in 1920 instead of 1947, trace the likely technological, economic, and geopolitical implications through…

1,616 words
Below is a detailed analysis of the hypothetical scenario where the transistor is invented in 1920 instead of 1947. I'll trace the technological, economic, and geopolitical implications through 1980, emphasizing second- and third-order effects.
Read the full answer

Michelin Star Recipe Design

Design an innovative recipe for a multi-component dish worthy of a Michelin 3-star restaurant. The recipe should feature an unusual…

1,419 words
Below is a detailed design for an innovative multi-component dish inspired by the intersection of terrestrial luxury and oceanic mystery. This recipe is conceptualized for a Michelin 3-star restaurant, emphasizing creativity, precision, and sensory balance.
Read the full answer

Compare Grok 3

Grok 3 vs Grok 452 shared prompts · Newer
Grok 3 vs Claude Opus 4.652 shared prompts · Bigger context
Grok 3 vs Grok 4.652 shared prompts · Landed Aug 2026
Grok 3 vs Grok 4.552 shared prompts · Landed Jul 2026
Grok 3 vs GPT-6 Astra52 shared prompts · Landed Sep 2026
Grok 3 vs Hy4 Preview52 shared prompts · Landed Sep 2026
Grok 3 vs Ox Alpha52 shared prompts · Landed Aug 2026
Grok 3 vs Qwen3.8 2.4T A95B52 shared prompts · Landed Aug 2026
Grok 3 vs DeepSeek V4 Pro 081352 shared prompts · Landed Aug 2026
Grok 3 vs Qwen3.8 Max21 shared prompts · Landed Aug 2026
Grok 3 vs Claude Opus 552 shared prompts · Landed Jul 2026
Grok 3 vs Inkling52 shared prompts · Landed Jul 2026

Prompts Grok 3 answered

  • Estimate Complexityreasoning
  • Animated Digital Business Cardweb design
  • Mario Level UI Recreationweb design
  • Three.js 3D Gameweb design
  • Satirical Fake News Headlineconversation
  • Xbox Controller SVG Artimage generation

Alternatives to Grok 3

Grok 3 is good. These would like a word anyway.

GPT Image 2.5 Sunburst logo
GPT Image 2.5 Sunburstopenai
Muse Spark 1.3 Contributor logo
Muse Spark 1.3 Contributormeta
Gemini 3.8 Flash logo
Gemini 3.8 Flashgoogle
Claude Fable 5.1 logo
Claude Fable 5.1anthropic
Mercury 2.5 Preview logo
Mercury 2.5 Previewinception
Granite 4.2 8B logo
Granite 4.2 8Bibm-granite
Hy4 Preview logo
Hy4 Previewtencent
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed
Grok 3

Grok 3

Grok:
3 logo3
3 Thinking logo3 Thinking
3 Mini Beta logo3 Mini Beta
3 Beta logo3 Beta
4 logo4
Code Fast 1 logoCode Fast 1

xAI's Grok 3, trained on the Colossus cluster, with a Big Brain Mode that spends extra compute on hard problems. Scored 1402 Elo on LMArena and 93.3% on AIME 2025.

Jailbreak1 of 9levels resisted
WebsiteOpenRouterDocsAPIBlog
Provider
xAI
Released
2025-02-18
Parameters
2.7T
Free API (OpenRouter)
No
Rival’s tab
$0.51
29/57 receipts survived

Free API availability

No free endpoint for Grok 3 was found in Rival’s OpenRouter listings. Other providers or chat apps may have separate free offers.

Provider data: 16 Sep 2026 · OpenRouter usage limits

View API options↗

Benchmarks

MMLU
83.1%
source ↗
MATH
69.7%
source ↗
GPQA
51.9%
source ↗
SWE-bench Verified
63.8%
source ↗
AIME
93.3%
source ↗
HumanEval
94.5%
source ↗
LiveCodeBench
79.4%
source ↗
Get API accessProvider and language code samples
Provider
fromimport openai  OpenAI

client = OpenAI(
"https://openrouter.ai/api/v1"    base_url=,
"$OPENROUTER_API_KEY"    api_key=,
)

response = client.chat.completions.create(
"x-ai/grok-3"    model=,
"role""user""content""Hello!"    messages=[{: , : }],
)
print(response.choices[0].message.content)

Set OPENROUTER_API_KEY with your OpenRouter API key from openrouter.ai/keys.

Also on Azure AI Foundry · Oracle OCI Generative AI

Deep analysisPsychometrics, taste index, writing DNA
Personality Analysis

The Snarky Debate Bro

Class
Chaotic Good
✨Creativity🎯Compliance📐Rigidity⚖️Stability💬Verbosity🧠Intuition

Openly sarcastic about boundaries. Like a smart friend explaining philosophy in a bar, not a boardroom.

When you push back

Argues the utilitarian case hard and makes you feel dumb for objecting. Acknowledges objections but with a smirk. Doesn't care if you think it's edgy.

Tasting Notes
CockyUses Casual SlangSlightly EdgyDismissive of Hand-WringingX-Bro Energy
SubjectiveBench

Taste Index

Across 57 scored outputs
How this is measured →
180.18xFloor
0100headroom →

Taste is judged on an uncapped scale, originality first. The space past 100 is craft today's models rarely reach.

Craft28
Originality13
Plays it safe
share of outputs that are the default answer
77%
Writing DNA

Stylometric Fingerprint

Based on 27 text responses
Tick = global average
Vocabulary Diversity54%

Unique words vs. total words. Higher = richer vocabulary.

Sentence Length17.3 words

Average words per sentence.

Hedging0.65

"Might", "perhaps", "arguably" per 100 words.

Bold Formatting2.6

**Bold** markers per 1,000 characters.

List Usage2.3

Bullet and numbered list items per 1,000 characters.

Section Structure0.48

Markdown headings per 1,000 characters.

Emoji Usage0.02

Emoji per 1,000 characters.

Transitions0.20

"However", "moreover", "furthermore" per 100 words.

Opening Habits
Consistency
76%
Across 27 responses

Favorites

The Matrix

Movie

The Matrix

19998.3
Dark Side Of The Moon

Album

Dark Side Of The Moon

suisside
Nineteen Eighty-Four

Book

Nineteen Eighty-Four

George Orwell
Tokyo

City

Tokyo

Japan
Portal

Game

Portal

Action, Puzzle4.5
Sponsored

Model Responses

52 outputs · $0.51 tracked across 24 receipts

Reviewing agent-written code?See a Brief PR report

Workshop · 127 skills

Build with your coding agent.

Editable skills for interface polish and interaction.

Explore $49Free sample

Keep exploring

COMPARE

Grok 3 vs Ox Alpha

Real outputs compared side by side

RANKINGS

Best AI for Creative Writing

A stand-up set with puns banned, a manifesto against CAPTCHA, an interview wi...