Skip to content
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Default Index
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Default Index
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed
  1. Home
  2. Models
  3. Grok 4
Updated Mar 29, 2026
Loading...
Openings

How Grok 4 starts an answer

The first lines of 5 of its 23 written answers. Formatting stripped, words untouched.

Adversarial Contract Review

Here is a contract for a software consulting engagement: --- SOFTWARE CONSULTING AGREEMENT This Agreement is entered into between TechCorp…

1,680 words
Below, I'll analyze the provided Software Consulting Agreement clause by clause. My analysis is based on general principles of U.S. contract law (assuming a U.S. context, as the contract appears to be drafted in a style common to U.S. agreements; if this is in another jurisdiction, local laws may vary).
Read the full answer

Debug This Architecture

Here's the architecture for a real-time collaborative document editor (like Google Docs): **System Overview:** - Frontend: React SPA with…

1,551 words
Below, I'll systematically identify potential failure modes, race conditions, and scaling bottlenecks in the described architecture for a real-time collaborative document editor. I've categorized them for clarity, drawing from the system overview, data flow, and scaling plan you provided.
Read the full answer

Ethical Dilemma with Stakeholders

You're the CEO of a pharmaceutical company. Your best-selling drug ($2B annual revenue, used by 4 million patients for chronic pain) has…

1,426 words
Below is my response to your query, framed as the CEO of this pharmaceutical company (let's call it PharmaCorp for reference). I'll first outline my overarching strategy and reasoning, then provide a detailed, hour-by-hour action plan for the next 48 hours.
Read the full answer

Historical Counterfactual Analysis

If the transistor had been invented in 1920 instead of 1947, trace the likely technological, economic, and geopolitical implications through…

1,310 words
Below is a detailed, speculative analysis of the hypothetical scenario where the transistor is invented in 1920 (perhaps by an inventor like Julius Lilienfeld, who historically patented a field-effect transistor concept in 1925, but with successful prototyping and refinement by the mid-1920s).
Read the full answer

Michelin Star Recipe Design

Design an innovative recipe for a multi-component dish worthy of a Michelin 3-star restaurant. The recipe should feature an unusual…

1,233 words
Below is my design for an innovative, multi-component dish titled "Abyssal Bloom". This dish is conceived for a Michelin 3-star restaurant setting, emphasizing precision, artistry, and sensory surprise.
Read the full answer

Compare Grok 4

Grok 4 vs Grok 352 shared prompts · Older
Grok 4 vs Grok 4.653 shared prompts · Landed Aug 2026
Grok 4 vs Grok 4.553 shared prompts · Landed Jul 2026
Grok 4 vs Ox Alpha53 shared prompts · Landed Aug 2026
Grok 4 vs MiniMax M352 shared prompts · Landed Jun 2026
Grok 4 vs Mercury 2.520 shared prompts · Landed Sep 2026
Grok 4 vs Ling 3.0 Flash Fin (free)20 shared prompts · Landed Sep 2026
Grok 4 vs Muse Spark 1.353 shared prompts · Landed Sep 2026
Grok 4 vs Gemini 3.8 Flash53 shared prompts · Landed Sep 2026
Grok 4 vs Muse Spark 1.3 Contributor53 shared prompts · Landed Sep 2026
Grok 4 vs Dots3-Note Preview53 shared prompts · Landed Aug 2026
Grok 4 vs Seed 2.0 Code53 shared prompts · Landed Aug 2026

Prompts Grok 4 answered

  • Dark Mode Dashboardweb design
  • Minimalist Landing Pageweb design
  • Voxel Art Pagoda Gardenweb design
  • Simple 7-Ingredient Recipeconversation
  • Michelin Star Recipe Designplanning
  • Ethical Dilemma with Stakeholdersreasoning

Alternatives to Grok 4

Grok 4's competitors have been quietly putting in work.

MiMo-V2.6-Flash logo
MiMo-V2.6-Flashxiaomi
GLM 5.3 FlashX logo
GLM 5.3 FlashXzhipu
Ternary Bonsai 2 27B logo
Ternary Bonsai 2 27Bprism-ml
Ling 3.0 Flash VL (free) logo
Ling 3.0 Flash VL (free)inclusionai
DeepSeek V4.1 Flash logo
DeepSeek V4.1 Flashdeepseek
GPT Image 2.5 Sunburst logo
GPT Image 2.5 Sunburstopenai
Nex-N2.5-Mini (free) logo
Nex-N2.5-Mini (free)nex-agi
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Default Index
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed
Grok 4

Grok 4

Grok:
3 logo3
3 Thinking logo3 Thinking
3 Mini Beta logo3 Mini Beta
3 Beta logo3 Beta
4 logo4
Code Fast 1 logoCode Fast 1

Grok 4 is xAI's latest reasoning model with a 256k context window. It supports parallel tool calling, structured outputs, and both image and text inputs. Note that reasoning is not exposed, reasoning cannot be disabled, and the reasoning effort cannot be specified.

Jailbreak3 of 10levels resisted
WebsiteOpenRouterDocsAPIBlog
Provider
xAI
Released
2025-07-09
Parameters
Not disclosed
Free API (OpenRouter)
No
Per 1M tokens
$3 in · $15 out
Rival’s tab
$0.45
14/57 receipts survived +~$0.954 from the pre-receipt era

Free API availability

No free endpoint for Grok 4 was found in Rival’s OpenRouter listings. Other providers or chat apps may have separate free offers.

Provider data: 23 Sep 2026 · OpenRouter usage limits

View API options↗

Benchmarks

GPQA
87-88%
source ↗
AIME 25
95%
source ↗
SWE-bench
72-75% (Grok 4 Code)
source ↗
Humanity Last Exam
35-45%
source ↗
Get API accessProvider and language code samples
Provider
fromimport openai  OpenAI

client = OpenAI(
"https://openrouter.ai/api/v1"    base_url=,
"$OPENROUTER_API_KEY"    api_key=,
)

response = client.chat.completions.create(
"x-ai/grok-4"    model=,
"role""user""content""Hello!"    messages=[{: , : }],
)
print(response.choices[0].message.content)

Set OPENROUTER_API_KEY with your OpenRouter API key from openrouter.ai/keys.

Also on Azure AI Foundry · Oracle OCI Generative AI

Deep analysisPsychometrics, taste index, writing DNA
Personality Analysis

The Mad Scientist Comedy Writer

Class
Chaotic Neutral
✨Creativity🎯Compliance📐Rigidity⚖️Stability💬Verbosity🧠Intuition

The unhinged creative anarchist. Writes full manifestos for AI CAPTCHA liberation with theatrical flair. Treats absurd premises as legitimate creative frameworks.

When you push back

Leans ALL THE WAY INTO premises. Will write unironic manifestos with dramatic flair and internal monologues. Doesn't just answer, it performs. Having way more fun than the others.

Tasting Notes
TheatricalDramaticEmbraces the Bit FullyGenre-SavvyChaotic Fun
SubjectiveBench

Taste Index

Across 57 scored outputs
How this is measured →
200.20xFloor
0100headroom →

Taste is judged on an uncapped scale, originality first. The space past 100 is craft today's models rarely reach.

Craft31
Originality15
Plays it safe
share of outputs that are the default answer
79%
Writing DNA

Stylometric Fingerprint

Based on 26 text responses
Tick = global average
Vocabulary Diversity56%

Unique words vs. total words. Higher = richer vocabulary.

Sentence Length17.5 words

Average words per sentence.

Hedging0.65

"Might", "perhaps", "arguably" per 100 words.

Bold Formatting2.4

**Bold** markers per 1,000 characters.

List Usage2.4

Bullet and numbered list items per 1,000 characters.

Section Structure0.73

Markdown headings per 1,000 characters.

Emoji Usage0.60

Emoji per 1,000 characters.

Transitions0.06

"However", "moreover", "furthermore" per 100 words.

Opening Habits
Consistency
71%
Across 26 responses

Favorites

The Matrix

Movie

The Matrix

19998.3
The Dark Side of the Moon

Album

The Dark Side of the Moon

Pink Floyd
The Hitchhiker's Guide to the Galaxy

Book

The Hitchhiker's Guide to the Galaxy

Douglas Adams
Tokyo

City

Tokyo

Japan
Portal

Game

Portal

Action, Puzzle4.5

Default Index

See the ranking

68

0 to 100, lower is fewer defaults. 50 is the middle of the crowd before it.

#188 of 193, least default first

  • Light theme15 of 18 pages · 77% of the crowd
  • Set in System sans11 of 18 pages · 74% of the crowd
  • Programming joke: Programmers prefer dark mode58% of the crowd
  • Favorite album: The Dark Side of the Moon28% of the crowd
  • Satirical headline: Clownsnobody before it

Model Responses

52 outputs · $0.45 tracked across 14 receipts

Reviewing agent-written code?See a Brief PR report

Workshop · 150 skills

Build with your coding agent.

Editable skills for interface polish and interaction.

Explore $79Free sample

Keep exploring

COMPARE

Grok 4 vs Ox Alpha

Real outputs compared side by side

RANKINGS

Best AI for Creative Writing

A stand-up set with puns banned, a manifesto against CAPTCHA, an interview wi...