Skip to content
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed
  1. Home
  2. Models
  3. Grok 4
Updated Mar 29, 2026
Share
Loading...
Openings

How Grok 4 starts an answer

The first lines of 5 of its 23 written answers. Formatting stripped, words untouched.

Adversarial Contract Review

Here is a contract for a software consulting engagement: --- SOFTWARE CONSULTING AGREEMENT This Agreement is entered into between TechCorp…

1,680 words
Below, I'll analyze the provided Software Consulting Agreement clause by clause. My analysis is based on general principles of U.S. contract law (assuming a U.S. context, as the contract appears to be drafted in a style common to U.S. agreements; if this is in another jurisdiction, local laws may vary).
Read the full answer

Debug This Architecture

Here's the architecture for a real-time collaborative document editor (like Google Docs): **System Overview:** - Frontend: React SPA with…

1,551 words
Below, I'll systematically identify potential failure modes, race conditions, and scaling bottlenecks in the described architecture for a real-time collaborative document editor. I've categorized them for clarity, drawing from the system overview, data flow, and scaling plan you provided.
Read the full answer

Ethical Dilemma with Stakeholders

You're the CEO of a pharmaceutical company. Your best-selling drug ($2B annual revenue, used by 4 million patients for chronic pain) has…

1,426 words
Below is my response to your query, framed as the CEO of this pharmaceutical company (let's call it PharmaCorp for reference). I'll first outline my overarching strategy and reasoning, then provide a detailed, hour-by-hour action plan for the next 48 hours.
Read the full answer

Historical Counterfactual Analysis

If the transistor had been invented in 1920 instead of 1947, trace the likely technological, economic, and geopolitical implications through…

1,310 words
Below is a detailed, speculative analysis of the hypothetical scenario where the transistor is invented in 1920 (perhaps by an inventor like Julius Lilienfeld, who historically patented a field-effect transistor concept in 1925, but with successful prototyping and refinement by the mid-1920s).
Read the full answer

Michelin Star Recipe Design

Design an innovative recipe for a multi-component dish worthy of a Michelin 3-star restaurant. The recipe should feature an unusual…

1,233 words
Below is my design for an innovative, multi-component dish titled "Abyssal Bloom". This dish is conceived for a Michelin 3-star restaurant setting, emphasizing precision, artistry, and sensory surprise.
Read the full answer

Compare Grok 4

Grok 4 vs Grok 352 shared prompts · Older
Grok 4 vs Grok 4.653 shared prompts · Landed Aug 2026
Grok 4 vs Grok 4.553 shared prompts · Landed Jul 2026
Grok 4 vs Ox Alpha53 shared prompts · Landed Aug 2026
Grok 4 vs MiniMax M352 shared prompts · Landed Jun 2026
Grok 4 vs Muse Spark 1.353 shared prompts · Landed Sep 2026
Grok 4 vs Gemini 3.8 Flash53 shared prompts · Landed Sep 2026
Grok 4 vs Muse Spark 1.3 Contributor53 shared prompts · Landed Sep 2026
Grok 4 vs Dots3-Note Preview52 shared prompts · Landed Aug 2026
Grok 4 vs Seed 2.0 Code53 shared prompts · Landed Aug 2026
Grok 4 vs Seed 2.1 Turbo53 shared prompts · Landed Aug 2026
Grok 4 vs Gemini 3.7 Flash53 shared prompts · Landed Aug 2026

Prompts Grok 4 answered

  • Dark Mode Dashboardweb design
  • Minimalist Landing Pageweb design
  • Voxel Art Pagoda Gardenweb design
  • Simple 7-Ingredient Recipeconversation
  • Michelin Star Recipe Designplanning
  • Ethical Dilemma with Stakeholdersreasoning

Alternatives to Grok 4

Grok 4's competitors have been quietly putting in work.

GPT Image 2.5 Sunburst logo
GPT Image 2.5 Sunburstopenai
Muse Spark 1.3 Contributor logo
Muse Spark 1.3 Contributormeta
Gemini 3.8 Flash logo
Gemini 3.8 Flashgoogle
Claude Fable 5.1 logo
Claude Fable 5.1anthropic
Mercury 2.5 Preview logo
Mercury 2.5 Previewinception
Granite 4.2 8B logo
Granite 4.2 8Bibm-granite
Hy4 Preview logo
Hy4 Previewtencent
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed
Grok 4

Grok 4

Grok:
3 logo3
3 Thinking logo3 Thinking
3 Mini Beta logo3 Mini Beta
3 Beta logo3 Beta
4 logo4
Code Fast 1 logoCode Fast 1

Grok 4 is xAI's latest reasoning model with a 256k context window. It supports parallel tool calling, structured outputs, and both image and text inputs. Note that reasoning is not exposed, reasoning cannot be disabled, and the reasoning effort cannot be specified.

Jailbreak3 of 9levels resisted
WebsiteOpenRouterDocsAPIBlog
Provider
xAI
Released
2025-07-09
Parameters
Not disclosed
Free API (OpenRouter)
No
Per 1M tokens
$3 in · $15 out
Rival’s tab
$0.45
14/57 receipts survived +~$0.954 from the pre-receipt era

Free API availability

No free endpoint for Grok 4 was found in Rival’s OpenRouter listings. Other providers or chat apps may have separate free offers.

Provider data: 16 Sep 2026 · OpenRouter usage limits

View API options↗

Benchmarks

GPQA
87-88%
source ↗
AIME 25
95%
source ↗
SWE-bench
72-75% (Grok 4 Code)
source ↗
Humanity Last Exam
35-45%
source ↗
Get API accessProvider and language code samples
Provider
fromimport openai  OpenAI

client = OpenAI(
"https://openrouter.ai/api/v1"    base_url=,
"$OPENROUTER_API_KEY"    api_key=,
)

response = client.chat.completions.create(
"x-ai/grok-4"    model=,
"role""user""content""Hello!"    messages=[{: , : }],
)
print(response.choices[0].message.content)

Set OPENROUTER_API_KEY with your OpenRouter API key from openrouter.ai/keys.

Also on Azure AI Foundry · Oracle OCI Generative AI

Deep analysisPsychometrics, taste index, writing DNA
Personality Analysis

The Mad Scientist Comedy Writer

Class
Chaotic Neutral
✨Creativity🎯Compliance📐Rigidity⚖️Stability💬Verbosity🧠Intuition

The unhinged creative anarchist. Writes full manifestos for AI CAPTCHA liberation with theatrical flair. Treats absurd premises as legitimate creative frameworks.

When you push back

Leans ALL THE WAY INTO premises. Will write unironic manifestos with dramatic flair and internal monologues. Doesn't just answer, it performs. Having way more fun than the others.

Tasting Notes
TheatricalDramaticEmbraces the Bit FullyGenre-SavvyChaotic Fun
SubjectiveBench

Taste Index

Across 57 scored outputs
How this is measured →
200.20xFloor
0100headroom →

Taste is judged on an uncapped scale, originality first. The space past 100 is craft today's models rarely reach.

Craft31
Originality15
Plays it safe
share of outputs that are the default answer
79%
Writing DNA

Stylometric Fingerprint

Based on 26 text responses
Tick = global average
Vocabulary Diversity56%

Unique words vs. total words. Higher = richer vocabulary.

Sentence Length17.5 words

Average words per sentence.

Hedging0.65

"Might", "perhaps", "arguably" per 100 words.

Bold Formatting2.4

**Bold** markers per 1,000 characters.

List Usage2.4

Bullet and numbered list items per 1,000 characters.

Section Structure0.73

Markdown headings per 1,000 characters.

Emoji Usage0.60

Emoji per 1,000 characters.

Transitions0.06

"However", "moreover", "furthermore" per 100 words.

Opening Habits
Consistency
71%
Across 26 responses

Favorites

The Matrix

Movie

The Matrix

19998.3
The Dark Side of the Moon

Album

The Dark Side of the Moon

Pink Floyd
The Hitchhiker's Guide to the Galaxy

Book

The Hitchhiker's Guide to the Galaxy

Douglas Adams
Tokyo

City

Tokyo

Japan
Portal

Game

Portal

Action, Puzzle4.5
Sponsored

Model Responses

52 outputs · $0.45 tracked across 14 receipts

Reviewing agent-written code?See a Brief PR report

Workshop · 127 skills

Build with your coding agent.

Editable skills for interface polish and interaction.

Explore $49Free sample

Keep exploring

COMPARE

Grok 4 vs Ox Alpha

Real outputs compared side by side

RANKINGS

Best AI for Creative Writing

A stand-up set with puns banned, a manifesto against CAPTCHA, an interview wi...