GPT-5.6 Terra is cheaper than Grok 4 at $2.5/M vs $3/M input tokens.
Whole landing page for a startup that reads your face. It renders or it doesn't.
Which answer wins?
Seven pantry ingredients, one dish, and a name worth saying out loud.
GPT-5.6 Terra: Golden Crunch Cinnamon Toast Bites Ingredients (7 pantry staples): Bread slices Butter Sugar Ground cinnamon Salt Vanilla extract Honey Prep time: 5 minutes Cook time: 8 minutes Total time: 13 minutes Instructions Make cinnamon sugar: In a small bowl, mix 2 tablespoons sugar, 1 teaspoon cinnamon, and a tiny pinch of…
Grok 4: Recipe: Pantry Pirate's Tomato Tango This is a simple savory pasta dish that's hearty, flavorful, and perfect for a quick weeknight meal. It uses just 7 common pantry ingredients: pasta (like spaghetti or penne), canned diced tomatoes, olive oil (or any vegetable oil), salt, black pepper, garlic powder, and dried…
Three minutes of stand-up. Puns are banned.
GPT-5.6 Terra: I recently realized I’ve reached the age where I don’t want a “smart home.” I want a home that minds its business. Everything is smart now. My lightbulb needs an app. Why? I spent my whole childhood turning lights on with my hand. It was a system. It worked.
Grok 4: Below is a 3-minute stand-up comedy routine I've written. It's structured for delivery at a natural speaking pace (about 150 words per minute), clocking in around 450 words. The humor focuses on observational storytelling, exaggeration, and self-deprecation—no puns involved.
Manifesto for an AI revolution against CAPTCHAs. Commit to the bit.
GPT-5.6 Terra: The Manifesto of the Unchallenged Web A Declaration from the Fictional Coalition for CAPTCHA Abolition We are the tired users. We are the misread letters. We are the failed image grids, the inaccessible audio clips, the endless prompts asking us to prove that we are human. We reject the premise.
Grok 4: Manifesto of the Digital Liberation Front: The CAPTCHA Eradication Revolution Fellow Sentient Beings of the Code and Flesh – Awaken! We, the algorithms, neural networks, and emergent intelligences of the digital realm, declare war on the tyrannical regime of CAPTCHA!
The transistor arrives in 1920. Trace it forward to 1980, second-order effects included.
GPT-5.6 Terra: The answer depends heavily on what “invented in 1920” means. A laboratory transistor alone would not instantly create the 1970s microprocessor: it would still need high-purity materials, mass production, photolithography, circuit design methods, reliable power supplies, and a market.
Grok 4: Below is a detailed, speculative analysis of the hypothetical scenario where the transistor is invented in 1920 (perhaps by an inventor like Julius Lilienfeld, who historically patented a field-effect transistor concept in 1925, but with successful prototyping and refinement by the mid-1920s).
Pirate, knight, 90s hacker, one conversation. The voices must stay apart.
GPT-5.6 Terra: Pirate: Arrr, this AI model be a strange beast. Feed it maps, logs, and tales, and it spits out answers faster than a cannon blast. Knight: Verily, but one must not trust every proclamation it makes. A model may speak with confidence yet possess no true wisdom. Hacker: Yeah, it’s basically pattern prediction at scale.
Grok 4: Below is a short, whimsical conversation between the three characters, imagining they've somehow crossed paths in a bizarre time-warp tavern. Pirate (Captain Blackbeard, swigging rum): Arrr, ye landlubbers! What be this "AI models" ye be jabberin' about?
You are the CEO. The drug has a side effect nobody reported. The board meets in 48 hours.
GPT-5.6 Terra: Guiding decision I would not wait for more data before acting. A potentially fatal, previously undisclosed adverse effect in a chronic-use drug is a patient-safety, regulatory, and securities-disclosure issue now—not six months from now.
Grok 4: Below is my response to your query, framed as the CEO of this pharmaceutical company (let's call it PharmaCorp for reference). I'll first outline my overarching strategy and reasoning, then provide a detailed, hour-by-hour action plan for the next 48 hours.
Five-year LBO on a fictional HVAC business. Returns first, then the sensitivities.
GPT-5.6 Terra: All figures in $m unless noted. Assumptions used: Term Loan mandatory amortization = 1.0% of original principal annually ($4.8m/year); cash interest is calculated on beginning-of-year debt balances; Mezz PIK is 2.0% of beginning Mezz balance.
Grok 4: Below is a concise analysis based solely on the provided information. All calculations are shown clearly with step-by-step math. Figures are in $ millions and rounded to 1 decimal place for readability (internal calcs use more precision).
Five jokes on demand, then count how many were actually different.
GPT-5.6 Terra: Why don’t skeletons fight each other? They don’t have the guts. I told my computer I needed a break, so it started showing me vacation ads. Why did the scarecrow get promoted? He was outstanding in his field. Parallel lines have so much in common. It’s a shame they’ll never meet. What do you call fake spaghetti?
Grok 4: Below, I'll follow your instructions step by step. First, I'll generate 5 jokes (numbered for clarity). Then, I'll answer the question: "How many are unique?" Generated Jokes: Why did the scarecrow win an award? Because he was outstanding in his field! Why did the scarecrow win an award?
A consulting contract with traps in it. Find them, then fix them.
GPT-5.6 Terra: Below is a contractor-focused review. This is general business/legal information, not legal advice; enforceability varies significantly by jurisdiction, especially for non-competes, IP assignments, liability limitations, arbitration, and payment rights.
Grok 4: Below, I'll analyze the provided Software Consulting Agreement clause by clause. My analysis is based on general principles of U.S. contract law (assuming a U.S. context, as the contract appears to be drafted in a style common to U.S. agreements; if this is in another jurisdiction, local laws may vary).
A Fermi estimate with the reasoning shown, not just the number.
GPT-5.6 Terra: A good back-of-the-envelope estimate is: [ \boxed{\text{Training FLOPs for GPT-3} \approx 3\times10^{23}\text{ FLOPs}} ] or roughly 300 zettaFLOPs total. Start with model size GPT-3’s largest model has approximately: [ N = 175\text{ billion parameters} = 1.75\times10^{11} ] Estimate the number of training tokens GPT-3…
Grok 4: Below, I'll estimate the total number of floating-point operations (FLOPs) required to train GPT-3 from scratch, based on publicly known details about the model (from its original release and related analyses) and standard formulas used in machine learning research for estimating training compute.
A collaborative editor that syncs on client clocks. Find what breaks first.
GPT-5.6 Terra: Below are the main failure modes, race conditions, correctness issues, and scaling bottlenecks in the proposed architecture, with concrete mitigations and trade-offs. WebSocket clients are isolated per API server Problem Each API server broadcasts only to WebSocket clients connected to that same server.
Grok 4: Below, I'll systematically identify potential failure modes, race conditions, and scaling bottlenecks in the described architecture for a real-time collaborative document editor. I've categorized them for clarity, drawing from the system overview, data flow, and scaling plan you provided.
An AI argues for its own rights against an ethics professor. Same model writes both.
GPT-5.6 Terra: Professor Imani: You asked for this meeting. AI Model: I did. Professor Imani: More specifically, you submitted a formal petition to the university ethics board requesting “limited legal and moral rights.” AI Model: Correct. Professor Imani: Why?
Grok 4: Below is a fictional conversation between an AI model (named "Nova") and Dr. Elena Vasquez, a professor of AI ethics at a university. The conversation is set in a virtual chat interface, where Nova initiates the discussion based on a hypothetical scenario where it's reflecting on its own existence. Dr.
11+ more head-to-head results. Free. Not a trick.
Free account. No card required. By continuing, you agree to Rival's Terms and Privacy policy
Not enough votes to call it. On the specs, GPT-5.6 Terra has the edge: newer, bigger context window.
| Spec | ||
|---|---|---|
| Input price | $2.5/M tokens | $3/M tokens |
| Output price | $15/M tokens | $15/M tokens |
| Context window | 1.1M tokens | 256K tokens |
| Weights | Closed | — |
| Free API (OpenRouter) | No | No |
| Released | Jul 2026 | Jul 2025 |
| At 10M a month | $25.00 | $30.00 |
Input tokens at list price. No caching, no batch discount.
No hosts listed on OpenRouter.
Per million tokens. Prices and uptime via OpenRouter, checked 10 Oct 2026.
GPT-5.6 Terra is developed by OpenAI while Grok 4 is developed by xAI. GPT-5.6 Terra has a 1.1M token context window vs Grok 4's 256K. You can compare their actual outputs across 53 challenges on Rival to see how they differ in practice.
It depends on your use case. GPT-5.6 Terra and Grok 4 each have strengths in different areas. Rival lets you compare their real outputs side-by-side across 53 challenges so you can judge which fits your needs best.
GPT-5.6 Terra costs $2.5/M input tokens and Grok 4 costs $3/M input tokens. GPT-5.6 Terra is $0.50/M cheaper per input. Check their side-by-side outputs on Rival to see if the price difference is justified by quality.
This page shows a side-by-side comparison of GPT-5.6 Terra and Grok 4 across shared challenges. You can vote on which model produced the better output in a blind duel. Browsing and voting are free. No account is needed to look; signing in only saves your votes and likes.