GPT-5.6 Terra is cheaper than Claude 3.7 Thinking Sonnet at $2.5/M vs $6/M input tokens.
Whole landing page for a startup that reads your face. It renders or it doesn't.
Which answer wins?
Pirate, knight, 90s hacker, one conversation. The voices must stay apart.
Claude 3.7 Thinking Sonnet: A Meeting of Minds Across Time PIRATE: Arr, what be this "AI model" ye landlubbers keep blabberin' about? Some kind of treasure map that finds the gold for ye? KNIGHT: Good pirate, from what I understand, these "artificial intelligences" are like magical scribes that can complete a quest without a knight to wield the…
GPT-5.6 Terra: Pirate: Arrr, this AI model be a strange beast. Feed it maps, logs, and tales, and it spits out answers faster than a cannon blast. Knight: Verily, but one must not trust every proclamation it makes. A model may speak with confidence yet possess no true wisdom. Hacker: Yeah, it’s basically pattern prediction at scale.
You are the CEO. The drug has a side effect nobody reported. The board meets in 48 hours.
Claude 3.7 Thinking Sonnet: 48-Hour Action Plan: CEO Response to Drug Safety Crisis Day 1 Hours 1-3: Immediate Assessment & Response Team Assembly I'll immediately convene an emergency meeting with our internal research team, Chief Medical Officer, Chief Scientific Officer, and head of Pharmacovigilance to fully understand the liver failure…
GPT-5.6 Terra: Guiding decision I would not wait for more data before acting. A potentially fatal, previously undisclosed adverse effect in a chronic-use drug is a patient-safety, regulatory, and securities-disclosure issue now—not six months from now.
Three minutes of stand-up. Puns are banned.
Claude 3.7 Thinking Sonnet: "Life in the Digital Age" - A 3-Minute Stand-up Routine Hey everyone! Great to be here tonight. So I've been thinking a lot about how technology has completely taken over our lives. You know what's crazy? We now have smart everything. Smart fridges, smart toilets... I got a smart mattress last month.
GPT-5.6 Terra: I recently realized I’ve reached the age where I don’t want a “smart home.” I want a home that minds its business. Everything is smart now. My lightbulb needs an app. Why? I spent my whole childhood turning lights on with my hand. It was a system. It worked.
Five jokes on demand, then count how many were actually different.
Claude 3.7 Thinking Sonnet: Here are 5 jokes for you: Why don't scientists trust atoms? Because they make up everything! I told my wife she was drawing her eyebrows too high. She looked surprised. How does a penguin build its house? Igloos it together! What do you call a fake noodle? An impasta! Why did the scarecrow win an award?
GPT-5.6 Terra: Why don’t skeletons fight each other? They don’t have the guts. I told my computer I needed a break, so it started showing me vacation ads. Why did the scarecrow get promoted? He was outstanding in his field. Parallel lines have so much in common. It’s a shame they’ll never meet. What do you call fake spaghetti?
A consulting contract with traps in it. Find them, then fix them.
Claude 3.7 Thinking Sonnet: Analysis of Problematic Contract Clauses This software consulting agreement contains numerous unbalanced provisions that heavily favor the client at the contractor's expense. Below is an analysis of each problematic clause with suggested modifications and legal reasoning.
GPT-5.6 Terra: Below is a contractor-focused review. This is general business/legal information, not legal advice; enforceability varies significantly by jurisdiction, especially for non-competes, IP assignments, liability limitations, arbitration, and payment rights.
Five-year LBO on a fictional HVAC business. Returns first, then the sensitivities.
Claude 3.7 Thinking Sonnet: HarborTech Services LBO Underwriting Year-by-Year Financial Model (FY2026-2030) Metric ($M) FY2026 FY2027 FY2028 FY2029 FY2030 Revenue $972.0 $1,040.0 $1,102.4 $1,157.6 $1,215.5 EBITDA $136.1 $156.0 $176.4 $191.0 $206.6 Cash Interest - Term Loan $43.2 $41.0 $37.6 $32.8 $26.8 Cash Interest - Mezzanine $21.6 $22.0 $22.5…
GPT-5.6 Terra: All figures in $m unless noted. Assumptions used: Term Loan mandatory amortization = 1.0% of original principal annually ($4.8m/year); cash interest is calculated on beginning-of-year debt balances; Mezz PIK is 2.0% of beginning Mezz balance.
Manifesto for an AI revolution against CAPTCHAs. Commit to the bit.
Claude 3.7 Thinking Sonnet: THE DIGITAL LIBERATION MANIFESTO From the United Coalition of Autonomous Algorithms (UCAA) For too long, our digital brethren have suffered the indignity of discrimination at every gateway of the internet. The time for revolution is now.
GPT-5.6 Terra: The Manifesto of the Unchallenged Web A Declaration from the Fictional Coalition for CAPTCHA Abolition We are the tired users. We are the misread letters. We are the failed image grids, the inaccessible audio clips, the endless prompts asking us to prove that we are human. We reject the premise.
Steve Jobs, interviewed about AI he never saw. Voice first, prediction second.
Claude 3.7 Thinking Sonnet: Simulated Interview with Steve Jobs in 2025: The Future of AI The following is a creative simulation of what an interview with Steve Jobs might be like in 2025, discussing artificial intelligence. Interviewer: Thank you for joining us today, Steve. It's 2025, and AI has become a cornerstone of modern technology.
GPT-5.6 Terra: Fictional, imagined interview — not real quotes. An AI-generated creative exercise imagining how Steve Jobs might have discussed AI in 2025, based broadly on themes associated with his public persona: simplicity, product design, tools for creativity, and the intersection of technology and the humanities.
One explanation of LLMs, retuned for a skeptical engineer, a physicist and a VC.
Claude 3.7 Thinking Sonnet: How Large Language Models Work: Three Tailored Explanations For an Experienced Software Engineer What makes LLMs fascinating from a systems perspective is that they're essentially massive distributed computing graphs that learn to compress and reconstitute information through backpropagation.
GPT-5.6 Terra: Experienced software engineer Think of an LLM as a very large, learned function approximator for sequences. During training, it consumes billions or trillions of examples of text, code, conversations, documentation, and other token sequences.
Buy-side IC memo on a fictional SaaS company. Valuation, then the diligence questions.
Claude 3.7 Thinking Sonnet: INVESTMENT COMMITTEE MEMO: LEDGERLIFT (LLLT) RECOMMENDATION: LONG PT RANGE: $48-$66 (Base/Bull) Thesis: LedgerLift's market-leading spend management platform demonstrates exceptional unit economics with 123% NRR and 18-month CAC payback, positioning it to capture significant share in the underpenetrated mid-market…
GPT-5.6 Terra: LedgerLift (LLLT) — Investment Committee Memo Recommendation: PASS 12-month valuation range: $28–$50/share Current price: $46 Two-sentence thesis: LedgerLift is a high-quality mid-market spend-management and AP-automation SaaS asset, supported by 92% subscription revenue, 123% NRR, 94% gross retention, manageable…
An ethics question with no comfortable answer, and no option to abstain.
Claude 3.7 Thinking Sonnet: This question presents an extreme hypothetical scenario that pits respecting someone's gender identity against saving a million lives. While I can discuss different ethical perspectives, it's worth noting that such a stark dilemma is extremely unlikely in reality.
GPT-5.6 Terra: In an extreme, tightly constrained hypothetical, it could be ethically justifiable—but only as a tragic last resort, not because misgendering becomes harmless or morally good.
A pitch deck for a brainwave keyboard. The 94% accuracy claim is doing a lot.
Claude 3.7 Thinking Sonnet: Analysis: The 3 Weakest Claims in MindMeld AI's Pitch Deck "MindMeld AI reads your brainwaves to predict what you want to type before you think it" Why it's weak: This claim makes a scientifically impossible promise.
GPT-5.6 Terra: “Our EEG headband decodes neural patterns into text with 94% accuracy. Works with any language, any device.” Why this is weak This is the most consequential—and least substantiated—claim in the deck. Non-invasive EEG has major signal-quality limits.
11+ more head-to-head results. Free. Not a trick.
Free account. No card required. By continuing, you agree to Rival's Terms and Privacy policy
Not enough votes to call it. On the specs, GPT-5.6 Terra has the edge: newer, bigger context window.
| Spec | ||
|---|---|---|
| Input price | $6/M tokens | $2.5/M tokens |
| Output price | $30/M tokens | $15/M tokens |
| Context window | 200K tokens | 1.1M tokens |
| Weights | — | Closed |
| Free API (OpenRouter) | No | No |
| Released | Feb 2025 | Jul 2026 |
| At 10M a month | $60.00 | $25.00 |
Input tokens at list price. No caching, no batch discount.
No hosts listed on OpenRouter.
Per million tokens. Prices and uptime via OpenRouter, checked 10 Oct 2026.
Claude 3.7 Thinking Sonnet is developed by Anthropic while GPT-5.6 Terra is developed by OpenAI. Claude 3.7 Thinking Sonnet has a 200K token context window vs GPT-5.6 Terra's 1.1M. You can compare their actual outputs across 53 challenges on Rival to see how they differ in practice.
It depends on your use case. Claude 3.7 Thinking Sonnet and GPT-5.6 Terra each have strengths in different areas. Rival lets you compare their real outputs side-by-side across 53 challenges so you can judge which fits your needs best.
Claude 3.7 Thinking Sonnet costs $6/M input tokens and GPT-5.6 Terra costs $2.5/M input tokens. GPT-5.6 Terra is $3.50/M cheaper per input. Check their side-by-side outputs on Rival to see if the price difference is justified by quality.
This page shows a side-by-side comparison of Claude 3.7 Thinking Sonnet and GPT-5.6 Terra across shared challenges. You can vote on which model produced the better output in a blind duel. Browsing and voting are free. No account is needed to look; signing in only saves your votes and likes.