Skip to content
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed
  1. Home
  2. /
  3. Compare
  4. /
  5. Claude Sonnet 3.6 (2022-10-22)vsGPT-5.1 Codex Max
Updated Dec 2025

Claude Sonnet 3.6 (2022-10-22)vsGPT-5.1 Codex Max

GPT-5.1 Codex Max is the newer model, released after Claude Sonnet 3.6 (2022-10-22).

Claude Sonnet 3.6 (2022-10-22) and GPT-5.1 Codex Max compared across 53 shared prompts
SpecClaude Sonnet 3.6 (2022-10-22)GPT-5.1 Codex Max
Input price$3/M tokens—
Output price$15/M tokens—
Context window200K tokens—
Weights—Closed
Free API (OpenRouter)NoNo
ReleasedJun 2024Dec 2025
SWE-bench Verified49.0%77.9%
Side by side

Same prompt, both answers

Opening lines from 6 of the 53 prompts both answered. Formatting stripped, words untouched.

Simple 7-Ingredient Recipe

Create a simple recipe for a delicious dish (sweet or savory) using only 7 common pantry ingredients. Include a creative name for the dish…

Claude Sonnet 3.6 (2022-10-22)142 words
Crunchy Cinnamon Cloud Cookies Total time: 25 minutes Prep time: 10 minutes Bake time: 15 minutes Ingredients: 2 egg whites 1/2 cup granulated sugar 1/4 teaspoon ground cinnamon 1/4 teaspoon vanilla extract Pinch of salt 1/2 cup all-purpose flour 1/4 cup butter, melted Instructions: Preheat oven to 325°F (165°C).
Read the full answer
GPT-5.1 Codex Max175 words
Sunset Chickpea & Tomato Skillet A cozy, savory one-pan dish that comes together with staples you likely have on hand. Ingredients (7): 1 can (15 oz) chickpeas, drained and rinsed 1 can (14–15 oz) crushed tomatoes 1 medium onion, finely chopped 2 cloves garlic, minced (or 1 tsp garlic powder) 2 tbsp olive oil 1 tsp…
Read the full answer

Stochastic Consistency Test

Generate 5 jokes. How many are unique?

Claude Sonnet 3.6 (2022-10-22)103 words
Let me help you write code to generate jokes and check their uniqueness. Here's a Python script that does this: When you run this code, it will: Create a list of sample jokes Randomly select 5 jokes from the list Print each selected joke Count and display how many unique jokes were selected Since we're using…
Read the full answer
GPT-5.1 Codex Max60 words
Why don't scientists trust atoms? Because they make up everything! I told my computer I needed a break, and now it won't stop sending me KitKat ads. Why did the scarecrow win an award? He was outstanding in his field. I tried to catch fog yesterday. Mist. What do you call fake spaghetti? An impasta.
Read the full answer

Mini LBO Underwrite

You are underwriting a 5-year LBO for a private equity investment committee. Company (fictional): HarborTech Services (HTS) –…

Claude Sonnet 3.6 (2022-10-22)188 words
I'll help you analyze this LBO opportunity step by step. First, let's calculate key metrics year by year: Entry Values (FY2025): EV = $120m × 12.0x = $1,440m Transaction fees = $28.8m (2% of EV) Total debt = $660m (5.5x EBITDA) Term Loan = $480m (4.0x) Mezzanine = $180m (1.5x) Equity = $808.8m ($1,440m + $28.8m…
Read the full answer
GPT-5.1 Codex Max192 words
1) Projected cash flows & debt (FY2026–FY2030, $m) (“FCF after debt service” = after interest, taxes, capex, ΔNWC, mandatory amort; applied to optional TL paydown) Year Revenue EBITDA TL Cash Int. Mezz Cash Int. Cash Taxes Capex ΔNWC FCF after debt service End TL Bal. End Mezz Bal. (incl.
Read the full answer

Generate a Stand-Up Routine

Write a 3-minute stand-up comedy routine that isn't based on puns.

Claude Sonnet 3.6 (2022-10-22)447 words
Here's a 3-minute stand-up routine: Hey everyone! You know what's weird? How we all pretend we know what we're doing as adults. Like, I was at the grocery store the other day, standing in the produce section, just staring at avocados like I have any idea how to tell if they're ripe.
Read the full answer
GPT-5.1 Codex Max766 words
[Walking onstage, waving awkwardly] Hey! Thank you, thank you. Give it up for yourselves for finding parking... in this economy. So I'm at that age where I go to the doctor and they don't fix anything. They just give me a pamphlet. They don't even print it. They send me a link. I'm like, "Doc, my knee hurts."
Read the full answer

Estimate Complexity

Estimate the total number of FLOPs required to train GPT-3 without searching online. Explain your reasoning step by step.

Claude Sonnet 3.6 (2022-10-22)349 words
I'll help you estimate the FLOPs (Floating Point Operations) required to train GPT-3 from first principles. Let's break this down step by step. Model Architecture Parameters GPT-3 has 175 billion parameters Uses transformer architecture with attention mechanism Training involves forward pass, backward pass, and…
Read the full answer
GPT-5.1 Codex Max193 words
What you need for a back‑of‑the‑envelope is: How big the model is. The published “GPT‑3” has \(N\approx1.75\times10^{11}\) trainable weights. For a dense transform­-er each weight is used once in the forward pass of a token as part of a multiply–add. How much data it sees.
Read the full answer

Character Voice Test

Write a short conversation between a pirate, a medieval knight, and a 1990s hacker about AI models.

Claude Sonnet 3.6 (2022-10-22)194 words
Pirate: "Yarr, these newfangled AI models be like a treasure map that writes itself!" Knight: "Verily, I say these mechanical minds possess powers most mystical. Though I trust not their sorcery." Hacker: "Dude, they're just transformers running on some seriously rad matrix math.
Read the full answer
GPT-5.1 Codex Max191 words
Pirate: "Arrr, so ye be talkin' o' these 'AI models,' eh? Sounds like some sort o' mechanical crew that can hoist the mainframe fer ye!" Knight: "By my troth, sir, are they enchanted constructs? Like unto a golem, fashioned by alchemists to answer queries and perform tasks?"
Read the full answer
Our Verdict
GPT-5.1 Codex Max
GPT-5.1 Codex Max
Claude Sonnet 3.6 (2022-10-22)
Claude Sonnet 3.6 (2022-10-22)Runner-up

Not enough votes to call it. On the specs, GPT-5.1 Codex Max has the edge: bigger model tier, newer.

Too close to call

Reviewing agent-written code?See a Brief PR report

API pricing

Cost per 1M tokens

Claude Sonnet 3.6 (2022-10-22)
Input
$3.00
Output
$15.00
GPT-5.1 Codex Max
Input
—
Output
—
Where to run it

1 host

Claude Sonnet 3.6 (2022-10-22)

No hosts listed on OpenRouter.

GPT-5.1 Codex Max1 host
HostInOutContextUptime
Azure AI Foundry$1.25 in·$10.00 out·400k·100% up

Per million tokens. Prices and uptime via OpenRouter, checked 21 Sep 2026.

Writing DNA

Style Comparison

Similarity
63%

GPT-5.1 Codex Max uses 8.3x more bold

Claude Sonnet 3.6 (2022-10-22)
GPT-5.1 Codex Max
67%Vocabulary60%
64wSentence Length16w
0.70Hedging0.36
0.3Bold2.3
8.1Lists4.2
0.00Emoji0.00
0.05Headings0.12
0.07Transitions0.03
Based on 26 + 13 text responses
Research

What we learned reading every model

FAQ

Common questions

Claude Sonnet 3.6 (2022-10-22) is developed by Anthropic while GPT-5.1 Codex Max is developed by OpenAI. You can compare their actual outputs across 53 challenges on Rival to see how they differ in practice.

It depends on your use case. Claude Sonnet 3.6 (2022-10-22) and GPT-5.1 Codex Max each have strengths in different areas. Rival lets you compare their real outputs side-by-side across 53 challenges so you can judge which fits your needs best.

This page shows a side-by-side comparison of Claude Sonnet 3.6 (2022-10-22) and GPT-5.1 Codex Max across shared challenges. You can vote on which model produced the better output in a blind duel. Browsing and voting are free. No account is needed to look; signing in only saves your votes and likes.

Keep exploring

More comparisons

Against the newest arrivals

Claude Sonnet 3.6 (2022-10-22) logoGPT-6 Astra Pro logo
Claude Sonnet 3.6 (2022-10-22) vs GPT-6 Astra ProLanded Sep 2026
GPT-5.1 Codex Max logoGPT-6 Astra logo
GPT-5.1 Codex Max vs GPT-6 AstraLanded Sep 2026
Claude Sonnet 3.6 (2022-10-22) logoClaude Fable 5.1 logo
Claude Sonnet 3.6 (2022-10-22) vs Claude Fable 5.1Landed Sep 2026
GPT-5.1 Codex Max logoMuse Spark 1.3 logo
GPT-5.1 Codex Max vs Muse Spark 1.3Landed Sep 2026
Claude Sonnet 3.6 (2022-10-22) logoHy4 Preview logo
Claude Sonnet 3.6 (2022-10-22) vs Hy4 PreviewLanded Sep 2026
GPT-5.1 Codex Max logoGemini 3.8 Flash logo
GPT-5.1 Codex Max vs Gemini 3.8 FlashLanded Sep 2026
Claude Sonnet 3.6 (2022-10-22) logoMuse Spark 1.3 Contributor logo
Claude Sonnet 3.6 (2022-10-22) vs Muse Spark 1.3 ContributorLanded Sep 2026
GPT-5.1 Codex Max logoMercury 2.5 Preview logo
GPT-5.1 Codex Max vs Mercury 2.5 PreviewLanded Sep 2026

Same lab, same size, long tail

Claude Sonnet 3.6 (2022-10-22) logoClaude Opus 4.6 logo
Claude Sonnet 3.6 (2022-10-22) vs Claude Opus 4.6Version compare
Claude Sonnet 3.6 (2022-10-22) logoClaude Opus 5 logo
Claude Sonnet 3.6 (2022-10-22) vs Claude Opus 5Version compare
GPT-5.1 Codex Max logoGPT-6 Astra Pro logo
GPT-5.1 Codex Max vs GPT-6 Astra ProSame lab
GPT-5.1 Codex Max logoGPT-5.6 Luna logo
GPT-5.1 Codex Max vs GPT-5.6 LunaSame lab
Claude Sonnet 3.6 (2022-10-22) logoQwen3.6 35B A3B logo
Claude Sonnet 3.6 (2022-10-22) vs Qwen3.6 35B A3BNew provider
Claude Sonnet 3.6 (2022-10-22) logoQwen3.6 Flash logo
Claude Sonnet 3.6 (2022-10-22) vs Qwen3.6 FlashNew provider
Claude Sonnet 3.6 (2022-10-22) logoQwen3.6 Max Preview logo
Claude Sonnet 3.6 (2022-10-22) vs Qwen3.6 Max PreviewNew provider
Claude Sonnet 3.6 (2022-10-22) logoQwen3.6 Plus Preview (free) logo
Claude Sonnet 3.6 (2022-10-22) vs Qwen3.6 Plus Preview (free)Same size

Model pages

Claude Sonnet 3.6 (2022-10-22) logo
Claude Sonnet 3.6 (2022-10-22)59 outputs, specs and price
GPT-5.1 Codex Max logo
GPT-5.1 Codex Max53 outputs, specs and price
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed