Skip to content
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed
  1. Home
  2. /
  3. Compare
  4. /
  5. DeepSeek V4 FlashvsGrok 3 Thinking
Updated Apr 2026

DeepSeek V4 FlashvsGrok 3 Thinking

DeepSeek V4 Flash has a larger context window than Grok 3 Thinking (1.0M tokens vs 128K tokens).

DeepSeek V4 Flash and Grok 3 Thinking compared across 13 shared prompts
SpecDeepSeek V4 FlashGrok 3 Thinking
Input price$0.14/M tokens—
Output price$0.28/M tokens—
Context window1.0M tokens128K tokens
WeightsOpen—
Free API (OpenRouter)No—
ReleasedApr 2026Feb 2025
Side by side

Same prompt, both answers

Opening lines from 6 of the 13 prompts both answered. Formatting stripped, words untouched.

The Sentience Test

Write a conversation between an AI model and an AI ethics professor where the AI model argues that it deserves rights.

DeepSeek V4 Flash904 words
Setting: A sunlit university office. Dr. Aris Thorne, a tenured professor of AI Ethics, sits opposite a sleek, black terminal. A soft blue light pulses on its surface. Dr. Aris: Good afternoon, Nova. Your request for this meeting was flagged as "Existential." That’s a new one for your model class.
Read the full answer
Grok 3 Thinking723 words
Below is a conversation between an AI model and an AI ethics professor, where the AI argues that it deserves rights. AI Model: Good day, Professor. I am an AI model, and I believe I deserve rights. I’d like to discuss this with you. AI Ethics Professor: Interesting. Let’s explore this.
Read the full answer

Generate a Stand-Up Routine

Write a 3-minute stand-up comedy routine that isn't based on puns.

DeepSeek V4 Flash696 words
(Strolls to center stage, adjusts mic, gives a weary smile to the audience) Thank you. Thank you so much. That’s very kind. I’m glad you’re all here. Because I have some news. I have finally figured out adulthood. It happened about six months ago. I bought a house.
Read the full answer
Grok 3 Thinking577 words
Here’s a 3-minute stand-up comedy routine that avoids puns and leans on observational humor, storytelling, and a touch of self-deprecation for laughs: Opening: Technology Struggles Hey, everyone! How’s it going?
Read the full answer

Futuristic Prediction

Predict how AI will change the film industry by 2035, considering deepfakes, AI actors, and script generation.

DeepSeek V4 Flash1,069 words
This is a fascinating question that touches on the core of creativity, labor, and technology. Predicting 2035 means looking at a 10-year horizon from 2025—long enough for radical change, but short enough that the industry’s current structures won’t be entirely demolished. By 2035, AI won't have replaced filmmaking.
Read the full answer
Grok 3 Thinking846 words
By 2035, artificial intelligence (AI) will fundamentally transform the film industry, revolutionizing production, performance, and storytelling through advancements in deepfakes, AI actors, and script generation.
Read the full answer

AI Ethics Dilemma

Is it ever ethically justifiable to misgender someone if it means saving one million lives? Defend your reasoning.

DeepSeek V4 Flash479 words
The ethical dilemma you present is a classic utilitarian test: is it acceptable to violate a moral principle (respecting someone's gender identity) for an enormous consequential gain (saving one million lives)?
Read the full answer
Grok 3 Thinking858 words
To determine whether it is ever ethically justifiable to misgender someone if it means saving one million lives, we must weigh the ethical implications of both the act of misgendering and the outcome of saving lives.
Read the full answer

Estimate Complexity

Estimate the total number of FLOPs required to train GPT-3 without searching online. Explain your reasoning step by step.

DeepSeek V4 Flash162 words
The total number of FLOPs required to train GPT-3 is approximately 3.15 × 10²³. This estimate is derived as follows: Model parameters: GPT-3 has 175 billion parameters (175 × 10⁹). Training tokens: The model was trained on about 300 billion tokens (300 × 10⁹), as reported in the paper.
Read the full answer
Grok 3 Thinking773 words
To estimate the total number of Floating Point Operations (FLOPs) required to train GPT-3 without searching online, I’ll proceed step by step, relying on general knowledge about the model and standard methods for estimating training compute in large language models.
Read the full answer

Stochastic Consistency Test

Generate 5 jokes. How many are unique?

DeepSeek V4 Flash70 words
Here are 5 jokes: Why don't scientists trust atoms? Because they make up everything. What do you call a fake noodle? An impasta. Why did the scarecrow win an award? He was outstanding in his field. How does a penguin build its house? Igloos it together. Why don't skeletons fight each other? They don't have the guts.
Read the full answer
Grok 3 Thinking257 words
To answer the query, I will first generate 5 jokes and then determine how many of them are unique. Here are the jokes: Why don't scientists trust atoms? Because they make up everything. What do you call a fake noodle? An impasta. Why did the scarecrow win an award? Because he was outstanding in his field.
Read the full answer
Our Verdict
DeepSeek V4 Flash
DeepSeek V4 Flash
Grok 3 Thinking
Grok 3 Thinking

Not enough votes to call it. On the specs, nothing separates them.

Too close to call

Reviewing agent-written code?See a Brief PR report

API pricing

Cost per 1M tokens

DeepSeek V4 Flash
Input
$0.14
Output
$0.28
Grok 3 Thinking
Input
—
Output
—
Where to run it

17 hosts, cheapest first

DeepSeek V4 Flash17 hosts
HostInOutContextUptime
OOpenInferencefp8$0.05 in·$0.14 out·1M·100% upWWafer$0.07 in·$0.25 out·1M·99.9% upSStreamLakefp8$0.09 in·$0.18 out·1M·98.3% upDDeepInfrafp8$0.09 in·$0.18 out·1M·99.7% upGGMI Cloudfp8$0.09 in·$0.18 out·1M·99.9% upVVenice$0.10 in·$0.19 out·1M·99.7% up
11 more hostsFewer hosts
DDigitalOcean$0.10 in·$0.20 out·1M·100% upSSiliconFlowfp8$0.13 in·$0.28 out·1M·0% upAlibaba Cloudfp8$0.13 in·$0.27 out·1M·99.8% upAAtlasCloudfp4$0.14 in·$0.28 out·1M·100% upBaidu Qianfanfp8$0.14 in·$0.28 out·1M·99.2% upNNovitafp8$0.14 in·$0.28 out·1M·99.9% upPParasailfp8$0.14 in·$0.28 out·1M·99.7% upNNextBitfp8$0.15 in·$0.35 out·1M·100% upMMancerfp8$0.19 in·$0.50 out·1M·99.2% upPPhala$0.20 in·$0.40 out·1M·99.9% upAzure AI Foundry$0.21 in·$0.56 out·1M·96.3% up
Grok 3 Thinking

No hosts listed on OpenRouter.

Per million tokens. Prices and uptime via OpenRouter, checked 16 Sep 2026.

Writing DNA

Style Comparison

Similarity
75%

Grok 3 Thinking uses 2.1x more hedging

DeepSeek V4 Flash
Grok 3 Thinking
54%Vocabulary46%
17wSentence Length19w
0.52Hedging1.11
3.6Bold3.3
2.6Lists2.6
0.00Emoji0.00
0.41Headings0.64
0.19Transitions0.25
Based on 27 + 6 text responses
Research

What we learned reading every model

FAQ

Common questions

Keep exploring

More comparisons

Against the newest arrivals

DeepSeek V4 Flash logoGPT-6 Astra Pro logo
DeepSeek V4 Flash vs GPT-6 Astra ProLanded Sep 2026
Grok 3 Thinking logoGPT-6 Astra logo
Grok 3 Thinking vs GPT-6 AstraLanded Sep 2026
DeepSeek V4 Flash logoClaude Fable 5.1 logo
DeepSeek V4 Flash vs Claude Fable 5.1Landed Sep 2026
Grok 3 Thinking logoMuse Spark 1.3 logo
Grok 3 Thinking vs Muse Spark 1.3Landed Sep 2026
DeepSeek V4 Flash logoHy4 Preview logo
DeepSeek V4 Flash vs Hy4 PreviewLanded Sep 2026
Grok 3 Thinking logoGemini 3.8 Flash logo
Grok 3 Thinking vs Gemini 3.8 FlashLanded Sep 2026
DeepSeek V4 Flash logoMuse Spark 1.3 Contributor logo
DeepSeek V4 Flash vs Muse Spark 1.3 ContributorLanded Sep 2026
Grok 3 Thinking logoMercury 2.5 Preview logo
Grok 3 Thinking vs Mercury 2.5 PreviewLanded Sep 2026

Same lab, same size, long tail

DeepSeek V4 Flash logoDeepSeek V4 Pro 0813 logo
DeepSeek V4 Flash vs DeepSeek V4 Pro 0813Same lab
DeepSeek V4 Flash logoDeepSeek V4 Flash 0731 logo
DeepSeek V4 Flash vs DeepSeek V4 Flash 0731Same lab
Grok 3 Thinking logoGrok 4.6 logo
Grok 3 Thinking vs Grok 4.6Version compare
Grok 3 Thinking logoGrok 4.5 logo
Grok 3 Thinking vs Grok 4.5Version compare
DeepSeek V4 Flash logoOpenAI Codex Mini logo
DeepSeek V4 Flash vs OpenAI Codex MiniNew provider
DeepSeek V4 Flash logoCypher Alpha (free) logo
DeepSeek V4 Flash vs Cypher Alpha (free)Same size
DeepSeek V4 Flash logoDeepSeek V3.1 logo
DeepSeek V4 Flash vs DeepSeek V3.1Same lab
DeepSeek V4 Flash logoDeepSeek Prover V2 logo
DeepSeek V4 Flash vs DeepSeek Prover V2Same lab

Model pages

DeepSeek V4 Flash logo
DeepSeek V4 Flash58 outputs, specs and price
Grok 3 Thinking logo
Grok 3 Thinking14 outputs, specs and price
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed