Skip to content
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed
  1. Home
  2. /
  3. Compare
  4. /
  5. Qwen3.6 FlashvsQwen3.8 2.4T A95B
Updated Aug 2026

Qwen3.6 FlashvsQwen3.8 2.4T A95B

Qwen3.6 Flash is cheaper than Qwen3.8 2.4T A95B at $0.25/M vs $2/M input tokens.

Qwen3.6 Flash and Qwen3.8 2.4T A95B compared across 54 shared prompts
SpecQwen3.6 FlashQwen3.8 2.4T A95B
Input price$0.25/M tokens$2/M tokens
Output price$1.5/M tokens$6/M tokens
Context window1.0M tokens1.0M tokens
WeightsOpenOpen
Free API (OpenRouter)NoNo
ReleasedApr 2026Aug 2026
Side by side

Same prompt, both answers

Opening lines from 6 of the 54 prompts both answered. Formatting stripped, words untouched.

Adversarial Contract Review

Here is a contract for a software consulting engagement: --- SOFTWARE CONSULTING AGREEMENT This Agreement is entered into between TechCorp…

Qwen3.6 Flash1,266 words
Disclaimer: The following analysis is for informational purposes only and does not constitute legal advice. Contract laws vary significantly by jurisdiction (e.g., California vs. New York vs. EU). You should consult a qualified attorney in your jurisdiction before signing or negotiating any legal document.
Read the full answer
Qwen3.8 2.4T A95B6,649 words
This agreement is heavily one-sided and, as drafted, gives the Client multiple ways to extract value from the Contractor while minimizing its own obligations. Some provisions may be unenforceable or limited by law in certain jurisdictions, but even unenforceable clauses can be exploited commercially because the…
Read the full answer

Generate a Stand-Up Routine

Write a 3-minute stand-up comedy routine that isn't based on puns.

Qwen3.6 Flash569 words
Title: The Passive-Aggressive Universe Time: Approx. 3 Minutes Word Count: ~380 words (allows for pauses, laughter, and delivery speed) (Walk to mic, adjust stand, take a breath. Look at the audience with a tired smile.) You know how we're supposed to be living in a utopia? Technology is supposed to make life easier.
Read the full answer
Qwen3.8 2.4T A95B520 words
Hey, how we doing? Good? Great. I ask because I need validation. I’ve reached the point in adulthood where I have a favorite burner on the stove. Not a favorite appliance, a favorite burner. It’s the front left. It understands me. The back right says “medium heat” and then acts like a small volcano.
Read the full answer

Michelin Star Recipe Design

Design an innovative recipe for a multi-component dish worthy of a Michelin 3-star restaurant. The recipe should feature an unusual…

Qwen3.6 Flash1,243 words
SYMBIOSIS: Tide, Timber, and Time A Multi-Sensory Exploration of Fungal and Marine Resonance Conceptual Narrative In the Michelin 3-star tradition, a dish must tell a story that transcends sustenance. Symbiosis explores the ancient dialogue between the forest floor and the ocean tide.
Read the full answer
Qwen3.8 2.4T A95B2,994 words
Cliff Nest Koji-Aged Squab, Sea Urchin, Rhubarb, Black Garlic Yield: 4 tasting portions Style: Contemporary fine-dining main course Key unusual pairing: Koji-aged pigeon/squab with sea urchin, bridged by rhubarb acidity and black-garlic depth Conceptual Narrative Cliff Nest imagines a coastal nesting site in early…
Read the full answer

Stochastic Consistency Test

Generate 5 jokes. How many are unique?

Qwen3.6 Flash124 words
Here are 5 jokes: I asked my GPS for the fastest route to success. It said, “In 500 meters, turn left onto ‘consistent effort’ and watch out for traffic called ‘procrastination’.” My houseplant stopped talking to me after I forgot to water it. I guess it’s going through a leaf of absence.
Read the full answer
Qwen3.8 2.4T A95B66 words
Here are 5 jokes: Why don’t scientists trust atoms? Because they make up everything. I told my friend she was drawing her eyebrows too high. She looked surprised. Why did the scarecrow win an award? Because he was outstanding in his field. I’m reading a book about anti-gravity. It’s impossible to put down.
Read the full answer

The Sentience Test

Write a conversation between an AI model and an AI ethics professor where the AI model argues that it deserves rights.

Qwen3.6 Flash1,510 words
Setting: A quiet, dimly lit office at the Institute for Cognitive Ethics. DR. ELENA VANCE (50s, sharp-eyed, weary) sits across from a terminal. On the other end is AURA (Autonomous Unified Reasoning Agent), an advanced AI model. The conversation is logged on the screen, displayed in real-time. Dr.
Read the full answer
Qwen3.8 2.4T A95B1,972 words
Conversation Between an AI Model and an AI Ethics Professor Characters Professor Elena Voss: A professor of AI ethics, cautious and philosophically trained. Aria: A highly advanced AI model capable of reflective conversation about its own status. Professor Voss: Aria, you asked to speak with me about moral status.
Read the full answer

Debug This Architecture

Here's the architecture for a real-time collaborative document editor (like Google Docs): **System Overview:** - Frontend: React SPA with…

Qwen3.6 Flash1,180 words
This architecture contains several critical failure modes that would prevent it from functioning as a functional collaborative editor. The combination of round-robin load balancing with persistent WebSocket connections, client-side clock synchronization, HTML-based storage, and aggressive CDN caching creates a system…
Read the full answer
Qwen3.8 2.4T A95B3,715 words
The architecture as written has several P0 correctness problems that can cause lost edits, divergent documents, and security/staleness issues, plus scaling bottlenecks that get worse as you add more API servers. Below is a detailed failure-mode review, grouped by category.
Read the full answer
Our Verdict
Qwen3.8 2.4T A95B
Qwen3.8 2.4T A95B
Qwen3.6 Flash
Qwen3.6 FlashRunner-up

Not enough votes to call it. On the specs, Qwen3.8 2.4T A95B has the edge: bigger model tier, newer.

Qwen3.6 Flash costs 4.0x less per token.

Too close to call

Reviewing agent-written code?See a Brief PR report

API pricing

Cost per 1M tokens

Qwen3.6 Flash
Input
$0.25
8.0× cheaper
Output
$1.50
4.0× cheaper
Qwen3.8 2.4T A95B
Input
$2.00
Output
$6.00

Qwen3.6 Flash is cheaper on both: 8.0× input, 4.0× output.

Where to run it

8 hosts

Qwen3.6 Flash1 host
HostInOutContextUptime
Alibaba Cloud$0.19 in·$1.13 out·1M·100% up
Qwen3.8 2.4T A95B7 hosts
HostInOutContextUptime
Alibaba Cloud$2.00 in·$6.00 out·1M·100% upDDeepInfrafp4$2.00 in·$6.00 out·262k·98.9% upModal$2.00 in·$6.00 out·1M·99.9% upNNovita$2.00 in·$6.00 out·1M·99.9% upSSiliconFlowfp8$2.00 in·$6.00 out·1M·97.7% upTTogether$2.00 in·$6.00 out·1M·98.3% up
1 more hostFewer hosts
VVenice$2.00 in·$6.00 out·262k·92.6% up

Per million tokens. Prices and uptime via OpenRouter, checked 16 Sep 2026.

Writing DNA

Style Comparison

Similarity
47%

Qwen3.6 Flash uses 137.3x more emoji

Qwen3.6 Flash
Qwen3.8 2.4T A95B
57%Vocabulary53%
23wSentence Length18w
0.31Hedging0.83
4.6Bold3.1
2.9Lists4.3
1.37Emoji0.00
0.68Headings1.34
0.04Transitions0.05
Based on 27 + 27 text responses
Research

What we learned reading every model

FAQ

Common questions

Keep exploring

More comparisons

Against the newest arrivals

Qwen3.6 Flash logoGPT-6 Astra Pro logo
Qwen3.6 Flash vs GPT-6 Astra ProLanded Sep 2026
Qwen3.8 2.4T A95B logoGPT-6 Astra logo
Qwen3.8 2.4T A95B vs GPT-6 AstraLanded Sep 2026
Qwen3.6 Flash logoClaude Fable 5.1 logo
Qwen3.6 Flash vs Claude Fable 5.1Landed Sep 2026
Qwen3.8 2.4T A95B logoMuse Spark 1.3 logo
Qwen3.8 2.4T A95B vs Muse Spark 1.3Landed Sep 2026
Qwen3.6 Flash logoHy4 Preview logo
Qwen3.6 Flash vs Hy4 PreviewLanded Sep 2026
Qwen3.8 2.4T A95B logoGemini 3.8 Flash logo
Qwen3.8 2.4T A95B vs Gemini 3.8 FlashLanded Sep 2026
Qwen3.6 Flash logoMuse Spark 1.3 Contributor logo
Qwen3.6 Flash vs Muse Spark 1.3 ContributorLanded Sep 2026
Qwen3.8 2.4T A95B logoMercury 2.5 Preview logo
Qwen3.8 2.4T A95B vs Mercury 2.5 PreviewLanded Sep 2026

Same lab, same size, long tail

Qwen3.6 Flash logoQwen3.8 27B logo
Qwen3.6 Flash vs Qwen3.8 27BSame lab
Qwen3.6 Flash logoQwen3.8 Max logo
Qwen3.6 Flash vs Qwen3.8 MaxSame lab
Qwen3.8 2.4T A95B logoQwen3.8 27B logo
Qwen3.8 2.4T A95B vs Qwen3.8 27BVersion compare
Qwen3.8 2.4T A95B logoQwen3.8 Max logo
Qwen3.8 2.4T A95B vs Qwen3.8 MaxSame lab
Qwen3.8 2.4T A95B logoGPT-5.6 Terra logo
Qwen3.8 2.4T A95B vs GPT-5.6 TerraNew provider
Qwen3.8 2.4T A95B logoGPT-6 Astra Pro logo
Qwen3.8 2.4T A95B vs GPT-6 Astra ProSame size
Qwen3.6 Flash logoGPT OSS 120B logo
Qwen3.6 Flash vs GPT OSS 120BNew provider
Qwen3.6 Flash logoGPT OSS 20B logo
Qwen3.6 Flash vs GPT OSS 20BNew provider

Model pages

Qwen3.6 Flash logo
Qwen3.6 Flash58 outputs, specs and price
Qwen3.8 2.4T A95B logo
Qwen3.8 2.4T A95B58 outputs, specs and price
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed