Skip to content
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Default Index
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed
  1. Home
  2. /
  3. Compare
  4. /
  5. Llama 3.1 405BvsQwen3.8 2.4T A95B
Updated Aug 2026

Llama 3.1 405BvsQwen3.8 2.4T A95B

Qwen3.8 2.4T A95B is cheaper than Llama 3.1 405B at $2/M vs $2.7/M input tokens.

Llama 3.1 405B and Qwen3.8 2.4T A95B compared across 12 shared prompts
SpecLlama 3.1 405BQwen3.8 2.4T A95B
Input price$2.7/M tokens$2/M tokens
Output price$3.1/M tokens$6/M tokens
Context window128K tokens1.0M tokens
WeightsOpenOpen
Free API (OpenRouter)NoNo
ReleasedJul 2024Aug 2026
Our Verdict
Llama 3.1 405B
Llama 3.1 405B
Qwen3.8 2.4T A95B
Qwen3.8 2.4T A95B

Not enough votes to call it. On the specs, nothing separates them.

Too close to call

Reviewing agent-written code?See a Brief PR report

API pricing

Cost per 1M tokens

Llama 3.1 405B
Input
$2.70
Output
$3.10
1.9× cheaper
Qwen3.8 2.4T A95B
Input
$2.00
1.4× cheaper
Output
$6.00

Qwen3.8 2.4T A95B wins input (1.4× cheaper)·Llama 3.1 405B wins output (1.9× cheaper)

Where to run it

7 hosts

Llama 3.1 405B

No hosts listed on OpenRouter.

Qwen3.8 2.4T A95B7 hosts
HostInOutContextUptime
Alibaba Cloud$2.00 in·$6.00 out·1M·99.9% upDDeepInfrafp4$2.00 in·$6.00 out·262k·100% upModal$2.00 in·$6.00 out·1M·99.7% upNNovita$2.00 in·$6.00 out·1M·100% upSSiliconFlowfp8$2.00 in·$6.00 out·1M·100% upTTogether$2.00 in·$6.00 out·1M·100% up
1 more hostFewer hosts
VVenice$2.00 in·$6.00 out·262k·99.7% up

Per million tokens. Prices and uptime via OpenRouter, checked 27 Sep 2026.

Writing DNA

Style Comparison

Similarity
45%

Qwen3.8 2.4T A95B uses 312.2x more bold

Llama 3.1 405B
Qwen3.8 2.4T A95B
55%Vocabulary53%
15wSentence Length18w
0.41Hedging0.83
0.0Bold3.1
1.9Lists4.3
0.03Emoji0.00
0.00Headings1.34
0.16Transitions0.05
Based on 5 + 27 text responses
Research

What we learned reading every model

FAQ

Common questions

Llama 3.1 405B is developed by Meta AI while Qwen3.8 2.4T A95B is developed by Qwen. Llama 3.1 405B has a 128K token context window vs Qwen3.8 2.4T A95B's 1.0M. You can compare their actual outputs across 12 challenges on Rival to see how they differ in practice.

It depends on your use case. Llama 3.1 405B and Qwen3.8 2.4T A95B each have strengths in different areas. Rival lets you compare their real outputs side-by-side across 12 challenges so you can judge which fits your needs best.

Llama 3.1 405B costs $2.7/M input tokens and Qwen3.8 2.4T A95B costs $2/M input tokens. Qwen3.8 2.4T A95B is $0.70/M cheaper per input. Check their side-by-side outputs on Rival to see if the price difference is justified by quality.

This page shows a side-by-side comparison of Llama 3.1 405B and Qwen3.8 2.4T A95B across shared challenges. You can vote on which model produced the better output in a blind duel. Browsing and voting are free. No account is needed to look; signing in only saves your votes and likes.

Keep exploring

More comparisons

Against the newest arrivals

Llama 3.1 405B logoSolar Mini 4 logo
Llama 3.1 405B vs Solar Mini 4Landed Sep 2026
Qwen3.8 2.4T A95B logoQwen3.8 Max Prime logo
Qwen3.8 2.4T A95B vs Qwen3.8 Max PrimeLanded Sep 2026
Llama 3.1 405B logoGLM 5.3 Prime logo
Llama 3.1 405B vs GLM 5.3 PrimeLanded Sep 2026
Qwen3.8 2.4T A95B logoQwen3.8 Omni Flash logo
Qwen3.8 2.4T A95B vs Qwen3.8 Omni FlashLanded Sep 2026
Llama 3.1 405B logoCommand A+ logo
Llama 3.1 405B vs Command A+Landed Sep 2026
Qwen3.8 2.4T A95B logoClaude Opus 5.5 logo
Qwen3.8 2.4T A95B vs Claude Opus 5.5Landed Sep 2026
Llama 3.1 405B logoGPT-6 Luna Pro logo
Llama 3.1 405B vs GPT-6 Luna ProLanded Sep 2026
Qwen3.8 2.4T A95B logoGPT-6 Sol Pro logo
Qwen3.8 2.4T A95B vs GPT-6 Sol ProLanded Sep 2026

Same lab, same size, long tail

Llama 3.1 405B logoMuse Glimmer 30B logo
Llama 3.1 405B vs Muse Glimmer 30BSame lab
Llama 3.1 405B logoMuse Spark 1.3 logo
Llama 3.1 405B vs Muse Spark 1.3Same lab
Qwen3.8 2.4T A95B logoQwen3.8 27B logo
Qwen3.8 2.4T A95B vs Qwen3.8 27BVersion compare
Qwen3.8 2.4T A95B logoQwen3.7 Flash logo
Qwen3.8 2.4T A95B vs Qwen3.7 FlashSame lab
Qwen3.8 2.4T A95B logoClaude 3.7 Thinking Sonnet logo
Qwen3.8 2.4T A95B vs Claude 3.7 Thinking SonnetNew provider
Qwen3.8 2.4T A95B logoClaude Sonnet 4.5 logo
Qwen3.8 2.4T A95B vs Claude Sonnet 4.5New provider
Llama 3.1 405B logoClaude Fable 5 logo
Llama 3.1 405B vs Claude Fable 5Same size
Llama 3.1 405B logoClaude Fable 5.1 logo
Llama 3.1 405B vs Claude Fable 5.1Same size

Model pages

Llama 3.1 405B logo
Llama 3.1 405B12 outputs, specs and price
Qwen3.8 2.4T A95B logo
Qwen3.8 2.4T A95B58 outputs, specs and price
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Default Index
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed