Skip to content
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Default Index
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed
  1. Home
  2. /
  3. Compare
  4. /
  5. Llama 3.1 405BvsMercury 2.5
Updated Sep 2026

Llama 3.1 405BvsMercury 2.5

Mercury 2.5 is cheaper than Llama 3.1 405B at $0.04/M vs $2.7/M input tokens.

Llama 3.1 405B and Mercury 2.5 compared across 5 shared prompts
SpecLlama 3.1 405BMercury 2.5
Input price$2.7/M tokens$0.04/M tokens
Output price$3.1/M tokens$0.15/M tokens
Context window128K tokens260K tokens
WeightsOpen—
Free API (OpenRouter)NoNo
ReleasedJul 2024Sep 2026
Our Verdict
Llama 3.1 405B
Llama 3.1 405B
Mercury 2.5
Mercury 2.5

Not enough votes to call it. On the specs, nothing separates them.

Mercury 2.5 costs 21x less per token.

Too close to call
API pricing

Cost per 1M tokens

Llama 3.1 405B
Input
$2.70
Output
$3.10
Mercury 2.5
Input
$0.04
68× cheaper
Output
$0.15
21× cheaper

Mercury 2.5 is cheaper on both: 68× input, 21× output.

Where to run it

1 host

Llama 3.1 405B

No hosts listed on OpenRouter.

Mercury 2.51 host
HostInOutContextUptime
Inception$0.04 in·$0.15 out·260k·100% up

Per million tokens. Prices and uptime via OpenRouter, checked 23 Sep 2026.

Research

What we learned reading every model

FAQ

Common questions

Llama 3.1 405B is developed by Meta AI while Mercury 2.5 is developed by Inception. Llama 3.1 405B has a 128K token context window vs Mercury 2.5's 260K. You can compare their actual outputs across 5 challenges on Rival to see how they differ in practice.

It depends on your use case. Llama 3.1 405B and Mercury 2.5 each have strengths in different areas. Rival lets you compare their real outputs side-by-side across 5 challenges so you can judge which fits your needs best.

Llama 3.1 405B costs $2.7/M input tokens and Mercury 2.5 costs $0.04/M input tokens. Mercury 2.5 is $2.66/M cheaper per input. Check their side-by-side outputs on Rival to see if the price difference is justified by quality.

This page shows a side-by-side comparison of Llama 3.1 405B and Mercury 2.5 across shared challenges. You can vote on which model produced the better output in a blind duel. Browsing and voting are free. No account is needed to look; signing in only saves your votes and likes.

Keep exploring

More comparisons

Against the newest arrivals

Llama 3.1 405B logoDeepSeek V4 Flash Vision Exp logo
Llama 3.1 405B vs DeepSeek V4 Flash Vision ExpLanded Sep 2026
Mercury 2.5 logoSolar Pro 4 logo
Mercury 2.5 vs Solar Pro 4Landed Sep 2026
Llama 3.1 405B logoHy3 logo
Llama 3.1 405B vs Hy3Landed Sep 2026
Mercury 2.5 logoQwen3.7 Flash logo
Mercury 2.5 vs Qwen3.7 FlashLanded Sep 2026
Llama 3.1 405B logoLing 3.0 Flash logo
Llama 3.1 405B vs Ling 3.0 FlashLanded Sep 2026
Mercury 2.5 logoMuse Glimmer 30B logo
Mercury 2.5 vs Muse Glimmer 30BLanded Sep 2026
Llama 3.1 405B logoGLM 5.3 logo
Llama 3.1 405B vs GLM 5.3Landed Sep 2026
Mercury 2.5 logoTernary Bonsai 2 27B logo
Mercury 2.5 vs Ternary Bonsai 2 27BLanded Sep 2026

Same lab, same size, long tail

Llama 3.1 405B logoMuse Glimmer 30B logo
Llama 3.1 405B vs Muse Glimmer 30BSame lab
Llama 3.1 405B logoMuse Spark 1.3 logo
Llama 3.1 405B vs Muse Spark 1.3Same lab
Mercury 2.5 logoMercury 2.5 Preview logo
Mercury 2.5 vs Mercury 2.5 PreviewSame lab
Mercury 2.5 logoMercury logo
Mercury 2.5 vs MercurySame lab
Mercury 2.5 logoGPT-6 Astra logo
Mercury 2.5 vs GPT-6 AstraNew provider
Mercury 2.5 logoGPT-6 Astra Pro logo
Mercury 2.5 vs GPT-6 Astra ProNew provider
Mercury 2.5 logoGPT OSS 120B logo
Mercury 2.5 vs GPT OSS 120BNew provider
Mercury 2.5 logoGPT OSS 20B logo
Mercury 2.5 vs GPT OSS 20BSame size

Model pages

Llama 3.1 405B logo
Llama 3.1 405B12 outputs, specs and price
Mercury 2.5 logo
Mercury 2.520 outputs, specs and price
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Default Index
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed