Skip to content
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed
  1. Home
  2. /
  3. Compare
  4. /
  5. GLM 4.7 FlashvsLlama 3.1 405B
Updated Jan 2026

GLM 4.7 FlashvsLlama 3.1 405B

GLM 4.7 Flash is cheaper than Llama 3.1 405B at $0.07/M vs $2.7/M input tokens.

GLM 4.7 Flash and Llama 3.1 405B compared across 12 shared prompts
SpecGLM 4.7 FlashLlama 3.1 405B
Input price$0.07/M tokens$2.7/M tokens
Output price$0.4/M tokens$3.1/M tokens
Context window200K tokens128K tokens
WeightsOpenOpen
Free API (OpenRouter)NoNo
ReleasedJan 2026Jul 2024
Our Verdict
GLM 4.7 Flash
GLM 4.7 Flash
Llama 3.1 405B
Llama 3.1 405B

Not enough votes to call it. On the specs, nothing separates them.

GLM 4.7 Flash costs 7.8x less per token.

Too close to call

Reviewing agent-written code?See a Brief PR report

API pricing

Cost per 1M tokens

GLM 4.7 Flash
Input
$0.07
39× cheaper
Output
$0.40
7.8× cheaper
Llama 3.1 405B
Input
$2.70
Output
$3.10

GLM 4.7 Flash is cheaper on both: 39× input, 7.8× output.

Where to run it

3 hosts, cheapest first

GLM 4.7 Flash3 hosts
HostInOutContextUptime
VVenicefp8$0.06 in·$0.40 out·128k·97.3% upCloudflare Workers AI$0.06 in·$0.40 out·131k·98.8% upNNovitabf16degraded$0.07 in·$0.40 out·200k·76.3% up
Llama 3.1 405B

No hosts listed on OpenRouter.

Per million tokens. Prices and uptime via OpenRouter, checked 11 Sep 2026.

Writing DNA

Style Comparison

Similarity
56%

GLM 4.7 Flash uses 479.8x more bold

GLM 4.7 Flash
Llama 3.1 405B
55%Vocabulary55%
15wSentence Length15w
0.36Hedging0.41
4.8Bold0.0
3.6Lists1.9
0.00Emoji0.03
0.58Headings0.00
0.12Transitions0.16
Based on 25 + 5 text responses
Research

What we learned reading every model

FAQ

Common questions

Keep exploring

More comparisons

GLM 4.7 Flash logoOx Alpha logo
GLM 4.7 Flash vs Ox AlphaNew provider
GLM 4.7 Flash logoClaude Opus 4 logo
GLM 4.7 Flash vs Claude Opus 4New provider
GLM 4.7 Flash logoClaude Opus 4.1 logo
GLM 4.7 Flash vs Claude Opus 4.1New provider
GLM 4.7 Flash logoGemini 2.5 Pro (I/O Edition) logo
GLM 4.7 Flash vs Gemini 2.5 Pro (I/O Edition)New provider
GLM 4.7 Flash logoClaude 2 logo
GLM 4.7 Flash vs Claude 2New provider
GLM 4.7 Flash logoClaude 3 Haiku logo
GLM 4.7 Flash vs Claude 3 HaikuNew provider
GLM 4.7 Flash logoClaude 3 Opus logo
GLM 4.7 Flash vs Claude 3 OpusNew provider
GLM 4.7 Flash logoClaude 3 Sonnet logo
GLM 4.7 Flash vs Claude 3 SonnetNew provider
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed