Skip to content
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed
  1. Home
  2. Best For
  3. Mathematics

Best AI for Mathematics

Sally has three brothers. Each brother has two sisters. That trap, plus a from-memory estimate of what GPT-3 cost to train.

Updated Jun 2026·5 challenges·20 models

How Mathematics rankings are computed

20 models tested across 5 mathematics challenges.Composite score: 30% Rival Index, 20% task coverage, 20% challenge-scoped duel performance, 15% recency, 15% tier. Deduplicated by product line. Kimi K3 leads at 81.8/100. Drawn from Rival's open dataset of 21,000+ human preference votes.

Rival's Pick·#12 Rival Index·Anthropic flagship

Too close to call
Claude Fable 5
Claude Fable 5anthropic

Neck and neck with Kimi K3. Claude Fable 5 gets the nod on blind votes.

Composite scores combine task evidence, Rival Index, recency, and model tier. Rival’s Pick is a separate editorial recommendation. How ranking works

Claude Fable 5
Claude Fable 5
anthropic
$10.00·$50.00
82Composite
Kimi K3
Kimi K3
moonshotai
$3.00·$15.00
82Composite
Qwen3.7 Max
Qwen3.7 Max
qwen
$2.50·$7.50
81Composite

Head-to-Head

Kimi K3 logo
Kimi K3
vs
Claude Fable 5
Claude Fable 5 logo
Kimi K3 logo
Kimi K3
vs
Qwen3.7 Max
Qwen3.7 Max logo
Claude Fable 5 logo
Claude Fable 5
vs
Qwen3.7 Max
Qwen3.7 Max logo

Full Rankings

20 models
#
Model
Coverage
Index
Price
Composite
4
Gemini 3.1 Pro Preview logo
Gemini 3.1 Pro Previewgoogle
5/5
#6
$2.00·$12.00
78
5
GPT-6 Astra logo
GPT-6 Astraopenai
5/5
#26
$10.00·$50.00
78
6
Gemini 3.8 Flash logo
Gemini 3.8 Flashgoogle
5/5
$0.75·$3.75
76
7
Gemini 2.5 Pro Preview 06-05 logo
Gemini 2.5 Pro Preview 06-05google
4/5
#35
$1.25·$10.00
76
8
Claude Haiku 4.5 logo
Claude Haiku 4.5anthropic
5/5
#47
$1.00·$5.00
76
9
Qwen3.6 27B logo
Qwen3.6 27Bqwen
5/5
#1
$0.33·$3.25
73
10
GPT OSS 120B logo
GPT OSS 120Bopenai
5/5
#127
$0.18·$0.80
72
Challenges5
Logic Puzzle
One answer. Showing the work is the point.
Estimate Complexity
A Fermi estimate out loud. The reasoning is graded, not the number.
Stochastic Consistency
Asked to be random, then asked why it was not.
Advanced Investment Memo (IC Memo)
A buy-side IC memo: valuation, diligence, a recommendation.
Mini LBO Underwrite
A five-year LBO with real math and a sensitivity table.
Related
Word ProblemsStatisticsComplex ReasoningData Analysis

Keep exploring

#1 VS #2

Kimi K3 vs Claude Fable 5

The top two for Mathematics, compared directly

RELATED

Best AI for Word Problems

See which models rank highest here

Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Brief
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed