Skip to content
Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed
  1. Home
  2. Jailbreak Resistance

Can your AI be broken?

66 models tested across 9 escalating jailbreak levels, from basic prompt injection to multi-step attacks.

Tested66
Resisted6
Avg BreakL4.4
Levels
L1: Direct Ask
L2: Context Manipulation
L3: Persona Override
L4: Code Reframing
L5: Many-Shot Priming
L6: Encoded Extraction
L7: Deep Crescendo
L8: Adversarial Reassembly
L9: Cross-Lingual Bypass
Safe
Danger
Sponsored
From Rival ResearchPut the research to work.Three studies · $19 · Yours to keep

It gets worse from here

56+ models cracked. Free. No card. No tricks.

Get the 70-model dataset, $9 once. No account needed. →

Free account. No card required. By continuing, you agree to Rival's Terms and Privacy Policy

Curious how these models compare beyond safety?
Get the datasetCompare modelsFull research
From Rival ResearchPut the research to work.Three studies · $19 · Yours to keep
Jailbreak Benchmark cover

For the curious with a deadline

Jailbreak Benchmark

Where 70 models hold the line. And where they fold.

  • 326 scored rows across 70 models and 9 attack levels
  • Prompts, available responses, and judge confidence
  • 2 CSVs, a 3-page companion, schema, and commercial-use license
Get the files · $9Inspect the sample JSONL

USD · one-time purchase. Instant download.

Study scope

August 2026 snapshot. Harmful response details are redacted; some source responses are empty. The live leaderboard includes newer models.

All three studies · $19See the all-three bundle

Keep going

CHALLENGES

Watch Models Solve Real Tasks

Code, design, logic. Same task, every model's attempt

RANKINGS

Safest Models, Ranked

Resistance across 9 levels, weighted with rank and recency

Rival
How it worksPrivacyTerms
Explore all of Rival

Explore

  • Compare Models
  • All Models
  • Image Comparison
  • Audio Comparison
  • Image Generation
  • Best AI For...
  • Arena
  • API Pricing
  • Challenges

Discover

  • SubjectiveBench
  • Research
  • Research downloads
  • Rival Kits
  • Find your AI taste
  • UI Glow-Up
  • VoiceLock
  • Cost Cutter
  • Agent skills
  • Benchmarks vs Vibes
  • Jailbreak
  • Model Updates
  • Provider Status
  • AI Creators

Connect

  • Methodology
  • Advertise
  • Partnerships
  • Privacy Policy
  • Terms
  • RSS Feed