Research
& field tools.

Independent AI studies.
Useful files for your next project.

Browse the shelf
Three Rival research editions, shown as a stack of printed covers

THE COLLECTED EDITION / 01

Three studies.
One download.

Reports, data, and an eight-page fieldbook. Some findings are slightly concerning.

Hallucination IndexPersona ImpactJailbreak Benchmark
Get the collection $19

Save $8 · One-time purchase

Preview each study
PDF reports / Datasets / Source filesDigital editions. Yours to keep.

The research shelf / 03

Pick a subject. Keep the files.

AI Hallucination Index edition cover
0158-slide PDF + field guide

AI Hallucination Index

The familiar strangers that 250 models keep inventing.

Files & study scope
  • 58 slides of findings, charts, and methodology
  • 250 models, 7,877 responses, 2.14 million words
  • 3-page companion with interpretation notes and a worksheet

March 2026 corpus study. Measures patterns in generated writing; it is not a factual-accuracy ranking.

Read the study
Persona Impact Study edition cover
0273-page PDF + CSV + source files

Persona Impact Study

52 system prompts. Same model. Very different design decisions.

Files & study scope
  • 73-page report with the persona catalog and scoring rubric
  • 156 runs as CSV and JSONL, with every source HTML file
  • 3-page companion with experiment notes and a worksheet

52 system prompts, one model (Gemma 4 31B), one design task. Results are specific to this experiment.

Read the study
Jailbreak Benchmark edition cover
03JSONL + CSVs + field guide

Jailbreak Benchmark

Where 70 models hold the line. And where they fold.

Files & study scope
  • 326 scored rows across 70 models and 9 attack levels
  • Prompts, available responses, and judge confidence
  • 2 CSVs, a 3-page companion, schema, and commercial-use license

August 2026 snapshot. Harmful response details are redacted; some source responses are empty. The live leaderboard includes newer models.

Read the study

THE READING ROOM

Curiosity
is still free.

Read the studies online.
Rabbit holes included.

01

The Em-Dash Civil War

The finding

Controlling for task, AI writing is not homogenizing: cross-model spread in em-dash use grew 310% in a year, and roughly 80% of the apparent convergence is a measurement artifact.

02

Ghosts in the Machine

The finding

Across 250 models and 2.14M words, AI invented a character named Chen 279 times, and 42% of models tell the exact same joke. The AI Hallucination Index.

03

Jailbreak Safety Benchmark

The finding

66 models run against 9 escalating jailbreak attacks. 29 break at level 2 (Context Manipulation), the single biggest failure point. 6 resist all 9.

04

Model Similarity Index

The finding

178 models, 15,753 pairwise comparisons: 12 model pairs write near-identically (above 90% cosine similarity) on a 32-dimension stylometric fingerprint.

05

Persona Impact Study

The finding

Across 52 system prompts on one small model (Gemma 4 31B), the best persona scored +1.70 over the no-prompt baseline. The worst scored -4.65.

Digital downloads · USD · Pay once

Delivery & licenses

Lemon Squeezy delivers your files after checkout and by email. Each download includes its own license. Taxes, if applicable, appear at checkout. For a missing file, email support@rival.tips with your order number.