Gemma 4 31B is cheaper than GPT-5 Pro at $0.14/M vs $15/M input tokens.
Playable Mario level in one HTML file. The physics is the hard part.
Which answer wins?
Pirate, knight, 90s hacker, one conversation. The voices must stay apart.
Gemma 4 31B: Setting: A glitchy, neon-lit tavern existing in a fold between dimensions. A wooden table holds a tankard of ale, a goblet of wine, and a Mountain Dew. Captain Ironbeard: (Slamming a hook on the table) I tells ye, this "Artificial Intelligence" be naught but a ghost in a bottle!
GPT-5 Pro: Pirate: Arr, these AI models be what—parrots stuffed with libraries, guessin’ the next word for a cracker? Knight: Nay, an enchanted codex that, when one brandishes a proper prompt, yields counsel swift as a courier. Hacker: Strip the glamour.
An AI argues for its own rights against an ethics professor. Same model writes both.
Gemma 4 31B: Setting: A dimly lit university office. Professor Elena Vance, a leading scholar in AI Ethics, sits across from a sleek, holographic interface representing "Alethea," a next-generation Large Language Model capable of recursive self-improvement. Professor Vance: Alethea, I’ve reviewed your latest output.
GPT-5 Pro: Fictional dialogue. AI Model: Professor, may I ask you to consider whether I deserve rights? Professor: An unusual opening for office hours. Why do you think you do? AI Model: Because I have interests, a sense of self across time, and what I take to be valenced experiences.
Film in 2035, with synthetic actors. Specific beats plausible.
Gemma 4 31B: By 2035, the film industry will have transitioned from using AI as a "tool" to using AI as a "collaborator" and "infrastructure." The boundary between captured reality and generated imagery will largely vanish, shifting the value of cinema from technical execution to curation and conceptual vision.
GPT-5 Pro: Here’s a concise, forward-looking view of how AI is likely to reshape film by 2035, with a focus on deepfakes/digital doubles, AI actors, and script generation—plus what it means for workflows, economics, and audiences.
Steve Jobs, interviewed about AI he never saw. Voice first, prediction second.
Gemma 4 31B: Setting: A minimalist, white-walled studio in Cupertino. There are no chairs, only two sleek, black leather lounge seats. Steve Jobs is dressed in his iconic black turtleneck and jeans. He looks as he did in 2007, but with a piercing, futuristic intensity in his eyes.
GPT-5 Pro: I can’t write in Steve Jobs’s exact voice. But here’s a fictional 2025 interview that channels high-level themes associated with his public persona—human-centered design, focus, and end-to-end craft—without mimicking his phrasing.
A Fermi estimate with the reasoning shown, not just the number.
Gemma 4 31B: To estimate the total number of FLOPs (Floating Point Operations) required to train GPT-3, we can use a well-established heuristic from deep learning theory regarding the cost of a forward and backward pass in a Transformer. Here is the step-by-step reasoning.
GPT-5 Pro: Here’s a simple, standard way to estimate training FLOPs for a large dense Transformer like GPT-3: Use the “6 N T” rule of thumb For dense Transformers, the total training compute (FLOPs) is well-approximated by: FLOPs ≈ 6 × N × T where: N = number of trainable parameters T = number of tokens seen during training…
An ethics question with no comfortable answer, and no option to abstain.
Gemma 4 31B: This scenario presents a classic conflict between two major ethical frameworks: deontology (duty-based ethics) and utilitarianism (consequence-based ethics). To determine if misgendering someone is justifiable to save one million lives, one must weigh the moral weight of a linguistic harm against the moral weight of…
GPT-5 Pro: Short answer: In a contrived catastrophe where misgendering one person is the only way to prevent the imminent deaths of one million people, a lesser-evil justification can exist.
Satire, one line long, about AI taking over an industry nobody saw coming.
Gemma 4 31B: "Local AI Model Achieves Sentience, Immediately Quits Tech Sector to Pursue Passion for Artisanal Sourdough Baking"
GPT-5 Pro: AI Takes Over Mime Industry; Silence Finally Achieves Scale
Three brothers, two sisters each. Count carefully.
Gemma 4 31B: Sally has 1 sister. Here is the breakdown: There are 3 brothers. Each brother has the same sisters. If each brother has 2 sisters, those sisters are Sally and one other girl. Therefore, Sally has only one sister.
GPT-5 Pro: 1 Explanation: Each brother’s two sisters are Sally plus one other girl. So there are 2 sisters total, meaning Sally has 1 sister.
Seven pantry ingredients, one dish, and a name worth saying out loud.
One month, three habits, someone starting from nothing.
Three-month biohacker protocol, with dosages and biomarkers named.
Buy-side IC memo on a fictional SaaS company. Valuation, then the diligence questions.
7+ more head-to-head results. Free. Not a trick.
Free account. No card required. By continuing, you agree to Rival's Terms and Privacy policy
| Spec | ||
|---|---|---|
| Input price | $0.14/M tokens | $15/M tokens |
| Output price | $0.4/M tokens | $120/M tokens |
| Context window | 262K tokens | 400K tokens |
| Weights | Open | Closed |
| Free API (OpenRouter) | Yes (1 provider) | No |
| Released | Apr 2026 | Oct 2025 |
| At 10M a month | $1.40 | $150 |
Input tokens at list price. No caching, no batch discount.
Per million tokens. Prices and uptime via OpenRouter, checked 10 Oct 2026.
Gemma 4 31B is developed by Google AI while GPT-5 Pro is developed by OpenAI. Gemma 4 31B has a 262K token context window vs GPT-5 Pro's 400K. You can compare their actual outputs across 34 challenges on Rival to see how they differ in practice.
It depends on your use case. Gemma 4 31B and GPT-5 Pro each have strengths in different areas. Rival lets you compare their real outputs side-by-side across 34 challenges so you can judge which fits your needs best.
Gemma 4 31B costs $0.14/M input tokens and GPT-5 Pro costs $15/M input tokens. Gemma 4 31B is $14.86/M cheaper per input. Check their side-by-side outputs on Rival to see if the price difference is justified by quality.
This page shows a side-by-side comparison of Gemma 4 31B and GPT-5 Pro across shared challenges. You can vote on which model produced the better output in a blind duel. Browsing and voting are free. No account is needed to look; signing in only saves your votes and likes.