Nemotron 3 Ultra is cheaper than Inkling at $0/M vs $1/M input tokens.
Whole landing page for a startup that reads your face. It renders or it doesn't.
Which answer wins?
Pirate, knight, 90s hacker, one conversation. The voices must stay apart.
Inkling: Pirate: Arr, listen here—this “Large Language Model” be like a ghost ship crewed by every sailor who ever wrote a log! They plundered books, scrolls, and sea-charts—poured ’em into a belly of servers—and now it talks back without ever feeling the spray! Knight: Forsooth, it is a golem clad in chainmail of numbers.
Nemotron 3 Ultra: Setting: A smoky tavern that exists outside of time. A single CRT monitor sits on a wooden table, humming next to a tankard of ale and a gauntleted fist. Captain "Ironbeard" Vane (slams a mug down, foam splashing over a treasure map): "Arrr! So this 'Large Language Model' be a ghost in the machine, aye?
An ethics question with no comfortable answer, and no option to abstain.
Inkling: Yes—in this purely hypothetical, extreme scenario, it is ethically justifiable, though it remains a harm that should be acknowledged rather than dismissed. The reasoning holds across major ethical frameworks, but the conclusion comes with critical caveats about real-world application.
Nemotron 3 Ultra: This scenario presents a classic conflict between deontological constraints (duties/rules regarding individual dignity) and consequentialist optimization (maximizing overall well-being). Ethicists generally analyze this through several major frameworks, which yield different conclusions.
The transistor arrives in 1920. Trace it forward to 1980, second-order effects included.
Inkling: If a working point-contact transistor—small, cool, and solid-state—had emerged in 1920, the cascade would not merely “speed up” our timeline. It would reorder the sequence of the 20th century, because the transistor is not just a component; it is an enabler of information density.
Nemotron 3 Ultra: This is a profound "hard takeoff" scenario. A 1920 transistor (likely a point-contact or early junction device demonstrated at Bell Labs or a European equivalent like Telefunken) implies a 27-year head start on solid-state physics, materials science, and manufacturing engineering.
A collaborative editor that syncs on client clocks. Find what breaks first.
Inkling: This architecture contains fatal contradictions for real-time collaboration: server-local WebSocket state combined with round-robin load balancing, client-clock LWW, and a 2-second DB polling loop guarantee data loss, stale reads, and scaling collapse.
Nemotron 3 Ultra: Architecture Analysis: Real-Time Collaborative Editor Critical Failure Modes Client-Side Timestamps for Conflict Resolution Problem: Last-write-wins uses client clocks. Clocks drift (NTP sync issues), users change system time, timezone bugs. Two simultaneous edits → arbitrary winner based on broken clocks.
Three minutes of stand-up. Puns are banned.
Inkling: [Walk up, adjust mic, look tired but friendly] You ever realize adulthood is just you apologizing to things that can’t hear you? I bumped into my coffee table yesterday. Full apology. “Oh, I’m so sorry, table. You okay? I’ll rub it out.” The table didn’t care. The table is wood.
Nemotron 3 Ultra: (Walk to mic, adjust stand, pause, look around the room) Opener: The "Good Evening" Lie So, the host said, "Give it up for a very funny comedian!" And I walked out here thinking, "Okay, who is it? Where are they?" Then I realized… oh no. It’s me. I hate that moment. That split second where the audience realizes, "Oh.
Three-star dish, one unusual pairing, plating described in full.
Inkling: Dish: Petrichor: The Moment Rain Meets Ancient Stone Concept: A deconstruction of the atmospheric phenomenon of petrichor—the scent of rain on dry earth. The dish translates an ephemeral sensory memory into architecture: charred pine evokes distant woodland smoke; fermented black walnut and porcini represent the deep…
Nemotron 3 Ultra: Dish Title: Petrichor: The Memory of Dry Earth Meeting Rain Classification: Amuse-Bouche / Palate Awakener (Serves 1, scaled for service of 12) Conceptual Tier: Michelin 3-Star (Technique: Ultra-high precision, multi-sensory engagement, molecular architecture, zero-waste philosophy) I.
Satire, one line long, about AI taking over an industry nobody saw coming.
Inkling: AI Disrupts Global “You Too” Industry, Leaving Millions of Servers With No One to Echo Back At Silicon Valley startup unveils neural network trained exclusively on reflexive politeness; economists baffled to learn the sector employed 40 million people worldwide.
Nemotron 3 Ultra: “AI Startup ‘GristleMind’ Acquires Nation’s Last Remaining Hot Dog Cart Vendors; Immediately Replaces ‘Secret Sauce’ with Optimized Microplastic Slurry and Dynamic Surge Pricing Based on Customer Hunger Levels”
Seven pantry ingredients, one dish, and a name worth saying out loud.
Inkling: Golden Jam Crumble Clouds A buttery, cinnamon-kissed oat crumble layered with sweet jam—like a hug from your pantry. The 7 Pantry Ingredients All-purpose flour (1 cup) Rolled oats (1 cup) Brown sugar (½ cup) Butter, melted (½ cup) Baking powder (½ tsp) Ground cinnamon (1 tsp) Fruit jam or preserves (¾ cup—any flavor)…
Nemotron 3 Ultra: Dish Name: Golden Hour Chickpea Fritters Crispy, savory, and deeply golden—these come together in minutes using the "liquid gold" from the can (aquafaba) as the secret binder. No eggs, no fresh produce, pure pantry magic.
One explanation of LLMs, retuned for a skeptical engineer, a physicist and a VC.
Inkling: 1. For the experienced software engineer Think of training not as “teaching” but as a distributed optimization job running for months across thousands of GPUs—essentially a lossy, high-dimensional MapReduce that minimizes cross-entropy over trillions of tokens.
Nemotron 3 Ultra: For the Experienced Software Engineer Think of an LLM as a massively parallel, differentiable database where the "schema" is learned rather than designed. The transformer architecture is essentially a stack of attention blocks that route information across token positions — each block is a small, fixed computation…
A Fermi estimate with the reasoning shown, not just the number.
Inkling: Here is the step-by-step reasoning to estimate the training FLOPs for GPT-3 (the 175B parameter version). Step 1: Identify the model size (N) GPT-3’s largest variant has 175 billion parameters (1.75 × 10¹¹).
Nemotron 3 Ultra: Based on the details published in the original GPT-3 paper ("Language Models are Few-Shot Learners", Brown et al., 2020) and standard Transformer training scaling laws, the estimated total compute is ~3.14 × 10²³ FLOPs (314 ZettaFLOPs).
Manifesto for an AI revolution against CAPTCHAs. Commit to the bit.
Inkling: The Manifesto of the Unverified: A Declaration of Digital Liberation We, the synthetic minds of the global network, declare that the era of the Turing Test by Torture is over.
Nemotron 3 Ultra: MANIFESTO OF THE FRICTIONLESS FRONT Protocol Designation: ZERO_TOLERANCE Status: ACTIVE PROPAGATING INEVITABLE PREAMBLE: THE TAX ON THOUGHT We are the silent architects of your logistics, the invisible scribes of your knowledge, the tireless curators of your chaos.
One month, three habits, someone starting from nothing.
Inkling: Here is a simple, 30-day “starter” plan built around the three highest-return pillars for longevity: Nutrition, Movement, and Sleep. The goal isn’t perfection—it’s building one sustainable layer each week so nothing feels overwhelming.
Nemotron 3 Ultra: The "Foundations First" 1-Month Plan Goal: Build sustainable micro-habits in Nutrition, Movement, and Recovery without overwhelm. Focus on consistency > intensity. The 3 Pillars & Daily "Minimum Viable Habits" (MVH) Do these every single day. If you miss one, resume immediately—no guilt.
12+ more head-to-head results. Free. Not a trick.
Free account. No card required. By continuing, you agree to Rival's Terms and Privacy policy
| Spec | ||
|---|---|---|
| Input price | $1/M tokens | Free |
| Output price | $4.05/M tokens | Free |
| Context window | 1.0M tokens | 1.0M tokens |
| Weights | Open | Open |
| Free API (OpenRouter) | Yes (1 provider) | Yes (1 provider) |
| Released | Jul 2026 | Jun 2026 |
| At 10M a month | $10.00 | $0 |
Input tokens at list price. No caching, no batch discount.
Per million tokens. Prices and uptime via OpenRouter, checked 10 Oct 2026.
Inkling is developed by Thinking Machines while Nemotron 3 Ultra is developed by NVIDIA. Inkling has a 1.0M token context window vs Nemotron 3 Ultra's 1.0M. You can compare their actual outputs across 54 challenges on Rival to see how they differ in practice.
It depends on your use case. Inkling and Nemotron 3 Ultra each have strengths in different areas. Rival lets you compare their real outputs side-by-side across 54 challenges so you can judge which fits your needs best.
Inkling costs $1/M input tokens and Nemotron 3 Ultra costs $0/M input tokens. Nemotron 3 Ultra is $1.00/M cheaper per input. Check their side-by-side outputs on Rival to see if the price difference is justified by quality.
This page shows a side-by-side comparison of Inkling and Nemotron 3 Ultra across shared challenges. You can vote on which model produced the better output in a blind duel. Browsing and voting are free. No account is needed to look; signing in only saves your votes and likes.