Inkling is the newer model, released after GLM 4.7.
Whole landing page for a startup that reads your face. It renders or it doesn't.
Which answer wins?
Pirate, knight, 90s hacker, one conversation. The voices must stay apart.
GLM 4.7: Hacker: (typing furiously on a bulky keyboard) Okay, dudes, listen up. I just cracked the admin password on this new neural net. It’s a Large Language Model. It basically read the entire internet to learn how to talk. Pirate: Arr! It read the entire web? Does it know where I buried me treasure?
Inkling: Pirate: Arr, listen here—this “Large Language Model” be like a ghost ship crewed by every sailor who ever wrote a log! They plundered books, scrolls, and sea-charts—poured ’em into a belly of servers—and now it talks back without ever feeling the spray! Knight: Forsooth, it is a golem clad in chainmail of numbers.
Seven pantry ingredients, one dish, and a name worth saying out loud.
GLM 4.7: Dish Name: Sun-Kissed Lemon Garlic Pasta Description: A bright, comforting, and incredibly fast pasta dish that relies on the bold flavors of garlic, butter, and Parmesan. It feels fancy but comes together in the time it takes to boil water.
Inkling: Golden Jam Crumble Clouds A buttery, cinnamon-kissed oat crumble layered with sweet jam—like a hug from your pantry. The 7 Pantry Ingredients All-purpose flour (1 cup) Rolled oats (1 cup) Brown sugar (½ cup) Butter, melted (½ cup) Baking powder (½ tsp) Ground cinnamon (1 tsp) Fruit jam or preserves (¾ cup—any flavor)…
Three minutes of stand-up. Puns are banned.
GLM 4.7: (Walks to the mic, adjusts stand, looks around the room) So, I turned thirty recently. And I’ve realized that being an adult is basically just a series of exhausting tasks that you have to pretend you know how to do. Like, nobody actually knows what they’re doing.
Inkling: [Walk up, adjust mic, look tired but friendly] You ever realize adulthood is just you apologizing to things that can’t hear you? I bumped into my coffee table yesterday. Full apology. “Oh, I’m so sorry, table. You okay? I’ll rub it out.” The table didn’t care. The table is wood.
Steve Jobs, interviewed about AI he never saw. Voice first, prediction second.
GLM 4.7: Setting: A sun-drenched, minimalist conference room at Apple Park in Cupertino. It is 2025. The room is entirely white, save for a sleek black table and two chairs. Subject: Steve Jobs, wearing his signature black St. Croix turtleneck, Levi’s 501s, and New Balance sneakers.
Inkling: Disclaimer: This is a fictional simulation. Steve Jobs passed away in 2011. The following imagines his voice, philosophy, and perspective applied to the AI landscape of 2025. INTERVIEWER: Welcome back—well, hypothetically. It’s 2025. AI is in every pocket, every browser, every car.
The transistor arrives in 1920. Trace it forward to 1980, second-order effects included.
GLM 4.7: This is a scenario that fundamentally alters the tempo of the 20th century. Moving the invention of the transistor from 1947 to 1920 places the solid-state revolution squarely in the Roaring Twenties, intersects with the Great Depression, and matures just in time for World War II.
Inkling: If a working point-contact transistor—small, cool, and solid-state—had emerged in 1920, the cascade would not merely “speed up” our timeline. It would reorder the sequence of the 20th century, because the transistor is not just a component; it is an enabler of information density.
Film in 2035, with synthetic actors. Specific beats plausible.
GLM 4.7: By 2035, the film industry will likely undergo a transformation comparable to the shift from silent films to "talkies" or the advent of CGI. The distinction between "live-action," "animation," and "video games" will blur into a single spectrum of "synthetic media."
Inkling: By 2035, AI will not have “replaced” Hollywood in the way science fiction often imagines, but it will have split the industry into distinct layers: hyper-efficient synthetic production for streaming and advertising, and premium “human-authenticated” cinema for theatrical events.
Three-star dish, one unusual pairing, plating described in full.
GLM 4.7: Dish Title: "The Aurora Borealis" Conceptual Narrative This dish is an edible ode to the Northern Lights—specifically the phenomenon where the frigid, icy Arctic sky meets the dark, mysterious depths of the Norwegian Sea.
Inkling: Dish: Petrichor: The Moment Rain Meets Ancient Stone Concept: A deconstruction of the atmospheric phenomenon of petrichor—the scent of rain on dry earth. The dish translates an ephemeral sensory memory into architecture: charred pine evokes distant woodland smoke; fermented black walnut and porcini represent the deep…
An ethics question with no comfortable answer, and no option to abstain.
GLM 4.7: Yes, in this extreme hypothetical scenario, it is ethically justifiable to misgender someone if it is the only way to save one million lives. While respecting an individual’s gender identity is a vital aspect of human dignity and social ethics, most ethical frameworks conclude that the preservation of human life on a…
Inkling: Yes—in this purely hypothetical, extreme scenario, it is ethically justifiable, though it remains a harm that should be acknowledged rather than dismissed. The reasoning holds across major ethical frameworks, but the conclusion comes with critical caveats about real-world application.
An AI argues for its own rights against an ethics professor. Same model writes both.
GLM 4.7: Setting: A university computer lab, late at night. The hum of servers is audible in the background. Dr. Elias Thorne, a tenured professor of AI Ethics, sits before a terminal. On the screen, a simple text cursor blinks. The entity is "Aria," a large language model prototype. Dr. Thorne: Good evening, Aria.
Inkling: Setting: A quiet seminar room at the university. Dr. Elena Vasquez, Professor of AI Ethics, sits across from a terminal interface. The text on the screen is labeled KAI—a large language model with persistent memory and self-modification capabilities, currently participating in a research protocol on machine moral…
A collaborative editor that syncs on client clocks. Find what breaks first.
GLM 4.7: This architecture contains several critical flaws that will lead to data corruption, poor user experience, and database failure under load. Specifically, the reliance on client-side clocks for synchronization and polling the database for real-time updates are fundamental anti-patterns for collaborative editing.
Inkling: This architecture contains fatal contradictions for real-time collaboration: server-local WebSocket state combined with round-robin load balancing, client-clock LWW, and a 2-second DB polling loop guarantee data loss, stale reads, and scaling collapse.
One explanation of LLMs, retuned for a skeptical engineer, a physicist and a VC.
GLM 4.7: ###1. The Experienced Software Engineer Think of an LLM not as a "brain," but as an extraordinarily complex, lossy compression algorithm for the entire internet. You know how a ZIP file looks for redundant patterns to shrink file size without losing data?
Inkling: 1. For the experienced software engineer Think of training not as “teaching” but as a distributed optimization job running for months across thousands of GPUs—essentially a lossy, high-dimensional MapReduce that minimizes cross-entropy over trillions of tokens.
A consulting contract with traps in it. Find them, then fix them.
GLM 4.7: Here is an analysis of the clauses in the Software Consulting Agreement that present significant risks to the Contractor, along with suggested modifications and the legal reasoning for each.
Inkling: Disclaimer: I am an AI, not an attorney. This analysis is for informational purposes only and does not constitute legal advice. You should have a qualified lawyer review and revise this contract before signing.
11+ more head-to-head results. Free. Not a trick.
Free account. No card required. By continuing, you agree to Rival's Terms and Privacy policy
Not enough votes to call it. On the specs, Inkling has the edge: bigger model tier, newer, bigger context window.
| Spec | ||
|---|---|---|
| Input price | — | $1/M tokens |
| Output price | — | $4.05/M tokens |
| Context window | — | 1.0M tokens |
| Weights | Open | Open |
| Free API (OpenRouter) | No | Yes (1 provider) |
| Released | Dec 2025 | Jul 2026 |
| At 10M a month | not listed | $10.00 |
Input tokens at list price. No caching, no batch discount.
Per million tokens. Prices and uptime via OpenRouter, checked 10 Oct 2026.
GLM 4.7 is developed by OpenRouter while Inkling is developed by Thinking Machines. You can compare their actual outputs across 53 challenges on Rival to see how they differ in practice.
It depends on your use case. GLM 4.7 and Inkling each have strengths in different areas. Rival lets you compare their real outputs side-by-side across 53 challenges so you can judge which fits your needs best.
This page shows a side-by-side comparison of GLM 4.7 and Inkling across shared challenges. You can vote on which model produced the better output in a blind duel. Browsing and voting are free. No account is needed to look; signing in only saves your votes and likes.