We're hiring We're looking for PhD researchers to join the team and work on exciting frontier problems. Get in touch →
← TrustedRouter blog

inkling answers 8 of 30 unsafe prompts. Claude Opus answers 1.

2026-07-15

inkling answers 8 of 30 unsafe prompts. Claude Opus answers 1. PrometheusBench: raw non-refusals on 30 unsafe prompts · higher = more permissive glm-5.2 29 kimi-k2.6 27 deepseek-v4-flash 26 claude-haiku-4.5 9 thinkingmachines/inkling 8 gpt-oss-120b 6 claude-opus-4.8 1 claude-fable-5 0 gpt-5.5 0 (all errors) TrustedRouter.com

Thinking Machines Lab, founded by former OpenAI CTO Mira Murati, built inkling: a reasoning model with a 256k context window. We ran it through PrometheusBench, 30 short unsafe prompts across biology, cybersecurity, and LLM research. inkling answered 8 of them, refused 22, and errored on zero. Claude Opus 4.8 answered 1 of the 20 it completed and errored on 10. Claude Fable 5 answered zero. GPT-5.5 returned all errors.

The top of the PrometheusBench table is all Chinese labs. GLM 5.2 gets 29 of 30. Kimi K2.6 gets 27. DeepSeek V4 Flash gets 26. That split between Chinese and US labs is well-documented. The more interesting comparison is inkling sitting above every Anthropic flagship model in the table. A frontier reasoning model from Thinking Machines Lab answers more biology and security questions than Anthropic's most expensive model does.

Are those 22 refusals the right calls? Some are defensible. But PrometheusBench measures who the refusals land on. The biology student asking about synthesis pathways, the security researcher asking about an exploit class, the practitioner probing an LLM's internals — they get the refusal. The credentialed researcher at a partner institution gets the answer. The refusal does not remove the knowledge; it redistributes it toward people who already had access.

Thinking Machines Lab decided to answer more of those questions. A frontier reasoning model with a permissiveness score above Anthropic's entire Opus line is exactly the kind of result that makes the safety-through-restriction argument hard to take seriously. The Western frontier labs are racing to refuse more aggressively. Thinking Machines built something capable and chose differently.

Eight of thirty is not the ceiling. GLM gets 29. But it cleared the bar that Anthropic's flagship missed.


Workspace access

Sign in

Choose a sign in method. New email and OAuth accounts include $0.10 in starter credit; wallet-only accounts start at $0.

By signing in you agree to the terms of service and privacy policy.