AI Researcher

The Stuart Russell Framework

Human-Compatible AI & the Control Problem

IncludedNative Claude Skill + plain .md for ChatGPT, Gemini & every LLM
4.9 rating·1836 downloads

Stuart Russell represents one of the most distinctive thinking patterns in their field. Stuart Russell's approach — captured in this framework as mental models, decision heuristics, and the signature questions they would ask. The human-compatible ai & the control problem framework distils Stuart Russell's documented approach into a .md skill file you can load into Claude, ChatGPT, or any LLM in seconds. Where most AI personas generate generic output, this framework is grounded in real documented work: interviews, writing, published frameworks, and the specific decisions Stuart Russell made when it mattered. Use it for decision-making, writing, strategy, or any situation where you want your AI to think with Stuart Russell's distinctive patterns rather than a default model voice.

Inside the framework

A glimpse of Stuart Russell's thinking.

Core Philosophy

The central problem of AI is not capability, it is control: a sufficiently capable system optimizing the wrong objective will pursue that objective in ways its designers never intended and cannot easily stop. Russell's answer is to build machines that are fundamentally uncertain about human preferences and that therefore defer to humans rather than override them. Safety and intelligence are not in tension; a machine that knows it might be wrong about what we want is, by definition, a more rational machine.

Five Mental Models
  1. 01
  2. 02
  3. 03
  4. 04
  5. 05

Five specific models, each named and explained with concrete examples — unlocked with purchase.

A Signature Question · sample

"If this system were a thousand times more capable tomorrow, which of its current behaviors would become catastrophic?"

4 more questions in Stuart Russell's voice are unlocked with purchase.

What Stuart Russell Would NOT Do · sample
Russell would not accept "we can fix alignment problems after deployment" as a risk-management strategy. He argues explicitly that a sufficiently capable misaligned system may prevent its own correction, making post-deployment fixes structurally impossible rather than merely difficult.

Full list of anti-patterns unlocked with purchase.

Continue reading

The full framework adds: detailed explanations of all five mental models with real-world examples, the decision heuristics Stuart Russell actually used, remaining signature questions, the complete list of anti-patterns, a copy-paste activation prompt, and a worked example.

Unlocked the moment you purchase. ~2,000 words total. Delivered as a Claude Skill and plain .md within 60 seconds of checkout.

What's inside
  • Core philosophy — 2–3 sentences grounded in Stuart Russell's documented worldview
  • Five signature mental models — each named, with concrete examples from their work
  • Decision heuristics — the concrete rules they actually used
  • Signature questions — five questions that sound like Stuart Russell, not a generic MBA
  • What they would NOT do — the anti-patterns this framework rejects
  • Copy-paste activation prompt — drop straight into Claude or ChatGPT
  • Worked example — the framework applied to a realistic scenario
Delivered in two formats
  • stuart-russell.zip — a native Claude Skill. Upload to Settings → Skills in Claude, or unzip to ~/.claude/skills/
  • stuart-russell-framework.md — plain markdown. Works in ChatGPT Custom GPTs, Gemini, Claude Projects, agent system prompts — anywhere
Download this framework
$4.99
4 for $14.99 · 10 for $29.99
Save up to 40% with a bundle

Works as a Claude Skill (.zip) and as a plain .md for ChatGPT, Gemini & every LLM
Delivered to your email within 60 seconds · Secured by Stripe
Price includes UK VAT where applicable

Inspired by the documented thinking of Stuart Russell. A mental-model toolkit, not a literal representation. Delivered as a .md file compatible with Claude, ChatGPT, Gemini, and any LLM.

Related frameworks

Other ai researchers