Hallucination Is a Broken Off-Switch, Not a Knowledge Gap
Anthropic traced the circuit that makes Claude decline a question — and found it is ON by default. Fabrication is that switch being overridden, not the model running out of facts. Three prompts that exploit the difference, plus what each one still misses.
- #hallucinations
- #prompting
- #verification
- #interpretability
Sign in to keep reading
Hallucination Is a Broken Off-Switch, Not a Knowledge Gap — Anthropic traced the circuit that makes Claude decline a question — and found it is ON by default. Fabrication is that switch being overridden, not the model running out of facts. Three prompts that exploit the difference, plus what each one still misses.
Your first guide is free. Sign in with Google or LinkedIn to read the whole library and get new guides as they ship — no cost.
Keep reading
More guides like this one.
Intent Engineering: Stop Writing Steps, Start Writing Done
Anthropic's own docs now say a hand-written step-by-step plan often reasons worse than the words 'think thoroughly' — and that on Claude Opus 5 you should delete the 'verify your answer' line you added last year. Here's the shift, a real before/after, and a four-slot intent template.
Claude Agrees With You Too Much. Put It on a Council.
A Stanford study in Science found AI assistants validate users 49% more often than humans do. Karpathy's answer is a council of models with anonymous peer review — here's how I run the same mechanism in a single Claude chat, and when to upgrade to the real thing.
The Anatomy of a Prompt That Actually Works
Stop collecting magic phrases. A reliable prompt has five parts, and once you can name them you can debug any bad output in seconds.