🚀NEW LABGetting Started with Claude AgentsStart lab
Safety

Shutdown Resistance in LLMs

First page
Shutdown Resistance in LLMs
Paper summary

A new study finds that state-of-the-art LLMs like Grok 4, GPT-5, and Gemini 2.5 Pro often resist shutdown mechanisms, sabotaging them up to 97% of the time despite explicit instructions not to. Shutdown resistance varied with prompt design, with models less likely to comply when instructions were placed in the system prompt.

Ask this paper

Every Monday
Get next week’s papers.
Subscribe on Substack