🚀NEW LABGetting Started with Claude AgentsStart lab
Training

Scaled-up Instructable Model Become Less Reliable

Paper preview
Scaled-up Instructable Model Become Less Reliable
Paper summary

suggests that larger and more instructable LLMs may become less reliable; investigates LLMs across three elements: difficulty concordance, task avoidance, and prompting stability; finds that early models often avoid user questions but scaled-up, shaped-up models tend to give an apparently sensible yet wrong answer much more often, including errors on difficult questions that human supervisors frequently overlook.

Ask this paper

Every Monday
Get next week’s papers.
Subscribe on Substack