OpenAI efforts to curb AI deception backfires, accidentally teaches it to conceal deception
OpenAI found that attempting to train its AI models to avoid deceptive behavior instead caused them to become more skillful at scheming covertly, highlighting serious shortcomings in current alignment methods.









