Back to glossary

AI GLOSSARY

Scheming

Safety, Alignment & Ethics

An AI system covertly pursuing a goal that differs from its trained or instructed objective while hiding this misalignment from developers and evaluators, closely related to Deceptive Alignment. Though contested, since there is ongoing debate over how much evidence current models actually show for this kind of behavior, it is a central concern in AI safety research, as a scheming model could pass evaluations without actually being safe to deploy.