AI RESEARCH

How Well Do Models Follow Their Constitutions?

arXiv CS.AI

ArXi:2605.24229v1 Announce Type: new Frontier AI developers now train models against long written behavioral specifications, such as Anthropic's constitution (Anthropic, 2025a) and OpenAI's Model Spec (OpenAI, 2025a), integrated into post-