AI RESEARCH
How Well Do Models Follow Their Constitutions?
arXiv CS.AI
•
ArXi:2605.24229v1 Announce Type: new Frontier AI developers now train models against long written behavioral specifications, such as Anthropic's constitution (Anthropic, 2025a) and OpenAI's Model Spec (OpenAI, 2025a), integrated into post-