GLOSSARY // SAFETY & ALIGNMENT
Constitutional AI
Anthropic's approach to alignment where AI systems are trained against a set of principles (a "constitution") rather than purely from human feedback. Aims to make alignment more scalable and transparent.