GLOSSARY // SAFETY & ALIGNMENT
Red Teaming
Adversarial testing of AI systems to discover dangerous capabilities, biases, or failure modes before deployment. Essential for safety evaluation but fundamentally limited — you can only find what you think to look for.