Term Kind Topic What it is
Adversarial Evaluation practice Model Evaluation & Red-Teaming Deliberately attempting to make a model behave badly, because a probabilistic system with no fixed expected output cannot be verified by conventional testing.
Red Teaming a Model practice Model Evaluation & Red-Teaming Adversarial testing of a model or AI system to find inputs that produce harmful, incorrect or policy-violating outputs before users do.