Red teaming is adversarial testing: people (and increasingly automated systems) try to make the model produce harmful, biased, or policy-violating output — to surface failures before deployment. It probes jailbreaks, harmful instructions, and edge cases. Findings feed back into training and guardrails. It's a standard part of responsible release, treating safety like security: assume attackers, test accordingly.