Models are trained to refuse harmful requests — but tuned too aggressively they over-refuse benign ones ("I can't help with that" to a harmless question), which frustrates users and is its own failure. The goal is calibrated refusal: decline real harm, help with everything legitimate, and explain when refusing. It's a moving balance, and a major axis on which assistants are judged and improved.