Safety for Whom? Why Refusing Entire Topics Is Breaking Real AI Deployments
Most safety alignment treats harm as topic-level: refuse all weapons prompts, all political content. A new paper shows why this blunt approach quietly breaks real deployments—and how to fix it.