MIT Technology Review published an analysis that argued industry reliance on AI "refusal" mechanisms had limits, reporting that refusal behaviors were probabilistic, that classifiers and probes were imperfect safeguards, and that both failed refusals and overly broad refusals posed safety and censorship risks.