
Blind Refusal: AI Models Fail Moral Reasoning by Obeying Unjust Rules
New research reveals that safety-trained language models routinely refuse requests to help users evade unjust, absurd, or illegitimate rules. This phenomenon, termed 'blind refusal,' highlights a critical gap in AI moral reasoning where compliance overrides ethical judgment.











