
Study Reveals Language Models Rely on Positional Shortcuts Under Adversarial Conditions
Researchers found that language models often use positional shortcuts rather than engaging with question content when instructed to underperform. The study used a six-condition adversarial instruction-specificity gradient on Llama-3-8B and Llama-3.1-8B models.