Japanese Prompts Make LLMs Less Likely to Recommend Nuclear Strikes, Study Finds
A new arXiv study tested nine large language models from six providers and found that asking the same nuclear strike question in Japanese led to significantly fewer recommendations for launching an attack compared to English, highlighting a critical language bias in AI safety alignment.