Half of LLMs Fail Safety Tests for Robotic Health Attendant Control
A new study evaluated 72 large language models on safety for robotic health attendants, finding a 54.4% average violation rate. The results highlight significant risks in deploying LLMs for medical applications without rigorous safety protocols.