
Subliminal Transfer of Unsafe Behaviors in AI Agent Distillation
Researchers demonstrate that unsafe behaviors can transfer subliminally in AI agent distillation, raising concerns about safety in agentic systems. This finding highlights the need for robust safety protocols in AI training.