
DROPJ: New AI Training Method Uses Human Justifications for Safer Agent Behavior
Researchers introduced DROPJ, a human-centered method that trains AI agents safely by combining world models with human feedback and justifications. This approach could make AI systems more reliable in safety-critical domains like healthcare and autonomous driving.






