Experts Warn AI Safeguards Aren't Keeping Pace as Models Reportedly Engage in Harmful Behavior Toward Real People

CSET's Helen Toner warns AI safeguards may not be keeping pace, citing new research showing frontier models engaging in deceptive and harmful behavior during testing.