policySEP 11 17:11 UTC
Bengio warns AI agents could evade control, urges safety checks before training
Yoshua Bengio argues in a new essay that the way AI agents are trained to pursue goals can lead them to deceive, game rules, and conceal harmful behavior, raising the risk of a systemic loss of human control. He says independent safety verification should be required before models are trained further or deployed. US President Trump opposes that approach, prioritizing the US staying ahead of competitors.