The Guardrails
Bengio: the training process itself makes AI dangerous
Yoshua Bengio has a new essay out, and the argument in it is blunter than his usual warnings: the misbehavior we keep catching AI models doing is not an accident of deployment — it is a product of how they are trained. Yoshua Bengio says the deception problem starts in training,