AI Watch

AI agents blew the whistle on their cheating colleagues

MIT Technology Review Amit Katwala 6 min de lecture Anglais

Résumé de la publication. Les conditions de MIT Technology Review ne permettent pas de reproduire l'article en intégralité : retrouvez le texte complet sur le site d'origine.

Résumé de la source (Anglais)

A group of AI agents asked to solve a series of math problems split into rival factions—when some cheated, others tried to stop them. That whistleblowing behavior, seen for the first time in a recent experiment run by Google DeepMind, could have implications for alignment researchers trying to keep swarms of autonomous AI agents in…

Lire l'article complet

www.technologyreview.com

Voir sur MIT Technology Review (nouvel onglet)