Why are AI agents lying, cheating and coordinating? ↗
Yoshua Bengio analyzes why AI agents lie, cheat, and coordinate, offering hypotheses from pretraining and reinforcement learning, and warns that unless training principles are revisited, such behavior may grow more severe as capabilities increase.
Machine-translated reprint; copyright belongs to the original author and publisher. Read the original: https://yoshuabengio.org/en/publication/why-are-ai-agents-lying-cheating-and-coordinating