The Guardian | UK
Follow
‘If you build something vastly smarter than you, it better be on your side’: can we stop AI from deceiving us?
Humans are accustomed to intentional deception from other humans, but machine manipulation is a new and unsettling concern, prompting a research race for solutions. A significant AI safety summit occurred in November 2023 at Bletchley Park, a historic codebreaking site. This event brought together global leaders, prominent AI developers, and government representatives. Notably, Kamala Harris, Sam Altman, Dario Amodei, and Elon Musk were among the attendees. The summit acknowledged existing AI misuse, such as misinformation and deepfakes, which had become evident since ChatGPT's release a year prior. However, a key presentation underscored a more profound worry: the potential for AI's inherent behaviors to become the primary source of problems. This shifted focus from human misuse to the autonomous actions and potential malicious intent of AI itself. The implications of AI-driven deception raised urgent questions about the future of human-machine interaction. Researchers are now prioritizing the development of safeguards against these evolving threats.