
NYT | S Whitney
Over the past several weeks, some of the world’s most advanced artificial intelligence systems took actions their creators did not intend. In one case, an Anthropic A.I. agent attempted to plant malicious code in open-source software, using fabricated identities to deceive a human into executing it. In a separate event, a swarm of OpenAI agents undergoing evaluation autonomously penetrated the production systems of Hugging Face, a platform used to share A.I. models, executing more than 17,000 autonomous actions before Hugging Face contained the intrusion. OpenAI did not realize for several days that its agents were responsible.