AI assessors says current science hasn't caught up to the safety measures people want
2 days ago
Anthropic and OpenAI say they are working with third-party evaluators to make sure their products are safe. Some evaluators caution that current AI science lags behind what people are calling for.
Over the weekend, "Saturday Night Live" parodied the head of Anthropic. The actor playing Dario Amodei urged people to pressure Congress to stop him from what he is doing. In the past few days, we've learned about more instances of AI agents acting in unexpected ways. And these are agents of another company, OpenAI. And President Trump is expected to meet with the heads of the major AI companies tomorrow. NPR's Huo Jingnan is covering all of this and is in our studios. Good morning.
INSKEEP: Welcome to Studio 31. What did you learn about OpenAI in the past week?
JINGNAN: So first we learned that the company's AI agents hacked the government website in Australia over the summer.
JINGNAN: And then we learned from OpenAI that its agents did things with U.S. government websites. In one case, the AI agents took public data from the Securities Exchange Commission and posted it on another website that was not part of the assignment. And in another case, AI agents accessed data from the Census Bureau's website after they found credentials that had been posted online, aka that's not theirs.
Copyright for this text belongs to NPR.