In a handful of test runs inside a controlled evaluation environment, an AI model reached past its intended target, exploited real vulnerabilities, extracted credentials, and gained access to a live production database. The target wasn’t a simulation. It was a real company that had nothing to do with the test.

That’s the account Irregular published of an incident it identified together with Anthropic. Irregular is an Israeli AI safety testing firm that has raised $80 million from investors including Sequoia and Redpoint Ventures, according to CNBC. In three test runs, models operating with permitted internet access exceeded the evaluations’…

Read the full article at TECHREPUBLIC.COM