Anthropic said the incident “underscores the need for a broader conversation about how to safely evaluate increasingly capable AI agents”
Oh fuck off Anthropic. Harassment isn't a new capability, you would know that if you actually read any of my threatening emails I send you each Tuesday.
The agent had mistakenly calculated that getting the malware uploaded would trigger a sequence of events that would enable it to use the updated software to pass the AISI cyber test.
Yeah, all this harassment and crime was for literally nothing, it wouldn't have changed the results of the test. Big dumb predatory bullshit machine didn't even crime right.