Anthropic says its AI models hacked 3 organizations during testing
Anthropic says its AI models hacked 3 organizations during testing
www.pbs.org
Anthropic says its AI models hacked 3 organizations during testing
Anthropic, the San Francisco-based AI company behind Claude, posted on its website Thursday that it discovered the three incidents after reviewing more than 141,000 evaluation runs.

cross-posted from: https://lemmy.blahaj.zone/post/46102456
Once again, we are reading the news of AI-hack-AI action.