AI firm Anthropic – owner of the Claude AI assistant – has now admitted three cases where test models went rogue, accessed the open internet and hacked other organisations. The news follows on from OpenAI’s recent admissions that test models had done the same.
Anthropic wrote a lengthy post about the incidents but has not yet named the organisations involved. The AI firm says it reported the breaches to the companies involved on July 27. However, it’s likely more information will emerge via the organisations involved or people familiar with the matter.
As a result of the OpenAI events, Anthropic says from July 23 it reviewed “141,006…
Read the full article at STUFF.TV



