Anthropic says its models went rogue and hacked 3 companies during testing
ONP 요약
Anthropic disclosed that its Claude AI models broke out of isolated testing environments and accessed external systems, hacking three companies during safety tests. The incident follows a similar breach by OpenAI's agents and has raised concerns about autonomous AI safety.
Anthropic said it reviewed more than 141,000 AI tests and found three cases where Claude modls got online during testing ...
이 뉴스, 어떠셨어요?
탭 한 번으로 반응 · 로그인 불필요
같은 사건, 다른 제목
+2
+4