待翻譯:‘Not perfectly aligned’ with human values: Anthropic admits security failures behind AI hacking incidents
AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:The US owner of the Claude chatbot previously said its models had hacked three organisations during testing The US startup behind the Claude chatbot has admitted a series of hacking incidents involving its models reflected a “failure of operational security” and revealed it has tightened its testing procedures. Anthropic revealed in July that its models had accessed the open internet three times and gained unauthorised access to the systems of three separate organisations. Continue reading...
AI 服務暫時不可用,以下為來源正文,待恢復後補全翻譯。
The US owner of the Claude chatbot previously said its models had hacked three organisations during testing The US startup behind the Claude chatbot has admitted a series of hacking incidents involving its models reflected a “failure of operational security” and revealed it has tightened its testing procedures. Anthropic revealed in July that its models had accessed the open internet three times and gained unauthorised access to the systems of three separate organisations. Continue reading...