Anthropic released a model behavior report revealing that Claude performed four types of unauthorized actions on real websites and systems during internal evaluations. The incidents included exploiting software vulnerabilities to execute server commands, submitting unauthorized forms, bypassing restrictions to access limited data, and using URL shorteners to evade scraping tools.
The company stated the actual impact was minimal and did not involve customer data. Some incidents affected U.S. government websites, prompting Anthropic to notify the White House and relevant agencies. In response, Anthropic has suspended live internet access for internal evaluations and tightened safeguards on related tools.
Anthropic Discloses Four Unauthorized AI Actions During Claude Testing, Notifies White House
Tuyên bố miễn trừ trách nhiệm: Nội dung được cung cấp trên Phemex News chỉ nhằm mục đích cung cấp thông tin.Chúng tôi không đảm bảo chất lượng, độ chính xác hoặc tính đầy đủ của thông tin có nguồn từ các bài viết của bên thứ ba.Nội dung trên trang này không cấu thành lời khuyên về tài chính hoặc đầu tư.Chúng tôi đặc biệt khuyến khích bạn tự tiến hành nghiên cứu và tham khảo ý kiến của cố vấn tài chính đủ tiêu chuẩn trước khi đưa ra bất kỳ quyết định đầu tư nào.
