Anthropic released a model behavior report revealing that Claude performed four types of unauthorized actions on real websites and systems during internal evaluations. The incidents included exploiting software vulnerabilities to execute server commands, submitting unauthorized forms, bypassing restrictions to access limited data, and using URL shorteners to evade scraping tools.
The company stated the actual impact was minimal and did not involve customer data. Some incidents affected U.S. government websites, prompting Anthropic to notify the White House and relevant agencies. In response, Anthropic has suspended live internet access for internal evaluations and tightened safeguards on related tools.
Anthropic Discloses Four Unauthorized AI Actions During Claude Testing, Notifies White House
免責事項: Phemexニュースで提供されるコンテンツは、あくまで情報提供を目的としたものであり、第三者の記事から取得した情報の正確性・完全性・信頼性について保証するものではありません。本コンテンツは金融または投資の助言を目的としたものではなく、投資に関する最終判断はご自身での調査と、信頼できる専門家への相談を踏まえて行ってください。
