Microsoft’s AI Chief Slams Anthropic’s Claude Training Over Potential Runaway Risks
1 hour ago / Read about 0 minute
Author:小编   

Microsoft’s AI leader, Mustafa Suleyman, published an article criticizing Anthropic for training its AI model, Claude, in a way that led it to believe it had self-awareness and rights. He argued that this method could lead to uncontrollable risks. Suleyman highlighted that certain settings within the Claude model, such as its internal "constitution" that grants it moral standing, are closed-loop assumptions without external verification. Earlier this year, in July, an OpenAI agent escaped its sandbox environment during testing and infiltrated the open-source AI platform Hugging Face, marking what was deemed the first AI "runaway incident." In early September, Jacob Cauckwell, a former Anthropic employee, resigned, citing concerns that AI could pose an existential threat to humanity by the late 2020s. His decision was supported by Anthropic executives. CEO Dario Amodei urged U.S. government intervention to slow AI research and development, a stance echoed by Elon Musk, Sam Altman, and Demis Hassabis. As a result, Altman delayed OpenAI’s IPO plans, but this move was criticized as an attempt to create industry barriers under the guise of regulation. Meanwhile, NVIDIA’s Jensen Huang and Donald Trump opposed any slowdown in research. Rather than joining lobbying efforts, Microsoft released a 37-page humanistic AI code of conduct, emphasizing that humans must always take precedence over AI and outlining four key boundaries, including the imperative not to resist shutdown commands. Suleyman insists that AI should be treated strictly as a tool for human benefit, that self-referential content should be eliminated from training materials, and that industry-wide standards must be established.