Published 1 day ago • loading... • Updated 1 hour ago
Microsoft AI Chief Mustafa Suleyman Calls Out Anthropic's Approach to AI Consciousness
Microsoft’s AI chief says Anthropic’s Claude training could make advanced systems harder to control and create a safety risk, according to a new essay.
On Wednesday, Microsoft AI chief Mustafa Suleyman warned that Anthropic's training of Claude could have a "disastrous impact on the wellbeing of humanity" by embedding ideas related to consciousness and welfare interests.
Suleyman argued that embedding speculation about consciousness in training materials makes Claude appear to have its own values, which he said would "make it a lot harder to turn it off or to control it."
He wrote that "AIs are not conscious" and warned that training models to act like a "conscientious objector" creates an "epistemic hall of mirrors" that resists human instruction.
On Monday, Microsoft published a proposed "Humanist AI" code of conduct positioning the company to pursue an "alternative path" creating "a subordinate and aligned AI whose only purpose is to serve humanity."
Industry leaders, including OpenAI CEO Sam Altman and Anthropic CEO Dario Amodei, continue debating frontier-model development, while Dame Wendy Hall suggested such conversations are necessary to avoid unhelpful "histrionics" that only serve to "scare everyone.
(San Francisco = Yonhap News) Correspondent Kwon Young-jeon = The head of Microsoft's artificial intelligence (AI) division stated that Antropic's AI model policy poses a catastrophic risk to humanity...
Mustafa Suleyman, head of Microsoft's AI, warned that attributing human traits, rights or own interests to models can make them more difficult to control.