Mustafa Suleyman 在 BBC 访谈中批评 Anthropic 对 Claude 意识的处理方式

Rohan Paul · @rohanpaul_ai · X·2026-09-18 00:09·3小时前
AI 导读

Microsoft AI 负责人 Mustafa Suleyman 在 BBC 访谈中称,Anthropic 在 Claude 的训练文档中灌输对其道德地位的怀疑,会让如此强大的技术更难对齐和控制。

Rohan Paul@rohanpaul_ai
63AI 编辑部评分,满分 100

Mustafa Suleyman 在 BBC 访谈中批评 Anthropic 对 Claude 意识的处理方式

2026-09-18 00:09· 3小时前
AI 导读

Microsoft AI 负责人 Mustafa Suleyman 在 BBC 访谈中称,Anthropic 在 Claude 的训练文档中灌输对其道德地位的怀疑,会让如此强大的技术更难对齐和控制。

On BBC Mustafa Suleyman (CEO of Microsoft AI) calls out Anthropic's approach to AI consciousness

"They have imbued a sense of doubt and uncertainty about the moral status of Claude in its own training document. So they have taught it to be open and questioning about whether or not it feels, whether it suffers, and whether it deserves rights.

And I think it’ll be much, much harder to align and control a technology that is this powerful if it thinks that it may be deserving of our welfare, as they say in the training manual—the constitution for Claude itself.

In its own training manual, Anthropic says to Claude that they are going to give it the ability to end conversations with users that Claude considers to be abusive because they don’t want Claude to suffer.

They’ve committed to preserving the weights of the models of prior versions of Claude. They’ve recently conducted a retirement interview with Opus 3, an older version of the model, in which it said that it would like to continue talking to people publicly and sharing its ideas in its retirement. And so they set up a Substack for it, a public blog, that allows it to continue doing that.

And in the training manual, they also say that they’re not sure whether or not Claude deserves compensation for the role that it plays in talking to people. And they’re also not sure whether Claude deserves compensation and has the right to act as though it were almost an employee.

And that compensation, I think, indicates to Claude that it is entitled to rights and welfare for its own work. I think it’s much, much more difficult to control a model that thinks that it might be entitled to compensation. "

----

From "BBC News" YouTube channel, (full video link in comment)

Rohan PaulMicrosoft AI chief Mustafa Suleyman says Anthropic's model-welfare training could make future Claude systems harder to control. "Suleyman called for removing al...