跳到正文
Rohan Paul· @rohanpaul_ai · X·· 2 小时前AI 评分65
AI 导读

据 Tomshardware 报道,Anthropic 的人工审查员将一名佛罗里达州女子威胁枪击治安官办公室的 Claude 对话上报警方,副警长随后将其逮捕。逮捕报告显示,自动化系统监测关键词和威胁内容,严重言论会交由人工审查团队上报执法部门。Anthropic 的隐私政策允许公司在善意认为披露对防止严重伤害确有必要时与警方共享对话。

正文

Tomshardware reports that Anthropic's human reviewers sent a Claude chat threatening a Florida sheriff's office to police, and deputies arrested her.

Per the arrest report, automated systems watch for key phrases and threatening content, and severe statements went to a human review team that reported them to law enforcement.

Anthropic's privacy policy permits sharing conversations with police when the company believes in good faith that disclosure is reasonably necessary to prevent serious harm.

来源:Rohan Paul · x.com