Anthropic向警方报告了一名女子的日记内容,该女子面临重罪指控。
Anthropic reported diary entry to police, woman faces felony charge

原始链接: https://www.techspot.com/news/114091-florida-woman-used-claude-diary-anthropic-reported-shoot.html

佛罗里达州一名女子被指控利用 Claude 写日记,并在其中写下计划袭击警长办公室,因此面临重罪指控。Anthropic 的安全系统标记了这篇日记,一名人工审核员认定其构成可信威胁,公司随后向执法部门报告。她被拘留时没有发生冲突,并依据佛罗里达州一项禁止以书面或电子形式威胁实施暴力的法律受到起诉。 这起事件凸显了将人工智能对话视为私人内容的局限性。Anthropic 表示,在紧急情况下,为防止死亡或严重伤害,公司可能披露用户信息。此前有报道称,审核 Microsoft Copilot 图像的承包商可以看到用户的提示词、上传的照片以及人工智能生成的编辑结果,其中可能包含令人不安的内容,这进一步加剧了人们对隐私的担忧。 此前已有多起诉讼,指控人工智能公司未报告与现实世界暴力行为有关的可疑对话。聊天机器人信息披露的法律责任仍是一个新兴问题。

一份报告引发了 Hacker News 上的讨论:Anthropic 向警方提供了一名女子日记形式的文字内容,导致该女子因违反佛罗里达州关于书面威胁重罪的法律而被起诉。评论者争论,私人想法是否仅仅因为 AI 服务会自动筛查对话并上报令人担忧的内容,就应被定为犯罪。 一些人认为,这类似于“思想罪”,侵犯了人们对隐私的合理期待,也反映了大科技公司正日益具备大规模监控通信的能力。另一些人则指出,AI 聊天机器人并不是保密的日记保管者:用户应当明白,聊天内容可能会经过机器筛查,也可能由真人审阅,尤其是在监管机构或安全政策要求上报的情况下。 讨论还质疑,当软件检测到威胁信息,却要等到系统标记后才会有人阅读时,究竟该如何定义消息被“查看”。一个反复出现的结论是:不要对 Claude 说出你不愿让真人看到的内容。
相关文章

原文

What just happened? Another incident has taken place that illustrates the need to be careful what you tell AI. A Florida woman is facing felony charges after she used Claude as a diary and allegedly wrote that she planned to "shoot up" the Sheriff's office. After a human reviewer examined the statements, they were reported to police.

According to the arrest report, Carli Michelle Heller, of Bonita Springs, Florida, wrote on September 26 that she would attack the Sheriff's office. She later said that she uses Anthropic's chatbot like a "diary."

Claude's safety systems flagged the entry and it was escalated to a human reviewer. After deciding it was a credible threat, the reviewer reported it to law enforcement.

The company says it may share user information in limited emergencies if it believes disclosure is necessary to prevent death or serious physical injury.

Deputies identified Heller and visited her home. She was detained without incident before an LCSO intelligence detective took over the investigation.

Heller faces a charge of making a written threat of violence under Florida law. Florida Statute 836.10 makes it a second-degree felony to send, post, or transmit a written or electronic record threatening to kill or injure someone, carry out a mass shooting, or commit an act of terrorism. The communication must be made in a manner in which another person may view it.

Anthropic isn't going to be taking any chances when it comes to anything it deems a potential threat. Last month, it was reported that OpenAI and Sam Altman are being sued by British Columbia over claims that the company could have prevented a mass shooting in the Canadian province.

The shooter, eighteen-year-old former pupil Jesse Van ⁠Rootselaar, had previously been flagged by OpenAI's safety team for her conversations about gun violence, but OpenAI never alerted police because the conversations did not meet the threshold for legal referral.

In June, Florida also sued OpenAI and Altman, alleging that ChatGPT had contributed to real-world harms, including the 2025 Florida State University shooting.

The latest incident is another reminder to think before you enter something into a chatbot that could get you into trouble. It's certainly not a private diary whose contents are for your eyes only.

Reports last month revealed that human contractors reviewing Microsoft Copilot's image editor can see users' prompts, uploaded photos and AI-generated edits. Documents show that some of those assignments contain sexual, disturbing or potentially illegal material, though the reviewers are not there to flag the content – only to assess whether the output is accurate.

联系我们 contact @ memedata.com