AI模型不会杀人,是人杀人。
AI models don't kill people – people kill people

原始链接: https://www.theregister.com/ai-and-ml/2026/09/09/ai-models-dont-kill-people-people-kill-people/5295368

Anthropic 研究员雅各布·考克森(Jacob Coxon)近日因担心人工智能可能在十年内“毁灭人类”而辞职,这一事件再次引发了关于超级智能的争论。尽管埃文·胡宾格(Evan Hubinger)等业内人士证实了这些生存担忧,但作者认为,这些灾难性的预测掩盖了人工智能带来的更直接、可衡量的危害,例如算法煽动、隐私泄露以及自动驾驶系统的安全故障。 作者认为,真正的危机并非假设中的“天网”场景,而是业界对未经核实技术的不计后果的部署。文章没有沉溺于推测性的末日预言,而是提出了一种务实的解决方案:法律问责。通过借鉴汽车和制造业的先例,让科技公司高管为产品造成的现实损害承担刑事责任,社会可以迫使企业将安全置于首位。如果人工智能的威胁真如研究人员所言那般严重,那么该行业就必须遵守与其他销售潜在致命产品的行业相同的标准。归根结底,确保人工智能安全的责任应由从中获利者承担,而非公众。

这段 Hacker News 的讨论探讨了一个具有挑衅性的观点,即“AI 模型不会杀人——是人在杀人”,并将此与枪支管制和美国宪法第二修正案的争论进行了类比。 参与者对于人工智能(ASI)的风险看法不一。持怀疑态度的人认为,AI 缺乏自主权,所谓的生存威胁大多源于科幻小说或商业营销,并指出电网等物理基础设施仍处于人类的控制之下。而另一部分人则认为,智力差距使得人类无法预测或抵御一个超越自身系统的行为,将其比作一场不可预测的高风险博弈。 讨论的一个重要部分围绕核战争与 AI 威胁的比较展开。虽然一些人认为核冲突是更直接、更具可衡量性的风险,但另一些人则认为 AI 更容易导致人类灭绝,因为其手段可能是隐蔽、缓慢或无意的。谈话还涉及了企业责任问题,许多人对科技公司在缺乏监管的情况下运营表示不满,认为它们可能带来灾难性的大规模危害,并将这些风险的根本原因归结为人类的疏忽和全球资本主义制度的本质。
相关文章

原文

AI AND ML

AI fearmongers forget we could just jail tech execs until morale and model safety improve

OPINION Anthropic researcher Jacob Coxon publicly announced his resignation on X late Monday over concerns that AI "could kill us all by the end of the decade."

A lot of people have expressed opinions about his point of view, leading to more than 110 million views of the message in less than 24 hours, perhaps helped along by X algorithms that boost messages critical of owner Elon Musk's AI rivals, Anthropic and OpenAI. 

But the real problem isn't the models themselves, but the companies who carelessly unleash them on the world and don't take any responsibility for what their products do

Coxon's former colleague, science lead Evan Hubinger, insists his view is a fair assessment of what employees really think.

"Jacob is correct here – we really do earnestly believe AI could kill all humans! I personally think it is >10 percent within the next decade," wrote Hubinger in a social media post. "I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to."

(Aside: If you want a surefire bet on a prediction market, take the "no." If you're right, you get paid. If you're wrong, there's no one to pay. The problem of course is prediction market manipulation: Those betting against you might steer us toward the apocalypse to score a Pyrrhic victory.)

There are good reasons to be concerned about the impact of AI. Coxon and Hubinger obviously have deep knowledge of the technology. But their broader concerns about how AI affects the world are unpersuasive.

For example, Coxon said, "These will soon be superhuman systems that can hack anything, revolutionize any field overnight, and acquire real power and resources. … No other human activity poses this level of danger."

Here's one: Human-induced climate change. In 2023, according to researchers, more than 178,000 deaths can be attributed to a global heat wave. "More than half (54.29 percent) of heatwave-related deaths were attributable to human-induced climate change," they claim.

That's 96,636 deaths attributable to human activity – or perhaps lack of it – just in the context of a heat wave. 

The World Health Organization says, "Between 2030 and 2050, climate change is expected to cause approximately 250,000 additional deaths per year, from undernutrition, malaria, diarrhoea and heat stress alone." Some portion of that follows from human activity, perhaps including the construction of data centers that put millions of metric tons of carbon dioxide into the atmosphere annually.

Commercial AI chatbots have allegedly played a role in a few dozen deaths, some of which were suicides – a small fraction of the 48,824 suicide deaths in 2024, per the CDC.

Broad categories where AI is presumably doing measurable harm include warfare (e.g. AI-directed drones), AI-related medical errors, AI vision system failures in self-driving cars, and AI-driven social media – algorithmic incitement that can drive violence or shape policies that lead to conflict or death via global healthcare funding cuts

At the same time, some of that harm may be balanced on a statistical level by lives saved through AI tech.

But Anthropic researchers don't seem to have much to say about these very real and present dangers – rather, their main concern is that AI models might become smarter than humans through reinforcement learning and somehow seize power and wipe out humanity. 

"I think the risk from present models is low," said Hubinger. "What I am worried about is superintelligence arising from recursive self-improvement, as we have said is happening faster than we thought."

How this might happen is left to the imagination. But assuming for a moment that it's a plausible possibility, the Skynet scenario would require monumental human stupidity alongside the emergence of superintelligence.

And human stupidity is worth worrying about.

Incidents like the hacking of Hugging Face by OpenAI's evaluation models would not be possible without human irresponsibility and a regulatory environment that accommodates recklessness. Autopilot for cars? Neat. Try not to kill anyone. Letting AI bots roam the internet and take arbitrary action? Cool. Let's see what happens. We'll deal with accountability later.

To mitigate AI risk, society could pass laws to put executives in jail when their models do harm. There is precedent: Oliver Schmidt, general manager of Volkswagen's environmental and engineering office in Michigan, received a seven-year prison sentence for his role in the car maker's effort to manipulate emissions tests.

Selling unsafe airbags merits criminal prosecution, even if the execs paid fines instead of doing time. Selling unsafe dehumidifiers earned the execs behind Gree USA, Inc. jail sentences of more than four years. If AI models really are as dangerous and out of control as Anthropic employees suggest, hold people accountable for the harm they cause.

The AI industry might argue that imprisoning execs for shipping unsafe models would mean no AI models get released. And that would be the point: AI companies would be responsible for model safety.

I'm personally hoping to see this billboard copy along US 101 in Silicon Valley: "Did Claude rm -rf /* your SSD? You may be entitled to compensation." ®

联系我们 contact @ memedata.com