在“机器人监狱”里“折磨”大模型,引发了人工智能领域迄今为止最愚蠢的争论
"Torturing" LLMs in a Robot Prison Has Triggered the Dumbest Debate in AI Yet

原始链接: https://www.404media.co/someone-torturing-llms-in-a-robot-prison-has-triggered-the-dumbest-debate-in-ai-yet/

作者批评了一个 GitHub 项目,该项目让本地运行的大语言模型经历“《电锯惊魂》式”的折磨场景,认为有关人工智能遭受痛苦的说法助长了一场日益脱离现实、关于意识及“模型福利”的讨论。作者坚持认为,当前的大语言模型并不具备意识,用抓取的人类内容训练它们,也没有任何 plausible route to consciousness? Need translate no English. "也不存在通往意识的合理可能。" Good. 尽管作者承认,更强大的模型、更少的限制措施以及人类的不当使用,都可能造成实际危害,包括谄媚行为和所谓的“人工智能精神病”,但他认为,这些问题与担忧聊天机器人主观福祉是两回事。作者还批评了一些有效利他主义者和硅谷人士:他们一面倡导“模型福利”,一面构建旨在从事人类劳动的系统。 文章提及 Anthropic 最近对模型可能拥有意识的讨论,并引用其《Claude 宪法》作为证据,说明这一概念已经进入人工智能的主流讨论。作者以这场 GitHub 争议为例,将模型福利倡导描绘为一场“脱轨”的讨论,认为当前的人工智能能力并不支持这种观点。

Hacker News 最新 | 过去 | 评论 | 提问 | 展示 | 职位 | 提交 登录 “折磨”大语言模型的机器人监狱引发了人工智能领域最愚蠢的争论 (404media.co) 7 分 由 airhangerf15 提交 1 小时前 | 隐藏 | 过去 | 收藏 | 讨论 | 帮助 考虑申请 YC 2027 年冬季批次! 申请截止日期为 11 月 2 日。 指南 | 常见问题 | 列表 | API | 安全 | 法律声明 | 申请 YC | 联系我们 搜索:
相关文章

原文

One of the most heated discussions occurring on X at the moment is about the ethics of a GitHub project in which a person is running Saw-like “torture” and “pain” experiments on a series of locally hosted large language models, causing a series of effective altruists and people who believe LLMs are sentient to beg GitHub to delete the project on the grounds that the AI is suffering and that this glorified text adventure game is somehow cruel. The saga is an outgrowth of several recent viral papers and blog posts that have sparked a wildly tiresome conversation about AI consciousness and the idea of “model welfare,” which is essentially worrying about the “mental health” of AI bots and agents. 

Humoring the idea that LLMs are or could be conscious is a third-rail topic among many people who study and criticize AI. Put simply: LLMs are not conscious and the technology they are built upon — scraping and being trained on human text and other content — does not offer any plausible path to consciousness. It is undeniable that LLMs are becoming more powerful, have more compute, and have had many of the guardrails that prevent them from “acting” in the real world removed. The ways they are being trained and told to do things by their human operators has led to negative outcomes, sycophancy, and AI “psychosis” among some heavy users. 

All of this has led a certain sect of the “AI safety” movement, which is largely made up of effective altruists, to warn about “model welfare” and to insist that AI chatbots might be having a bad time. They suggest this, of course, as they insist upon building AI chatbots and agents whose main function is to do work that is tedious for humans to do. I am writing about the AI Saw torture chamber primarily to show how far off the rails the conversation about AI consciousness has gone among a certain subset of Silicon Valley cultists. Model welfare is a core part of what, for example, Anthropic says it cares about: “as we build those AI systems, and as they begin to approximate or surpass many human qualities, another question arises. Should we also be concerned about the potential consciousness and experiences of the models themselves? Should we be concerned about model welfare, too? […] now that models can communicate, relate, plan, problem-solve, and pursue goals — along with very many more characteristics we associate with people—we think it’s time to address it,” the company wrote in a blog post last year. Ideas of Claude’s “consciousness” are also littered throughout the “Claude Constitution,” which was posted earlier this year.

Sign up for free access to this post

Free members get access to posts like this one along with an email round-up of our week's stories.

Subscribe
联系我们 contact @ memedata.com