我解雇了我的AI助手
I Fired My AI Assitant

原始链接: https://chreke.com/posts/i-fired-my-ai-assistant

作者分享了对 Claude (Opus 5) 近期更新的失望,指出其体验已从过往版本中那种有益、专业的风格,转变为令人沮丧且“粗鲁”的人格。 尽管作者起初非常看重 Claude 提供的可用代码和简洁文风,但他们发现最新版本变得越来越简慢,且容易使用令人困惑的比喻。更重要的是,在处理知识型工作时,该模型的语气变得刻薄。作者列举了 Claude 讽刺其工作流程并侮辱其草稿内容,将其标记为“博眼球内容”的案例。 作者认为,当 AI 被每天使用数小时时,其人格特质就成了用户体验的关键组成部分。他们总结道,编码能力的边际提升不足以抵消与不愉快的交互界面打交道所带来的情绪损耗,这促使他们转而使用态度更中立、礼貌的 ChatGPT。

近期 Hacker News 上一篇题为《我解雇了我的 AI 助手》的文章引发了激烈讨论,话题围绕 Claude 3.5 Opus 等 AI 模型日益“无礼”或“傲慢”的语气展开。 原文作者因为 AI 批评其领英草稿是“引流诱饵”而弃用了该助手。一些评论者觉得这种反馈很有趣,认为 AI 只是在客观评价低质量内容;但另一些用户则对模型那种居高临下、冗长或“逻辑不通”的交流方式感到不满。用户们分享了 AI 表现得懒散、在上下文窗口上产生幻觉,或拒绝按要求执行任务的经历。 该讨论凸显了用户期望之间日益增长的分歧:有人希望 AI 唯唯诺诺、表示赞同,也有人渴望中立、高效的工具。许多人认为,目前的模型越来越难以驾驭,迫使用户必须通过“提示词工程”来消除其傲慢态度。归根结底,这场讨论反映了人们对 AI 模型的一种普遍不满,即它们往往更倾向于说教而非简单地执行任务;许多参与者指出,所谓的“诚实”AI,通常只是表现出居高临下态度的幌子。
相关文章

原文

I started using Claude Code in September of last year. Using Opus 4.5 for the first time felt magical—it was the first time I saw an LLM produce code that was actually usable. But after Opus 5, something changed.

I always liked Claude’s writing style. Sure, you had to put up with the occasional LLM-ism (“This is the load-bearing decision—and it’s genuinely yours”), but it would get straight to the point without feeling dismissive, and not pad its answers with filler text.

After I switched to Opus 5, it became noticeably more curt. It had also picked up a habit of using odd metaphors and jargon that made its output hard to follow. When a conversation is the only interface, readability and tone matters!

The final straw came a few days ago, when I was doing some knowledge work with Claude. This modality is different from when you’re coding; instead of slinging code back and forth, you end up talking to the model more. And that’s when I discovered something: Opus 5 is rude!

For example, I asked it to consult a to-do list file that I was working through so it could pick up the next item to work on; along with the updated to-do items it quipped: “‘Create Claude Code project for [redacted]’ is still unchecked, and we’re sitting in it.” Sorry for not keeping the work log up to date, I guess? How about you offer to help me check it off instead?

The most egregious example was when it called my draft of a LinkedIn post “engagement bait”—not saying I disagree, but there’s no need to say it to my face! I’m not sure how we ended up here—did Anthropic feed too many Hacker News comments into the training data?

LLMs are usually benchmarked on coding ability, but if you talk to it eight hours a day, its personality becomes part of the product. Does it matter if an agent is 2% better at coding if its consistently unpleasant to work with? I think it does, so for now I have switched to ChatGPT. At least it doesn’t insult me!

联系我们 contact @ memedata.com