如果我拥有 Claude 的输出内容,为什么我不能用它们来训练我自己的模型?
If I own Claude's outputs why can't I train my own model on them?

原始链接: https://support.claude.com/en/articles/12326764-can-i-use-my-outputs-to-train-an-ai-model

尽管您拥有 Claude 生成的输出内容,但 Anthropic 禁止在未经书面许可的情况下将其用于训练或开发其他 AI 模型。 此政策旨在维护安全性并防止竞争性损害。由于 Anthropic 的安全协议不会随之转移到使用 Claude 输出训练的模型上,因此第三方开发可能会产生未经审查且具有潜在危害的系统。此外,Anthropic 亦禁止利用其基础设施构建直接竞争产品。 **允许的行为:** 您可以将输出内容用于创建非竞争性工具,例如内容分类器、信息提取系统,或为您的应用程序、工作流程及内部生产力提供支持的功能。 **禁止的行为:** 严禁将输出内容用于训练与 Anthropic 竞争的模型,或协助他人进行此类行为。具体包括:构建旨在进行开放式文本生成的模型、将输出内容作为训练目标,或试图对 Anthropic 的训练方法进行逆向工程。

最近 Hacker News 上的一场讨论凸显了社区对 Anthropic 的强烈抵制,原因在于该公司禁止用户利用 Claude 的输出结果来训练竞争性 AI 模型。 Anthropic 辩称这些限制是出于安全风险、缺乏监管以及保护其基础设施投资的考虑。然而,许多用户认为这些说法是“AI 公司典型的伪善”。批评者指出,AI 公司在构建自身模型时,往往在未经许可的情况下抓取了海量公共数据(包括版权作品和开源项目),而现在却试图对其声称属于用户“所有”的输出结果强制实施专有控制。 这场辩论还涉及了 AI 生成内容的法律模糊性。许多参与者认为,由于这些输出并非由人类创作,它们可能属于公共领域,因此不受限制性服务条款的约束。另一些人则将这种限制视为企业进行“看门人”式管理但无法执行的尝试,并指出一旦数据被发布或传输,原始公司就失去了对其使用方式的实际控制权。总的来说,该讨论反映了人们对当前 AI 商业实践中“只许州官放火,不许百姓点灯”做法的深深不满。
相关文章

原文

Understanding our policies on using Claude's Outputs for model training and development

When you use Claude, you own the Outputs generated from your Inputs. However, there are important restrictions on using these Outputs to train AI models which are standard practice across the AI industry. We prohibit customers from using our services to train or develop AI models without our written permission. This article explains what uses are permitted, what uses are prohibited, and why these policies exist.

Why we restrict model training

Anthropic invests significantly in making Claude safe, helpful, and harmless. We conduct rigorous pre-release testing, implement multiple safety layers, and continuously monitor our models' behavior. When Outputs are used to train new models without our oversight, additional risks emerge. Safety controls may be lost – models trained on Claude's Outputs won't have our safety measures, potentially leading to harmful or dangerous AI systems. We also have no visibility into deployment, meaning we cannot monitor how these distilled models are used or prevent misuse.

When customers use Claude to generate Outputs that then train competing models, they're essentially using our infrastructure and investment to build direct competitors to our service. Like other software and service providers, we expect that our services won't be used to undermine our product offerings.

What you can do with Outputs

You can use Claude's Outputs to train models that don't compete with Anthropic's own models. This includes creating specialized classifiers and tools such as:

  • Content categorization systems

  • Information extraction tools

Outputs can also be integrated into your applications to power features within your products, generate content for your customers, analyze and structure your data, or improve internal workflows and productivity.

What's prohibited

Our Terms do not allow the use of Outputs to train models that are competitive with Anthropic's own. It is also a violation of our Terms to support a third party's attempt to do the same.

Uses that are prohibited include:

  • Models designed for open-ended text generation

  • Using Outputs as training targets for models

  • Reverse engineering training methods

联系我们 contact @ memedata.com