Codeberg:禁止大语言模型数据抓取的服务条款扩展
Codeberg: ToU extension to prohibit LLM-extrusions

原始链接: https://codeberg.org/Codeberg/org/pulls/1253

讨论的核心在于在开源项目中使用大语言模型(LLM)的风险,特别是关于版权的“法律灰色地带”。支持更严格监管的人士认为,依赖 AI 生成的代码会使开源项目容易受到未来的诉讼或来自 AI 公司的掠夺性索赔,而大多数项目缺乏应对这些诉讼的资源。 除了法律担忧之外,目前还存在一种推动力,旨在界定合乎道德的 AI 使用边界。虽然有些人将 LLM 视为“辅助输入”工具,但对于完全自动化、无需人工干预的“感觉编程”(vibecoding)存在重大担忧。潜在的解决方案包括禁止 AI 自动提交代码或合并请求(PR),并限制在人与人之间的项目沟通中使用 LLM,这类似于 Godot 团队目前的立场。归根结底,社区正在讨论如何在有益的 AI 辅助与威胁开源工作完整性及法律安全的行为之间划清界限。

Codeberg 已更新其使用条款,禁止使用由大语言模型(LLM)生成的内容,这一举动在 Hacker News 上引发了激烈的争论。 支持这一变动的人认为,该政策是针对版权纠纷、潜在恶意软件风险以及维护自由软件完整性等未决问题所采取的必要举措。反对者则认为该禁令目光短浅,他们主张人工智能生成的代码质量终将超越人类,平台应当适应而非抵制。 批评该禁令的人还指出,执行此类规定难度极大;而支持者则强调维护“纯人类”代码库的价值。对人工智能的主导地位持怀疑态度的人则认为,该技术目前尚缺乏必要的可靠性和验证机制,无法在没有人工监督的情况下被信任。这场讨论凸显了软件开发领域中,在拥抱人工智能的快速整合与维持传统的问责、署名及安全标准之间所存在的行业矛盾。
相关文章

原文

Why is the focus exclusively on "copyright status"? Codeberg is meant to be the ethical alternative
@hsza

hi! i'm not on the codeberg team at all, but i wanted to reply my thoughts. The short answer is because, in meme terms "we live in a society". Copyright is an unavoidable fact and LLMs are right now working in a legal gray area which can be problematic, SPECIALLY for open source projects! What if suddenly the AI companies lobby laws to make sure the copyright of LLM-coded projects goes towards them? Or something like asking for percentages of profit, exclusive usage, etc. Even if you were to make your case via the judicial route, how many Open Source projects have the money to fight that? Working in a legal gray area is pracitcally a blank check that allows these companies to force projects to god-knows-what in the future.

also, "mostly" is up to too-lenient interpretation

To the rest of everyone here, I totally agree and am also worried about this! I've been thinking that the best scenario for llms is basically assisted typing. While the worst case is agentic hands-off vibecoding. The bad bit is this is a spectrum definitely, so where do we draw the line? I think for starters LLM-automated commits or PRs could be prohibited. The Godot team is prohibiting the usage of LLMs for human-to-human communication (as well as just generally in the project basically) which I think it's a good idea. I'm quite interested to see how this discussion evolves and particularly where do we draw the line? as I like seeing the possibility of good usages of LLMs. But again, the legal concerns are real and unavoidable

> Why is the focus exclusively on "copyright status"? Codeberg is meant to be the ethical alternative @hsza hi! i'm not on the codeberg team at all, but i wanted to reply my thoughts. The short answer is because, in meme terms "we live in a society". Copyright is an unavoidable fact and LLMs are right now working in a legal gray area which can be problematic, SPECIALLY for open source projects! What if suddenly the AI companies lobby laws to make sure the copyright of LLM-coded projects goes towards them? Or something like asking for percentages of profit, exclusive usage, etc. Even if you were to make your case via the judicial route, how many Open Source projects have the money to fight that? Working in a legal gray area is pracitcally a blank check that allows these companies to force projects to god-knows-what in the future. > also, "mostly" is up to too-lenient interpretation To the rest of everyone here, I totally agree and am also worried about this! I've been thinking that the best scenario for llms is basically assisted typing. While the worst case is agentic hands-off vibecoding. The bad bit is this is a spectrum definitely, so where do we draw the line? I think for starters LLM-automated commits or PRs could be prohibited. The Godot team is prohibiting the usage of LLMs for human-to-human communication (as well as just generally in the project basically) which I think it's a good idea. I'm quite interested to see how this discussion evolves and particularly **where do we draw the line?** as I like seeing the possibility of good usages of LLMs. But again, the legal concerns are real and unavoidable

联系我们 contact @ memedata.com