美国政府就版权材料训练 LLM 争议表态支持 OpenAI

内容摘要
美国政府支持OpenAI在未经许可的情况下使用版权材料训练其大型语言模型(LLM)。在纽约时报对OpenAI提起的诉讼中,特朗普政府提交了一份20页的简报,为ChatGPT等聊天机器人背后的LLM训练提供了辩护。简报强调,美国有强烈的兴趣继续发展强大且具有竞争力的AI产业,并保持在全球AI领域的领导地位。许多出版商,包括纽约时报,认为AI公司如OpenAI在未经许可的情况下使用其版权材料是非法的。然而,简报指出,对公平使用的误解可能会阻碍创新和科学进步,同时阻碍美国的经济繁荣和流动性。目前,关于AI训练和版权侵权的案件大多对AI公司有利。特朗普政府的简报虽然不是裁决,但可能会对案件产生重要影响。
美国政府支持OpenAI在未经许可的情况下使用版权材料训练其大型语言模型(LLM)。在纽约时报对OpenAI提起的诉讼中,特朗普政府提交了一份20页的简报,为ChatGPT等聊天机器人背后的LLM训练提供了辩护。简报强调,美国有强烈的兴趣继续发展强大且具有竞争力的AI产业,并保持在全球AI领域的领导地位。许多出版商,包括纽约时报,认为AI公司如OpenAI在未经许可的情况下使用其版权材料是非法的。然而,简报指出,对公平使用的误解可能会阻碍创新和科学进步,同时阻碍美国的经济繁荣和流动性。目前,关于AI训练和版权侵权的案件大多对AI公司有利。特朗普政府的简报虽然不是裁决,但可能会对案件产生重要影响。

In a lawsuit that The New York Times filed against OpenAI, the Trump administration has contributed a 20-page brief in defense of the ChatGPT maker’s unlicensed use of copyrighted material to train its LLMs.

“The United States has a strong interest in continuing to develop a robust and competitive artificial intelligence industry that sets the standard for the practice and procedure of AI use globally… As such, it is critical for the United States to ‘retain global leadership in artificial intelligence,'” the brief reads, referencing an executive order that President Donald Trump signed last year.

The LLMs powering chatbots like ChatGPT, Claude, and Gemini are trained on incomprehensibly massive databases of published works, including copyrighted books, articles, and other media that AI companies feed into these databases without permission. Many publishers, including The New York Times in this case, have sought to argue that it is illegal for AI companies like OpenAI to train AI models on their copyrighted material.

This question — can you use copyrighted material to train an AI? — isn’t black and white, hence the extensive legal debate around the subject. These conversations often center on fair use, a carve out of copyright law that makes exceptions for certain scenarios when it can be ruled legal to use someone else’s copyrighted work without permission. In this case, the fair use debate addresses whether AI companies’ use of copyrighted work is “transformative” enough for a judge to rule it legal.

“Constraining LLM development under a misunderstanding of fair use doctrine would thwart such creative and scientific progress while hindering American prosperity and economic mobility,” the brief says.

So far, cases about AI training and copyright infringement have largely been favorable to AI companies. Last year, Judge William Alsup ordered Anthropic to pay a $1.5 billion copyright settlement to a group of writers whose works were used to train the company’s AI models; but Anthropic wasn’t dinged for its AI training. Rather, the company was fined for using illegal shadow libraries to pirate the books it used for training.

“Like any reader aspiring to be a writer, Anthropic’s LLMs trained upon works not to race ahead and replicate or supplant them — but to turn a hard corner and create something different,” Judge Alsup wrote, comparing the LLM’s training to a human reading a book.

This new Trump administration brief is not a ruling, as the case is being tried in the U.S. District Court for the Southern District of New York, and the authors of the brief do not have jurisdiction. However, this intervention by the Trump administration could still carry weight.

原始发布方:TechCrunch:AI(RSS)

原文时间:2026-09-03 01:09:06 +08:00

阅读原文 · 数据来源:AIHOT

提示

本文用于信息整理与经验分享。第三方订阅、支付及账号服务可能调整,实际规则、价格和可用性请以下单页面及服务方最新说明为准。

咨询 GPT 充值咨询充值