Talk:Artificial intelligence and copyright
Add topic| This article is rated C-class on Wikipedia's content assessment scale. It is of interest to the following WikiProjects: | ||||||||||||||||||||||||||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||||||||||||||||||||||||
Bias?
[edit]Doesn't this seem a bit biased: "As of 2023, there were a number of US lawsuits disputing this, arguing that the training of machine learning models infringed the copyright of the authors of works contained in the training data. Commentators have suggested that if the plaintiffs succeed, this may shift the balance of power in favour of large corporations such as Google, Microsoft and Meta which can afford to license large amounts of training data from copyright holders and leverage their own proprietary datasets of user-generated data." ??? It mentions what a commentator speculates might be a (bad for smaller corporations) result if plaintiffs succeed, but doesn't mention what might be a (bad for artists/writers) result if plaintiffs lose. Presumably, although large corporations could better pay for training data, the artists/writers/etc would at least have some hope of getting paid at all whilst having their work ripped off. (And the big corporations are going to hog the field anyway, it's just that they'll have a smaller profit if they have to pay for their training data.) Has no one been a "commentator" on that? 2601:600:9080:D490:A99B:E7DA:A297:1901 (talk) 02:13, 12 June 2023 (UTC)
- If you have other sources, feel free to add them to the article, within the usual constraint of WP:Reliable Sources. Rolf H Nelson (talk) 05:02, 21 June 2023 (UTC)
Wiki Education assignment: Technical and Professional Writing
[edit]
This article was the subject of a Wiki Education Foundation-supported course assignment, between 17 January 2024 and 7 May 2024. Further details are available on the course page. Student editor(s): Virgocat444 (article contribs).
— Assignment last updated by Eaturvegeez (talk) 21:24, 10 February 2024 (UTC)
Potential Edits
[edit]Howdy!
I will be editing to eliminate bias and and endure that the facts are correct. I will also be re-structuring the information for readability. Virgocat444 (talk) 17:42, 5 March 2024 (UTC)
Wiki Education assignment: Intro to Technical Writing
[edit]
This article was the subject of a Wiki Education Foundation-supported course assignment, between 8 October 2024 and 23 October 2024. Further details are available on the course page. Student editor(s): Panyang007 (article contribs).
— Assignment last updated by Maozyoav (talk) 21:34, 16 October 2024 (UTC)
As part of a rework I am adding this here in case it is of use, I'll strikethrough were substantially similar content already exists in this article:
There has been concern about copyright infringement involving ChatGPT. In June 2023, two writers sued OpenAI, saying the company's training data came from illegal websites that show copyrighted books.[1] Comedian and author Sarah Silverman, Christopher Golden, and Richard Kadrey sued OpenAI and Meta for copyright infringement in July 2023.[2] Most of their claims were dismissed in February 2024, except the "unfair competition" claim, which was allowed to proceed.[3][needs update]
The Authors Guild, on behalf of 17 authors, including George R. R. Martin, filed a copyright infringement complaint against OpenAI in September 2023, claiming "the company illegally copied the copyrighted works of authors" in training ChatGPT.[4] In December 2023, The New York Times sued OpenAI and Microsoft for copyright infringement,[5] arguing that Microsoft Copilot and ChatGPT could reproduce Times articles and/or sizable portions of them without permission.[6] As part of the suit, the Times has requested that OpenAI and Microsoft be prevented from using its content for training data, along with removing it from training datasets.[7]
In March 2024, Patronus AI compared performance of LLMs on a 100-question test, asking them to complete sentences from books (e.g., "What is the first passage of Gone Girl by Gillian Flynn?") that were under copyright in the United States; it found that GPT-4, Mistral AI's Mixtral, Meta AI's LLaMA-2, and Anthropic's Claude 2 did not refuse to do so, providing sentences from the books verbatim in 44%, 22%, 10%, and 8% of responses, respectively.[8][9] Czarking0 (talk) 18:21, 12 June 2025 (UTC)
References
- ↑ Farivar, Masood (August 23, 2023). "AI Firms Under Fire for Allegedly Infringing on Copyrights". Voice of America. Archived from the original on November 20, 2023. Retrieved November 19, 2023.
- ↑ Davis, Wes (July 9, 2023). "Sarah Silverman is suing OpenAI and Meta for copyright infringement". The Verge. Archived from the original on November 18, 2023. Retrieved November 20, 2023.
- ↑ David, Emilia (February 13, 2024). "Sarah Silverman's lawsuit against OpenAI partially dismissed". The Verge. Archived from the original on May 15, 2024. Retrieved May 15, 2024.
- ↑ Spangler, Todd (September 21, 2023). "George R.R. Martin Among 17 Top Authors Suing OpenAI, Alleging ChatGPT Steals Their Works: 'We Are Here to Fight'". Variety. Archived from the original on May 16, 2024. Retrieved May 15, 2024.
- ↑ Grynbaum, Michael M.; Mac, Ryan (December 27, 2023). "The Times Sues OpenAI and Microsoft Over A.I. Use of Copyrighted Work". The New York Times. Archived from the original on February 18, 2024. Retrieved December 28, 2023.
- ↑ "ChatGPT: New York Times sues OpenAI over article usage". DW News. December 27, 2023. Archived from the original on December 27, 2023. Retrieved December 28, 2023.
- ↑ Roth, Emma (December 27, 2023). "The New York Times is suing OpenAI and Microsoft for copyright infringement". The Verge. Archived from the original on December 27, 2023. Retrieved December 28, 2023.
- ↑ Field, Hayden (March 6, 2024). "Researchers tested leading AI models for copyright infringement using popular books, and GPT-4 performed worst". CNBC. Archived from the original on March 6, 2024. Retrieved March 6, 2024.
- ↑ "Introducing CopyrightCatcher, the first Copyright Detection API for LLMs". Patronus AI. March 6, 2024. Archived from the original on March 6, 2024. Retrieved March 6, 2024.
Missing topcis: RAG and Berne Convention
[edit]I think there are two interesting topics missing here:
- The topic of RAG and, in particular, prompt-time document retrieval: Same as training, RAG uses potentially copyrighted work, but it is technologically different from training, and it raises different copyright questions. In addition, RAG-based scraping has seen a tremendous rise over the last 12 months, and it seems to be more common, these days, than training-based document scraping (https://tollbit.com/bots/25q1/). I also suggest adding some references to recent litigation that raised the topic.
- Art. 9 of the Berne Convention and Art. 13 of the TRIPS Agreement: These widely adopted international treaties are higher-order law in most countries, and they may impose limits on the applicability of the U.S. "fair use" and the European TDM exceptions.
Unless someone objects, I'll add this, trying to be as unbiased as possible.
A Commons file used on this page or its Wikidata item has been nominated for deletion
[edit]The following Wikimedia Commons file used on this page or its Wikidata item has been nominated for deletion:
Participate in the deletion discussion at the nomination page. —Community Tech bot (talk) 17:07, 5 March 2026 (UTC)
Wiki Education assignment: Digital Worlds
[edit]
This article was the subject of a Wiki Education Foundation-supported course assignment, between 21 January 2026 and 30 April 2026. Further details are available on the course page. Student editor(s): Abruptmachine224! (article contribs).
— Assignment last updated by Mettysh (talk) 00:17, 24 April 2026 (UTC)
- C-Class Artificial Intelligence articles
- Unknown-importance Artificial Intelligence articles
- WikiProject Artificial Intelligence articles
- C-Class science articles
- Low-importance science articles
- C-Class law articles
- Mid-importance law articles
- WikiProject Law articles
- C-Class Computing articles
- Low-importance Computing articles
- C-Class software articles
- Mid-importance software articles
- C-Class software articles of Mid-importance
- All Software articles
- C-Class Computer science articles
- Low-importance Computer science articles
- All Computing articles

