{"id":887300,"date":"2026-09-07T23:46:24","date_gmt":"2026-09-07T23:46:24","guid":{"rendered":"https:\/\/www.newsbeep.com\/ca\/887300\/"},"modified":"2026-09-07T23:46:24","modified_gmt":"2026-09-07T23:46:24","slug":"openais-chatgpt-was-built-on-concealed-mass-piracy-authors-tell-court-torrentfreak","status":"publish","type":"post","link":"https:\/\/www.newsbeep.com\/ca\/887300\/","title":{"rendered":"OpenAI&#8217;s ChatGPT Was Built on Concealed &#8216;Mass Piracy&#8217;, Authors Tell Court * TorrentFreak"},"content":{"rendered":"<p><img fetchpriority=\"high\" decoding=\"async\" src=\"https:\/\/www.newsbeep.com\/ca\/wp-content\/uploads\/2026\/09\/openai-logo.jpg\" alt=\"openai logo\" width=\"300\" height=\"188\" class=\"alignright size-full wp-image-246884\"  \/>Over the past three years, authors have filed a series of lawsuits accusing AI companies of training their models on pirated books.<\/p>\n<p>Some of those cases have already produced rulings, with a bittersweet victory for <a href=\"https:\/\/torrentfreak.com\/meta-secures-bittersweet-fair-use-victory-in-ai-piracy-case-250626\/\" rel=\"nofollow noopener\" target=\"_blank\">Meta<\/a> in California for example. <\/p>\n<p>In New York, several other cases were bundled into a single proceeding where Judge Sidney Stein is overseeing claims against OpenAI and Microsoft. <\/p>\n<p>This includes the Authors Guild\u2019s class action, a case filed by a group of nonfiction writers who were the first to name Microsoft as a defendant, and the <a href=\"https:\/\/torrentfreak.com\/authors-accuse-openai-of-using-pirate-sites-to-train-chatgpt-230630\/\" rel=\"nofollow noopener\" target=\"_blank\">Tremblay and Silverman lawsuit<\/a>, which started in California in 2023 and <a href=\"https:\/\/torrentfreak.com\/court-dismisses-authors-copyright-infringement-claims-against-openai-240213\/\" rel=\"nofollow noopener\" target=\"_blank\">survived a partial dismissal<\/a> before moving to New York.<\/p>\n<p>This week, these authors filed a motion for summary judgment. Ahead of any trial, they want Judge Stein to rule that OpenAI copied their work without permission, and that this can\u2019t qualify as fair use. The motion covers 194 titles and asks for a finding of liability, not damages.<\/p>\n<p>\u201cOpenAI\u2019s GPT models pose an existential threat to those who write and publish books,\u201d the brief states, while adding that \u201cAI-generated books of all types are already flooding the market.\u201d<\/p>\n<p>Built on Mass Piracy<\/p>\n<p>The authors start by accusing OpenAI of obtaining the book copies through unauthorized sources. While the filing is heavily redacted, OpenAI stands accused of using torrented copies downloaded from LibGen,<\/p>\n<p>\u201cOpenAI did not even buy the books it used. Instead, it began by torrenting [REDACTED] books from the notorious and illegal pirate library Library Genesis, also known as LibGen,\u201d the motion reads.<\/p>\n<p>At the time, LibGen had already been featured in the U.S. Trade Representative\u2019s <a href=\"https:\/\/torrentfreak.com\/us-government-targets-pirate-bay-and-other-piracy-havens-161221\/\" rel=\"nofollow noopener\" target=\"_blank\">list of notorious piracy markets<\/a>. According to the authors, OpenAI was well aware of the controversial nature of the site.<\/p>\n<p> OpenAI \u201ctook steps to conceal their piracy from the public,\u201d the motion notes, pointing to the paper that introduced GPT-3. In that paper, OpenAI relabeled book compilations it previously called \u201cLibgen1\u201d and \u201cLibgen 2\u201d as the more \u201cnondescript\u201d \u201cBooks1\u201d and \u201cBooks2.\u201d <\/p>\n<p>\u201cOpenAI employees understood at the time that they had sourced books from an illegal site,\u201d the filing reads.<\/p>\n<p>Concealed<br \/><img loading=\"lazy\" decoding=\"async\" src=\"https:\/\/www.newsbeep.com\/ca\/wp-content\/uploads\/2026\/09\/conceal-1.png\" alt=\"concealed\" width=\"600\" height=\"433\" class=\"alignnone size-full wp-image-280535\"  \/><\/p>\n<p>The renaming was not the end of it. OpenAI \u201cdeleted its LibGen files in the summer of 2022 due to legal concerns,\u201d the motion notes, adding that these are \u201cthe only two training corpuses OpenAI has ever deleted.\u201d<\/p>\n<p>Before deleting the books, OpenAI allegedly used them to train the early GPT models. Or as the authors write, the company \u201cbuilt the foundations of its business on mass piracy.\u201d<\/p>\n<p>Replacing George R.R. Martin<\/p>\n<p>The torrenting and piracy angle is one part of the filing. The motion also alleged that OpenAI built its models to replace the human writers it copied, and as evidence it highlights controversial tweets from a key employee. <\/p>\n<p>In 2022, OpenAI hired Tarun Gogineni to lead its work on the writing quality of its models. According to the motion, Gogineni knew the models he was training would displace authors but considered that \u201cacceptable economic disruption.\u201d<\/p>\n<p>This is notable because Gogineni specifically mentioned one of the plaintiffs, author George R.R. Martin, known for writing A Song of Ice and Fire which the HBO series Game of Thrones was based on. <\/p>\n<p>In 2025, nearly two years after Martin sued, Gogineni tweeted that his \u201cresearch mission\u201d was to have GPT models write the \u201clast two books of [Martin\u2019s] A Song of Ice and Fire.\u201d<\/p>\n<p>Even if\u2026<br \/><img loading=\"lazy\" decoding=\"async\" src=\"https:\/\/www.newsbeep.com\/ca\/wp-content\/uploads\/2026\/09\/rrmartin.png\" alt=\"martin\" width=\"600\" height=\"337\" class=\"alignnone size-full wp-image-280539\"  \/><\/p>\n<p>Even if Martin \u201cdies early, GPT-5 will autocomplete his series,\u201d he added, suggesting that AI can replace the author.<\/p>\n<p>Not Fair Use<\/p>\n<p>OpenAI and other AI companies argue that training models on books is fair use. Courts have partly agreed with this, but with an important caveat.<\/p>\n<p>The authors cite Bartz v. Anthropic, the 2025 California ruling that classified model training as potentially fair use, while stressing that downloading from a pirate library was not. Pirating books that can be purchased legally is \u201cinherently, irredeemably infringing,\u201d that court found.<\/p>\n<p>The authors also argue that the copying was avoidable for training purposes, as their books were not per se necessary to create a general-purpose model.<\/p>\n<p>Broader Claims<\/p>\n<p>The motion is not limited to OpenAI. It also asks the court to hold that Microsoft is vicariously liable for OpenAI\u2019s copyright infringement, since Microsoft could supervise the conduct and profited from it. <\/p>\n<p>Microsoft invested roughly $13 billion across three agreements signed in 2019, 2021, and 2023, the authors stress.<\/p>\n<p><img loading=\"lazy\" decoding=\"async\" src=\"https:\/\/www.newsbeep.com\/ca\/wp-content\/uploads\/2026\/09\/conclusion-1.png\" alt=\"conclusion\" width=\"600\" height=\"447\" class=\"alignnone size-full wp-image-280536\"  \/><\/p>\n<p>OpenAI has yet to respond to the authors directly, but it clearly believes that the evidence points in its favor. <\/p>\n<p>In a cross-motion for summary judgment, filed on the same day, the company argues that its use of the books was fair use as a matter of law and that any regurgitation is vanishingly rare.<\/p>\n<p>The filings highlighted here are part of a much broader push. Over the past days, plaintiffs <a href=\"https:\/\/torrentfreak.com\/newspapers-sue-openai-for-copyright-infringement-and-fake-news-240501\/\" rel=\"nofollow noopener\" target=\"_blank\">including The New York Times<\/a>, Daily News, and the Center for Investigative Reporting all submitted a combined summary judgment motion of their own against OpenAI and Microsoft.<\/p>\n<p>With many millions of dollars at stake, as well as the future of AI training, these cases will be fought tooth and nail, so we certainly haven\u2019t heard the last of it. <\/p>\n<p>\u2014<\/p>\n<p>A copy of the authors\u2019 redacted motion for partial summary judgment is available <a href=\"https:\/\/torrentfreak.com\/images\/motionsforpart.pdf\" rel=\"nofollow noopener\" target=\"_blank\">here (pdf)<\/a>, filed at the U.S. District Court for the Southern District of New York.<\/p>\n","protected":false},"excerpt":{"rendered":"Over the past three years, authors have filed a series of lawsuits accusing AI companies of training their&hellip;\n","protected":false},"author":2,"featured_media":887301,"comment_status":"","ping_status":"","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[20],"tags":[62,276,277,49,48,63,278,61],"class_list":["post-887300","post","type-post","status-publish","format-standard","has-post-thumbnail","category-artificial-intelligence","tag-ai","tag-artificial-intelligence","tag-artificialintelligence","tag-ca","tag-canada","tag-microsoft","tag-openai","tag-technology"],"_links":{"self":[{"href":"https:\/\/www.newsbeep.com\/ca\/wp-json\/wp\/v2\/posts\/887300","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.newsbeep.com\/ca\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.newsbeep.com\/ca\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/ca\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/ca\/wp-json\/wp\/v2\/comments?post=887300"}],"version-history":[{"count":0,"href":"https:\/\/www.newsbeep.com\/ca\/wp-json\/wp\/v2\/posts\/887300\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.newsbeep.com\/ca\/wp-json\/wp\/v2\/media\/887301"}],"wp:attachment":[{"href":"https:\/\/www.newsbeep.com\/ca\/wp-json\/wp\/v2\/media?parent=887300"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.newsbeep.com\/ca\/wp-json\/wp\/v2\/categories?post=887300"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.newsbeep.com\/ca\/wp-json\/wp\/v2\/tags?post=887300"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}