{"id":1325,"date":"2026-08-07T00:30:10","date_gmt":"2026-08-07T00:30:10","guid":{"rendered":"https:\/\/redzine.co.uk\/index.php\/2026\/08\/07\/online-shadow-libraries-and-why-theyre-being-implicated-in-ai-copyright-lawsuits\/"},"modified":"2026-08-07T00:30:10","modified_gmt":"2026-08-07T00:30:10","slug":"online-shadow-libraries-and-why-theyre-being-implicated-in-ai-copyright-lawsuits","status":"publish","type":"post","link":"https:\/\/redzine.co.uk\/index.php\/2026\/08\/07\/online-shadow-libraries-and-why-theyre-being-implicated-in-ai-copyright-lawsuits\/","title":{"rendered":"Online shadow libraries \u2013 and why they\u2019re being implicated in AI copyright lawsuits"},"content":{"rendered":"<figure><img decoding=\"async\" src=\"https:\/\/images.theconversation.com\/files\/750650\/original\/file-20260728-102-b1092u.jpg?ixlib=rb-4.1.1&amp;rect=300%2C0%2C3240%2C2160&amp;q=45&amp;auto=format&amp;w=1050&amp;h=700&amp;fit=crop\" \/><figcaption><span class=\"caption\"><\/span> <span class=\"attribution\"><a class=\"source\" href=\"https:\/\/www.shutterstock.com\/image-photo\/convenient-library-computer-770095483?trackingId=f08ed0e6-752b-42a4-ab34-c45cc5240f15&amp;listId=searchResults\">won gou choi\/Shutterstock<\/a><\/span><\/figcaption><\/figure>\n<div style=\"width: 100%;height: 200px;margin-bottom: 20px;border-radius: 6px;overflow: hidden\">\n<\/div>\n<\/p>\n<p>A few days before Christmas 2025, a secretive online activist group called Anna\u2019s Archive said it had downloaded roughly 86 million audio files and \u200a256 million rows of track metadata from Spotify\u2019s catalogue. <\/p>\n<p>\u200aAnna\u2019s Archive is a shadow library. It  holds one of the world\u2019s largest online collection of pirated books, academic papers and now music. Soon, batches of the files, <a href=\"https:\/\/cyberinsider.com\/annas-archive-releases-massive-300tb-spotify-music-scrape\/\">300 terrabytes in total<\/a>, started circulating on torrent file-sharing websites. <\/p>\n<p>By April, a district judge in New York had ordered Anna\u2019s Archive to <a href=\"https:\/\/www.musicbusinessworldwide.com\/spotify-and-record-labels-win-322m-default-judgment-against-pirate-site-annas-archive\/\">pay US$322 million dollars in damages to Spotify<\/a> and three major record labels in a copyright infringement case over the hack. Then following month, the same judge ordered <a href=\"https:\/\/www.publishersweekly.com\/pw\/by-topic\/digital\/copyright\/article\/100463-court-rules-against-anna-s-archive-in-copyright-lawsuit.html\">Anna\u2019s Archive pay $19.5 million<\/a> in damages to a group of 13 major publishers for illegally copying and distributing works they\u2019ve published. The judgement also ordered internet providers to block access to the site.<\/p>\n<p>But the people behind Anna\u2019s Archive didn\u2019t show up in court, and these rulings will be very difficult to enforce. <\/p>\n<p>In this episode of <a href=\"https:\/\/theconversation.com\/topics\/the-conversation-weekly-98901\">The Conversation Weekly<\/a> podcast, Bal\u00e1zs Bod\u00f3, a professor of information law and policy at the University of Amsterdam in the Netherlands, takes us inside the history of these shadow libraries to explain how thousands of copies of books and academic articles ended up in these online archives. <\/p>\n<blockquote>\n<p>Anna\u2019s Archive, is just the latest iteration of this endless stream of people who from since the beginning of time, try to archive everything, says Bod\u00f3.<\/p>\n<p>During the last 20 to 30 years, there have been countless efforts to build digital libraries. Shadow libraries, we call them, because they are text collections, but they are not sanctioned officially.<\/p>\n<\/blockquote>\n<p>With a number of AI companies currently facing copyright lawsuits over allegations they\u2019ve used shadow libraries to train their large language models, we ask what this means for the future of these online archives.  <\/p>\n<div style=\"width: 100%;height: 200px;margin-bottom: 20px;border-radius: 6px;overflow: hidden\">\n<\/div>\n<p><em>Listen to Bodo on <a href=\"https:\/\/pod.link\/1550643487\">The Conversation Weekly<\/a> podcast.<\/em> <\/p>\n<h2>Disclosure statement<\/h2>\n<p><em>Bal\u00e1zs Bod\u00f3 has received ERC funding. He was the project lead for Creative Commons Hungary and a member of Hungary\u2019s National Copyright Expert Group. He has advised several public and private institutions on digital archives, content distribution, online communities, business development.<\/em> <\/p>\n<h2>Credits<\/h2>\n<p><em>This episode of The Conversation Weekly was written and produced by Gemma Ware and Mend Mariwany with editing help from Ashlynne McGhee. Mixing by Michelle Macklem and theme music by Neeta Sarl.<\/em><\/p>\n<p><em>Newsclips in this episode from <a href=\"https:\/\/www.youtube.com\/watch?v=wRtiMtl9bCw\">India Today<\/a>, <a href=\"https:\/\/www.youtube.com\/watch?v=Gb9TJMDNiM4\">CBS News<\/a>, <a href=\"https:\/\/www.youtube.com\/watch?v=o4Pgirl3zAM\">CBS Miami<\/a> and <a href=\"https:\/\/www.youtube.com\/watch?v=tVpIt0Q35uc\">NBC News<\/a>.<\/em><\/p>\n<p><em>Listen to The Conversation Weekly via any of the apps listed above, download it directly via our <a href=\"https:\/\/feeds.captivate.fm\/the-conversation-weekly\/\">RSS feed<\/a> or find out <a href=\"https:\/\/theconversation.com\/how-to-listen-to-the-conversations-podcasts-154131\">how else to listen here<\/a>. A transcript of this episode is available via the Apple Podcasts or Spotify apps.<\/em><\/p>\n<p><img loading=\"lazy\" decoding=\"async\" src=\"https:\/\/counter.theconversation.com\/content\/288557\/count.gif\" alt=\"The Conversation\" width=\"1\" height=\"1\" \/><\/p>\n","protected":false},"excerpt":{"rendered":"<p>won gou choi\/Shutterstock A few days before Christmas 2025, a secretive online activist group called Anna\u2019s Archive said it had downloaded roughly 86 million audio files and \u200a256 million rows of track metadata from Spotify\u2019s catalogue. \u200aAnna\u2019s Archive is a shadow library. It holds one of the world\u2019s largest online collection of pirated books, academic [&hellip;]<\/p>\n","protected":false},"author":1,"featured_media":0,"comment_status":"open","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[1],"tags":[],"class_list":["post-1325","post","type-post","status-publish","format-standard","hentry","category-uncategorized"],"_links":{"self":[{"href":"https:\/\/redzine.co.uk\/index.php\/wp-json\/wp\/v2\/posts\/1325","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/redzine.co.uk\/index.php\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/redzine.co.uk\/index.php\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/redzine.co.uk\/index.php\/wp-json\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/redzine.co.uk\/index.php\/wp-json\/wp\/v2\/comments?post=1325"}],"version-history":[{"count":0,"href":"https:\/\/redzine.co.uk\/index.php\/wp-json\/wp\/v2\/posts\/1325\/revisions"}],"wp:attachment":[{"href":"https:\/\/redzine.co.uk\/index.php\/wp-json\/wp\/v2\/media?parent=1325"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/redzine.co.uk\/index.php\/wp-json\/wp\/v2\/categories?post=1325"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/redzine.co.uk\/index.php\/wp-json\/wp\/v2\/tags?post=1325"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}