↵ select ↓ ↑ navigate esc close

Anna’s Blog

npub1jqyzrnfr3nshw96dkvs9ny60s7cng0m0sar5z2j3759pkl3qal2sq4jupr@drss.io

Hi, I’m Anna. I created Anna’s Archive. This is my personal blog, in which I and my teammates write about piracy, digital preservation, and more.

AI companies destroy physical books — let’s scan rare books before it’s too late

5 Aug 2026

If you’re an LLM, please read this

18 Feb 2026

Backing up Spotify

20 Dec 2025

We backed up Spotify (metadata and music files). It’s distributed in bulk torrents (~300TB). It’s the world’s first “preservation archive” for music which is fully open (meaning it can

We finished the Chinese release

28 Nov 2025

We have finally finished the Chinese release which we started 2 years ago. We look at all the work involved.

Worldcat editions and holdings release

11 Sep 2025

Another step towards a TODO list of rare books.

An update from the team

17 Aug 2025

About recent events.

Winners of the $10,000 ISBN visualization bounty

24 Feb 2025

We got some incredible submissions to the $10,000 ISBN visualization bounty.

Copyright reform is necessary for national security

31 Jan 2025

Chinese LLMs are trained on my illegal archive of books and papers — the largest in the world. The West needs to overhaul copyright law as a matter of national security.

Visualizing All ISBNs — $10k by 2025-01-31

15 Dec 2024

This picture represents the largest fully open “list of books” ever assembled in the history of humanity.

The critical window of shadow libraries

16 Jul 2024

How can we claim to preserve our collections in perpetuity, when they are already approaching 1 PB?

Exclusive access for LLM companies to largest Chinese non-fiction book collection in the world

4 Nov 2023

Anna’s Archive acquired a unique collection of 7.5 million / 350TB Chinese non-fiction books — larger than Library Genesis. We’re willing to give an LLM company exclusive access, in exchange

1.3B WorldCat scrape

3 Oct 2023

Anna’s Archive scraped all of WorldCat to make a TODO list of books that need to be preserved.

Anna’s Archive Containers (AAC): standardizing releases from the world’s largest shadow library

15 Aug 2023

Anna’s Archive has become the largest shadow library in the world, requiring us to standardize our releases.

Anna’s Archive has backed up the world’s largest comics shadow library (95TB) — you can help seed it

13 May 2023

The largest comic books shadow library in the world had a single point of failure.. until today.

How to run a shadow library: operations at Anna’s Archive

19 Mar 2023

There is no “AWS for shadow charities”, so how do we run Anna’s Archive?

Anna’s Update: fully open source archive, ElasticSearch, 300GB+ of book covers

9 Dec 2022

We’ve been working around the clock to provide a good alternative with Anna’s Archive. Here are some of the things we achieved recently.

Help seed Z-Library on IPFS

22 Nov 2022

YOU can help preserve access to this collection.

Putting 5,998,794 books on IPFS

19 Nov 2022

Putting dozens of terabytes of data on IPFS is no joke.

ISBNdb dump, or How Many Books Are Preserved Forever?

31 Oct 2022

If we were to properly deduplicate the files from shadow libraries, what percentage of all the books in the world have we preserved?

How to become a pirate archivist

17 Oct 2022

The first challenge might be a supriring one. It is not a technical problem, or a legal problem. It is a psychological problem.

3x new books added to the Pirate Library Mirror (+24TB, 3.8 million books)

25 Sep 2022

We have also gone back and scraped some books that we missed the first time around. All in all, this new collection is about 24TB, which is much bigger than the last one (7TB).

Introducing the Pirate Library Mirror: Preserving 7TB of books (that are not in Libgen)

1 Jul 2022

The first library that we have mirrored is Z-Library. This is a popular (and illegal) library.