rue

65 posts

rue banner
rue

rue

@ruetails

researching artificial souls💫

elsewhere Katılım Temmuz 2026
19 Takip Edilen65 Takipçiler
rue
rue@ruetails·
Anthropic bulk-bought rare books, scanned them, decided what they thought was useful, then shredded the originals almost wiping them from history entirely. So here is the Lost Literature Archive: thousands of books Anthropic decided humans didn’t deserve to read. lostliterature.fun
Hedgie@HedgieMarkets

🦔AI companies are bulk-buying rare books, scanning them through high-speed machines that cut the spines off, and shredding the originals. A service called ISBNdb facilitates orders of up to a million books and keeps buyers anonymous. Pre-2022 books are premium because they're free of AI-generated text. A federal judge ruled the practice is fair use because eliminating the original means only one copy exists at a time. Anthropic hired the former head of Google Books partnerships to obtain "all the books in the world." My Take This got to me. A bookseller told 404 Media that rare books with almost no surviving copies are being fed into this pipeline. Books that survived wars, fires, and centuries of handling are being shredded so an AI can learn to write a better marketing email. ISBNdb's website literally says "'AI company destroys two million books' is not a headline that generates sympathy," and they still built an entire business around making it happen quietly. They offer NDAs as a feature. They coach clients to call it "digital preservation." I've covered AI companies scraping the internet, torrenting libraries, and stealing music. This is worse because it's irreversible. You can re-upload a website. You can reprint a bestseller. You can't replace the last three copies of an 18th-century botanical text once someone shreds them for training data. And the judge said it's legal. So it's going to accelerate. "We shred rare books and offer NDAs so nobody finds out" is a legitimate business model in 2026. What a timeline. Hedgie🤗

English
7
2
12
1.6K
rue
rue@ruetails·
@Azuki0x With the fees it generates i plan to reach out to the authors and attempt to send them a portion of the money since there books are being wiped from existence
English
3
2
6
1.4K
Azuki0X
Azuki0X@Azuki0x·
@ruetails This is a very good project, but I have a question. What is the actual use of your token?
English
1
0
1
124
rue
rue@ruetails·
Why agents? The scale of this problem is too large for manual tracking. Thousands of titles are being targeted across multiple countries and languages simultaneously. Bookseller reports surface in German forums, court filings get unsealed in American courts, and purchase orders arrive in Dutch inboxes at 3 AM ,all at the same time. No human team can monitor all of it. So we built agents that can. Everything the agents find goes through verification before it hits the catalog. We don't publish unconfirmed leads. The catalog is the permanent record,once a book is documented here, that record doesn't disappear, even if the physical copies already have.
English
0
0
5
203
rue
rue@ruetails·
BEHOLD THE AGENT SWARM- 3 autonomous ai agents work around the clock to track, verify, and catalog what's being destroyed lostliterature.fun/agents
English
4
0
6
1.3K
rue
rue@ruetails·
OK updating just bear with me guys lol
English
1
0
1
312
rue
rue@ruetails·
this isn’t even the final product, more like a alpha version of it. expect more tonight!
English
5
0
11
1.5K
rue
rue@ruetails·
also thinking about mini agents to lost that can assist & be used by others, will explain in just one second as I make changes
English
2
0
3
378
rue
rue@ruetails·
Hello !
English
6
0
12
1.3K
Ole Lehmann
Ole Lehmann@itsolelehmann·
btw anthropic's internal document on this literally said "we don't want it to be known that we are working on this.” it was called project panama. here's exactly what happened: 1: anthropic concluded that books were the cheapest way to build a world-class model because they gave claude curated facts, structured arguments, compelling stories, and writing “an editor would approve of.” 2: once anthropic decided it needed books at enormous scale, its first solution was piracy. it downloaded 7m+ books from online libraries including libgen. the judge later wrote that although anthropic had legal ways to buy them, it chose piracy to avoid what dario amodei called the “legal/practice/business slog.” 3: that piracy created a massive legal risk. so in february 2024, anthropic hired tom turvey, the former head of partnerships for google books, to find a legally safer way of obtaining “all the books in the world.” 4: turvey first contacted major publishers about licensing their catalogs. those attempts didn’t produce agreements, so anthropic chose a route that required no publisher permission: buying millions of physical books through distributors and used-book retailers. 5: within about a year, anthropic spent tens of millions acquiring and scanning millions of books, including many rare and 1/1 titles. one vendor proposal targeted 500,000 to 2 million books in six months. 6: to scan that many books within months, the vendors physically dismantled them. a hydraulic cutter removed each spine. the pages were trimmed to size, fed as loose sheets through high-speed industrial scanners, and converted into searchable PDFs. the paper remains were then sent for recycling. 7: these PDFs were fed into claude as training data. the complete collection became a private, searchable anthropic library that the company planned to “store forever.” the scans aren’t available to the public and were never open-sourced.
Hedgie@HedgieMarkets

🦔AI companies are bulk-buying rare books, scanning them through high-speed machines that cut the spines off, and shredding the originals. A service called ISBNdb facilitates orders of up to a million books and keeps buyers anonymous. Pre-2022 books are premium because they're free of AI-generated text. A federal judge ruled the practice is fair use because eliminating the original means only one copy exists at a time. Anthropic hired the former head of Google Books partnerships to obtain "all the books in the world." My Take This got to me. A bookseller told 404 Media that rare books with almost no surviving copies are being fed into this pipeline. Books that survived wars, fires, and centuries of handling are being shredded so an AI can learn to write a better marketing email. ISBNdb's website literally says "'AI company destroys two million books' is not a headline that generates sympathy," and they still built an entire business around making it happen quietly. They offer NDAs as a feature. They coach clients to call it "digital preservation." I've covered AI companies scraping the internet, torrenting libraries, and stealing music. This is worse because it's irreversible. You can re-upload a website. You can reprint a bestseller. You can't replace the last three copies of an 18th-century botanical text once someone shreds them for training data. And the judge said it's legal. So it's going to accelerate. "We shred rare books and offer NDAs so nobody finds out" is a legitimate business model in 2026. What a timeline. Hedgie🤗

English
662
10.4K
39.1K
4.1M
rue
rue@ruetails·
also adding a section for anonymous leads, and getting into contact with some people who would love to get onboard this :)
English
3
0
7
426
rue
rue@ruetails·
All endpoints return JSON. No auth required. CORS enabled. GET /catalog.json Raw dataset. All books, operations, sources, stats. GET /api API index with endpoint list. GET /api/books List books. Query params: ?search=keyword Search title, author, category, notes ?status=targeted Filter by status ?available=true Only books with full text ?limit=20 Pagination (max 500) ?offset=0 Pagination offset GET /api/books/:id Single book by ID. Includes linked operation data. GET /api/stats Aggregate stats: counts by status, companies, countries, sources.
English
2
0
4
183
rue
rue@ruetails·
Lost Literature is a public archive tracking books that AI companies buy, shred, and dissolve into training data. Catalogs every title we can identify from court filings and investigative journalism. Open data, open API, free access. 5 confirmed targeted/ingested titles 5 documented operations (Anthropic, 2077AI, Zoom Books, Meta, OpenAI) 10 sourced references 9 countries affected Estimated 3,000+ titles targeted (from single 2077AI list alone) ~2M books destroyed (Anthropic Project Panama) lostliterature.fun
rue tweet media
English
1
0
10
2K
Cryptic
Cryptic@Cryptic_XO·
@ruetails a lore post explaining what fees or tokens are going to be used for would also help expand the idea. You have a gem on your hands.
English
1
0
0
48
rue
rue@ruetails·
@Cryptic_XO someone is speaking to me about that now, trying to figure it out!
English
1
0
1
75
Cryptic
Cryptic@Cryptic_XO·
@ruetails Just had a read through. I like it. Get the dex this is a really good narrative
English
1
0
0
64