Themata.AI
Themata.AI

Popular tags:

#developer-tools#ai-agents#discussion#llms#trending#claude#ai-ethics#code-generation#ai-safety#openai

AI is changing the world. Don't stay behind. Clear summaries, community insight, delivered without the noise. Subscribe to never miss a beat.

© 2026 Themata.AI • All Rights Reserved

Archive

|

Topics

|

Privacy

|

Cookies

|

Contact
🕒 Latest🔥 Top

Filtering by tag:

digital-preservationClear
AI companies destroy physical books — let’s scan rare books before it’s too late
ai-ethicsdigital-preservationllmscorporate-responsibility
Opinion

AI companies destroy physical books – let's scan rare books before it's too late

Anna’s Archive volunteer “u” called for worldwide volunteers to scan and upload physical books, periodicals, rare works, and archival materials to shadow libraries before copies are lost or become inaccessible. The August 5, 2026 guest post says small contributors may receive recognition and lifetime Anna’s Archive membership, while the organization can help cover scanning costs and offer other rewards for large-scale book uploads. The post alleges that several AI companies have bought large quantities of secondhand books through intermediaries, scanned them for training data, and destroyed the originals. It says Anthropic launched a confidential initiative called Project Panama in early 2024, spending tens of millions of dollars to purchase and scan millions of paper books for Claude training before destroying them; the post says the project emerged in a $1.5 billion copyright settlement. The volunteer argues that destruction can prevent competitors from obtaining the same books, reduce legal exposure, and cost less than lossless digitization. The post further claims that AI-generated material has accounted for more than half of newly published internet content since the beginning of 2025, and warns that digital preservation of human-created works is becoming more urgent. It calls for contributions through scanning, purchasing material for scanning, or donations.

annas-archive.gl

🔥🔥🔥🔥🔥

3 min

8/21/2026

Blocking Internet Archive Won't Stop AI, but Will Erase Web's Historical Record

The Internet Archive is the largest digital library, preserving over one trillion archived web pages through its Wayback Machine. Recent actions by publishers to block the Internet Archive threaten the preservation of the web's historical record.

eff.org

🔥🔥🔥🔥🔥

3 min

3/21/2026

News publishers limit Internet Archive access due to AI scraping concernsNews

News publishers limit Internet Archive access due to AI scraping concerns

News publishers are restricting access to the Internet Archive due to concerns over AI scraping of their content. The Internet Archive's crawlers capture webpage snapshots, which are accessible via the Wayback Machine, potentially exposing publishers' material to unauthorized use by AI models.

niemanlab.org

🔥🔥🔥🔥🔥

9 min

2/14/2026

AI companies destroy physical books – let's scan rare books before it's too late

Anna’s Archive volunteer “u” called for worldwide volunteers to scan and upload physical books, periodicals, rare works, and archival materials to shadow libraries before copies are lost or become inaccessible. The August 5, 2026 guest post says small contributors may receive recognition and lifetime Anna’s Archive membership, while the organization can help cover scanning costs and offer other rewards for large-scale book uploads. The post alleges that several AI companies have bought large quantities of secondhand books through intermediaries, scanned them for training data, and destroyed the originals. It says Anthropic launched a confidential initiative called Project Panama in early 2024, spending tens of millions of dollars to purchase and scan millions of paper books for Claude training before destroying them; the post says the project emerged in a $1.5 billion copyright settlement. The volunteer argues that destruction can prevent competitors from obtaining the same books, reduce legal exposure, and cost less than lossless digitization. The post further claims that AI-generated material has accounted for more than half of newly published internet content since the beginning of 2025, and warns that digital preservation of human-created works is becoming more urgent. It calls for contributions through scanning, purchasing material for scanning, or donations.

annas-archive.gl

🔥🔥🔥🔥🔥

3 min

8/21/2026

News publishers limit Internet Archive access due to AI scraping concerns

News publishers are restricting access to the Internet Archive due to concerns over AI scraping of their content. The Internet Archive's crawlers capture webpage snapshots, which are accessible via the Wayback Machine, potentially exposing publishers' material to unauthorized use by AI models.

niemanlab.org

🔥🔥🔥🔥🔥

9 min

2/14/2026

Blocking Internet Archive Won't Stop AI, but Will Erase Web's Historical Record

The Internet Archive is the largest digital library, preserving over one trillion archived web pages through its Wayback Machine. Recent actions by publishers to block the Internet Archive threaten the preservation of the web's historical record.

eff.org

🔥🔥🔥🔥🔥

3 min

3/21/2026

AI companies destroy physical books – let's scan rare books before it's too late

Anna’s Archive volunteer “u” called for worldwide volunteers to scan and upload physical books, periodicals, rare works, and archival materials to shadow libraries before copies are lost or become inaccessible. The August 5, 2026 guest post says small contributors may receive recognition and lifetime Anna’s Archive membership, while the organization can help cover scanning costs and offer other rewards for large-scale book uploads. The post alleges that several AI companies have bought large quantities of secondhand books through intermediaries, scanned them for training data, and destroyed the originals. It says Anthropic launched a confidential initiative called Project Panama in early 2024, spending tens of millions of dollars to purchase and scan millions of paper books for Claude training before destroying them; the post says the project emerged in a $1.5 billion copyright settlement. The volunteer argues that destruction can prevent competitors from obtaining the same books, reduce legal exposure, and cost less than lossless digitization. The post further claims that AI-generated material has accounted for more than half of newly published internet content since the beginning of 2025, and warns that digital preservation of human-created works is becoming more urgent. It calls for contributions through scanning, purchasing material for scanning, or donations.

annas-archive.gl

🔥🔥🔥🔥🔥

3 min

8/21/2026

Blocking Internet Archive Won't Stop AI, but Will Erase Web's Historical Record

The Internet Archive is the largest digital library, preserving over one trillion archived web pages through its Wayback Machine. Recent actions by publishers to block the Internet Archive threaten the preservation of the web's historical record.

eff.org

🔥🔥🔥🔥🔥

3 min

3/21/2026

News publishers limit Internet Archive access due to AI scraping concerns

News publishers are restricting access to the Internet Archive due to concerns over AI scraping of their content. The Internet Archive's crawlers capture webpage snapshots, which are accessible via the Wayback Machine, potentially exposing publishers' material to unauthorized use by AI models.

niemanlab.org

🔥🔥🔥🔥🔥

9 min

2/14/2026

No more articles to load