Simon Reading Lessons
2.28K subscribers
1 photo
10 videos
7 files
1 link
Download Telegram
Channel created
Media is too big
VIEW IN TELEGRAM
IELTS-Simon-Reading-part-1
119
Media is too big
VIEW IN TELEGRAM
IELTS-Simon-Reading-part-2
104
Media is too big
VIEW IN TELEGRAM
IELTS-Simon-Reading-part-3
67
IELTS-Simon-Reading-Practice-exercise-3.pdf
56.7 KB
IELTS-Simon-Reading-Practice-exercise-3
56
Media is too big
VIEW IN TELEGRAM
IELTS-Simon-Reading-part-4
49
PIELTS-Simon-Reading-arargraph-headings-worksheet-4.pdf
70.3 KB
PIELTS-Simon-Reading-arargraph-headings-worksheet-4
43
Media is too big
VIEW IN TELEGRAM
IELTS-Simon-Reading-part-5
46
IELTS-Simon-Reading-Which-paragraph-worksheet-part5.pdf
73 KB
IELTS-Simon-Reading-Which-paragraph-worksheet-part5
42
Media is too big
VIEW IN TELEGRAM
IELTS-Simon-Reading-part-6
40
IELTS-Simon-Reading-Multiple-Choice-Worksheet-part6.pdf
52.7 KB
IELTS-Simon-Reading-Multiple-Choice-Worksheet-part6
35
Media is too big
VIEW IN TELEGRAM
IELTS-Simon-Reading-part-7
35
IELTS-Simon-Reading-Matching-names-worksheet-part7.pdf
58.2 KB
IELTS-Simon-Reading-Matching-names-worksheet-part7
34
Media is too big
VIEW IN TELEGRAM
IELTS-Simon-Reading-part-8
36
IELTS-Simon-Reading-Short-answers-worksheet-part8.pdf
50.6 KB
IELTS-Simon-Reading-Short-answers-worksheet-part8
40
IELTS-Simon-Reading-Sentence-endings-worksheet-part9.pdf
122.3 KB
IELTS-Simon-Reading-Sentence-endings-worksheet-part9
40
Media is too big
VIEW IN TELEGRAM
IELTS-Simon-Reading-part-9
50
Media is too big
VIEW IN TELEGRAM
IELTS-Simon-Reading-part-10
77
Forwarded from Crypto Head
📖AI Companies Are Buying Books in Bulk – and Destroying Them

404 Media has published a report raising serious questions about how modern AI companies obtain training data.

According to the report, AI companies have begun purchasing vast quantities of old physical books. The reason is simple: books are one of the largest sources of “clean” text written by humans rather than generated by neural networks.

Books published before 2022 are particularly valuable because they predate the widespread availability of AI-generated content. Using them helps prevent AI models from being trained on text that was itself generated by other AI systems.

For example, ISBNdb offers AI labs a service that can source anywhere from 1,000 to 1 million books in a single order. The books are then scanned and converted into datasets for AI training. To digitize such enormous quantities of books quickly, the books are effectively destroyed. Hydraulic cutting equipment removes the binding, the pages are separated into individual sheets, and high-speed industrial scanners convert them into PDFs. The remaining paper is sent for recycling.

Anthropic is cited as one example. Court documents reportedly show that the company purchased large quantities of physical books, cut them apart, and scanned them to train its Claude model. The company allegedly attempted to keep the process confidential.

OpenAI, Google, xAI, and other companies are not named in the report. According to 404 Media, ISBNdb’s customers are reluctant to disclose that they are purchasing and destroying millions of books for AI training.

Elon Musk, however, responded to the report:

“I asked the SpaceX AI team to preserve all the rare books in the library and digitize them using traditional methods rather than simply cutting off their bindings and scanning them.”

Turning physical books into datasets for AI training raises an unsettling question: perhaps we are already living in the dystopian future we once imagined.
Please open Telegram to view this post
VIEW IN TELEGRAM