Meta illegally downloaded 80TB of torrent data for AI training

Recently unsealed Meta emails contain “incriminating evidence” against the company in a copyright lawsuit filed submitted by book authors, and claim that Meta illegally trained its AI models with pirated books.

Last month, Meta admitted to seeding a torrent containing a large dataset known as LibGen, which includes tens of millions of pirated books.

Discover more articles in search results.

However, the details surrounding torrenting were unclear until yesterday, when the Meta emails were first made public.

The new evidence showed that Meta torrented “at least 81,7 terabytes of data from multiple illegal libraries through the site.” Anna's Archive, including at least 35,7 terabytes of data from Z-Library and its LibGen“, the authors’ lawsuit stated. “Meta had also previously downloaded a torrent containing 80,6 terabytes of data from LibGen.”

“The scale of Meta’s illegal torrenting is staggering,” the authors’ lawsuit claims, insisting that “much smaller acts of data hacking lead to convictions.

The authors' works are protected by copyright, which led authorities to refer the lawsuit to the US attorney's office.


Google preferences

Leave a Comment

Your email address is not published. Required fields are mentioned with *

Your message will not be published if:
1. Contains insulting, defamatory, racist, offensive or inappropriate comments.
2. Causes harm to minors.
3. It interferes with the privacy and individual and social rights of other users.
4. Advertises products or services or websites.
5. Contains personal information (address, phone, etc.).