@FinanceYF5: Meta illegally downloaded over 80 TB of books from LibGen, Anna's Archive, and Z-Library to train its AI models. Aaron Swartz downloaded 70 GB of papers from JSTOR in 2010 (only equivalent to...

X AI KOLs Following News

Summary

Meta is accused of illegally downloading over 80 TB of books from LibGen, Anna's Archive, and Z-Library to train AI models. The article contrasts the case of Aaron Swartz, who faced severe charges for downloading a much smaller amount of papers, highlighting the double standard in copyright enforcement.

Meta illegally downloaded over 80 TB of books from LibGen, Anna's Archive, and Z-Library to train its AI models. Aaron Swartz downloaded 70 GB of papers from JSTOR in 2010 (only 0.0875% of Meta's download), yet faced a $1 million fine and 35 years in prison, and took his own life in 2013. https://t.co/OOyX8LmzeS
Original Article
View Cached Full Text

Cached at: 05/08/26, 10:46 AM

Meta illegally downloaded over 80 TB of books from LibGen, Anna’s Archive, and Z-Library to train its AI models.

In 2010, Aaron Swartz downloaded 70 GB of papers from JSTOR (just 0.0875% of Meta’s download), yet faced charges of $1 million in fines and 35 years in prison. He died by suicide in 2013. https://t.co/OOyX8LmzeS

Similar Articles

@Phoenixyin13: This latest blockbuster paper from Meta FAIR aims to tell the AI industry an important bellwether: "Large model data is ushering in the era of intelligent scientists." In this paper, a 4B small model precisely refined by Autodata not only crushes the same-scale models trained with traditional synthetic data on legal reasoning tasks, but also...

X AI KOLs Timeline

Meta FAIR's latest paper proposes the Autodata method, which uses an intelligent data scientist Agent to autonomously generate and optimize high-quality data, enabling a 4B small model to defeat a 397B large model on legal reasoning tasks. This indicates that data quality can bridge the gap in parameter count, providing new insights for data pipelines and scaling.