Meta torrented and seeded a large dataset containing copyrighted data, totaling 81.7 TB. The dataset is believed to contain a large number of books, potentially in the tens of millions. The discussion revolves around the implications of this action, including potential copyright infringement and the use of the data for training large language models. The exact details of the article are not available, but the comments suggest that the dataset was obtained from a torrent and may have been used for training AI models.