Comment by silver_sun
11 days ago
The difference here is that the scientific papers he downloaded weren't freely available to the public, like those scraped webpages would be. Corporate scrapers have been sued[0] in the past for scraping pages from behind a login page / paywall.
Neither are the pirated books Meta is using for model training.
I totally agree with you, and didn't mean to endorse that.