GKRootWire
AI Google Adds 'Preferred Source' Button to Help Publishers Fight AI Traffic LossesGadgets Linkdaze Launches a Smart Calendar Aimed at Running Your Whole HouseholdSecurity Popular Rust Crate arrayref Hijacked to Spread Infostealer MalwareCloud & Sysadmin GitHub Details Cause of August 17 Outage, Outlines Reliability FixesDev Tools Show HN: 'Huzzah' Proposes a Fresh Take on AI-Assisted CodingCloud & Sysadmin The Weird Science of Cooling Data Centers With UrineAI Google Adds 'Preferred Source' Button to Help Publishers Fight AI Traffic LossesGadgets Linkdaze Launches a Smart Calendar Aimed at Running Your Whole HouseholdSecurity Popular Rust Crate arrayref Hijacked to Spread Infostealer MalwareCloud & Sysadmin GitHub Details Cause of August 17 Outage, Outlines Reliability FixesDev Tools Show HN: 'Huzzah' Proposes a Fresh Take on AI-Assisted CodingCloud & Sysadmin The Weird Science of Cooling Data Centers With Urine
AI

A Decade Later, Meta's Mass Scraping Draws a Fraction of Aaron Swartz's Legal Heat

A viral blog post revisits how the same act that led to Aaron Swartz's prosecution is now standard practice for AI-era data harvesting.

A widely discussed blog post is reigniting debate over how selectively the U.S. legal system treats data scraping. It points out that Aaron Swartz, co-creator of RSS, was aggressively prosecuted under the Computer Fraud and Abuse Act for downloading academic papers from JSTOR, facing charges that carried decades of potential prison time before his death in 2013. Meanwhile, Meta has faced lawsuits and settlements over scraping copyrighted books and other content to train its AI models, but nothing resembling criminal prosecution.

The post argues this gap reflects who scraping laws are actually enforced against: individuals versus well-resourced corporations that can absorb civil penalties as a cost of doing business. Commenters on Hacker News largely agreed, framing it as evidence that data access laws protect corporate AI ambitions more than they protect the public interest Swartz was trying to serve.

Why it matters: As AI companies increasingly rely on scraped data to train models, this disparity matters because it shapes what kind of scraping gets normalized versus criminalized. Weak, inconsistent enforcement effectively favors large firms with legal budgets over independent researchers or activists doing similar things for public benefit.

Sources: Hacker News