By Cory Doctorow (GPG 0xBF3D9110957E5F4C) @doctorow.pluralistic.net · 18/08/2026From the OED to search engines to the Internet Archive, so many beneficial activities rely on the fact that copyright permits unlicensed collection and analysis of every copyrighted work as a single, massive corpus. 6/ 1377
By Cory Doctorow (GPG 0xBF3D9110957E5F4C) @doctorow.pluralistic.net · 18/08/2026They also depend on the fact that copyright allows the publication of that analysis without permission from the creators of the works it analyzes. A lot of people who are (rightfully) very angry about AI dispute this. 7/ 1232
By Cory Doctorow (GPG 0xBF3D9110957E5F4C) @doctorow.pluralistic.net · 18/08/2026They believe that they can craft an "AI training" law that would ban scraping, analysis and publication when these activities are part of AI training, but not when they're undertaken for a benign purpose. I am very, very skeptical of this. 8/ 1344