Reposted by Max Kaiser
A couple of interesting bits in this article about a large public domain book dataset, but perhaps the most important thing that is maybe not obvious to a lot of folks: the embargoes on the scans made for the Google Books Project are *finally* starting to end!: www.wired.com/story/harvar... 📜📚
wired.com
Harvard Is Releasing a Massive Free AI Training Dataset Funded by OpenAI and Microsoft
The project’s leader says that allowing everyone to access the collection of public-domain books will help “level the playing field” in the AI industry.