‘Unlicensed, unrestricted AI training could destroy the ecosystem for books’ — quote of the day by the Authors Guild on the sourcing of training data

0



To achieve any level of competency, large language models (LLMs) need ample data for sufficient training. AI companies have looked to various sources to mine this information, including content publicly available on the internet, synthetic data generated from other AI models, and printed literature.


“Unlicensed, unrestricted AI training could destroy the ecosystem for books in the long run, and copyright exceptions do not extend to undermining the very purpose of copyright law.”


Reading difficulties

Prompted by news that AI companies were allegedly using books from pirate ebook sites to build their LLMs, writers, authors, and publishers publicly called out this deeply worrying process.

Quote of the day

This article is part of TechRadar Pro’s QOTD project to provide an insight into the minds of the brightest and most recognized figures in the technology industry today and in years gone by. Read the full series here.



Source
Las Vegas News Magazine

Leave A Reply

Your email address will not be published.


This website uses cookies to improve your experience. We'll assume you're ok with this, but you can opt-out if you wish. Accept Read More