AI companies used vast troves of published text to train their large language models, and executives at those firms were aware that doing so constituted a theft of copyrighted materials, according to comments cited by news publishers in a court filing that was unredacted this week.
The brief was initially filed earlier this month in a court proceeding that brings together several related copyright infringement cases against ChatGPT maker OpenAI and its partner Microsoft, including lawsuits brought by The New York Times and by Ziff Davis, which owns CNET.
Microsoft’s director of applied science, Brent Hecht, called it “an astonishing theft of…
Read the full article at CNET.COM


