It is said that the data set of generative AI such as training ChatGPT uses copyrighted content, and several publishers jointly claim compensation.
CTOnews.com March 25 news, the rise of chat robots such as ChatGPT and Bard is based on huge data sets. And there is a lot of copyright content from major publishers in these data sets, and a number of publishers have joined forces to promote relevant laws and regulations to protect their legitimate rights and interests.
According to the Wall Street Journal, several publishers have been investigating evidence of data sets training artificial intelligence to prove the existence of copyrighted content in these data sets. CTOnews.com learned from the report that these publishers have formed an alliance to pressure the companies that develop generative AI to pay adequate compensation and compensation through the publishing trade federation News Media Alliance.
"this valuable content is protected by copyright, and we have to be compensated for the fact that it is constantly used to generate revenue for others," said Danielle Coffey, a News Media Alliance executive.