The New York Times blocks OpenAI crawlers and forbids the use of its content in AI training
CTOnews.com, August 22 (Xinhua) the New York Times has blocked OpenAI's web crawlers, which means that OpenAI cannot use the content of the publication to train its artificial intelligence model.
Check the robots.txt page of the New York Times and you can see that the New York Times has blocked GPTBot, a crawler launched by OpenAI earlier this month. It is reported that the New York Times blocked the crawler as early as August 17.
It is worth mentioning that the New York Times updated its terms of service earlier this month, which forbids the use of its content to train artificial intelligence models, and is also considering a lawsuit against OpenAI for intellectual property infringement.
CTOnews.com noted that actor Sarah Silverman and two other writers sued the company in July over OpenAI's use of Books3 to train ChatGPT, a data set used to train ChatGPT that may contain thousands of copyrighted works, and that a programmer and lawyer, Matthew Butterick, accused the company of data grabbing that constituted software piracy.