Get the App
SLTechnology News&Howtos  ›  IT Information  › 

Twitter and other companies charge too much, and AI companies such as OpenAI and Cohere turn to synthetic data

Shulou Source: shulou.com Published: 2023-11-24 18:12:15 10月02日 Update

CTOnews.com, July 20 (Xinhua) Aiden Gomez, CEO of artificial intelligence company Cohere, recently revealed that AI, including Microsoft, OpenAI and Cohere, has used synthetic data to train AI models because data collection fees from companies such as Reddit and Twitter are too high.

Gomez said that the composite data can be applied to many training scenarios, but it has not been fully promoted yet.

CTOnews.com attached an example from Gomez: if an enterprise wants to train a model in advanced mathematics, it can create two artificial intelligence models that play the roles of teacher and student, and let them discuss topics such as trigonometry. The manual is mainly responsible for observation, and if you see any mistakes, you can correct them.

CTOnews.com Note:

Synthetic data (synthetic data) is data generated manually by computer technology, rather than data generated by real events.

However, the synthetic data has "availability" and can reflect the attributes of the original data mathematically or statistically, so it can be used as a substitute for the original data to train, test and verify large models.

Tags: Data artificial model training company primitive artificial intelligence mathematics intelligence trigonometry two events enterprises examples just usability scenarios students not yet attributes Apple Docker Huawei Linux macOS MariaDB Microsoft MySQL NVidia OPPO Reno Redmi MariaDB Shulou Information Apple vpn