Get the App
SLTechnology News&Howtos  ›  IT Information  › 

Microsoft launches artificial intelligence model CoDi, which can interact and generate multimodal content

Shulou Source: shulou.com Published: 2023-11-24 18:00:46 10月04日 Update

CTOnews.com, July 11, Microsoft recently released a press release called composable Diffusion Model (CoDi), a unique artificial intelligence model based on composable diffusion that is designed to interact and generate multimodal content.

Microsoft's goal of designing CoDi is to solve the limitations of the traditional single-mode AI model. In the case of synchronized video and audio, there may be inconsistency and alignment when independently generated information streams are spliced together.

CoDi adopts a unique combinable generation strategy to align multiple modes in the diffusion process to generate intertwined patterns. More importantly, CoDi can handle arbitrary input patterns and generate arbitrary modal content.

CoDi was developed by the Microsoft Azure Cognitive Services research team in collaboration with the University of North Carolina at Chapel Hill and is part of the Microsoft project i-Code, which uses artificial intelligence to enhance human-computer interaction.

CTOnews.com attached a link to the official introduction of the CoDi project, which can be read in depth by interested users.

Tags: Modes generation models projects combinations content Microsoft uniqueness patterns goals design artificial intelligence intelligence Cheng duo important man-machine traditional users information flow Apple Docker Huawei Linux macOS MariaDB Microsoft MySQL NVidia OPPO Reno Huawei OPPO Reno Docker Xiaomi Shulou Technology