Breaking cross-modal boundaries in multimodal AI: Introducing CoDi, composable diffusion for any-to-any generation
Imagine an AI model that can seamlessly generate high-quality content across text, images, video, and audio, all at once. Such a model would more accurately capture the multimodal nature of the world and human comprehension,…
Research Focus: Week of June 19, 2023
In this issue: Our new Responsible AI Maturity Model; FoundWright helps “re-find” web pages; Trace-guided Inductive Synthesis of Recursive Functional Programs; a wait-free algorithm for weak reference counting; and new research on concurrency testing.
DeepSpeed ZeRO++: A leap in speed for LLM and chat model training with 4X less communication
Large AI models are transforming the digital world. Generative language models like Turing-NLG, ChatGPT, and GPT-4, powered by large language models (LLMs), are incredibly versatile, capable of performing tasks like summarization, coding, and translation. Similarly,…
Collaborators: Renewable energy storage with Bichlien Nguyen and David Kwabi
Researcher Bichlien Nguyen is an organic electrochemist turned technologist. Professor David Kwabi is a mechanical engineer. Their work uses ML to help discover organic compounds for renewable energy storage. Learn about their collaboration.
Foundation of AGI
We focus on the fundamental research of Foundation of AGI as part of our mission-focused research on advancing AGI for humanity. Our research involves pushing #TheBigConvergence of foundation models across tasks, languages, and modalities, developing…