VideoPoet is an AI model developed by Google Research, Carnegie Mellon University (CMU) and Google DeepMind (United States), first published in December 2023. It works in the video, language and audio domain, on tasks such as video generation, audio generation, text-to-video and 2 more.
Epoch AI has no training-compute estimate for this model. The model has 8,000,000,000 parameters. It was trained on roughly 2T datapoints.
Access: Unreleased. Its weights are not openly released. The reference paper has 474 citations. Epoch AI rates the confidence of this record as confident.