1 story in this blend
An open weights generative foundation model that creates high quality video clips and aligned audio from text or image inputs. It suits video editors, digital avatar developers, and robotics engineers needing low latency visual predictions.