We present Movie Gen, a cast of foundation models that generates high-quality, 1080p HD videos with different aspect ratios and synchronized audio. We also show additional capabilities such as precise instruction-based video editing and generation of personalized videos based on a user's image. Our models set a new state-of-the-art on multiple tasks: text-to-video synthesis, video personalization, video editing, video-to-audio generation, and text-to-audio generation. Our largest video generation model is a 30B parameter transformer trained with a maximum context length of 73K video tokens, corresponding to a generated video of 16 seconds at 16 frames-per-second. We show multiple technical innovations and simplifications on the architecture, latent spaces, training objectives and recipes, data curation, evaluation protocols, parallelization techniques, and inference optimizations that allow us to reap the benefits of scaling pre-training data, model size, and training compute for training large scale media generation models. We hope this paper helps the research community to accelerate progress and innovation in media generation models. All videos from this paper are available at https://go.fb.me/MovieGenResearchVideos.
Movie Gen: A Cast of Media Foundation Models
Movie Gen, a suite of foundation models, generates high-quality videos with synchronized audio, excelling in various tasks through architectural, training, and technical innovations.
- Year
- 2024
- Venue
- arXiv 2024
- Authors
- 88
- Hosting
- Abstract onlyARXIV-DEFAULT
Cite
Notes
Only stored in your browser.
Attribution
- Abstract & full text
- arxiv.org/abs/2410.13720v2ARXIV-DEFAULT
- TL;DR
- Semantic Scholar
Abstract
Authors
88Tao XuBowen ShiAndros TjandraJohn HoffmanYi-Chiao WuLuya GaoMatt LeApoorv VyasSanyuan ChenWei-Ning HsuAnn LeeCe LiuJi HouSam TsaiJialiang WangZijian HePeter VajdaIshan MisraRohit GirdharXiaoliang DaiPeizhao ZhangAdam PolyakMannat SinghQuentin DuvalShikai LiCarleigh WoodMary WilliamsonBaishan GuoDhruv ChoudharyYue ZhaoLicheng YuKunpeng LiFelix Juefei-XuXi YinMarkos GeorgopoulosSamaneh AzadiVladan PetrovicChing-Yao ChuangAli ThabetRashel MoritzArun MallyaGuan PangSai Saketh RambhatlaYuval KirstainChih-Yao MaEdgar SchönfeldYaqiao LuoLuxin ZhangHaoyu MaJonas KöhlerRoshan SumbalyZecheng HeTingbo HouAnimesh SinhaArtsiom SanakoyeuAlbert PumarolaAmit ZoharAndrew BrownDavid YanDingkang WangGeet SethiKiran JagadeeshMatthew YuMitesh Kumar SinghSamyak DattaSean BellSharadh RamaswamyShelly SheyninSiddharth BhattacharyaSimran MotwaniTianhe LiYaniv TaigmanYen-Cheng LiuBoris ArayaBreena KerrCen PengDimitry VengertsevElliot BlanchardFraylie NordJeff LiangKaolin FireKarthik SivakumarLawrence ChenSara K. SampsonSimone ParmeggianiSteve FineTara FowlerYuming Du