An open source community implementation of the model from the paper: "Movie Gen: A Cast of Media Foundation Models". Join our community to help implement this model!
☆61Aug 29, 2026Updated this week
Alternatives and similar repositories for movie-gen
Users that are interested in movie-gen are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- CogNetX is an advanced, multimodal neural network architecture inspired by human cognition. It integrates speech, vision, and video proce…☆21Updated this week
- The code and weight for LoVA. LoVA is a novel model for Long-form Video-to-Audio generation. Based on the Diffusion Transformer (DiT) arc…☆16Feb 27, 2025Updated last year
- Simple Implementation of a Transformer in the new framework MLX by Apple☆19Nov 18, 2024Updated last year
- A swarm of LLM agents that will help you test, document, and productionize your code!☆20Updated this week
- ☆19Apr 16, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Movie Gen Bench - two media generation evaluation benchmarks released with Meta Movie Gen☆442Mar 8, 2025Updated last year
- A sophisticated multi-agent system designed for real-time market analysis of HTX (formerly Huobi) exchange data. This swarm combines spec…☆11Mar 18, 2025Updated last year
- Pytorch Implementation of the Model from "MIRASOL3B: A MULTIMODAL AUTOREGRESSIVE MODEL FOR TIME-ALIGNED AND CONTEXTUAL MODALITIES"☆26Jan 27, 2025Updated last year
- [arXiv 2024] I4VGen: Image as Free Stepping Stone for Text-to-Video Generation☆24Oct 6, 2024Updated last year
- Implementation of "Audio xLSTMs: Learning Self-supervised audio representations with xLSTMs" in PyTorch☆20Updated this week
- This is a [forked version] for author's debugging. Please jump to https://github.com/QualityAssessment/DOVER for stable version to use.☆14Oct 29, 2023Updated 2 years ago
- My personal implementation of the model from "Qwen-VL: A Frontier Large Vision-Language Model with Versatile Abilities", they haven't rel…☆13Jan 29, 2024Updated 2 years ago
- OmegaViT (ΩViT) is a cutting-edge vision transformer architecture that combines multi-query attention, rotary embeddings, state space mod…☆15Updated this week
- An interactive companion toy that engages kids with storytelling, singing, and encouragement for physical activities using advanced AI t…☆10Oct 15, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A Large-Scale Dataset and Cinematic Narrative Benchmark for Multi-Shot Subject-to-Video Generation☆34Jun 9, 2026Updated 2 months ago
- Low level software graphics library by ErrorSoft (ESLGL)☆19Apr 7, 2018Updated 8 years ago
- Minimum viable code for the Decodable Information Bottleneck paper. Pytorch Implementation.☆12Oct 20, 2020Updated 5 years ago
- Reward Guided Latent Consistency Distillation☆26Oct 9, 2024Updated last year
- A daemon that makes a desktop OS accessible to AI agents☆42May 29, 2025Updated last year
- ☆18Jan 5, 2025Updated last year
- Implementation of the paper: "BRAVE : Broadening the visual encoding of vision-language models"☆26Jun 22, 2026Updated 2 months ago
- ☆20Apr 26, 2024Updated 2 years ago
- A light-weight and high-efficient training framework for accelerating diffusion tasks.☆53Apr 23, 2026Updated 4 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- This repo contains evaluation code for the paper "AV-Odyssey: Can Your Multimodal LLMs Really Understand Audio-Visual Information?"☆31Dec 23, 2024Updated last year
- A framework making it effortless to convert any llm model into a reasoning agent like o1 or DeepSeek's r1☆24Oct 13, 2025Updated 10 months ago
- Ultra Fast Multi-Modality Vector Database☆18Feb 21, 2024Updated 2 years ago
- This repository contains the code for the paper "TaylorShift: Shifting the Complexity of Self-Attention from Squared to Linear (and Back)…☆15Feb 25, 2026Updated 6 months ago
- An open source community implementation of the model MELLE from the paper: "Autoregressive Speech Synthesis without Vector Quantization"☆16Updated this week
- [ECCV 2024 Oral] Audio-Synchronized Visual Animation☆60Mar 15, 2026Updated 5 months ago
- PyTorch implementation of Wide Activation for Efficient and Accurate Image Super-Resolution (CVPR Workshop 2018)☆16Jun 20, 2019Updated 7 years ago
- Measuring the Signal to Noise Ratio in Language Model Evaluation☆31Aug 19, 2025Updated last year
- ☆15Mar 30, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Talking Head from Speech Audio using a Pre-trained Image Generator☆22May 7, 2024Updated 2 years ago
- ☆12Dec 7, 2024Updated last year
- python code for Independent Component Analysis☆14Jan 8, 2018Updated 8 years ago
- An effective image quality assessment framework combining Segment Anything (SAM). This is the official implementation of our paper.☆24Jun 29, 2023Updated 3 years ago
- [ECCV'24 Oral] PiTe: Pixel-Temporal Alignment for Large Video-Language Model☆17Feb 13, 2025Updated last year
- An open source replication of the stawberry method that leverages Monte Carlo Search with PPO and or DPO☆31Updated this week
- Implementation of the paper: "Leave No Context Behind: Efficient Infinite Context Transformers with Infini-attention" from Google in pyTO…☆59Updated this week