An open source community implementation of the model from the paper: "Movie Gen: A Cast of Media Foundation Models". Join our community to help implement this model!
☆61Sep 14, 2026Updated this week
Alternatives and similar repositories for movie-gen
Users that are interested in movie-gen are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- CogNetX is an advanced, multimodal neural network architecture inspired by human cognition. It integrates speech, vision, and video proce…☆21Aug 29, 2026Updated 3 weeks ago
- OmniByteFormer is a generalized Transformer model that can process any type of data by converting it into byte sequences, bypassing tradi…☆17Updated this week
- Unofficial implementation of Meta's MovieGen models☆16Nov 25, 2025Updated 9 months ago
- Simple Implementation of a Transformer in the new framework MLX by Apple☆19Nov 18, 2024Updated last year
- A swarm of LLM agents that will help you test, document, and productionize your code!☆20Aug 29, 2026Updated 3 weeks ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆19Apr 16, 2025Updated last year
- Movie Gen Bench - two media generation evaluation benchmarks released with Meta Movie Gen☆442Mar 8, 2025Updated last year
- A sophisticated multi-agent system designed for real-time market analysis of HTX (formerly Huobi) exchange data. This swarm combines spec…☆11Mar 18, 2025Updated last year
- A curated list of resources, libraries, tools, and communities for working with Local Large Language Models (LLMs).☆11Dec 20, 2024Updated last year
- [arXiv 2024] I4VGen: Image as Free Stepping Stone for Text-to-Video Generation☆24Oct 6, 2024Updated last year
- Implementation of "Audio xLSTMs: Learning Self-supervised audio representations with xLSTMs" in PyTorch☆20Aug 28, 2026Updated 3 weeks ago
- ☆16Aug 7, 2026Updated last month
- Traditional operating systems are reactive - they wait for user input or system events before taking action. SwarmOS breaks this paradigm…☆16Dec 6, 2024Updated last year
- This is a [forked version] for author's debugging. Please jump to https://github.com/QualityAssessment/DOVER for stable version to use.☆14Oct 29, 2023Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- My personal implementation of the model from "Qwen-VL: A Frontier Large Vision-Language Model with Versatile Abilities", they haven't rel…☆13Jan 29, 2024Updated 2 years ago
- OmegaViT (ΩViT) is a cutting-edge vision transformer architecture that combines multi-query attention, rotary embeddings, state space mod…☆15Aug 28, 2026Updated 3 weeks ago
- Official code of the paper: Draw an Audio: Leveraging Multi-Instruction for Video-to-Audio Synthesis.☆45Sep 11, 2024Updated 2 years ago
- A Large-Scale Dataset and Cinematic Narrative Benchmark for Multi-Shot Subject-to-Video Generation☆34Jun 9, 2026Updated 3 months ago
- Low level software graphics library by ErrorSoft (ESLGL)☆20Apr 7, 2018Updated 8 years ago
- Minimum viable code for the Decodable Information Bottleneck paper. Pytorch Implementation.☆12Oct 20, 2020Updated 5 years ago
- Reward Guided Latent Consistency Distillation☆26Oct 9, 2024Updated last year
- The official codebase for Reflected Flow Matching (ICML 2024)☆24Jun 19, 2024Updated 2 years ago
- A daemon that makes a desktop OS accessible to AI agents☆45May 29, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆18Jan 5, 2025Updated last year
- ☆20Apr 26, 2024Updated 2 years ago
- Multi-Modal Tree of thoughts for DALLE-3 like auto self improvement☆17Nov 11, 2024Updated last year
- This repo contains evaluation code for the paper "AV-Odyssey: Can Your Multimodal LLMs Really Understand Audio-Visual Information?"☆31Dec 23, 2024Updated last year
- A framework making it effortless to convert any llm model into a reasoning agent like o1 or DeepSeek's r1☆24Oct 13, 2025Updated 11 months ago
- Ultra Fast Multi-Modality Vector Database☆18Feb 21, 2024Updated 2 years ago
- ☆20Aug 11, 2025Updated last year
- ☆12Oct 21, 2019Updated 6 years ago
- Only for education purposes☆24Nov 15, 2021Updated 4 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Neuroscience Inspired Agent Reasoning Framework☆31May 19, 2025Updated last year
- This repository contains the code for the paper "TaylorShift: Shifting the Complexity of Self-Attention from Squared to Linear (and Back)…☆15Feb 25, 2026Updated 6 months ago
- An open source community implementation of the model MELLE from the paper: "Autoregressive Speech Synthesis without Vector Quantization"☆16Sep 14, 2026Updated last week
- Measuring the Signal to Noise Ratio in Language Model Evaluation☆32Aug 19, 2025Updated last year
- ☆15Mar 30, 2025Updated last year
- Talking Head from Speech Audio using a Pre-trained Image Generator☆22May 7, 2024Updated 2 years ago
- J2ME 3D Game engine based on Iceberg3D☆22May 1, 2026Updated 4 months ago