☆228Mar 7, 2024Updated 2 years ago
Alternatives and similar repositories for train_your_own_sora
Users that are interested in train_your_own_sora are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [TMLR 2025] Latte: Latent Diffusion Transformer for Video Generation.☆1,949Aug 10, 2026Updated last month
- ☆10Apr 24, 2024Updated 2 years ago
- team Doggeee's solution to Ego4D LTA challenge@CVPRW23'☆14Nov 4, 2023Updated 2 years ago
- Official implementation of "VSTAR: Generative Temporal Nursing for Longer Dynamic Video Synthesis"☆21Jan 26, 2025Updated last year
- Code for Paper 'Redefining Temporal Modeling in Video Diffusion: The Vectorized Timestep Approach'☆36Jan 2, 2026Updated 8 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- NeurIPS 2024☆394Sep 26, 2024Updated 2 years ago
- [CVPR 2024] Panda-70M: Captioning 70M Videos with Multiple Cross-Modality Teachers☆707Oct 25, 2024Updated last year
- A one-stop library to standardize the inference and evaluation of all the conditional video generation models.☆50Feb 13, 2025Updated last year
- ☆23Jan 23, 2026Updated 8 months ago
- Code and data for "AnyV2V: A Tuning-Free Framework For Any Video-to-Video Editing Tasks" [TMLR 2024]☆660Oct 29, 2024Updated last year
- [AAAI 2025] Official pytorch implementation of "VideoElevator: Elevating Video Generation Quality with Versatile Text-to-Image Diffusion …☆163Apr 7, 2024Updated 2 years ago
- Official implementation of FIFO-Diffusion: Generating Infinite Videos from Text without Training (NeurIPS 2024)☆486Oct 18, 2024Updated last year
- [CVPR 2025] StreamingT2V: Consistent, Dynamic, and Extendable Long Video Generation from Text☆1,627Mar 27, 2025Updated last year
- Lumina-T2X is a unified framework for Text to Any Modality Generation☆2,250Feb 16, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- This project aim to reproduce Sora (Open AI T2V model), we wish the open source community contribute to this project.☆12,209Mar 8, 2026Updated 6 months ago
- Video-Infinity generates long videos quickly using multiple GPUs without extra training.☆191Aug 4, 2024Updated 2 years ago
- [ICLR 2025] Pyramidal Flow Matching for Efficient Video Generative Modeling☆3,212Dec 21, 2024Updated last year
- Let's finetune video generation models!☆554Sep 15, 2025Updated last year
- T2VScore: Towards A Better Metric for Text-to-Video Generation☆81Apr 10, 2024Updated 2 years ago
- ☆468Feb 12, 2024Updated 2 years ago
- A simple script that reads a directory of videos, grabs a random frame, and automatically discovers a prompt for it☆142Jan 22, 2024Updated 2 years ago
- Codes for ID-Specific Video Customized Diffusion☆459Feb 22, 2024Updated 2 years ago
- LLM Reasoning Benchmark & Chain-of-Thoughts Dataset for Chemistry☆57Oct 9, 2025Updated 11 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- This respository contains the code for the CVPR 2024 paper AVID: Any-Length Video Inpainting with Diffusion Model.☆177Feb 27, 2024Updated 2 years ago
- Papers and codes collection for customized, personalized and editable generative models☆28Oct 1, 2024Updated last year
- [CVPR 2025] Consistent and Controllable Image Animation with Motion Diffusion Models☆296May 17, 2025Updated last year
- Allegro is a powerful text-to-video model that generates high-quality videos up to 6 seconds at 15 FPS and 720p resolution from simple te…☆1,138Feb 7, 2025Updated last year
- [TPAMI 2025🔥] MagicTime: Time-lapse Video Generation Models as Metamorphic Simulators☆1,337Apr 14, 2026Updated 5 months ago
- [IJCV 2024] LaVie: High-Quality Video Generation with Cascaded Latent Diffusion Models☆952Nov 13, 2024Updated last year
- ✨ Hotshot-XL: State-of-the-art AI text-to-GIF model trained to work alongside Stable Diffusion XL☆1,112Jan 23, 2024Updated 2 years ago
- [ICML 2024 Spotlight] FiT: Flexible Vision Transformer for Diffusion Model☆436Nov 10, 2024Updated last year
- Open-Sora: Democratizing Efficient Video Production for All☆29,843Apr 9, 2026Updated 5 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [ECCV 2024 Oral] MotionDirector: Motion Customization of Text-to-Video Diffusion Models.☆1,054Aug 21, 2024Updated 2 years ago
- RankGAN: A Maximum Margin Ranking GAN for Generating Faces☆14May 9, 2019Updated 7 years ago
- [TMM 2025] StableIdentity: Inserting Anybody into Anywhere at First Sight☆257Dec 26, 2024Updated last year
- Stable Video Diffusion Training Code and Extensions.☆732Jul 25, 2024Updated 2 years ago
- Official repo for VGen: a holistic video generation ecosystem for video generation building on diffusion models☆3,155Jan 10, 2025Updated last year
- [ICLR & NeurIPS 2025] Repository for Show-o series, One Single Transformer to Unify Multimodal Understanding and Generation.☆1,978Jan 8, 2026Updated 8 months ago
- ☆16Mar 25, 2024Updated 2 years ago