☆58May 6, 2025Updated last year
Alternatives and similar repositories for IPO
Users that are interested in IPO are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Cockatiel: Ensembling Synthetic and Human Preferenced Training for Detailed Video Caption☆38May 21, 2025Updated last year
- ☆14Jul 17, 2024Updated 2 years ago
- ☆37Aug 20, 2026Updated last week
- Official implementation of LiFT: Leveraging Human Feedback for Text-to-Video Model Alignment.☆85May 4, 2025Updated last year
- ☆19Oct 23, 2024Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [ICLR 2026] Uni-CoT: Towards Unified Chain-of-Thought Reasoning Across Text and Vision☆236May 31, 2026Updated 3 months ago
- Uni-ViGU: Towards Unified Video Generation and Understanding via A Diffusion-Based Video Generator☆34Apr 15, 2026Updated 4 months ago
- [CVPR 2026] Fine-Grained GRPO for Precise Preference Alignment in Flow Models☆66Jun 1, 2026Updated 3 months ago
- [CVPR-2026] DiverseDiT: Towards Diverse Representation Learning in Diffusion Transformers☆23Updated this week
- Training-free Guidance in Text-to-Video Generation via Multimodal Planning and Structured Noise Initialization☆28Apr 14, 2025Updated last year
- [AAAI 2026] VisionReward: Fine-Grained Multi-Dimensional Human Preference Learning for Image and Video Generation☆424Mar 26, 2025Updated last year
- ☆162Feb 28, 2026Updated 6 months ago
- ☆69Aug 16, 2024Updated 2 years ago
- An official implementation of EvoSearch: Scaling Image and Video Generation via Test-Time Evolutionary Search☆107Oct 3, 2025Updated 10 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [NeurIPS 2025] Improving Video Generation with Human Feedback☆503Sep 24, 2025Updated 11 months ago
- ☆25Jul 20, 2025Updated last year
- Official Implementation of VideoDPO☆169Jun 1, 2025Updated last year
- Video models as pure visual reasoners for high-quality text-to-image generation via Chain-of-Frame reasoning.☆40Jan 16, 2026Updated 7 months ago
- Image Tokenizer Needs Post-Training☆24Oct 4, 2025Updated 10 months ago
- BranchGRPO: Stable and Efficient GRPO with Structured Branching in Diffusion Models☆50Oct 30, 2025Updated 10 months ago
- Official implementation of "Token Perturbation Guidance for Diffusion Models" [NeurIPS 2025]☆18May 19, 2026Updated 3 months ago
- JoVA: Unified Multimodal Learning for Joint Video-Audio Generation☆33Dec 22, 2025Updated 8 months ago
- [CVPR 2026] MixFlow Training: Alleviating Exposure Bias with Slowed Interpolation Mixture☆22Dec 23, 2025Updated 8 months ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Official Code of "UniMMVSR: A Unified Multi-Modal Framework for Cascaded Video Super-Resolution"☆26Oct 12, 2025Updated 10 months ago
- ☆24Oct 11, 2025Updated 10 months ago
- ☆55Feb 9, 2026Updated 6 months ago
- Official implementation of the paper "Koala-36M: A Large-scale Video Dataset Improving Consistency between Fine-grained Conditions and Vi…☆257Mar 19, 2025Updated last year
- ☆15Oct 27, 2025Updated 10 months ago
- Official Repo for Self-Forcing++ High Quality Long Video Generation☆270Oct 13, 2025Updated 10 months ago
- ☆30Jun 9, 2025Updated last year
- [NeurIPS 2024] The official implement of research paper "FreeLong : Training-Free Long Video Generation with SpectralBlend Temporal Atten…☆67Jul 2, 2025Updated last year
- ☆15Jun 21, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [NeurIPS'25] Towards Interpretability Without Sacrifice: Faithful Dense Layer Decomposition with Mixture of Decoders☆16May 28, 2025Updated last year
- The Best of Both Worlds: Integrating Language Models and Diffusion Models for Video Generation☆41May 4, 2025Updated last year
- Code for: "Long-Context Autoregressive Video Modeling with Next-Frame Prediction"☆314Apr 23, 2025Updated last year
- Concat-ID: Towards Universal Identity-Preserving Video Synthesis☆65May 7, 2025Updated last year
- ☆21Jan 1, 2026Updated 8 months ago
- Official code for VINCIE: Unlocking In-context Image Editing from Video☆61Jun 19, 2026Updated 2 months ago
- Official implementation of UnifiedReward & [NeurIPS 2025] UnifiedReward-Think & UnifiedReward-Flex☆806Jun 18, 2026Updated 2 months ago