Code for full fintuing Mochi model with FSDP (and CP)
☆29Jul 15, 2025Updated last year
Alternatives and similar repositories for Mochi-Full-Finetuner
Users that are interested in Mochi-Full-Finetuner are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Pusa: Thousands Timesteps Video Diffusion Model☆686Feb 13, 2026Updated 7 months ago
- Formal implementation of Robust Domain Misinformation Detection via Multi-modal Feature Alignment☆13Dec 8, 2023Updated 2 years ago
- ☆130Jun 24, 2025Updated last year
- Code for Paper 'Redefining Temporal Modeling in Video Diffusion: The Vectorized Timestep Approach'☆36Jan 2, 2026Updated 8 months ago
- Unofficial extension implementation of CausVid☆79Apr 28, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆14Feb 20, 2024Updated 2 years ago
- Unofficial implementation for Sigmoid Loss for Language Image Pre-Training☆11Sep 26, 2023Updated 3 years ago
- Community trainer for Lightricks' LTX Video model 🎬 ⚡️☆469Jan 6, 2026Updated 8 months ago
- ☆12Sep 28, 2024Updated last year
- This is the official implementation of SG-I2V: Self-Guided Trajectory Control in Image-to-Video Generation.☆116Nov 26, 2024Updated last year
- The offical code implementation of paper "Interpretable Multimodal Misinformation Detection with Logic Reasoning", accepted by Finding of…☆31Feb 5, 2026Updated 7 months ago
- CVPRW 2025 paper Progressive Autoregressive Video Diffusion Models: https://arxiv.org/abs/2410.08151☆89May 12, 2025Updated last year
- ☆40Apr 27, 2024Updated 2 years ago
- RLHF for Video Diffusion Models☆27Aug 19, 2026Updated last month
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Weird autoencoder experiments☆25May 20, 2026Updated 4 months ago
- Code of RealisHuman[AAAI 2025]: A Two-Stage Approach for Refining Malformed Human Parts in Generated Images☆90Jul 7, 2026Updated 2 months ago
- ☆12Dec 19, 2020Updated 5 years ago
- [NeurIPS 2025] Improving Video Generation with Human Feedback☆512Sep 24, 2025Updated last year
- Neuro-Symbolic Hierarchical Rule Induction☆15Dec 31, 2022Updated 3 years ago
- A minimalistic, hackable code base to finetune Wan video generation model☆49Feb 22, 2026Updated 7 months ago
- This is a python implementation of Hierarchical Image Matting Model for Segmentation.☆11Jun 21, 2022Updated 4 years ago
- ☆18Apr 21, 2026Updated 5 months ago
- [ICLR 2025, AAAI 2026] official implementation of "Diffusion-NPO: Negative Preference Optimization for Better Preference Aligned Generati…☆39Jan 26, 2026Updated 8 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [NeurIPS 2024] The official implementation of "Image Copy Detection for Diffusion Models"☆18Oct 1, 2024Updated last year
- Official implementation of LiFT: Leveraging Human Feedback for Text-to-Video Model Alignment.☆85May 4, 2025Updated last year
- Official Pytorch implementation of the paper Learning Input-agnostic Manipulation Directions in StyleGAN with Text Guidance (accepted to …☆28May 13, 2023Updated 3 years ago
- ☆11Feb 28, 2019Updated 7 years ago
- musubi-tuner modified to tune image2video/video infilling☆33Jan 30, 2025Updated last year
- (ICCV'25) TF-TI2I: Training-Free Text-and-Image-to-Image Generation via Multi-Modal Implicit-Context Learning in Text-to-Image Models (Au…☆17Aug 22, 2025Updated last year
- ☆33Jan 6, 2025Updated last year
- Official codes of VEnhancer: Generative Space-Time Enhancement for Video Generation☆579Sep 16, 2024Updated 2 years ago
- ☆15Oct 27, 2025Updated 11 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Sparse Autoencoders (SAE) vs CLIP fine-tuning fun.☆18Dec 19, 2024Updated last year
- Keyframe Interpolation with CogvideoX☆140Oct 31, 2024Updated last year
- ☆34Aug 9, 2024Updated 2 years ago
- Use MHFormer [CVPR 2022] to do pose estimation and use Unity to control rig of model. (not real-time)☆18Sep 14, 2022Updated 4 years ago
- [AAAI 2025] Follow-Your-Canvas: This repo is the official implementation of "Follow-Your-Canvas: Higher-Resolution Video Outpainting with…☆169Aug 26, 2025Updated last year
- [NeurIPS 2024] CV-VAE: A Compatible Video VAE for Latent Generative Video Models☆286Dec 4, 2024Updated last year
- Towards Variable and Coordinated Holistic Co-Speech Motion Generation, CVPR 2024☆59Jun 27, 2024Updated 2 years ago