[ICCV 2025] Enhancing spatial understanding in text-to-Image diffusion models
☆94Sep 11, 2025Updated 10 months ago
Alternatives and similar repositories for CoMPaSS
Users that are interested in CoMPaSS are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ECCV 2026] CustomX: Unified Character, Action, and Scene Customization in Video World Models☆96Jun 25, 2026Updated 3 weeks ago
- Official Implementation of DRA-Ctrl (Dimension-Reduction Attack! Video Generative Models are Experts on Controllable Image Synthesis)☆119Aug 15, 2025Updated 11 months ago
- [CVPR 2026] 🔥🔥 Official Repo of USO: Unified Style and Subject-Driven Generation via Disentangled and Reward Learning☆1,226Sep 12, 2025Updated 10 months ago
- ☆90May 13, 2026Updated 2 months ago
- [3DV 2026 Oral] VoxHammer: Training-Free Precise and Coherent 3D Editing in Native 3D Space☆234Nov 25, 2025Updated 7 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Compact image interpolation model for generating in-between frames, with ComfyUI support and the same core model used in FrameFusion Mo…☆25Apr 17, 2026Updated 3 months ago
- ViSAudio: End-to-End Video-Driven Binaural Spatial Audio Generation☆117Dec 11, 2025Updated 7 months ago
- Automated video dataset creator for Windows using WhisperX and Qwen2-VL☆18May 9, 2026Updated 2 months ago
- [SIGGRAPH Asia 25] Voost: A Unified and Scalable Diffusion Transformer for Bidirectional Virtual Try-On and Try-Off☆340May 15, 2026Updated 2 months ago
- Official inference code and LongText-Bench benchmark for our paper X-Omni (https://arxiv.org/pdf/2507.22058).☆426Aug 26, 2025Updated 10 months ago
- Feed-forward model for predicting 3D physics with 3DGS + NeRF☆297Mar 5, 2026Updated 4 months ago
- [ICLR 2026] Official implementation of "SeC: Advancing Complex Video Object Segmentation via Progressive Concept Construction"☆288Mar 27, 2026Updated 3 months ago
- [NeurIPS 2025] Official implementation of "XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulatio…☆626Oct 22, 2025Updated 9 months ago
- ToonOut, a fork of BiRefNet focused on background removal for anime images. We open-source our dataset & our weights. See our paper at: h…☆96Jun 12, 2026Updated last month
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- FantasyPortrait: Enhancing Multi-Character Portrait Animation with Expression-Augmented Diffusion Transformers☆511Aug 20, 2025Updated 11 months ago
- Official repository of paper "ProEdit: Inversion-based Editing From Prompts Done Right"☆116Feb 5, 2026Updated 5 months ago
- [ICML2026] Official Implementation of "TAG: Tangential Amplifying Guidance for Hallucination-Resistant Sampling"☆42Jul 6, 2026Updated 2 weeks ago
- ☆110Sep 3, 2025Updated 10 months ago
- OmniInsert: Mask-Free Video Insertion of Any Reference via Diffusion Transformer Models☆162Mar 4, 2026Updated 4 months ago
- Official implementation for "Story2Board: A Training‑Free Approach for Expressive Storyboard Generation"☆263Aug 22, 2025Updated 11 months ago
- Krea 2 & Klein 9B LoRA Studio — train, profile, repair, and extract Krea 2 & Flux 2 Klein 9B LoRAs☆67Jul 11, 2026Updated last week
- The official repository of EditCrafter: Tuning-free High-Resolution Image Editing via Pretrained Diffusion Model (CVPRW 2026)☆50Apr 19, 2026Updated 3 months ago
- [ECCV 2026] Generate high resolution videos with a custom voice and appearance, based on LTX-2/LTX-2.3 + Identity In-Context LoRA☆347Jun 24, 2026Updated 3 weeks ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Pusa: Thousands Timesteps Video Diffusion Model☆685Feb 13, 2026Updated 5 months ago
- DreamStyle: A Unified Framework for Video Stylization☆124Jan 7, 2026Updated 6 months ago
- [CVPR 2026] When Numbers Speak: Aligning Textual Numerals and Visual Instances in Text-to-Video Diffusion Models☆68Apr 11, 2026Updated 3 months ago
- Official Code of NAVA: Native Audio-Visual Alignment for Generation.☆212Jun 30, 2026Updated 3 weeks ago
- iMontage: Unified, Versatile, Highly Dynamic Many-to-many Image Generation☆188Dec 1, 2025Updated 7 months ago
- ☆192Jul 31, 2025Updated 11 months ago
- [ICML 2026] ReCo: In-Context Generation with Regional Constraints for Instructional Video Editing☆170May 26, 2026Updated last month
- [NeurIPS'25 Spotlight] Official implementation of "JavisGPT: A Unified Multi-modal LLM for Sounding-Video Comprehension and Generation"☆75Feb 26, 2026Updated 4 months ago
- UniMesh: Unifying 3D Mesh Understanding and Generation☆57Jul 14, 2026Updated last week
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- [IEEE TIP] Official implementation of Progressive Detail Injection for Training-Free Semantic Binding in Text-to-Image Generation☆33Aug 3, 2025Updated 11 months ago
- [ICML 2026] Official implementation for "DyPE: Dynamic Position Extrapolation for Ultra High Resolution Diffusion".☆356May 18, 2026Updated 2 months ago
- [CVPR 2026] Offical implementation of the paper "HiFi-Inpaint: Towards High-Fidelity Reference-Based Inpainting for Generating Detail-Pre…☆105Jun 7, 2026Updated last month
- [ICCV 2025] Inpaint4Drag: Repurposing Inpainting Models for Drag-Based Image Editing via Bidirectional Warping☆94Nov 30, 2025Updated 7 months ago
- ComfyUI custom nodes for Foundation-1 | Structured Text-to-Sample Diffusion for Music Production☆105Mar 22, 2026Updated 4 months ago
- 🔥 [CVPR 2024] The official repo for Zero-Painter!☆70Jun 8, 2024Updated 2 years ago
- [Arxiv 2025] Official PyTorch Implementation of "SVG-T2I: Scaling up Text-to-Image Latent Diffusion Model Without Variational Autoencoder…☆152Dec 18, 2025Updated 7 months ago