[ICCV 2025] Enhancing spatial understanding in text-to-Image diffusion models
☆94Sep 11, 2025Updated 11 months ago
Alternatives and similar repositories for CoMPaSS
Users that are interested in CoMPaSS are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ECCV 2026] CustomX: Unified Character, Action, and Scene Customization in Video World Models☆96Jun 25, 2026Updated last month
- Official Implementation of DRA-Ctrl (Dimension-Reduction Attack! Video Generative Models are Experts on Controllable Image Synthesis)☆119Aug 15, 2025Updated 11 months ago
- [CVPR 2026] 🔥🔥 Official Repo of USO: Unified Style and Subject-Driven Generation via Disentangled and Reward Learning☆1,234Sep 12, 2025Updated 11 months ago
- ☆91May 13, 2026Updated 2 months ago
- [3DV 2026 Oral] VoxHammer: Training-Free Precise and Coherent 3D Editing in Native 3D Space☆237Nov 25, 2025Updated 8 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Compact image interpolation model for generating in-between frames, with ComfyUI support and the same core model used in FrameFusion Mo…☆25Apr 17, 2026Updated 3 months ago
- ViSAudio: End-to-End Video-Driven Binaural Spatial Audio Generation☆117Dec 11, 2025Updated 8 months ago
- Automated video dataset creator for Windows using WhisperX and Qwen2-VL☆19Jul 30, 2026Updated last week
- [SIGGRAPH Asia 25] Voost: A Unified and Scalable Diffusion Transformer for Bidirectional Virtual Try-On and Try-Off☆340May 15, 2026Updated 2 months ago
- Official inference code and LongText-Bench benchmark for our paper X-Omni (https://arxiv.org/pdf/2507.22058).☆428Aug 26, 2025Updated 11 months ago
- Feed-forward model for predicting 3D physics with 3DGS + NeRF☆299Mar 5, 2026Updated 5 months ago
- [ICLR 2026] Official implementation of "SeC: Advancing Complex Video Object Segmentation via Progressive Concept Construction"☆290Mar 27, 2026Updated 4 months ago
- [NeurIPS 2025] Official implementation of "XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulatio…☆627Oct 22, 2025Updated 9 months ago
- ToonOut, a fork of BiRefNet focused on background removal for anime images. We open-source our dataset & our weights. See our paper at: h…☆101Jun 12, 2026Updated 2 months ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- FantasyPortrait: Enhancing Multi-Character Portrait Animation with Expression-Augmented Diffusion Transformers☆511Aug 20, 2025Updated 11 months ago
- Official repository of paper "ProEdit: Inversion-based Editing From Prompts Done Right"☆116Feb 5, 2026Updated 6 months ago
- [ICML2026] Official Implementation of "TAG: Tangential Amplifying Guidance for Hallucination-Resistant Sampling"☆43Jul 6, 2026Updated last month
- ☆109Sep 3, 2025Updated 11 months ago
- OmniInsert: Mask-Free Video Insertion of Any Reference via Diffusion Transformer Models☆162Mar 4, 2026Updated 5 months ago
- Official implementation for "Story2Board: A Training‑Free Approach for Expressive Storyboard Generation"☆266Aug 22, 2025Updated 11 months ago
- Krea 2 & Klein 9B LoRA - LoKR Studio — train, profile, repair, and extract Krea 2 & Flux 2 Klein 9B LoRAs & LoKRs☆144Updated this week
- The official repository of EditCrafter: Tuning-free High-Resolution Image Editing via Pretrained Diffusion Model (CVPRW 2026)☆51Apr 19, 2026Updated 3 months ago
- [ECCV 2026] Generate high resolution videos with a custom voice and appearance, based on LTX-2/LTX-2.3 + Identity In-Context LoRA☆349Jun 24, 2026Updated last month
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Pusa: Thousands Timesteps Video Diffusion Model☆686Feb 13, 2026Updated 5 months ago
- DreamStyle: A Unified Framework for Video Stylization☆124Jan 7, 2026Updated 7 months ago
- [CVPR 2026] When Numbers Speak: Aligning Textual Numerals and Visual Instances in Text-to-Video Diffusion Models☆68Apr 11, 2026Updated 4 months ago
- Official Code of NAVA: Native Audio-Visual Alignment for Generation.☆214Jun 30, 2026Updated last month
- iMontage: Unified, Versatile, Highly Dynamic Many-to-many Image Generation☆188Dec 1, 2025Updated 8 months ago
- ☆192Jul 31, 2025Updated last year
- [ICML 2026] ReCo: In-Context Generation with Regional Constraints for Instructional Video Editing☆175May 26, 2026Updated 2 months ago
- [NeurIPS'25 Spotlight] Official implementation of "JavisGPT: A Unified Multi-modal LLM for Sounding-Video Comprehension and Generation"☆75Feb 26, 2026Updated 5 months ago
- UniMesh: Unifying 3D Mesh Understanding and Generation☆58Jul 14, 2026Updated 3 weeks ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [IEEE TIP] Official implementation of Progressive Detail Injection for Training-Free Semantic Binding in Text-to-Image Generation☆33Aug 3, 2025Updated last year
- [ICML 2026] Official implementation for "DyPE: Dynamic Position Extrapolation for Ultra High Resolution Diffusion".☆357May 18, 2026Updated 2 months ago
- [CVPR 2026] Offical implementation of the paper "HiFi-Inpaint: Towards High-Fidelity Reference-Based Inpainting for Generating Detail-Pre…☆107Jun 7, 2026Updated 2 months ago
- [ICCV 2025] Inpaint4Drag: Repurposing Inpainting Models for Drag-Based Image Editing via Bidirectional Warping☆97Nov 30, 2025Updated 8 months ago
- ComfyUI custom nodes for Foundation-1 | Structured Text-to-Sample Diffusion for Music Production☆107Mar 22, 2026Updated 4 months ago
- 🔥 [CVPR 2024] The official repo for Zero-Painter!☆70Jun 8, 2024Updated 2 years ago
- [Arxiv 2025] Official PyTorch Implementation of "SVG-T2I: Scaling up Text-to-Image Latent Diffusion Model Without Variational Autoencoder…☆153Dec 18, 2025Updated 7 months ago