(AAAI 2025)MUSES: 3D-Controllable Image Generation via Multi-Modal Agent Collaboration
☆37May 21, 2025Updated last year
Alternatives and similar repositories for MUSES
Users that are interested in MUSES are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [NeurIPS 2024] TransAgent: Transfer Vision-Language Foundation Models with Heterogeneous Agent Collaboration☆25Oct 17, 2024Updated last year
- [ICLR2026] Video-GPT via Next Clip Diffusion.☆46Jun 2, 2025Updated last year
- Omni Controllable Video Diffusion☆46Dec 22, 2025Updated 7 months ago
- ☆29Updated this week
- Official implementation for the AAAI2025 paper "PIXELS - Progressive Image Xemplar-based Editing with Latent Surgery"☆11Dec 17, 2024Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆17Jun 29, 2025Updated last year
- An official pytorch implementation of the paper: [MV-Adapter: Multimodal Video Transfer Learning for Video Text Retrieval].☆14Jul 27, 2024Updated last year
- ☆157Feb 28, 2026Updated 4 months ago
- HyperMotion is a pose guided human image animation framework based on a large-scale video diffusion Transformer.☆154May 27, 2026Updated last month
- ☆16Jul 8, 2026Updated 2 weeks ago
- Representing 3D Shapes with 64 Latent Vectors for 3D Diffusion Models☆26Sep 15, 2025Updated 10 months ago
- [ICML26] AVGen-Bench is a task-driven benchmark for multi-granular evaluation of Text-to-Audio-Video (T2AV) generation.☆22Jul 2, 2026Updated 3 weeks ago
- ☆27Apr 25, 2025Updated last year
- [AAAI 2024] MESED: A Multi-modal Entity Set Expansion Dataset with Fine-grained Semantic Classes and Hard Negative Entities☆15Apr 26, 2024Updated 2 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- The official code of ’AdaFuse: Adaptive Medical Image Fusion Based on Spatial-Frequential Cross Attention‘.☆12Dec 11, 2024Updated last year
- MTVCraft: An Open Veo3-style Audio-Video Generation Demo☆98Oct 8, 2025Updated 9 months ago
- Multi-Sensor Place Recognition with Visual and Text Semantics☆23May 27, 2025Updated last year
- [TMM 2026] Rouse the Background in Human Image Animation☆30Apr 24, 2025Updated last year
- Official code for ICCV 2025 paper, X2I: Seamless Integration of Multimodal Understanding into Diffusion Transformer via Attention Distill…☆89Jun 26, 2025Updated last year
- [NN 2024] Code Release of Unsupervised Distribution-aware Keypoints Generation from 3D Point Clouds☆11Feb 20, 2024Updated 2 years ago
- [ICLR 2024] Neural Processing of Tri-Plane Hybrid Neural Fields☆15Feb 21, 2026Updated 5 months ago
- A 3rd-party implemented Face-Xray for deepfake detection.☆13Jun 2, 2020Updated 6 years ago
- The Pytorch implementation for the paper "Fusion is Not Enough: Single Modal Attack on Fusion Models for 3D Object Detection"☆20Mar 9, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆14May 27, 2024Updated 2 years ago
- ☆14Mar 23, 2024Updated 2 years ago
- Automated CAD assembly generation based on Combinatory Logic Synthesis.☆15Jul 7, 2026Updated 2 weeks ago
- ☆17Jun 17, 2020Updated 6 years ago
- ☆16Sep 26, 2024Updated last year
- Using Kolmogorov Arnold Networks (KANs) instead of MLPs in PointNet for Classification and Segmentation of 3D Point Sets☆15Apr 23, 2026Updated 3 months ago
- Re-implementation of VertexRegen [ICCV 25]☆41Jan 25, 2026Updated 6 months ago
- Benchmarking MLLMs for Parametric 3D Generation and Structural Reasoning (Text-to-3D, Image-to-3D, Assembly-3D)☆46Updated this week
- Implementation of "Robust Zero Level-Set Extraction from Unsigned Distance Fields Based on Double Covering"☆44Jun 3, 2026Updated last month
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆19Nov 18, 2024Updated last year
- Mesh to SDF implemented with CUDA.☆49Aug 2, 2024Updated last year
- ☆67Oct 15, 2024Updated last year
- ☆29Mar 30, 2025Updated last year
- UniVid: The Open-Source Unified Video Model☆32Oct 13, 2025Updated 9 months ago
- [CVPR'26] AdapTok: Learning Adaptive and Temporally Causal Video Tokenization in a 1D Latent Space☆29Mar 15, 2026Updated 4 months ago
- This repo includes code for the paper "DynStatF: An Efficient Feature Fusion Strategy for LiDAR 3D Object Detection"☆24Dec 20, 2023Updated 2 years ago