[ACMMM2025 Oral π] Weaving Any Visual Modalities to Enhance Multimodal Semantic Segmentation
β61Aug 25, 2025Updated last year
Alternatives and similar repositories for StitchFusion
Users that are interested in StitchFusion are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [Pattern Recognition 2025 π]Unbiased Multiscale Modal Fusion Model for Multimodal Semantic Segmentationβ10Jun 12, 2024Updated 2 years ago
- [CVPR 2025] This is the official implementation of Keep the Balance: A Parameter-Efficient Symmetrical Framework for RGB+X Semantic Segmeβ¦β26Aug 7, 2025Updated last year
- We propose a novel fusion strategy that can effectively fuse information from different modality combinations. We also propose a new modeβ¦β33Apr 18, 2024Updated 2 years ago
- β74Nov 29, 2023Updated 2 years ago
- An Open-Source Benchmark and Baseline for Multi-temporal Referring Segmentationβ57Jun 4, 2026Updated 2 months ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [CVPR2026 π] The first attempt to Marine Open Vocabulary Instance Segmentationβ56Jun 16, 2026Updated 2 months ago
- β18Jun 3, 2026Updated 2 months ago
- β34Mar 23, 2024Updated 2 years ago
- [AAAI2026 Oralπ] Towards Open Vocabulary Semantic Segmentation In Remote Sensing.β86Aug 18, 2026Updated last week
- [ACM MM 2025] LIDAR: Lightweight Adaptive Cue-Aware Fusion Vision Mamba for Multimodal Segmentation of Structural Cracksβ24Nov 18, 2025Updated 9 months ago
- #ICCV, #MoE, #Trackingβ38Jul 11, 2025Updated last year
- MemorySAM: Memorize Modalities and Semantics with Segment Anything Model 2 for Multi-modal Semantic Segmentationβ45Nov 4, 2025Updated 9 months ago
- A Cross-Modality Feature Adaptive Interaction Approach for RGB-Infrared Object Detection in Aerial Imageryβ22Mar 12, 2026Updated 5 months ago
- A new dataset for fusion network training and evaluationβ20Jul 18, 2025Updated last year
- Managed Database hosting by DigitalOcean β’ AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- β96Aug 18, 2024Updated 2 years ago
- Repository of DELIVER dataset and CMNeXt models (CVPR 2023)β211Aug 16, 2024Updated 2 years ago
- [CVPR2026 π] Training-Free Underwater World Segmentationβ43Jun 8, 2026Updated 2 months ago
- Multimodal image fusion and segmentationβ30Dec 20, 2025Updated 8 months ago
- β21May 3, 2025Updated last year
- β21Jul 18, 2025Updated last year
- Controllable-LPMoE: Adapting to Challenging Object Segmentation via Dynamic Local Priors from Mixture-of-Experts (ICCV, 2025)β26Dec 8, 2025Updated 8 months ago
- β439Sep 2, 2024Updated last year
- ζ·±εΊ¦ε¦δΉβ28Sep 15, 2025Updated 11 months ago
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Official implementation of MaeFuse οΌTIP 2025οΌβ55Apr 3, 2026Updated 4 months ago
- β29May 20, 2025Updated last year
- RGB-T Fusion, RGB-T SOD, RGB-T Vehicle Detection, RGB-T Crowd Counting, RGB-T Pedestrian Detection, RGB-T Semantic Segmeantaion, RGB-T Trβ¦β236Aug 9, 2026Updated 2 weeks ago
- [ACM MM 2023 Oral] Multispectral Object Detection via Cross-Modal Conflict-Aware Learning.β68Mar 13, 2024Updated 2 years ago
- [ECCV 2024] SDK for MUSES: The Multi-Sensor Semantic Perception Dataset for Driving under Uncertaintyβ43Feb 16, 2026Updated 6 months ago
- BRAVO Challenge Toolkit and Evaluation Codeβ21Apr 30, 2025Updated last year
- β31Oct 8, 2024Updated last year
- Utility to convert the NYU Depth V2 dataset into point clouds for advanced 3D visualization and analysis.β15Nov 27, 2024Updated last year
- Towards Open-Vocabulary Learing for Remote Sensing: A surveyβ35Jul 5, 2026Updated last month
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Fine-Grained Pixel-Text Alignment for Open-Vocabulary Semantic Segmentationβ16Mar 28, 2026Updated 4 months ago
- β17Jul 28, 2025Updated last year
- Official implementation of "Pix2Next: Leveraging Vision Foundation Models for RGB to NIR Image Translation"β16Nov 13, 2025Updated 9 months ago
- β16Nov 7, 2025Updated 9 months ago
- Exploring Deeper! Segment Anything Model with Depth Perception for Camouflaged Object Detection, ACM Multimedia (MM), 2024β25Oct 15, 2024Updated last year
- β15Feb 17, 2025Updated last year
- β73Jul 23, 2024Updated 2 years ago