[Pattern Recognition 2025 π]Unbiased Multiscale Modal Fusion Model for Multimodal Semantic Segmentation
β10Jun 12, 2024Updated 2 years ago
Alternatives and similar repositories for U3M
Users that are interested in U3M are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ACMMM2025 Oral π] Weaving Any Visual Modalities to Enhance Multimodal Semantic Segmentationβ61Aug 25, 2025Updated last year
- Jin, Xiao, et al. "FCMNet: Frequency-aware cross-modality attention networks for RGB-D salient object detection." Neurocomputing 491 (202β¦β11Apr 11, 2024Updated 2 years ago
- Towards Open-Vocabulary Learing for Remote Sensing: A surveyβ36Jul 5, 2026Updated last month
- Fine-Grained Pixel-Text Alignment for Open-Vocabulary Semantic Segmentationβ16Mar 28, 2026Updated 5 months ago
- [CVPR2026 π] Training-Free Underwater World Segmentationβ43Jun 8, 2026Updated 2 months ago
- GPUs on demand by Runpod - Special Offer Available β’ AdRun AI, ML, and HPC workloads on powerful cloud GPUsβwithout limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- SpectralX: Parameter-efficient Domain Generalization for Spectral Remote Sensing Foundation Models, ISPRS, 2026.β29Jul 3, 2026Updated last month
- β16Jan 21, 2025Updated last year
- [CVPR 2026 Findings] FineCog-Nav: Integrating Fine-grained Cognitive Modules for Zero-shot Multimodal UAV Navigationβ20Aug 10, 2026Updated 2 weeks ago
- TRACE, a framework for turn-aware credit assignment for multi-turn jailbreak optimizationβ22Updated this week
- β15Dec 20, 2024Updated last year
- Paper list for LLM/MLLM-based image segmentationβ48Dec 24, 2025Updated 8 months ago
- Implementation of "DIME-FM: DIstilling Multimodal and Efficient Foundation Models"