We propose a novel fusion strategy that can effectively fuse information from different modality combinations. We also propose a new model named Multi-Modal Segmentation TransFormer (MMSFormer) that incorporates the proposed fusion strategy to perform multimodal material and semantic segmentation tasks.
β33Apr 18, 2024Updated 2 years ago
Alternatives and similar repositories for MMSFormer
Users that are interested in MMSFormer are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ACMMM2025 Oral π] Weaving Any Visual Modalities to Enhance Multimodal Semantic Segmentationβ62Aug 25, 2025Updated 11 months ago
- β74Nov 29, 2023Updated 2 years ago
- β34Mar 23, 2024Updated 2 years ago
- Repository of DELIVER dataset and CMNeXt models (CVPR 2023)β211Aug 16, 2024Updated last year
- Breaking Modality Gap in RGBT Tracking: Coupled Knowledge Distillation ACMMM2024β23Oct 16, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- β17May 22, 2025Updated last year
- β14Jun 29, 2024Updated 2 years ago
- β19Nov 11, 2024Updated last year
- Neural Transmitted Radiance Fieldsβ12Apr 11, 2024Updated 2 years ago
- [Pattern Recognition 2025 π]Unbiased Multiscale Modal Fusion Model for Multimodal Semantic Segmentationβ10Jun 12, 2024Updated 2 years ago
- β22Jun 23, 2025Updated last year
- CVPR 2024 Official Repositoryβ13Mar 27, 2024Updated 2 years ago
- RGBD Pretraining code used in DFormer [ICLR 2024]β21Jul 8, 2025Updated last year
- β12Jan 31, 2026Updated 5 months ago
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- β29Jul 8, 2025Updated last year
- Official repository of "Efficient Fire Segmentation for Internet-of-Things-Assisted Intelligent Transportation Systems" [IEEE TITS 2022]β15Dec 22, 2024Updated last year
- Code for the paper "OTRE: Where Optimal Transport Guided Unpaired Image-to-Image Translation Meets Regularization by Enhancing"β11Aug 2, 2025Updated 11 months ago
- This is the repository for FCN and Transformer based object segmentation that relies on the fusion of camera and LiDAR data.β38Feb 6, 2026Updated 5 months ago
- Implementation Code for paper "Efficient Multimodal Fusion via Interactive Prompting" in CVPR2023β16Jul 24, 2023Updated 3 years ago
- [TCSVT2023] [LASNet] RGB-T Semantic Segmentation with Location, Activation, and Sharpeningβ32Jan 13, 2026Updated 6 months ago
- β17Jun 3, 2026Updated last month
- Official code for "ConTSG-Bench: A Unified Benchmark for Conditional Time Series Generation" οΌICML 2026οΌβ17May 2, 2026Updated 2 months ago
- β41Jun 30, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available β’ AdRun AI, ML, and HPC workloads on powerful cloud GPUsβwithout limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- β87Jun 27, 2024Updated 2 years ago
- Pytorch implementation of our WACV 2023 paper "Image-Consistent Detection of Road Anomalies As Unpredictable Patches"β12May 29, 2024Updated 2 years ago
- Source Code for the JAIR Paper "Does CLIP Know my Face?" (Demo: https://huggingface.co/spaces/AIML-TUDA/does-clip-know-my-face)β15Jul 9, 2024Updated 2 years ago
- β15May 5, 2025Updated last year
- β21Aug 29, 2022Updated 3 years ago
- Official Implementation of "OVS Meets Continual Learning: Towards Sustainable Open-Vocabulary Segmentation" (NeurIPS 2025).β16Feb 27, 2026Updated 4 months ago
- β16Aug 17, 2021Updated 4 years ago
- Superpixel-enhanced Deep Neural Forest for Remote Sensing Image Semantic Segmentationβ15Oct 14, 2020Updated 5 years ago
- MuCR is a benchmark designed to evaluate Multimodal Large Language Models' (MLLMs) ability to discern causal links across modalitiesβ20May 27, 2025Updated last year
- End-to-end encrypted email - Proton Mail β’ AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Reflection Removal Using a Dual-Pixel Sensor, CVPR 2019β17Jun 14, 2019Updated 7 years ago
- [preprint] π COP-GEN: Latent Diffusion Transformer for Copernicus Earth Observation Dataβ18Apr 28, 2026Updated 2 months ago
- MemorySAM: Memorize Modalities and Semantics with Segment Anything Model 2 for Multi-modal Semantic Segmentationβ45Nov 4, 2025Updated 8 months ago
- β90Jan 10, 2025Updated last year
- β14Jun 20, 2023Updated 3 years ago
- Multi Task Learning for Semantic Segmentation, Instance Segmentation and Depth Estimationβ12Jun 12, 2022Updated 4 years ago
- Text-to-face implementation using AttnGan architecture.β17Feb 27, 2022Updated 4 years ago