Repository for "CAFF-DINO: Multi-spectral object detection transformers with cross-attention features fusion" [Helvig et al.], accepted in the 20 th IEEE Workshop Perception Beyond the Visible Spectrum [CVPR 2024]. Propose an adaptation of DETRs models for IR-visible features fusion.
☆36Jan 27, 2025Updated last year
Alternatives and similar repositories for CAFF-DETR
Users that are interested in CAFF-DETR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆15Jun 21, 2022Updated 4 years ago
- [WACV2025] MiPa: Mixed Patch Infrared-Visible Modality Agnostic Object Detection☆29Dec 9, 2024Updated last year
- ☆73Jul 23, 2024Updated 2 years ago
- GM-DETR: Generalized Muiltispectral DEtection TRansformer with Efficient Fusion Encoder for Visible-Infrared Detection (Paddle&Torch)☆52Aug 27, 2024Updated 2 years ago
- ☆36Dec 31, 2024Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- [MICCAI 2024] Feature Fusion Based on Mutual-Cross-Attention Mechanism for EEG Emotion Recognition☆60Dec 14, 2024Updated last year
- [Pattern Recognition 2025 🌟]Unbiased Multiscale Modal Fusion Model for Multimodal Semantic Segmentation☆10Jun 12, 2024Updated 2 years ago
- ICAFusion: Iterative Cross-Attention Guided Feature Fusion for Multispectral Object Detection, Pattern Recognition☆273Jun 3, 2026Updated 2 months ago
- ☆23Jun 18, 2024Updated 2 years ago
- ☆11May 16, 2025Updated last year
- A human-annotated, fine-grained dataset for Vision-and-Language Navigation☆16Jan 20, 2022Updated 4 years ago
- Bio-inspired Motion Integration DETR for Infrared Small Target Detection.☆23Mar 11, 2026Updated 5 months ago
- This is a laboratory code of paper---MMDRFuse: Distilled Mini-Model with Dynamic Refresh for Multi-Modality Image Fusion☆28Sep 3, 2024Updated 2 years ago
- ☆14Jun 24, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- RGBT Tracking via All-layer Multimodal Interactions with Mamba☆19May 7, 2025Updated last year
- ☆17Dec 11, 2023Updated 2 years ago
- KRF: Keypoint Refinement with Fusion Network for 6D Pose Estimation☆16Nov 9, 2024Updated last year
- DEYOLO: Dual-Feature-Enhancement YOLO for Cross-Modality Object Detection☆182Jun 20, 2025Updated last year
- [TIP 2025] Official implementation for "MAFS: Masked Autoencoder for Infrared-Visible Image Fusion and Semantic Segmentation"☆23Jun 26, 2026Updated 2 months ago
- ☆24Dec 22, 2023Updated 2 years ago
- Official repository of our work: MS-DETR: Multispectral Pedestrian Detection Transformer with Loosely Coupled Fusion and Modality-Balance…☆29Sep 8, 2024Updated last year
- M-SpecGene: Generalized Foundation Model for RGBT Multispectral Vision (ICCV 2025)☆39Nov 19, 2025Updated 9 months ago
- A2FSeg: Adaptive Multi-Modal Fusion Network for Medical Image Segmentation☆50Nov 7, 2023Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Benchmark for Asphalt Pavement Crack Segmentation Using RGB and Infrared Images☆17Oct 13, 2024Updated last year
- ☆12Sep 3, 2021Updated 4 years ago
- Calibrated and Complementary Transformer for RGB-Infrared Object Detection☆126May 9, 2024Updated 2 years ago
- ☆29Jul 24, 2024Updated 2 years ago
- Winning 3rd Place solution for HubMap - Hacking the Human Vasculature hosted on Kaggle☆16Aug 10, 2023Updated 3 years ago
- A collection of deep learning based RGB-T-Fusion methods, codes, and datasets. The main directions involved are Multispectral Pedestrian …☆753Apr 17, 2026Updated 4 months ago
- Code for A Dual Domain Multi-exposure Image Fusion Network Based on the Spatial-frequency Integration.☆12Jul 25, 2024Updated 2 years ago
- Dual convolutional neural network with attention for image blind denoising (Multimedia Systems, 2024)☆15Oct 25, 2024Updated last year
- ☆14May 23, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Official code of "HS-FPN: High Frequency and Spatial Perception FPN for Tiny Object Detection".☆107Jun 16, 2025Updated last year
- [arXiv 2024] PyTorch implementation of RRD: https://arxiv.org/abs/2407.12073☆15Dec 2, 2025Updated 9 months ago
- ☆17Jun 28, 2024Updated 2 years ago
- Adapt or Perish: Adaptive Sparse Transformer with Attentive Feature Refinement for Image Restoration☆179Jul 16, 2025Updated last year
- Abstract. Person search is a challenging problem with various real- world applications, that aims at joint person detection and re-identi…☆13Feb 28, 2024Updated 2 years ago
- Pose Refinement Graph Convolutional Network for Skeleton-based Action Recognition(RA-L with ICRA 2021)☆22Aug 30, 2022Updated 4 years ago
- [ICDAR 2024] (Best Student Paper🏆) Exploring Knowledge Distillation Towards Document Object Detection with Structured Graph Creation☆14Sep 6, 2024Updated last year