[ICLR 2026] Official implementation of "Decomposed Attention Fusion in MLLMs for Training-Free Video Reasoning Segmentation"
☆35Jan 26, 2026Updated 5 months ago
Alternatives and similar repositories for DecAF
Users that are interested in DecAF are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Learning to Enhance Aperture Phasor Field for Non-Line-of-Sight Imaging☆15Dec 31, 2024Updated last year
- [WACV-2025] Exploring Scalability of Self-Training for Open-Vocabulary Temporal Action Localization☆17May 28, 2025Updated last year
- [ICCV2025] ExploreGS: Explorable 3D Scene Reconstruction with Virtual Camera Samplings and Diffusion Priors☆58May 17, 2026Updated 2 months ago
- [ICCV'25] CCMNet: Leveraging Calibrated Color Correction Matrices for Cross-Camera Color Constancy☆25Jun 22, 2026Updated 3 weeks ago
- [ICCV 2025] CoMoGaussian: Continuous Motion-Aware Gaussian Splatting from Motion-Blurred Images☆55Jul 15, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Representing 3D Shapes with 64 Latent Vectors for 3D Diffusion Models☆26Sep 15, 2025Updated 10 months ago
- [ECCV 2024] VISAGE: Video Instance Segmentation with Appearance-Guided Enhancement☆38Jul 29, 2024Updated last year
- Official Code for the paper "HieraMamba: Video Temporal Grounding via Hierarchical Anchor-Mamba Pooling"☆15Apr 30, 2026Updated 2 months ago
- [AAAI 2026] 4D Scaffold Gaussian Splatting with Dynamic-Aware Anchor Growing for Efficient and High-Fidelity Dynamic Scene Reconstruction☆27Jan 15, 2026Updated 6 months ago
- ☆12Aug 7, 2024Updated last year
- Official Pytorch Implementation of 'BAM-DETR: Boundary-Aligned Moment Detection Transformer for Temporal Sentence Grounding in Videos'☆36Feb 26, 2025Updated last year
- [ICLR 2026] Official implementation for CoT-RVS☆23Mar 17, 2026Updated 4 months ago
- [CoRL 2025] UniSkill: Imitating Human Videos via Cross-Embodiment Skill Representations☆87Dec 18, 2025Updated 7 months ago
- ☆42May 7, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [CVPR 2026 Findings] SwiftVGGT: A Scalable Visual Geometry Grounded Transformer for Large-Scale Scenes☆93Nov 25, 2025Updated 7 months ago
- Official PyTorch Implementation of the Paper "Test-Time Adaptation of 3D Point Clouds via Denoising Diffusion Models"☆23Apr 24, 2025Updated last year
- [ECCV 2024] Official PyTorch implementation of "Classification Matters: Improving Video Action Detection with Class-Specific Attention"☆18Nov 8, 2024Updated last year
- Official Code for the paper "UniversalVTG: A Univeral and Lightweight Foundation Model for Video Temporal Grounding"☆15Apr 15, 2026Updated 3 months ago
- [AAAI 2025] Video Diffusion Models are Strong Video Inpainter☆17Jul 21, 2025Updated last year
- [CVPR 2024] Guided Slot Attention for Unsupervised Video Object Segmentation☆66Dec 23, 2024Updated last year
- Official code of Veason-R1☆15Jul 14, 2026Updated last week
- DisTime: Distribution-based Time Representation for Video Large Language Models.☆21Jul 10, 2025Updated last year
- (ICML 2026) Seg-ReSearch: Segmentation with Interleaved Reasoning and External Search☆47May 1, 2026Updated 2 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [CVPR 2025] CoCoGaussian: Leveraging Circle of Confusion for Gaussian Splatting from Defocused Images☆46Nov 25, 2025Updated 7 months ago
- [ICCVW 2025] TransFlow: Motion Knowledge Transfer from Video Diffusion Models to Video Salient Object Detection☆15Oct 22, 2025Updated 8 months ago
- Repository for the CVPR23 paper Re^2TAL☆13Nov 21, 2025Updated 7 months ago
- 🔮 UniPixel: Unified Object Referring and Segmentation for Pixel-Level Visual Reasoning (NeurIPS 2025)☆247Jan 4, 2026Updated 6 months ago
- [ECCV 2024 Oral] Official implementation of the paper "DEVIAS: Learning Disentangled Video Representations of Action and Scene"☆29Nov 15, 2025Updated 8 months ago
- [CVPR 2026] Accelerating Streaming Video Large Language Models via Hierarchical Token Compression☆70Jun 8, 2026Updated last month
- [CVPR'24] Official repo of "Attentive Illumination Decomposition Model for Multi-Illuminant White Balancing"☆19Jul 3, 2025Updated last year
- [DATE 2023] Pipe-BD: Pipelined Parallel Blockwise Distillation☆12Jul 13, 2023Updated 3 years ago
- [NeurIPS 25] InfiniPot-V: Memory-Constrained KV Cache Compression for Streaming Video Understanding☆20Jan 25, 2026Updated 5 months ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- [CVPR 2026] Wavelet-based Frame Selection by Detecting Semantic Boundary for Long Video Understanding☆31Apr 12, 2026Updated 3 months ago
- A Benchmark and Agentic Framework for Omni-Modal Reasoning and Tool Use in Long Videos☆21Jun 20, 2026Updated last month
- This repository contains the implementation of FAPM (2023 ICASSP).☆25Jun 19, 2023Updated 3 years ago
- ☆57Sep 13, 2024Updated last year
- [ AAAI 2026 ] The official implementation of 'MonoCLUE: Object-Aware Clustering Enhances Monocular 3D Object Detection'☆21Mar 23, 2026Updated 3 months ago
- [NeurIPS 2023] The official implementation of SOC: Semantic-Assisted Object Cluster for Referring Video Object Segmentation☆33Mar 16, 2024Updated 2 years ago
- VITA: Video Instance Segmentation via Object Token Association (NeurIPS 2022)☆107Jan 4, 2024Updated 2 years ago