Improving Mamaba performance on Video Understanding task
☆49Dec 30, 2025Updated 7 months ago
Alternatives and similar repositories for VideoMambaPro
Users that are interested in VideoMambaPro are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The suite of modeling video with Mamba☆294May 14, 2024Updated 2 years ago
- ☆27Jun 4, 2024Updated 2 years ago
- [ECCV2024] VideoMamba: State Space Model for Efficient Video Understanding☆1,124Jul 6, 2024Updated 2 years ago
- ☆84Feb 27, 2025Updated last year
- Collaborative Learning of Anomalies with Privacy (CLAP) for Unsupervised Video Anomaly Detection: A New Baseline☆23Sep 30, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [AAAI'25 Oral] NightReID: A Large-Scale Nighttime Person Re-Identification Benchmark☆11Jun 10, 2025Updated last year
- [PRCV-2024] State Space Model based Frame-Event Tracking☆53Dec 6, 2025Updated 8 months ago
- [AAAI-2024] Structural Information Guided Multimodal Pre-training for Vehicle-centric Perception, Xiao Wang, Wentao Wu, Chenglong Li, Zhi…☆29Jul 29, 2024Updated 2 years ago
- [ACCV 2024] PyTorch Implementation of the Paper 'VideoPatchCore': Official Version☆34Sep 23, 2025Updated 10 months ago
- ☆27Oct 15, 2024Updated last year
- TrackGPT: Track What You Need in Videos via Text Prompts☆25May 16, 2023Updated 3 years ago
- We have implemented Track # 1 for ICME 2024: Spatial Action Localization on Chaotic World dataset. Our mAP on the validation set reaches …☆14Nov 11, 2024Updated last year
- [NeurIPS2024] Multi-Scale VMamba: Hierarchy in Hierarchy Visual State Space Model☆84Dec 25, 2024Updated last year
- ☆25Dec 23, 2024Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- The official implementation of "2024NeurIPS Dynamic Tuning Towards Parameter and Inference Efficiency for ViT Adaptation"☆54Dec 30, 2024Updated last year
- [NeurIPS24 Spotlight] Voxel Mamba: Group-Free State Space Models for Point Cloud based 3D Object Detection☆164Sep 26, 2024Updated last year
- [Official Repo] Visual Mamba: A Survey and New Outlooks☆742Jul 7, 2026Updated last month
- PyTorch implementation of "Efficient Motion Prompt Learning for Robust Visual Tracking" (ICML2025)☆31Dec 17, 2025Updated 7 months ago
- [ECCV 2024 Workshop Best Paper Award] Famba-V: Fast Vision Mamba with Cross-Layer Token Fusion☆34Sep 30, 2024Updated last year
- This is the official repository of Daily-Omni: Towards Audio-Visual Reasoning with Temporal Alignment across Modalities☆47Jul 26, 2026Updated 2 weeks ago
- Mamba in Vision: A Comprehensive Survey of Techniques and Applications☆145Oct 10, 2024Updated last year
- Official Implementation of Video-MA2MBA☆12Dec 3, 2024Updated last year
- ☆12Aug 7, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- The repo for "On-the-fly Modulation for Balanced Multimodal Learning", T-PAMI 2024☆19Sep 29, 2024Updated last year
- Papers of "A Survey on Multimodal LLMs from the Perspective of Input-Output Space Extension"☆20Feb 4, 2026Updated 6 months ago
- This is the repo for "Adaptive Unimodal Regulation for Balanced Multimodal Information Acquisition", CVPR2025.☆26Dec 22, 2025Updated 7 months ago
- The code of paper "O-Mamba: O-shape State-Space Model for Underwater Image Enhancement"☆14Oct 18, 2024Updated last year
- PySlowFast: video understanding codebase from FAIR for reproducing state-of-the-art video models.☆12Jul 26, 2024Updated 2 years ago
- ☆14Sep 12, 2020Updated 5 years ago
- Official PyTorch implementation of the paper "Revisiting Temporal Modeling for CLIP-based Image-to-Video Knowledge Transferring"☆107Jan 28, 2024Updated 2 years ago
- The official implementation for SETA (TIP 2024).☆12Feb 17, 2025Updated last year
- ☆12Jul 26, 2022Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ICCV2023: Disentangling Spatial and Temporal Learning for Efficient Image-to-Video Transfer Learning☆41Sep 25, 2023Updated 2 years ago
- A Triton Kernel for incorporating Bi-Directionality in Mamba2☆83Dec 18, 2024Updated last year
- This is a simple toolkit to view and crop image patches for image/video super-resolution tasks.☆11Jan 6, 2023Updated 3 years ago
- OmniStyle: Filtering High Quality Style Transfer Data at Scale (CVPR 2025)☆35Aug 9, 2025Updated last year
- ☆15Feb 18, 2024Updated 2 years ago
- [CVPR 2024] Adapting Short-Term Transformers for Action Detection in Untrimmed Videos☆11Jun 11, 2024Updated 2 years ago
- ☆12Apr 19, 2024Updated 2 years ago