Improving Mamaba performance on Video Understanding task
☆49Dec 30, 2025Updated 8 months ago
Alternatives and similar repositories for VideoMambaPro
Users that are interested in VideoMambaPro are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The suite of modeling video with Mamba☆294May 14, 2024Updated 2 years ago
- ☆27Jun 4, 2024Updated 2 years ago
- [ECCV2024] VideoMamba: State Space Model for Efficient Video Understanding☆1,125Jul 6, 2024Updated 2 years ago
- ☆84Feb 27, 2025Updated last year
- Collaborative Learning of Anomalies with Privacy (CLAP) for Unsupervised Video Anomaly Detection: A New Baseline☆23Sep 30, 2024Updated last year
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- [AAAI'25 Oral] NightReID: A Large-Scale Nighttime Person Re-Identification Benchmark☆11Jun 10, 2025Updated last year
- [PRCV-2024] State Space Model based Frame-Event Tracking☆53Dec 6, 2025Updated 8 months ago
- [AAAI-2024] Structural Information Guided Multimodal Pre-training for Vehicle-centric Perception, Xiao Wang, Wentao Wu, Chenglong Li, Zhi…☆29Jul 29, 2024Updated 2 years ago
- In this repository, a simple implementation of Video augmentation is provided to augment videos for machine learning training tasks.☆20Dec 4, 2024Updated last year
- [ACCV 2024] PyTorch Implementation of the Paper 'VideoPatchCore': Official Version☆34Sep 23, 2025Updated 11 months ago
- [ACMMM 2023] BMMAL: Towards Balanced Active Learning for Multimodal Classification☆17Sep 25, 2023Updated 2 years ago
- ☆27Oct 15, 2024Updated last year
- Official Implementation of Attentive Mask CLIP (ICCV2023, https://arxiv.org/abs/2212.08653)☆38May 29, 2024Updated 2 years ago
- TrackGPT: Track What You Need in Videos via Text Prompts☆25May 16, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [NeurIPS2024] Multi-Scale VMamba: Hierarchy in Hierarchy Visual State Space Model☆84Dec 25, 2024Updated last year
- The official implementation of "2024NeurIPS Dynamic Tuning Towards Parameter and Inference Efficiency for ViT Adaptation"☆55Dec 30, 2024Updated last year
- [Official Repo] Visual Mamba: A Survey and New Outlooks☆744Aug 17, 2026Updated 2 weeks ago
- [ECCV 2024 Workshop Best Paper Award] Famba-V: Fast Vision Mamba with Cross-Layer Token Fusion☆34Sep 30, 2024Updated last year
- Code for "Purify Unlearnable Examples via Rate-Constrained Variational Autoencoders" at ICML 2024☆11Sep 18, 2025Updated 11 months ago
- Mamba in Vision: A Comprehensive Survey of Techniques and Applications☆145Oct 10, 2024Updated last year
- MobileSAM のエンコーダー/デコーダーをONNXに変換し、推論するサンプル☆12Apr 11, 2024Updated 2 years ago
- ☆12Aug 7, 2024Updated 2 years ago
- The repo for "On-the-fly Modulation for Balanced Multimodal Learning", T-PAMI 2024☆19Sep 29, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Papers of "A Survey on Multimodal LLMs from the Perspective of Input-Output Space Extension"☆21Feb 4, 2026Updated 6 months ago
- This is the repo for "Adaptive Unimodal Regulation for Balanced Multimodal Information Acquisition", CVPR2025.☆27Dec 22, 2025Updated 8 months ago
- The code of paper "O-Mamba: O-shape State-Space Model for Underwater Image Enhancement"☆14Oct 18, 2024Updated last year
- PySlowFast: video understanding codebase from FAIR for reproducing state-of-the-art video models.☆12Jul 26, 2024Updated 2 years ago
- ☆14Sep 12, 2020Updated 5 years ago
- Placeholder☆10Jul 17, 2023Updated 3 years ago
- Official PyTorch implementation of the paper "Revisiting Temporal Modeling for CLIP-based Image-to-Video Knowledge Transferring"☆106Jan 28, 2024Updated 2 years ago
- ICCV2023: Disentangling Spatial and Temporal Learning for Efficient Image-to-Video Transfer Learning☆41Sep 25, 2023Updated 2 years ago
- A Triton Kernel for incorporating Bi-Directionality in Mamba2☆83Dec 18, 2024Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- This is a simple toolkit to view and crop image patches for image/video super-resolution tasks.☆11Jan 6, 2023Updated 3 years ago
- ☆18Aug 23, 2022Updated 4 years ago
- Generalized and Incremental Few-Shot Learning by Explicit Learning and Calibration without Forgetting, (ICCV'21)☆14Aug 4, 2022Updated 4 years ago
- ☆15Feb 18, 2024Updated 2 years ago
- [CVPR 2024] Adapting Short-Term Transformers for Action Detection in Untrimmed Videos☆11Jun 11, 2024Updated 2 years ago
- published in IEEE Transactions on Image Processing (TIP), 2023☆27Mar 4, 2023Updated 3 years ago
- ☆12Apr 19, 2024Updated 2 years ago