[CVPR 2026] Official code and models for Video Encoder-only Mask Transformer (VidEoMT).
☆253Jul 30, 2026Updated this week
Alternatives and similar repositories for videomt
Users that are interested in videomt are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [CVPR 2026 Workshop] Official code and models for Plain Mask Transformer (PMT).☆54Jul 23, 2026Updated last week
- [CVPR 2025 Highlight] Official code and models for Encoder-only Mask Transformer (EoMT).☆615Jul 22, 2026Updated last week
- Code for the paper "Attention Meets Post-hoc Interpretability: A Mathematical Perspective", ICML 2024☆22Nov 10, 2025Updated 8 months ago
- Official Repository for "Communication Efficient Federated Learning with Generalized Heavy-Ball Momentum", accepted at TMLR 2025☆28Jul 14, 2025Updated last year
- [ICML2026] From Statics to Dynamics: Physics-Aware Image Editing with Latent Transition Priors☆92Apr 30, 2026Updated 3 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Official implementation of "HiERO: understanding the hierarchy of human behavior enhances reasoning on egocentric videos", accepted at IC…☆17May 22, 2026Updated 2 months ago
- Code for the paper "A Sea of Words: An In-Depth Analysis of Anchors for Text Data", AISTATS 2023☆14Oct 26, 2024Updated last year
- We propose a novel modular framework that learns to dynamically mix low-rank adapters (LoRAs) to improve visual analogy learning, enablin…☆75Updated this week
- Code for the paper "SMACE: A New Method for the Interpretability of Composite Decision Systems", ECML 2022☆15Apr 17, 2023Updated 3 years ago
- [ICCVW 2025] Simplifying Traffic Anomaly Detection with Video Foundation Models☆18Dec 4, 2025Updated 7 months ago
- [CVPR 2026 Oral] "INSID3: Training-Free In-Context Segmentation with DINOv3"☆706Jun 26, 2026Updated last month
- Official code for "To Match or Not to Match: Revisiting Image Matching for Reliable Visual Place Recognition" CVPR IMW 2025☆38Oct 4, 2025Updated 10 months ago
- Official implementation of "A Backpack Full of Skills: Egocentric Video Understanding with Diverse Task Perspectives", accepted at CVPR 2…☆24Jun 13, 2024Updated 2 years ago
- Code for the paper "AMEGO: Active Memory from long EGOcentric videos" published at ECCV 2024☆45Dec 7, 2024Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Interface to stable-baselines3 APIs for training RL policies on gym-registered environments☆12Jan 24, 2024Updated 2 years ago
- Code repo for EffectMaker: Unifying Reasoning and Generation for Customized Visual Effect Creation☆42Mar 6, 2026Updated 4 months ago
- [CVPR 2026] Adaptive Spectral Feature Forecasting for Diffusion Sampling Acceleration☆125Apr 30, 2026Updated 3 months ago
- ☆12Jul 22, 2025Updated last year
- Official repository for ICCV23 paper "Divide&Classify: Fine-Grained Classification for City-Wide Visual Place Recognition"☆24Nov 9, 2023Updated 2 years ago
- [CVPR 2026 Oral] "MARCO: Navigating the Unseen Space of Semantic Correspondence"☆146Apr 21, 2026Updated 3 months ago
- [ECCV2026] Official repo for paper "SK-Adapter: Skeleton-Based Structural Control for Native 3D Generation".☆63Jun 26, 2026Updated last month
- [CVPR'26] VecGlypher: Unified Vector Glyph Generation with Language Models☆137Feb 26, 2026Updated 5 months ago
- ViCaS: A Dataset for Combining Holistic and Pixel-level Video Understanding using Captions with Grounded Segmentation (CVPR'25)☆21Apr 2, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ICLR 2026] Deforming Videos to Masks: Flow Matching for Referring Video Segmentation (FlowRVS)☆104Mar 7, 2026Updated 4 months ago
- Depth-Guided Scale-Aware Global Structure-from-Motion☆23Jul 10, 2026Updated 3 weeks ago
- [CVPR 2026 Highlight] MatAnyone 2: Scaling Video Matting via a Learned Quality Evaluator☆798Jul 28, 2026Updated last week
- ALGM applied to Segmenter☆33May 27, 2024Updated 2 years ago
- Domain Randomization via Entropy Maximization☆25Apr 18, 2024Updated 2 years ago
- ☆26Apr 4, 2025Updated last year
- [ECCV 2026] CustomX: Unified Character, Action, and Scene Customization in Video World Models☆96Jun 25, 2026Updated last month
- Implementation of <Streaming Autoregressive Video Generation via Diagonal Distillation> in ICLR 2026☆129Mar 18, 2026Updated 4 months ago
- ☆56Jun 7, 2026Updated last month
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [SIGGRAPH ASIA 2026]: End-to-End Motion Capture for Arbitrary Skeletons☆311Jul 21, 2026Updated last week
- 🌋LavaSR: Fast Speech restoration and enhancement☆567Jun 19, 2026Updated last month
- [CVPR 2025] Official PyTorch implementation of "EdgeTAM: On-Device Track Anything Model"☆955Jan 27, 2026Updated 6 months ago
- [CVPR 2025 Highlight] "SAMWISE: Infusing Wisdom in SAM2 for Text-Driven Video Segmentation"☆386Sep 25, 2025Updated 10 months ago
- DROPO: Sim-to-Real Transfer with Offline Domain Randomization☆26Jul 8, 2025Updated last year
- [CVPR 2026 Highlight] GlyphPrinter: Region-Grouped Direct Preference Optimization for Glyph-Accurate Visual Text Rendering☆103Apr 9, 2026Updated 3 months ago
- World Modeling by Forecasting Vision Foundation Model Features☆51Jul 25, 2026Updated last week