[CVPR 2025 Highlight] "SAMWISE: Infusing Wisdom in SAM2 for Text-Driven Video Segmentation"
☆386Sep 25, 2025Updated 10 months ago
Alternatives and similar repositories for SAMWISE
Users that are interested in SAMWISE are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [NeurIPS 2025 Spotlight] "SANSA: Unleashing the Hidden Semantics in SAM2 for Few-Shot Segmentation."☆204Dec 17, 2025Updated 7 months ago
- Official Repository for "Communication Efficient Federated Learning with Generalized Heavy-Ball Momentum", accepted at TMLR 2025☆28Jul 14, 2025Updated last year
- [CVPR 2026 Oral] "MARCO: Navigating the Unseen Space of Semantic Correspondence"☆148Apr 21, 2026Updated 3 months ago
- ☆26Apr 4, 2025Updated last year
- [CVPR 2025, IJCV 2026] "A Distractor-Aware Memory for Visual Object Tracking with SAM2", "Distractor-Aware Memory-Based Visual Object Tra…☆492Apr 7, 2026Updated 4 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [CVPR 2026 Oral] "INSID3: Training-Free In-Context Segmentation with DINOv3"☆716Jun 26, 2026Updated last month
- Official implementation of https://arxiv.org/abs/2106.03496☆15Jul 27, 2022Updated 4 years ago
- [ICCV 2025] MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation☆23Sep 5, 2025Updated 11 months ago
- Official implementation of "A Backpack Full of Skills: Egocentric Video Understanding with Diverse Task Perspectives", accepted at CVPR 2…☆24Jun 13, 2024Updated 2 years ago
- [CVPR 2025] Official PyTorch Implementation of GLUS: Global-Local Reasoning Unified into A Single Large Language Model for Video Segmenta…☆70Jun 23, 2025Updated last year
- (ICCV 2025) ReferDINO: Referring Video Object Segmentation with Visual Grounding Foundations☆142Nov 14, 2025Updated 8 months ago
- [CVPR 2026 Workshop] Official code and models for Plain Mask Transformer (PMT).☆56Jul 23, 2026Updated 2 weeks ago
- Code for the paper "Exploring Pre-trained Text-to-Video Diffusion Models for Referring Video Object Segmentation", ECCV 2024☆48Sep 28, 2024Updated last year
- Code for the paper "Attention Meets Post-hoc Interpretability: A Mathematical Perspective", ICML 2024☆22Nov 10, 2025Updated 9 months ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- [CVPR 2025 Highlight] Official code and models for Encoder-only Mask Transformer (EoMT).☆615Jul 22, 2026Updated 3 weeks ago
- List of papers wrote by Focoos AI research team!☆12Jun 3, 2025Updated last year
- Interface to stable-baselines3 APIs for training RL policies on gym-registered environments☆12Jan 24, 2024Updated 2 years ago
- Official implementation of "HiERO: understanding the hierarchy of human behavior enhances reasoning on egocentric videos", accepted at IC…☆17May 22, 2026Updated 2 months ago
- DROPO: Sim-to-Real Transfer with Offline Domain Randomization☆26Jul 8, 2025Updated last year
- [CVPR-2024] Decoupling Static and Hierarchical Motion Perception for Referring Video Segmentation☆83Jul 24, 2024Updated 2 years ago
- [CVPR 2024] PEM: Prototype-based Efficient MaskFormer for Image Segmentation☆129Mar 10, 2025Updated last year
- Official Repo For Pixel-LLM Codebase: Sa2VA (T-PAMI-26), SAMTok (CVPR-26), VRT (Arxiv-25), SaSaSa2VA (1-st solution for LSVOS)☆1,655Aug 4, 2026Updated last week
- 🔥 Latest advances in Video Object Segmentation (VOS) – papers, datasets, and projects.☆517Jul 13, 2026Updated 3 weeks ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [ICCV 2025] Official implementation of "InstructSeg: Unifying Instructed Visual Segmentation with Multi-modal Large Language Models"☆56Feb 10, 2025Updated last year
- Code for the paper "SMACE: A New Method for the Interpretability of Composite Decision Systems", ECML 2022☆15Apr 17, 2023Updated 3 years ago
- 🚀 Lightning-fast computer vision models. Fine-tune SOTA models with just a few lines of code. Ready for cloud ☁️ and edge 📱 deployment.…☆352Dec 11, 2025Updated 8 months ago
- ☆19May 20, 2022Updated 4 years ago
- This repo aims to include materials (papers, codes, slides) about SAM2 (segment anything in images and videos). We are continuously impro…☆153Oct 1, 2025Updated 10 months ago
- ☆11Sep 24, 2021Updated 4 years ago
- Domain Randomization via Entropy Maximization☆25Apr 18, 2024Updated 2 years ago
- ☆31Oct 27, 2022Updated 3 years ago
- Official code for WACV 2024 paper, "Annotation-free Audio-Visual Segmentation"☆38Oct 11, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [CVPR 2025] Official PyTorch implementation of "EdgeTAM: On-Device Track Anything Model"☆961Jan 27, 2026Updated 6 months ago
- Official Implementation of "Open-Vocabulary Audio-Visual Semantic Segmentation" [ACM MM 2024 Oral].☆37Nov 2, 2024Updated last year
- A list of referring video object segmentation papers☆65Jun 28, 2026Updated last month
- [ICCV 2025] SAM2Long: Enhancing SAM 2 for Long Video Segmentation with a Training-Free Memory Tree☆567Jul 29, 2025Updated last year
- Official code of "EVF-SAM: Early Vision-Language Fusion for Text-Prompted Segment Anything Model"☆506Mar 17, 2025Updated last year
- [CVPR 2026] Official code and models for Video Encoder-only Mask Transformer (VidEoMT).