[ECCV2026] MomentSeg: Moment-Centric Sampling for Enhanced Video Pixel Understanding
☆25Jun 19, 2026Updated 2 months ago
Alternatives and similar repositories for MomentSeg
Users that are interested in MomentSeg are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICCV2025] PropVG: End-to-End Proposal-Driven Visual Grounding with Multi-Granularity Discrimination☆32Oct 13, 2025Updated 10 months ago
- [AAAI2025 selected as oral] - Multi-task Visual Grounding with Coarse-to-Fine Consistency Constraints☆45Jul 2, 2025Updated last year
- [ICCV2025] DeRIS: Decoupling Perception and Cognition for Enhanced Referring Image Segmentation through Loopback Synergy☆48Nov 21, 2025Updated 9 months ago
- [PR2026] Drone Referring Localization: An Efficient Heterogeneous Spatial Feature Interaction Method For UAV Self-Localization☆96Feb 19, 2026Updated 6 months ago
- [NeurIPS2024] - SimVG: A Simple Framework for Visual Grounding with Decoupled Multi-modal Fusion☆104Oct 29, 2025Updated 10 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Official repository of OS-FPI☆17Dec 22, 2024Updated last year
- [ECCV2024]FALIP: Visual Prompt as Foveal Attention Boosts CLIP Zero-Shot Performance☆18Sep 11, 2024Updated last year
- 「TCSVT2021」A Transformer-Based Feature Segmentation and Region Alignment Method For UAV-View Geo-Localization☆122Mar 7, 2024Updated 2 years ago
- [ICCV 2025] MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation☆23Sep 5, 2025Updated 11 months ago
- Official code of Veason-R1☆16Jul 14, 2026Updated last month
- [CVPR 2026] Refer-Agent: A Collaborative Multi-Agent System with Reasoning and Reflection for Referring Video Object Segmentation☆37Mar 12, 2026Updated 5 months ago
- ☆18May 18, 2026Updated 3 months ago
- VPTracker: Global Vision-Language Tracking via Visual Prompt and MLLM☆16Mar 10, 2026Updated 5 months ago
- Video Reasoning Segmentation☆26Nov 29, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆23Aug 20, 2024Updated 2 years ago
- [CVPR 2025] Official PyTorch Implementation of GLUS: Global-Local Reasoning Unified into A Single Large Language Model for Video Segmenta…☆70Jun 23, 2025Updated last year
- [CVPR'25 Highlight] A VQA benchmark for 6D spatial reasoning.☆20Apr 29, 2026Updated 4 months ago
- LEO: A powerful Hybrid Multimodal LLM☆20Jan 18, 2025Updated last year
- Stabilizing an Inverted Pendulum on a cart using Deep Reinforcement Learning☆10Jul 8, 2018Updated 8 years ago
- paper list on Video Moment Retrieval (VMR), or Temporal Video Grounding (TVG), Video Grounding (VG), or Temporal Sentence Grounding in Vi…☆43Jul 30, 2026Updated last month
- [NeurIPS 2024] Repository for the paper "OVT-B: A New Large-Scale Benchmark for Open-Vocabulary Multi-Object Tracking".☆29Nov 9, 2024Updated last year
- ☆10Jan 6, 2025Updated last year
- Official Implementation of ECCV2024 paper: SLAck☆29Sep 18, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official codes of "Sketch-in-Latents: Eliciting Unified Reasoning in MLLMs"☆17Feb 15, 2026Updated 6 months ago
- (CVPR 2026) Long-RVOS: A Comprehensive Benchmark for Long-term Referring Video Object Segmentation☆39Feb 28, 2026Updated 6 months ago
- Transactions on Multimedia (TMM25)☆21Apr 8, 2025Updated last year
- The evaluation tool (Matlab version) for saliency maps.☆10Mar 18, 2022Updated 4 years ago
- A guide to render bokeh images by Blender 2.93☆13Mar 20, 2022Updated 4 years ago
- [ICCV 2025] Official implementation of "InstructSeg: Unifying Instructed Visual Segmentation with Multi-modal Large Language Models"☆56Feb 10, 2025Updated last year
- ☆33Oct 23, 2024Updated last year
- CurriculumLoc for Visual Geo-localization☆16Nov 23, 2023Updated 2 years ago
- [ICCV 2023] ADNet: Lane Shape Prediction via Anchor Decomposition☆37Oct 11, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [CVPR'24] Code for Emergent Open-Vocabulary Semantic Segmentation from Off-the-shelf Vision-Language Models☆18Jul 22, 2024Updated 2 years ago
- AAAI 25' Flexible Image Reflection Removal with Sparse Human Guidance☆12Jul 7, 2025Updated last year
- ☆12Dec 10, 2019Updated 6 years ago
- Project Page for ICLR'26: CoPRS, offering training overview, inference code, and downloadable links.☆24Mar 17, 2026Updated 5 months ago
- Dynamic Selective Network for RGB-D Salient Object Detection☆12Jan 22, 2025Updated last year
- [AAAI 2026] Segment Anything Across Shots: A Method and Benchmark☆30Nov 16, 2025Updated 9 months ago
- [ECCV 2026] Real-Time Interactive Multi-Target Video Segmentation☆61Jul 10, 2026Updated last month