☆19May 8, 2025Updated last year
Alternatives and similar repositories for CVPR2025-STOP
Users that are interested in CVPR2025-STOP are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆23May 8, 2025Updated last year
- [CVPR 2025] DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval☆22Jun 23, 2025Updated last year
- ☆20Aug 7, 2025Updated last year
- [ACMMM 2025] Officially implement of the paper "Seg-Wild: Interactive Segmentation based on 3D Gaussian Splatting for Unconstrained Image…☆22Jul 29, 2025Updated last year
- Text Proxy: Decomposing Retrieval from a 1-to-N Relationship into N 1-to-1 Relationships for Text-Video Retrieval -- AAAI2025☆22May 8, 2026Updated 4 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ICLR2023] Video Scene Graph Generation from Single-Frame Weak Supervision☆12Sep 17, 2023Updated 3 years ago
- ☆13Apr 9, 2026Updated 5 months ago
- DisTime: Distribution-based Time Representation for Video Large Language Models.☆22Jul 10, 2025Updated last year
- ☆13Dec 2, 2024Updated last year
- Official implementation of the ECCV2024 paper: Generalizable Facial Expression Recognition☆22Sep 20, 2024Updated 2 years ago
- [ICLR 2025] TempMe: Video Temporal Token Merging for Efficient Text-Video Retrieval☆27Feb 13, 2025Updated last year
- [CVPR 2026] LongVideo-R1: Smart Navigation for Low-cost Long Video Understanding☆56Jul 7, 2026Updated 2 months ago
- ☆13Jul 10, 2023Updated 3 years ago
- ☆19Oct 8, 2023Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Official implementation of "ImagineFSL: Self-Supervised Pretraining Matters on Imagined Base Set for VLM-based Few-shot Learning" [CVPR 2…☆30Sep 1, 2025Updated last year
- [CVPR2025] VideoICL: Confidence-based Iterative In-context Learning for Out-of-Distribution Video Understanding☆24Mar 24, 2025Updated last year
- The official implementation of "Cross-modal Causal Relation Alignment for Video Question Grounding. (CVPR 2025 Highlight)"☆53Apr 27, 2025Updated last year
- [CVPR 2024] Do you remember? Dense Video Captioning with Cross-Modal Memory Retrieval☆67Jun 19, 2024Updated 2 years ago
- COMP9321 Data Services Engineering Lab☆10Apr 23, 2018Updated 8 years ago
- [ECCV 2024] Official PyTorch implementation of TC-CLIP "Leveraging Temporal Contextualization for Video Action Recognition"☆103Feb 25, 2025Updated last year
- ☆23Mar 7, 2025Updated last year
- LI-FPN is an excellent model for depression recognition based on facial expression.☆15Apr 5, 2024Updated 2 years ago
- Pytorch Code for "Unified Coarse-to-Fine Alignment for Video-Text Retrieval" (ICCV 2023)☆66Jun 7, 2024Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆33Feb 22, 2026Updated 6 months ago
- [CVPR2024] FCS: Feature Calibration and Separation for Non-Exemplar Class Incremental Learning☆20Apr 18, 2025Updated last year
- Agentic Keyframe Search for Video Question Answering☆18Jun 30, 2026Updated 2 months ago
- SpaceVLLM: Endowing Multimodal Large Language Model with Spatio-Temporal Video Grounding Capability☆17May 8, 2025Updated last year
- ☆21Feb 10, 2026Updated 7 months ago
- ☆18Jul 3, 2025Updated last year
- [CVPR 2025 Highlight] Meta LoRA / MetaPEFT: Meta-Learning Hyperparameters for Parameter-Efficient Fine-Tuning (LoRA, Adapter, Prompt Tuni…☆19Sep 12, 2026Updated last week
- The first unofficial implementation of CLIP4Caption: CLIP for Video Caption (ACMMM 2021)☆16Jan 2, 2023Updated 3 years ago
- Code for EMNLP25 paper "Video-RTS: Rethinking Reinforcement Learning and Test-Time Scaling for Efficient and Enhanced Video Reasoning"☆24Feb 18, 2026Updated 7 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Official PyTorch implementation for "ZeroIDIR: Zero-Reference Illumination Degradation Image Restoration with Perturbed Consistency Diffu…☆25May 19, 2026Updated 4 months ago
- IntentVCNet: Bridging Spatio-Temporal Gaps for Intention-Oriented Controllable Video Captioning☆19Aug 16, 2025Updated last year
- [ICLR 2026] "Co-rewarding: Stable Self-supervised RL for Eliciting Reasoning in Large Language Models"☆60Feb 4, 2026Updated 7 months ago
- [ICLR 2026] Official implementation of "Decomposed Attention Fusion in MLLMs for Training-Free Video Reasoning Segmentation"☆36Jan 26, 2026Updated 7 months ago
- [CVPR'2025] Narrating the Video: Boosting Text-Video Retrieval via Comprehensive Utilization of Frame-Level Captions☆19Jan 16, 2026Updated 8 months ago
- AI 游戏助手 — 洛克王国世界 自动导航 + 采矿工具☆35Jun 9, 2026Updated 3 months ago
- [EMNLP25 Main]The official code of "Gradient-Attention Guided Dual-Masking Synergetic Framework for Robust Text-based Person Retrieval"☆26Mar 30, 2026Updated 5 months ago