☆19May 8, 2025Updated last year
Alternatives and similar repositories for CVPR2025-STOP
Users that are interested in CVPR2025-STOP are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆23May 8, 2025Updated last year
- [CVPR 2025] DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval☆22Jun 23, 2025Updated last year
- ☆19Aug 7, 2025Updated last year
- Text Proxy: Decomposing Retrieval from a 1-to-N Relationship into N 1-to-1 Relationships for Text-Video Retrieval -- AAAI2025☆21May 8, 2026Updated 3 months ago
- [ICLR2023] Video Scene Graph Generation from Single-Frame Weak Supervision☆12Sep 17, 2023Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Official implementation of paper "Vision Graph Prompting via Semantic Low-Rank Decomposition", ICML 2025☆16Dec 25, 2025Updated 7 months ago
- DisTime: Distribution-based Time Representation for Video Large Language Models.☆21Jul 10, 2025Updated last year
- ☆13Dec 2, 2024Updated last year
- [ICCV 2025] "Player-Centric Multimodal Prompt Generation for Large Language Model Based Identity-Aware Basketball Video Captioning".☆23Dec 11, 2025Updated 7 months ago
- Official implementation of the ECCV2024 paper: Generalizable Facial Expression Recognition☆23Sep 20, 2024Updated last year
- [ICLR 2025] TempMe: Video Temporal Token Merging for Efficient Text-Video Retrieval☆27Feb 13, 2025Updated last year
- ☆15May 5, 2025Updated last year
- ☆13Jul 10, 2023Updated 3 years ago
- ☆18Oct 8, 2023Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [CVPR2025] VideoICL: Confidence-based Iterative In-context Learning for Out-of-Distribution Video Understanding☆24Mar 24, 2025Updated last year
- The official implementation of "Cross-modal Causal Relation Alignment for Video Question Grounding. (CVPR 2025 Highlight)"☆52Apr 27, 2025Updated last year
- [CVPR 2024] Do you remember? Dense Video Captioning with Cross-Modal Memory Retrieval☆66Jun 19, 2024Updated 2 years ago
- COMP9321 Data Services Engineering Lab☆10Apr 23, 2018Updated 8 years ago
- [ECCV 2024] Official PyTorch implementation of TC-CLIP "Leveraging Temporal Contextualization for Video Action Recognition"☆102Feb 25, 2025Updated last year
- ☆22Mar 7, 2025Updated last year
- Official code repository of "UPP: Unified Point-Level Prompting for Robust Point Cloud Analysis", ICCV 2025.☆17Oct 17, 2025Updated 9 months ago
- Official implementation for paper "CEPrompt: Cross-Modal Emotion-Aware Prompting for Facial Expression Recognition" (accepted to IEEE TC…☆17Oct 20, 2025Updated 9 months ago
- Pytorch Code for "Unified Coarse-to-Fine Alignment for Video-Text Retrieval" (ICCV 2023)☆66Jun 7, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Official PyTorch implementation of CVPR'26 paper "Unleashing Vision-Language Semantics for Deepfake Video Detection".☆17May 15, 2026Updated 2 months ago
- [CVPR2024] FCS: Feature Calibration and Separation for Non-Exemplar Class Incremental Learning☆20Apr 18, 2025Updated last year
- Agentic Keyframe Search for Video Question Answering☆18Jun 30, 2026Updated last month
- SpaceVLLM: Endowing Multimodal Large Language Model with Spatio-Temporal Video Grounding Capability☆17May 8, 2025Updated last year
- [CVPR 2025 Highlight] Meta LoRA / MetaPEFT: Meta-Learning Hyperparameters for Parameter-Efficient Fine-Tuning (LoRA, Adapter, Prompt Tuni…☆18Mar 4, 2026Updated 5 months ago
- ☆21Feb 10, 2026Updated 6 months ago
- ☆18Sep 30, 2025Updated 10 months ago
- The first unofficial implementation of CLIP4Caption: CLIP for Video Caption (ACMMM 2021)☆16Jan 2, 2023Updated 3 years ago
- Official PyTorch implementation for "ZeroIDIR: Zero-Reference Illumination Degradation Image Restoration with Perturbed Consistency Diffu…☆20May 19, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Implementation of CV-SLT (Conditional Variational Autoencoder for Sign Language Translation with Cross-Modal Alignment).☆20Mar 13, 2025Updated last year
- [ICLR 2026] Official implementation of "Decomposed Attention Fusion in MLLMs for Training-Free Video Reasoning Segmentation"☆36Jan 26, 2026Updated 6 months ago
- IntentVCNet: Bridging Spatio-Temporal Gaps for Intention-Oriented Controllable Video Captioning☆19Aug 16, 2025Updated 11 months ago
- [CVPR'2025] Narrating the Video: Boosting Text-Video Retrieval via Comprehensive Utilization of Frame-Level Captions☆19Jan 16, 2026Updated 6 months ago
- AI 游戏助手 — 洛克王国世界 自动导航 + 采矿工具☆25Jun 9, 2026Updated 2 months ago
- [EMNLP25 Main]The official code of "Gradient-Attention Guided Dual-Masking Synergetic Framework for Robust Text-based Person Retrieval"☆25Mar 30, 2026Updated 4 months ago
- [NeurIPS 2024] Mixture of Experts for Audio-Visual Learning☆25Jan 19, 2025Updated last year