前沿论文持续更新--视频时刻定位 or 时域语言定位 or 视频片段检索。
☆14Nov 8, 2023Updated 2 years ago
Alternatives and similar repositories for Awesome-Cross-Modal-Video-Moment-Retrieval
Users that are interested in Awesome-Cross-Modal-Video-Moment-Retrieval are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ACM MULTIMEDIA CONFERENCE 2020☆11Jul 28, 2020Updated 6 years ago
- Temporal Sentence Grounding in Videos / Natural Language Video Localization / Video Moment Retrieval的相关工作☆31Mar 4, 2022Updated 4 years ago
- 前沿论文持续更新--视频时刻定位 or 时域语言定位 or 视频片段检索。☆266Aug 26, 2023Updated 2 years ago
- Temporal Moment(Action) Localization via Language / Temporal Language Grounding / Video Moment Retrieval☆100Jan 23, 2022Updated 4 years ago
- Video Feature Extraction Code for EMNLP 2020 paper "HERO: Hierarchical Encoder for Video+Language Omni-representation Pre-training"☆118Jun 9, 2021Updated 5 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Generating Structured Pseudo Labels for Noise-resistant Zero-shot Video Sentence Localization☆16Jul 20, 2023Updated 3 years ago
- ☆19May 14, 2025Updated last year
- ☆16Aug 28, 2024Updated last year
- Official pytorch implementation of "Explore-And-Match: Bridging Proposal-Based and Proposal-Free With Transformer for Sentence Grounding …☆42Aug 5, 2022Updated 3 years ago
- Dense Regression Network for Video Grounding (CVPR2020)☆53Jan 28, 2021Updated 5 years ago
- PyTorch implementation of "TALL: Temporal Activity Localization via Language Query. Gao et al. ICCV2017."☆14Apr 20, 2019Updated 7 years ago
- "Video Moment Retrieval from Text Queries via Single Frame Annotation" in SIGIR 2022☆68Jun 27, 2022Updated 4 years ago
- This repo takes the initial step towards leveraging text learning for online action detection without explicit human supervision.☆15Jul 13, 2026Updated 2 weeks ago
- Span-based Localizing Network for Natural Language Video Localization (ACL 2020)☆113Oct 15, 2021Updated 4 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆28Jul 18, 2025Updated last year
- ☆19Jul 28, 2025Updated last year
- [ICLR 2025] Knowing Your Target: Target-Aware Transformer Makes Better Spatio-Temporal Video Grounding☆44Mar 18, 2025Updated last year
- 📦 A lightweight machine learning toolkit for researchers, providing common model design & learning functionalities.☆29Jul 9, 2026Updated 3 weeks ago
- Code for Panoramic Semantic Segmentation☆16Apr 26, 2024Updated 2 years ago
- Project GaitSystem is a Gait recognition system based Windows Visual Studio 2017. The algorithm based the paper Gait optical flow image d…☆12Jun 3, 2019Updated 7 years ago
- A reading list of papers about Visual Grounding.☆31Aug 24, 2022Updated 3 years ago
- The code of DIffpose video setting☆16Dec 28, 2023Updated 2 years ago
- [ICLR 2025] This repo is the official implementation of our paper "Learning Fine-Grained Representations through Textual Token Disentangl…☆23Jul 28, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- This is a repository contains the implementation of our NeurIPS'24 paper "Temporal Sentence Grounding with Relevance Feedback in Videos"☆13Aug 22, 2025Updated 11 months ago
- ☆10Jan 4, 2022Updated 4 years ago
- 【ICME2025 Oral】Offical Pytorch Code for "Fraesormer: Learning Adaptive Sparse Transformer for Efficient Food Recognition"☆13Mar 21, 2025Updated last year
- Repository for the CVPR-20 paper "Local-Global Video-Text Interactions for Temporal Grounding"☆132Jul 5, 2021Updated 5 years ago
- The benchmark experiments of paper "ReSGait: The real scene gait dataset".☆12Jul 25, 2024Updated 2 years ago
- Cross-Modal Interaction Networks for Query-Based Moment Retrieval in Videos☆87Nov 22, 2020Updated 5 years ago
- The top conferences on video retrieval libraries in recent years, synchronized with my blog.☆14Nov 27, 2021Updated 4 years ago
- [IJCAI-24] Explore Internal and External Similarity for Single Image Deraining with Graph Neural Networks☆11Sep 2, 2024Updated last year
- Keyword spotting for audio with attention (KWS model for audio)☆18Jul 15, 2021Updated 5 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- An implementation of online data mixing for the Pile dataset, based on the GPT-NeoX library.☆14Jan 9, 2024Updated 2 years ago
- Code for the paper "Zero-shot Natural Language Video Localization" (ICCV2021, Oral).☆48Mar 15, 2023Updated 3 years ago
- Using machine learning techniques for prediction and modelling non linear dynamic systems.☆10Jun 29, 2018Updated 8 years ago
- VideoX: a collection of video cross-modal models☆1,071Jun 3, 2024Updated 2 years ago
- ☆17Jun 15, 2022Updated 4 years ago
- The evaluation tool (Matlab version) for saliency maps.☆10Mar 18, 2022Updated 4 years ago
- React 沙盒 📦,可理解为 React 版的 eval() 。该沙盒运行机制可使基于 React 实现的小程序框架「如 Taro3 等」拥有 🚀 热更新能力。☆15Mar 7, 2022Updated 4 years ago