showing how to use CLIP-Vip to do video search
☆16Nov 16, 2023Updated 2 years ago
Alternatives and similar repositories for clip-vip_video_search
Users that are interested in clip-vip_video_search are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official This-Is-My Dataset published in CVPR 2023☆16Jul 18, 2024Updated 2 years ago
- Sapsucker Woods 60 Audiovisual Dataset☆19Oct 7, 2022Updated 4 years ago
- A large scale dataset for Video Captioning in Italian☆13May 16, 2023Updated 3 years ago
- [ICCV 2023] The official PyTorch implementation of the paper: "Localizing Moments in Long Video Via Multimodal Guidance"☆23Sep 26, 2024Updated 2 years ago
- [ECCV'24] Official Implementation of Autoregressive Visual Entity Recognizer.☆14Mar 2, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- VPEval Codebase from Visual Programming for Text-to-Image Generation and Evaluation (NeurIPS 2023)☆45Nov 29, 2023Updated 2 years ago
- An official implementation for "X-CLIP: End-to-End Multi-grained Contrastive Learning for Video-Text Retrieval"☆191Apr 6, 2024Updated 2 years ago
- Dataset splits and evaluation code for the paper "Benchmark for Compositional Text-to-Image Synthesis" (NeurIPS 2021)☆45May 3, 2022Updated 4 years ago
- Code and datasets for "Text encoders are performance bottlenecks in contrastive vision-language models". Coming soon!☆11May 24, 2023Updated 3 years ago
- Code for the paper Joint Discovery of Object States and Manipulation Actions, ICCV 2017☆14Aug 7, 2018Updated 8 years ago
- This repository contains the Adverbs in Recipes (AIR) dataset and the code published at the CVPR 23 paper: "Learning Action Changes by Me…☆13May 25, 2023Updated 3 years ago
- ☆11Jul 31, 2022Updated 4 years ago
- ☆54Jul 31, 2022Updated 4 years ago
- Online visual analytics tool designed to investigate how attention maps in transformer models behaves, and build hypothesis on those mode…☆10Nov 10, 2021Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- An Evaluation Framework for Temporal Information Extraction Systems☆22Feb 19, 2026Updated 7 months ago
- NNVisBuilder and some cases including KD-t☆13Nov 18, 2023Updated 2 years ago
- IRFL: Image Recognition of Figurative Language☆12Nov 30, 2023Updated 2 years ago
- Rotation equivariance meets local feature matching☆18Oct 20, 2022Updated 3 years ago
- ☆11May 5, 2022Updated 4 years ago
- PyTorch Implementation of the paper "Defining and Quantifying the Emergence of Sparse Concepts in DNNs" (CVPR 2023)☆12Dec 24, 2023Updated 2 years ago
- ☆13Jul 20, 2024Updated 2 years ago
- ☆19Jan 30, 2023Updated 3 years ago
- Android application that localizes people in indoor environments, using wifi fingerprinting methods☆18Sep 30, 2017Updated 9 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Official code for the paper: "Metadata Archaeology"☆19May 10, 2023Updated 3 years ago
- ☆11Jul 19, 2024Updated 2 years ago
- BADLAD: Bengali Document Layout Analysis Dataset☆15Aug 11, 2026Updated last month
- ☆107Dec 23, 2022Updated 3 years ago
- ☆12Aug 14, 2018Updated 8 years ago
- ViLMA: A Zero-Shot Benchmark for Linguistic and Temporal Grounding in Video-Language Models (ICLR 2024, Official Implementation)☆16Jan 18, 2024Updated 2 years ago
- ☆12Mar 5, 2024Updated 2 years ago
- ☆18Jul 10, 2024Updated 2 years ago
- ☆17Oct 24, 2024Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- End-to-end Multi-modal Video Temporal Grounding, NeurIPS 2021☆18Oct 24, 2021Updated 4 years ago
- [CBMI 2024 Best Paper] Official repository of the paper "Is CLIP the main roadblock for fine-grained open-world perception?".☆31May 12, 2025Updated last year
- Labeled Movie Trailer Dataset☆16Mar 23, 2018Updated 8 years ago
- [TPAMI2024] Codes and Models for VALOR: Vision-Audio-Language Omni-Perception Pretraining Model and Dataset☆310Dec 25, 2024Updated last year
- [ICLR'24] Efficient Video Diffusion Models via Content-Frame Motion-Latent Decomposition☆55May 14, 2024Updated 2 years ago
- A library of speech gadgets.☆16Oct 15, 2022Updated 3 years ago
- Code and dataset release for "PACS: A Dataset for Physical Audiovisual CommonSense Reasoning" (ECCV 2022)☆18Dec 20, 2022Updated 3 years ago