showing how to use CLIP-Vip to do video search
☆16Nov 16, 2023Updated 2 years ago
Alternatives and similar repositories for clip-vip_video_search
Users that are interested in clip-vip_video_search are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- SMILE: A Multimodal Dataset for Understanding Laughter☆13Jun 15, 2023Updated 3 years ago
- Sapsucker Woods 60 Audiovisual Dataset☆19Oct 7, 2022Updated 3 years ago
- [ICCV 2023] The official PyTorch implementation of the paper: "Localizing Moments in Long Video Via Multimodal Guidance"☆23Sep 26, 2024Updated last year
- Container Specification, Tutorials and Examples for the AI4EU Experiments docker/grpc format for models☆13Jul 10, 2022Updated 4 years ago
- VPEval Codebase from Visual Programming for Text-to-Image Generation and Evaluation (NeurIPS 2023)☆45Nov 29, 2023Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- An official implementation for "X-CLIP: End-to-End Multi-grained Contrastive Learning for Video-Text Retrieval"☆189Apr 6, 2024Updated 2 years ago
- Dataset splits and evaluation code for the paper "Benchmark for Compositional Text-to-Image Synthesis" (NeurIPS 2021)☆45May 3, 2022Updated 4 years ago
- Code for the paper Joint Discovery of Object States and Manipulation Actions, ICCV 2017☆14Aug 7, 2018Updated 8 years ago
- This repository contains the Adverbs in Recipes (AIR) dataset and the code published at the CVPR 23 paper: "Learning Action Changes by Me…☆13May 25, 2023Updated 3 years ago
- Un semplice Chatbot in italiano usando Tensorflow☆14Mar 4, 2019Updated 7 years ago
- ☆54Jul 31, 2022Updated 4 years ago
- Online visual analytics tool designed to investigate how attention maps in transformer models behaves, and build hypothesis on those mode…☆10Nov 10, 2021Updated 4 years ago
- An Evaluation Framework for Temporal Information Extraction Systems☆21Feb 19, 2026Updated 7 months ago
- NNVisBuilder and some cases including KD-t☆13Nov 18, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- IRFL: Image Recognition of Figurative Language☆12Nov 30, 2023Updated 2 years ago
- Temporal Compact Bilinear Pooling (TCBP)☆11May 27, 2020Updated 6 years ago
- Rotation equivariance meets local feature matching☆18Oct 20, 2022Updated 3 years ago
- ☆11May 5, 2022Updated 4 years ago
- ☆13Jul 20, 2024Updated 2 years ago
- ☆19Jan 30, 2023Updated 3 years ago
- Official project page of CLIPDrawX☆36Nov 12, 2025Updated 10 months ago
- Official code for the paper: "Metadata Archaeology"☆19May 10, 2023Updated 3 years ago
- [CVPRW 2021] DUVE network for NTIRE 2021 Quality enhancement of heavily compressed videos - Track 3 Fixed bit-rate☆10Oct 17, 2024Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆11Jul 19, 2024Updated 2 years ago
- BADLAD: Bengali Document Layout Analysis Dataset☆15Aug 11, 2026Updated last month
- ☆15Aug 19, 2023Updated 3 years ago
- ☆14Oct 23, 2017Updated 8 years ago
- Pinterest dataset released with the paper "Learning Image and User Features for Recommendation in Social Networks" by Xue Geng et al. in …☆12Jan 4, 2023Updated 3 years ago
- ViLMA: A Zero-Shot Benchmark for Linguistic and Temporal Grounding in Video-Language Models (ICLR 2024, Official Implementation)☆16Jan 18, 2024Updated 2 years ago
- ☆12Mar 5, 2024Updated 2 years ago
- ☆18Jul 10, 2024Updated 2 years ago
- ☆17Oct 24, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Multi-modality pre-training☆511Aug 28, 2026Updated 3 weeks ago
- [CBMI 2024 Best Paper] Official repository of the paper "Is CLIP the main roadblock for fine-grained open-world perception?".☆31May 12, 2025Updated last year
- Labeled Movie Trailer Dataset☆16Mar 23, 2018Updated 8 years ago
- ☆15Jun 2, 2025Updated last year
- [TPAMI2024] Codes and Models for VALOR: Vision-Audio-Language Omni-Perception Pretraining Model and Dataset☆310Dec 25, 2024Updated last year
- pycalibrate is a Python library to visually analyze model calibration in Jupyter Notebooks☆17Jul 2, 2022Updated 4 years ago
- Ad-hoc Video Search☆29Feb 18, 2021Updated 5 years ago