Video Feature Extractor for S3D-HowTo100M
☆29Apr 30, 2021Updated 5 years ago
Alternatives and similar repositories for VideoFeatureExtractor
Users that are interested in VideoFeatureExtractor are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- code for downloading videos from HowTo100M dataset☆18May 13, 2021Updated 5 years ago
- ☆15May 23, 2023Updated 3 years ago
- An official implementation for " UniVL: A Unified Video and Language Pre-Training Model for Multimodal Understanding and Generation"☆365Jul 25, 2024Updated last year
- Easy to use video deep features extractor☆322Jul 5, 2020Updated 6 years ago
- Repository for Multimodal AutoML Benchmark☆67Dec 7, 2021Updated 4 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ESIM model with lanuage model☆27Nov 10, 2018Updated 7 years ago
- TyDiP Multilingual Politeness dataset and code☆12Oct 15, 2023Updated 2 years ago
- Code for the HowTo100M paper☆303Mar 10, 2020Updated 6 years ago
- [ACL 2021] mTVR: Multilingual Video Moment Retrieval☆27Aug 20, 2022Updated 3 years ago
- ☆24Feb 15, 2022Updated 4 years ago
- PyTorch code for "Perceiver-VL: Efficient Vision-and-Language Modeling with Iterative Latent Attention" (WACV 2023)☆34Feb 5, 2023Updated 3 years ago
- Python implementation of extraction of several visual features representations from videos☆23Jul 19, 2021Updated 5 years ago
- Cross-model active contrastive coding☆22Mar 17, 2021Updated 5 years ago
- 华为digix 2021 赛题1☆29Nov 10, 2021Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Data Release for VALUE Benchmark☆30Feb 16, 2022Updated 4 years ago
- Code for NeurIPS 2022 Datasets and Benchmarks paper - EgoTaskQA: Understanding Human Tasks in Egocentric Videos.☆44Apr 17, 2023Updated 3 years ago
- Official codes for the paper "Learning Hierarchical Discrete Linguistic Units from Visually-Grounded Speech"☆28Feb 22, 2022Updated 4 years ago
- The Code for ICME2019 Grand Challenge: Short Video Understanding (Single Model Ranks 6th)☆91Sep 1, 2019Updated 6 years ago
- ☆22Jun 6, 2020Updated 6 years ago
- 2019中国高校计算机大赛——大数据挑战赛 第三名解决方案☆122Feb 16, 2020Updated 6 years ago
- Text-Image Relationships (ACL 2019)☆23Sep 15, 2023Updated 2 years ago
- ☆31Mar 2, 2023Updated 3 years ago
- Source code of the paper titled *Improving Video Captioning with Temporal Composition of a Visual-Syntactic Embedding*☆30Apr 16, 2021Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆10Oct 7, 2023Updated 2 years ago
- ☆60Jun 16, 2023Updated 3 years ago
- ☆14Dec 25, 2020Updated 5 years ago
- smplify code for point cloud based HMR☆10Jan 11, 2022Updated 4 years ago
- [CVPR23 Highlight] CREPE: Can Vision-Language Foundation Models Reason Compositionally?☆35Apr 27, 2023Updated 3 years ago
- Training and evaluation codes for the BertGen paper (ACL-IJCNLP 2021)☆11Sep 17, 2023Updated 2 years ago
- Caffe++: assemble new features to enhance Caffe☕️☆11Dec 24, 2018Updated 7 years ago
- 首届电子商务AI算法大赛TOP2开源代码☆13Aug 31, 2021Updated 4 years ago
- 📖The Big-&-Extending-Repository-of-Transformers: Pretrained PyTorch models for Google's BERT, OpenAI GPT & GPT-2, Google/CMU Transformer…☆10Dec 4, 2020Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Pytorch Tutorial for M1 students. This repository include Encoder Deocder model and Classification model building code.☆12Jun 1, 2022Updated 4 years ago
- Code and datasets for "Text encoders are performance bottlenecks in contrastive vision-language models". Coming soon!☆11May 24, 2023Updated 3 years ago
- ☆13Jun 26, 2022Updated 4 years ago
- This repository contains the Adverbs in Recipes (AIR) dataset and the code published at the CVPR 23 paper: "Learning Action Changes by Me…☆13May 25, 2023Updated 3 years ago
- Annotations for the Mistake Detection benchmark of Assembly101☆12Aug 3, 2023Updated 2 years ago
- Turning to Video for Transcript Sorting☆49Aug 27, 2023Updated 2 years ago
- Learning Algebraic Representation for Systematic Generalization in Abstract Reasoning☆11Jul 20, 2022Updated 4 years ago