Video Feature Extractor for S3D-HowTo100M
☆29Apr 30, 2021Updated 5 years ago
Alternatives and similar repositories for VideoFeatureExtractor
Users that are interested in VideoFeatureExtractor are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- code for downloading videos from HowTo100M dataset☆18May 13, 2021Updated 5 years ago
- S3D Text-Video model trained on HowTo100M using MIL-NCE☆200Jul 3, 2020Updated 6 years ago
- An official implementation for " UniVL: A Unified Video and Language Pre-Training Model for Multimodal Understanding and Generation"☆366Jul 25, 2024Updated 2 years ago
- Easy to use video deep features extractor☆321Jul 5, 2020Updated 6 years ago
- Repository for Multimodal AutoML Benchmark☆67Dec 7, 2021Updated 4 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [ICIP 2022 oral] VLCap: Vision-Language with Contrastive Learning for Coherent Video Paragraph Captioning☆28Jun 28, 2023Updated 3 years ago
- Source code of the paper titled *Attentive Visual Semantic Specialized Network for Video Captioning*☆15Apr 6, 2021Updated 5 years ago
- ☆24Dec 16, 2022Updated 3 years ago
- [ICML 2022] Code and data for our paper "IGLUE: A Benchmark for Transfer Learning across Modalities, Tasks, and Languages"☆49Dec 7, 2022Updated 3 years ago
- Code for the HowTo100M paper☆305Mar 10, 2020Updated 6 years ago
- ☆25Mar 4, 2022Updated 4 years ago
- The implementation of our CIKM 2021 paper titled as: "Cross-Market Product Recommendation"☆20Nov 30, 2021Updated 4 years ago
- [ACL 2021] mTVR: Multilingual Video Moment Retrieval☆28Aug 20, 2022Updated 4 years ago
- PyTorch code for "Perceiver-VL: Efficient Vision-and-Language Modeling with Iterative Latent Attention" (WACV 2023)☆34Feb 5, 2023Updated 3 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Python implementation of extraction of several visual features representations from videos☆23Jul 19, 2021Updated 5 years ago
- 华为digix 2021 赛题1☆28Nov 10, 2021Updated 4 years ago
- ☆21Feb 18, 2022Updated 4 years ago
- Data Release for VALUE Benchmark☆30Feb 16, 2022Updated 4 years ago
- The Code for ICME2019 Grand Challenge: Short Video Understanding (Single Model Ranks 6th)☆91Sep 1, 2019Updated 7 years ago
- 2019中国高校计算机大赛——大数据挑战赛 第三名解决方案☆122Feb 16, 2020Updated 6 years ago
- ☆31Mar 2, 2023Updated 3 years ago
- ☆39Sep 23, 2021Updated 5 years ago
- ☆10Oct 7, 2023Updated 3 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆14Dec 25, 2020Updated 5 years ago
- ☆13Aug 11, 2026Updated last month
- [CVPR23 Highlight] CREPE: Can Vision-Language Foundation Models Reason Compositionally?☆36Apr 27, 2023Updated 3 years ago
- Training and evaluation codes for the BertGen paper (ACL-IJCNLP 2021)☆11Sep 17, 2023Updated 3 years ago
- Caffe++: assemble new features to enhance Caffe☕️☆11Dec 24, 2018Updated 7 years ago
- 首届电子商务AI算法大赛TOP2开源代码☆13Aug 31, 2021Updated 5 years ago
- Code and datasets for "Text encoders are performance bottlenecks in contrastive vision-language models". Coming soon!☆11May 24, 2023Updated 3 years ago
- ☆14Jun 26, 2022Updated 4 years ago
- Turning to Video for Transcript Sorting☆49Aug 27, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Code for "SlotLifter: Slot-guided Feature Lifting for Learning Object-centric Radiance Fields" (ECCV 2024)☆12Oct 30, 2024Updated last year
- This is tensorflow 2.2 based SCAMET framework for remote sensing image captioning.☆13Aug 10, 2023Updated 3 years ago
- ChangeIt dataset with more than 2600 hours of video with state-changing actions published at CVPR 2022☆12Mar 23, 2022Updated 4 years ago
- Latex template for CUHK PhD Thesis☆14Jun 29, 2025Updated last year
- code for composite in situ imaging (cisi) analysis☆12Oct 26, 2020Updated 5 years ago
- ☆44Mar 8, 2021Updated 5 years ago
- 微信大数据2021 1st,qq浏览器2021 3rd,mind新闻推荐2020 1st,NAIC2020 AI+遥感影像 2nd☆161Oct 16, 2022Updated 3 years ago