Code for Learning to Learn Language from Narrated Video
☆33Oct 3, 2023Updated 2 years ago
Alternatives and similar repositories for expert
Users that are interested in expert are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for the Globetrotter project☆23Mar 17, 2022Updated 4 years ago
- Project page for "Visual Grounding in Video for Unsupervised Word Translation" CVPR 2020☆43Apr 26, 2020Updated 6 years ago
- Code for the paper Learning the Predictability of the Future (CVPR 2021)☆172Jul 31, 2023Updated 3 years ago
- Data Release for VALUE Benchmark☆30Feb 16, 2022Updated 4 years ago
- Official code for the paper "Scaling Multilingual Visual Speech Recognition"☆20Aug 15, 2025Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- The implementation of CVPR2021 paper Temporal Query Networks for Fine-grained Video Understanding☆64Mar 9, 2022Updated 4 years ago
- CVPR 2021 Official Pytorch Code for UC2: Universal Cross-lingual Cross-modal Vision-and-Language Pre-training☆34Nov 9, 2021Updated 4 years ago
- Inferring and Executing Programs for Visual Reasoning☆21Jan 4, 2019Updated 7 years ago
- A one-stop shop for YouCook2 info such as leaderboard and recent advances on (cooking) video retrieval and captioning.☆41Jun 29, 2022Updated 4 years ago
- Shapley values for assessing the importance of each frame in a video☆17Mar 1, 2021Updated 5 years ago
- [ACL 2019] Visually Grounded Neural Syntax Acquisition☆90Feb 24, 2024Updated 2 years ago
- Website-based resource monitor for Slurm system☆39Apr 6, 2023Updated 3 years ago
- Support library for the MaskRCNN masks extracted on EPIC-KITCHENS-100☆14Dec 1, 2020Updated 5 years ago
- The 1st place solution of 2022 Ego4d Natural Language Queries.☆32Sep 5, 2022Updated 4 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- RareAct: A video dataset of unusual interactions☆35Aug 4, 2020Updated 6 years ago
- [CVPR'22 Oral] Temporal Alignment Networks for Long-term Video. Tengda Han, Weidi Xie, Andrew Zisserman.☆122Oct 9, 2023Updated 2 years ago
- Localize objects in images using referring expressions☆37Nov 1, 2016Updated 9 years ago
- Source code for "Weakly-Supervised Video Object Grounding from Text by Loss Weighting and Object Interaction"☆47Jun 22, 2024Updated 2 years ago
- PyTorch GPU distributed training code for MIL-NCE HowTo100M☆221Jul 5, 2022Updated 4 years ago
- Self-supervised learning through the eyes of a child☆145Jul 20, 2021Updated 5 years ago
- [ECCV'20 Spotlight] Memory-augmented Dense Predictive Coding for Video Representation Learning. Tengda Han, Weidi Xie, Andrew Zisserman.☆167Apr 29, 2021Updated 5 years ago
- PyTorch 3D video classification models pre-trained on 65 million Instagram videos☆264Dec 7, 2019Updated 6 years ago
- Video Representation Learning by Dense Predictive Coding. Tengda Han, Weidi Xie, Andrew Zisserman.☆256Oct 8, 2021Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ECCV 2020] PyTorch code of MMT (a multimodal transformer captioning model) on TVCaption dataset☆91Sep 6, 2023Updated 3 years ago
- ☆23May 11, 2026Updated 3 months ago
- ☆98Feb 14, 2022Updated 4 years ago
- Undoing the Damage of Dataset Bias☆28Sep 5, 2013Updated 13 years ago
- ☆11Feb 9, 2026Updated 6 months ago
- Starter Code for VALUE benchmark☆79Aug 23, 2022Updated 4 years ago
- Efficient Training for Multilingual Visual Speech Recognition: Pre-training with Discretized Visual Speech Representation (ACM MM 2024)☆20Mar 17, 2025Updated last year
- A collection of videos annotated with timelines where each video is divided into segments, and each segment is labelled with a short free…☆30Jan 15, 2022Updated 4 years ago
- Code for Oops! Predicting Unintentional Action in Video☆80Apr 13, 2020Updated 6 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- A repo for processing the raw hand object detections to produce releasable pickles + library for using these☆40Oct 26, 2024Updated last year
- Code for "Counterfactual Variable Control for Robust and Interpretable Question Answering"☆14Oct 13, 2020Updated 5 years ago
- The second version of the interface for Abstract Scenes research project.☆23May 16, 2022Updated 4 years ago
- ☆16Apr 10, 2022Updated 4 years ago
- ☆28Jul 18, 2025Updated last year
- Self-supervised Learning for Video Correspondence Flow (BMVC 2019)☆268Sep 15, 2019Updated 6 years ago
- PaperBot: Learning to Design Real-World Tools Using Paper☆13Mar 15, 2024Updated 2 years ago