☆97Feb 14, 2022Updated 4 years ago
Alternatives and similar repositories for CrossTask
Users that are interested in CrossTask are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for the HowTo100M paper☆304Mar 10, 2020Updated 6 years ago
- ☆19May 2, 2020Updated 6 years ago
- Source code for paper "Towards Automatic Learning of Procedures from Web Instructional Videos"☆34Jan 6, 2019Updated 7 years ago
- ☆131Jun 27, 2021Updated 5 years ago
- Code for the paper "Unsupervised Learning from Narrated Instruction Videos", CVPR2016☆20Jul 27, 2016Updated 10 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A one-stop shop for YouCook2 info such as leaderboard and recent advances on (cooking) video retrieval and captioning.☆41Jun 29, 2022Updated 4 years ago
- PyTorch GPU distributed training code for MIL-NCE HowTo100M☆221Jul 5, 2022Updated 4 years ago
- Weakly-supervised action segmentation in video☆16Feb 13, 2022Updated 4 years ago
- Annotations for the Mistake Detection benchmark of Assembly101☆12Aug 3, 2023Updated 3 years ago
- Annotations for the public release of the EPIC-KITCHENS-100 dataset☆173Aug 1, 2022Updated 4 years ago
- S3D Text-Video model trained on HowTo100M using MIL-NCE☆200Jul 3, 2020Updated 6 years ago
- Video narrator written in Python/GTK using vlc-lib☆26Jun 22, 2022Updated 4 years ago
- Code implementation for our ECCV, 2022 paper titled "My View is the Best View: Procedure Learning from Egocentric Videos"☆35Feb 5, 2024Updated 2 years ago
- Source code for "Weakly-Supervised Video Object Grounding from Text by Loss Weighting and Object Interaction"☆47Jun 22, 2024Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Code for the paper Action Sets: Weakly Supervised Action Segmentation without Ordering Constraints☆29Nov 12, 2020Updated 5 years ago
- ☆78Aug 16, 2021Updated 4 years ago
- Code for CVPR 2023 paper "Procedure-Aware Pretraining for Instructional Video Understanding"☆50Jun 2, 2026Updated 2 months ago
- EPIC-KITCHENS-55 dataset python library☆31Jun 21, 2022Updated 4 years ago
- 🍴 Annotations for the EPIC KITCHENS-55 Dataset.☆155Mar 17, 2021Updated 5 years ago
- EPIC-KITCHENS-55 baselines for Action Recognition☆75Jul 14, 2020Updated 6 years ago
- What Can You Learn from Your Muscles? Learning Visual Representation from Human Interactions (https://arxiv.org/pdf/2010.08539.pdf)☆39Mar 30, 2021Updated 5 years ago
- Project and dataset webpage:☆296Oct 12, 2023Updated 2 years ago
- EPIC-Kitchens-100 Action Recognition baselines: TSN, TRN, TSM☆33Mar 15, 2022Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Code for Oops! Predicting Unintentional Action in Video☆80Apr 13, 2020Updated 6 years ago
- Pytorch version of VidLanKD: Improving Language Understanding viaVideo-Distilled Knowledge Transfer (NeurIPS 2021))☆56Feb 6, 2023Updated 3 years ago
- Code for Learning to Learn Language from Narrated Video☆33Oct 3, 2023Updated 2 years ago
- HACS: Human Action Clips and Segments Dataset☆198Apr 23, 2020Updated 6 years ago
- Referring expression comprehension on ReferIt(RefClef)☆10Nov 28, 2016Updated 9 years ago
- [NeurIPS'20] Self-supervised Co-Training for Video Representation Learning. Tengda Han, Weidi Xie, Andrew Zisserman.☆288Oct 10, 2021Updated 4 years ago
- A video retrieval dataset How2R and a video QA dataset How2QA☆24Oct 15, 2020Updated 5 years ago
- Code for the paper Joint Discovery of Object States and Manipulation Actions, ICCV 2017☆14Aug 7, 2018Updated 8 years ago
- ☆48Apr 27, 2020Updated 6 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Dataset generated by the methods in "What's Cookin'? Interpreting Cooking Videos using Text, Speech and Vision"☆21May 27, 2015Updated 11 years ago
- Support library for the MaskRCNN masks extracted on EPIC-KITCHENS-100☆14Dec 1, 2020Updated 5 years ago
- ☆260Nov 13, 2023Updated 2 years ago
- [ECCV 2020] PyTorch code of MMT (a multimodal transformer captioning model) on TVCaption dataset☆91Sep 6, 2023Updated 2 years ago
- RareAct: A video dataset of unusual interactions☆35Aug 4, 2020Updated 6 years ago
- Self-supervised algorithm for learning representations from ego-centric video data. Code is tested on EPIC-Kitchens-100 and Ego4D in PyTo…☆13Oct 23, 2022Updated 3 years ago
- Shows visual grounding methods can be right for the wrong reasons! (ACL 2020)☆23Jun 26, 2020Updated 6 years ago