EgoCom: A Multi-person Multi-modal Egocentric Communications Dataset
☆63Nov 23, 2020Updated 5 years ago
Alternatives and similar repositories for EgoCom-Dataset
Users that are interested in EgoCom-Dataset are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [CVPR 2024] Code and datasets for 'Learning Spatial Features from Audio-Visual Correspondence in Egocentric Videos'☆14Jun 16, 2024Updated 2 years ago
- ☆21Feb 15, 2022Updated 4 years ago
- Code implementation for our ECCV, 2022 paper titled "My View is the Best View: Procedure Learning from Egocentric Videos"☆35Feb 5, 2024Updated 2 years ago
- This repository contains the source code of the CVPR 2020 paper: "Multimodal Future Localization and Emergence Prediction for Objects in …☆35Dec 8, 2020Updated 5 years ago
- Inferring Body Pose in Egocentric Video via First and Second Person Interactions☆51Aug 31, 2021Updated 5 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Implementation of FixMatch in PyTorch and experimentations☆12Aug 9, 2020Updated 6 years ago
- The Easy Communications (EasyCom) dataset is a world-first dataset designed to help mitigate the *cocktail party effect* from an augmente…☆144Dec 4, 2023Updated 2 years ago
- Agent-as-a-Judge grading framework for evaluating AI outputs/deliverables☆54Aug 7, 2026Updated last month
- The official repository for the CVPR 2019 paper "Overcoming Limitations of Mixture Density Networks: A Sampling and Fitting Framework for…☆48Jan 7, 2021Updated 5 years ago
- AI Benchmark for Investment Banking Workflows☆46Jun 30, 2026Updated 2 months ago
- Code release for the paper "Egocentric Video Task Translation" (CVPR 2023 Highlight)☆34Jun 12, 2023Updated 3 years ago
- Multi-Target Embodied Question Answering☆26Jul 17, 2020Updated 6 years ago
- Simple baseline model for the HEAR benchmark☆23Feb 17, 2026Updated 6 months ago
- [ICLR'25] Do Egocentric Video-Language Models Truly Understand Hand-Object Interactions?☆14Apr 11, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Permutation invariant training in PyTorch☆13Oct 2, 2020Updated 5 years ago
- ☆16Apr 10, 2019Updated 7 years ago
- The official implementation of paper: Estimating Egocentric 3D Human Pose in Global Space.☆13Sep 23, 2023Updated 2 years ago
- IEEE Computer Society Keywords to Organize Knowledge☆13Jan 13, 2020Updated 6 years ago
- ☆13Jul 6, 2022Updated 4 years ago
- The official PyTorch implementation of the IEEE/CVF Computer Vision and Pattern Recognition (CVPR) '24 paper PREGO: online mistake detect…☆35Jun 9, 2025Updated last year
- [AAAI-24] VVS : Video-to-Video Retrieval With Irrelevant Frame Suppression☆21May 14, 2024Updated 2 years ago
- ConvBench: A Multi-Turn Conversation Evaluation Benchmark with Hierarchical Ablation Capability for Large Vision-Language Models☆18Sep 27, 2024Updated last year
- Deploy automl models for tabular tasks on AWS Sagemaker with AutoGluon☆13Feb 28, 2020Updated 6 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Example workflow for our data-centric speech benchmark☆17Jul 6, 2023Updated 3 years ago
- CatNet: Class Incremental 3D ConvNets for Lifelong Egocentric Gesture Recognition☆12Apr 21, 2020Updated 6 years ago
- Implements the loss used in A. Furnari, S. Battiato, G. M. Farinella (2018). Leveraging Uncertainty to Rethink Loss Functions and Evaluat…☆12May 22, 2019Updated 7 years ago
- Deep learning for pedestrians: backpropagation in CNNs. Latex and PyTorch code to verify theoretical derivations.☆13Jun 21, 2022Updated 4 years ago
- Unsupervised Any-to-many Audiovisual Synthesis via Exemplar Autoencoders☆122Nov 21, 2022Updated 3 years ago
- This repository provides a small Python wrapper for the Matlab tool SNR Eval provided by Labrosa: https://labrosa.ee.columbia.edu/project…☆12Jun 22, 2022Updated 4 years ago
- Notebooks showing some examples of DSSATTools usage☆15Apr 13, 2025Updated last year
- [ICLR 2019] Learning Factorized Multimodal Representations☆72Aug 4, 2020Updated 6 years ago
- Code for the ICASSP-2021 paper: Don't shoot butterfly with rifles: Multi-channel Continuous Speech Separation with Early Exit Transformer☆12Sep 2, 2021Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- This repo covers the implementation for Labelling unlabelled videos from scratch with multi-modal self-supervision, which learns clusters…☆118Apr 26, 2021Updated 5 years ago
- Home Action Genome: Cooperative Contrastive Action Understanding☆22Nov 8, 2021Updated 4 years ago
- Deep Multi-Speech model☆11Jul 25, 2018Updated 8 years ago
- We rank the 1st in DSTC8 Audio-Visual Scene-Aware Dialog competition. This is the source code for our IEEE/ACM TASLP (AAAI2020-DSTC8-AVSD…☆56Jun 12, 2023Updated 3 years ago
- Saving nerds eyes.☆16Jan 15, 2022Updated 4 years ago
- ☆13Apr 24, 2024Updated 2 years ago
- ☆13Nov 28, 2021Updated 4 years ago