Code on selecting an action based on multimodal inputs. Here in this case inputs are voice and text.
☆73Jun 7, 2021Updated 5 years ago
Alternatives and similar repositories for Multimodal-action-recognition
Users that are interested in Multimodal-action-recognition are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official Pytorch Implementation for Continual Learning For On-Device Environmental Sound Classification☆14Jul 19, 2022Updated 4 years ago
- Multimodal late fusion for deepfake detection using video and audio data☆12May 7, 2019Updated 7 years ago
- Central repository for all public AIDA resources☆13Mar 1, 2021Updated 5 years ago
- Generalized cross-modal NNs; new audiovisual benchmark (IEEE TNNLS 2019)☆31Apr 13, 2020Updated 6 years ago
- Pytorch implementation of DSR-RL for Video Summarization Task☆12Aug 30, 2021Updated 4 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Chinese BERT classification with tf2.0 and audio classification with mfcc☆14Dec 2, 2020Updated 5 years ago
- collection of skeleton-based human action recognition☆10Jun 28, 2020Updated 6 years ago
- A collection of multimodal datasets, and visual features for VQA and captionning in pytorch. Just run "pip install multimodal"☆85Feb 25, 2022Updated 4 years ago
- Emotion analysis on DREAMER dataset using various Deep Learning Techniques☆13Jan 1, 2021Updated 5 years ago
- ☆14Aug 13, 2020Updated 6 years ago
- A Pytorch implementation of emotion recognition from videos☆18Sep 15, 2020Updated 5 years ago
- This repository contains various models targetting multimodal representation learning, multimodal fusion for downstream tasks such as mul…☆923Mar 15, 2023Updated 3 years ago
- ☆26Nov 23, 2021Updated 4 years ago
- Predicting Political Instability and Social Conflicts Using Multimodal Data☆10Jun 6, 2016Updated 10 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Code for paper "Robust SENSE reconstruction of simultaneous multislice EPI with low-rank enhanced coil sensitivity calibration and slice-…☆12Apr 14, 2024Updated 2 years ago
- Transformer for Action Recognition in PyTorch☆39Mar 14, 2020Updated 6 years ago
- video summarization lstm-gan pytorch implementation☆27Dec 6, 2019Updated 6 years ago
- Implementation of "With a Little Help from my Temporal Context: Multimodal Egocentric Action Recognition, BMVC, 2021" in PyTorch☆20Dec 16, 2021Updated 4 years ago
- Zicx's Notebook.☆10Nov 7, 2025Updated 9 months ago
- Official PyTorch implementation of paper Leveraging Unimodal Self Supervised Learning for Multimodal Audio-Visual Speech Recognition (ACL…☆67Jul 13, 2022Updated 4 years ago
- Attacks against proposed image encryption schemes☆10Apr 27, 2020Updated 6 years ago
- Self-Supervised Learning by Cross-Modal Audio-Video Clustering (NeurIPS 2020)☆91Oct 24, 2022Updated 3 years ago
- ACAV100M: Automatic Curation of Large-Scale Datasets for Audio-Visual Video Representation Learning. In ICCV, 2021.☆64Nov 18, 2021Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Build Your Own Delivery Robot - Teleoperated, autonomous, lightweight and weatherproof. Free for personal use.☆13Oct 18, 2022Updated 3 years ago
- Official pytorch implementation of I2I translation with low resolution conditioning☆23Sep 2, 2021Updated 4 years ago
- (Competition) 6th -- Scene-Text-Detection-and-Recognition.☆11Jun 14, 2022Updated 4 years ago
- Exploiting temporal redundancies of multi-coil cine cardiac data for MRI reconstruction with unrolled cross-domain networks.☆18Nov 9, 2022Updated 3 years ago
- Pytorch implementation of Meta-Learning for Short Utterance Speaker Recognition with Imbalance Length Pairs (Interspeech, 2020)☆73Sep 16, 2020Updated 5 years ago
- Neural Synchronization: A Generalizable Model-and-Data Driven Approach for Open-Set RFF Authentication☆21Nov 22, 2024Updated last year
- Detect objects from the image, integrated with FLASK for front-end.☆11Jan 30, 2021Updated 5 years ago
- A "gym" style toolkit for building lightweight NAS systems.☆13Jun 13, 2022Updated 4 years ago
- A multimodal UAV assistant dataset.☆11Jun 14, 2021Updated 5 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- An official implementation for " UniVL: A Unified Video and Language Pre-Training Model for Multimodal Understanding and Generation"☆365Jul 25, 2024Updated 2 years ago
- Hand Gesture Controlled Tello Drone using Python and OpenCV 2021☆12Jun 6, 2022Updated 4 years ago
- Important notes on scientific papers☆21Mar 1, 2021Updated 5 years ago
- Pytorch implementation of RNN, CNN, BiGRU and LSTM for text classifcation☆10Apr 30, 2021Updated 5 years ago
- Codes for IJCAI2020 paper "Unsupervised Representation Learning by Predicting Random Distances” https://arxiv.org/abs/1912.12186☆30Apr 25, 2020Updated 6 years ago
- Simple phoenix setup for padded window management☆13Apr 25, 2018Updated 8 years ago
- SGQuant: Squeezing the Last Bit on Graph Neural Networks with Specialized Quantization☆11Aug 12, 2020Updated 6 years ago