Code for the paper: Audio-Visual Model Distillation Using Acoustic Images
☆21Mar 24, 2023Updated 3 years ago
Alternatives and similar repositories for acoustic-images-distillation
Users that are interested in acoustic-images-distillation are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Tensorflow code for the paper 'Modality Distillation with Multiple Stream Networks for Action Recognition', ECCV 2018☆19May 2, 2019Updated 7 years ago
- End to End Multiview Lip Reading☆10Jan 26, 2018Updated 8 years ago
- Repository for Weak Label Learning for Audio Events - A closer look. Uses Audioset subset data provided for reproducibility.☆32Sep 13, 2023Updated 2 years ago
- Generalized cross-modal NNs; new audiovisual benchmark (IEEE TNNLS 2019)☆31Apr 13, 2020Updated 6 years ago
- ☆17Jul 17, 2017Updated 9 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Implementation of "With a Little Help from my Temporal Context: Multimodal Egocentric Action Recognition, BMVC, 2021" in PyTorch☆20Dec 16, 2021Updated 4 years ago
- Transformer-based online speech recognition system with TensorFlow 2☆26Jan 22, 2021Updated 5 years ago
- Simple C++ code for converting OpenCV Mat objects to Python and easily reusing OpenCV C++ code in Python.☆11Dec 11, 2015Updated 10 years ago
- ☆10Jun 1, 2023Updated 3 years ago
- ☆23Dec 5, 2023Updated 2 years ago
- (Unofficial) Code for the paper "Certifying Some Distributional Robustness with Principled Adversarial Training"☆13May 31, 2018Updated 8 years ago
- Important notes on scientific papers☆21Mar 1, 2021Updated 5 years ago
- Code for ''A Simple Baseline for Audio-Visual Scene-Aware Dialog``☆27May 26, 2020Updated 6 years ago
- Baseline of dcase 2019 task 4☆61Sep 2, 2022Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Road extraction with deep learning from high resolution satellite images.☆13Sep 16, 2021Updated 4 years ago
- ☆25Feb 20, 2024Updated 2 years ago
- 给定一张身份证正、反面,识别身份证上的所有文字信息☆10Sep 4, 2019Updated 6 years ago
- Evaluation metrics and submission file creation scripts the Action Recognition challenge☆15Feb 9, 2026Updated 6 months ago
- Audio-Visual Speech Recognition using Sequence to Sequence Models☆84Jul 10, 2020Updated 6 years ago
- Pytorch implementation of audio-visual fusion video captioning model☆27Jul 26, 2018Updated 8 years ago
- Excitation Backprop for RNNs☆15Jul 25, 2018Updated 8 years ago
- Jupyter notebook for DCASE 2020 challenge Task 1☆20Jun 24, 2020Updated 6 years ago
- FreiPose: A Deep Learning Framework for Precise Animal Motion Capture in 3D Spaces☆18Apr 11, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Repository for the paper Enhancing Land Subsidence Awareness via InSAR Data and Deep Transformers☆16Mar 15, 2022Updated 4 years ago
- ☆17Nov 9, 2023Updated 2 years ago
- Code for "Lifting Monocular Events to 3D Human Poses" - CVPRw 2021☆17Aug 30, 2024Updated last year
- Acoustic Scene Classification Using Deep Residual Networks with Late Fusion of Separated High and Low Frequency Paths - McDonnell and Gao…☆22Jul 3, 2024Updated 2 years ago
- UPC Deep Learning for Speech and Language 2018☆17Feb 26, 2018Updated 8 years ago
- An implementation of capsule routing for sound event detection☆15Jan 29, 2019Updated 7 years ago
- Multimodal Speech Recognition for phoneme level prediction using Audio-Visual data from TCDTIMIT dataset implementing RNNs with LSTMs for…☆15Jul 27, 2023Updated 3 years ago
- ☆54Jun 3, 2020Updated 6 years ago
- Theano-based Deep Learning library (convnets, recurrent neural networks, and more).☆14Aug 2, 2017Updated 9 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Dual cross modality attention audio-visual speech recognition model based on vgg transformer with hybrid CTC/attention architecture using…☆15Jul 2, 2020Updated 6 years ago
- Tensorflow implementation of "Hide-and-Seek: Forcing a Network to be Meticulous for Weakly-supervised Object and Action Localization"[ICC…☆13Mar 29, 2019Updated 7 years ago
- Code for the paper Learning Unbiased Representations via Mutual Information Backpropagation☆21Mar 23, 2020Updated 6 years ago
- Another implementation of topological loss☆38May 12, 2021Updated 5 years ago
- Web app created to collect audios for course project☆10Apr 6, 2018Updated 8 years ago
- ☆17Feb 14, 2020Updated 6 years ago
- Official PyTorch implementation of "Attention-Free Keyword Spotting", Mashrur. M. Morshed & Ahmad Omar Ahsan, PML4DC @ ICLR 2022.☆15Nov 5, 2022Updated 3 years ago