Implementation and reviews of Audio & Computer vision related papers in python using keras and tensorflow.
☆40Nov 1, 2018Updated 7 years ago
Alternatives and similar repositories for Audio-Vision
Users that are interested in Audio-Vision are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- MNSS (Music Noise Segmentation on a Spectrogram) is a deep-neural network based preprocessing technique that pre-filters unnecessary nois…☆11Dec 14, 2015Updated 10 years ago
- Scene Classification using Audio in the nearby Environment.☆19Sep 4, 2019Updated 7 years ago
- A multi-channel neural network audio classifier using Keras☆270Jul 29, 2021Updated 5 years ago
- Parallelize your computations in parallel-apply fashion.☆33Jul 19, 2019Updated 7 years ago
- Code for CVPR 2021 paper Exploring Heterogeneous Clues for Weakly-Supervised Audio-Visual Video Parsing☆24Dec 29, 2021Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Named-Entity Recognition model to extract "food" entities - Python☆45May 1, 2017Updated 9 years ago
- Neural network to classify some styles of Electronic music☆23Apr 4, 2019Updated 7 years ago
- audio classification using TensorFlow☆15Feb 9, 2017Updated 9 years ago
- Web framework for GeoSolver☆13Feb 18, 2017Updated 9 years ago
- Whisp - Environmental Sound Classifier☆13Aug 14, 2023Updated 3 years ago
- ☆11Mar 15, 2017Updated 9 years ago
- Implementation of an attack/decay model for piano transcription☆11Feb 1, 2018Updated 8 years ago
- Reproducible research code for the experiments presented in our article "Kara1k: a karaoke dataset for cover song identification and sing…☆10Jan 9, 2018Updated 8 years ago
- ☆14Nov 13, 2023Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Android sound localization and classification app.☆16Jul 4, 2025Updated last year
- Text pair classification☆13Jun 24, 2017Updated 9 years ago
- PANiC - PAraphrasing Noun-Compounds☆15Apr 6, 2018Updated 8 years ago
- Repository for Weak Label Learning for Audio Events - A closer look. Uses Audioset subset data provided for reproducibility.☆32Sep 13, 2023Updated 3 years ago
- Google Speech Command Dataset Classification Neural Network, CNN, RNN☆26Aug 29, 2017Updated 9 years ago
- Mycroft Skills Manager☆13Feb 1, 2022Updated 4 years ago
- This repository consists of all the work done regarding Heart sound classification employing ANN, CNN and other methods, Android Applicat…☆18Jun 7, 2019Updated 7 years ago
- Bag-of-Features Acoustic Event Detection☆14Oct 5, 2016Updated 10 years ago
- Codes for "Event localization in music auto-tagging"☆30Mar 15, 2017Updated 9 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- keras project for audio deep learning☆40Apr 10, 2018Updated 8 years ago
- SAAVN Code release for paper "Sound Adversarial Audio-Visual Navigation,ICLR2022" (In PyTorch)☆21Nov 9, 2022Updated 3 years ago
- A streaming Speech to Text server using DeepSpeech☆16May 10, 2020Updated 6 years ago
- ofmUI is a lightweight mobile UI library specifically crafted for the openFrameworks mobile framework.☆10Mar 7, 2024Updated 2 years ago
- Deep Learning on Music Information Retrieval Tutorial☆51Oct 4, 2017Updated 9 years ago
- ☆15Sep 11, 2026Updated 3 weeks ago
- SoundNet, built in Keras with pre-trained 8-layer model.☆29Oct 15, 2019Updated 6 years ago
- ☆12Aug 10, 2023Updated 3 years ago
- ☆14Oct 7, 2021Updated 5 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Voice Controlled Flappy Bird☆16Jan 26, 2018Updated 8 years ago
- Master thesis: Structured Auto-Encoder with application to Music Genre Recognition (code)☆16Apr 18, 2020Updated 6 years ago
- Code for https://arxiv.org/abs/1712.00254☆18Dec 6, 2017Updated 8 years ago
- [ECCV 2022] Joint-Modal Label Denoising for Weakly-Supervised Audio-Visual Video Parsing☆27Jul 15, 2022Updated 4 years ago
- 使用keras&tensorflow 框架,GTZAN数据集。☆15Feb 20, 2019Updated 7 years ago
- Combines the SSL Method MixMatch with a pre-trained model (EfficientNet) on a chest x-ray dataset.☆11Jun 22, 2019Updated 7 years ago
- Cochlear.ai submission for dcase2018 task2☆15Sep 14, 2018Updated 8 years ago