Implementation and reviews of Audio & Computer vision related papers in python using keras and tensorflow.
☆40Nov 1, 2018Updated 7 years ago
Alternatives and similar repositories for Audio-Vision
Users that are interested in Audio-Vision are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- MNSS (Music Noise Segmentation on a Spectrogram) is a deep-neural network based preprocessing technique that pre-filters unnecessary nois…☆11Dec 14, 2015Updated 10 years ago
- python template private service☆17Oct 20, 2020Updated 5 years ago
- https://www.kaggle.com/c/flavours-of-physics☆23Oct 23, 2015Updated 10 years ago
- A multi-channel neural network audio classifier using Keras☆270Jul 29, 2021Updated 5 years ago
- Scene Classification using Audio in the nearby Environment.☆19Sep 4, 2019Updated 6 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official implementation for MGN☆20Dec 22, 2022Updated 3 years ago
- Parallelize your computations in parallel-apply fashion.☆33Jul 19, 2019Updated 7 years ago
- Code for CVPR 2021 paper Exploring Heterogeneous Clues for Weakly-Supervised Audio-Visual Video Parsing☆24Dec 29, 2021Updated 4 years ago
- Named-Entity Recognition model to extract "food" entities - Python☆45May 1, 2017Updated 9 years ago
- creating audio preprocessing features in TensorFlow keras layers,☆14Jul 13, 2021Updated 5 years ago
- my notebooks☆39Dec 1, 2021Updated 4 years ago
- This is my Masters thesis project titled "Speaker Detection and Conversation Analysis on Mobile Devices".☆15May 21, 2017Updated 9 years ago
- An N-gram punctuator for Chinese and English.☆20Oct 14, 2025Updated 9 months ago
- A self-hosted drag-and-drop, nosql yet fully-featured file-scanning server.☆30Apr 21, 2022Updated 4 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- A json version of the OpenCyc-latest.owl Ontology☆13Oct 27, 2011Updated 14 years ago
- audio classification using TensorFlow☆15Feb 9, 2017Updated 9 years ago
- RaspberryPi SenseHat☆10Sep 21, 2015Updated 10 years ago
- Web framework for GeoSolver☆13Feb 18, 2017Updated 9 years ago
- Whisp - Environmental Sound Classifier☆13Aug 14, 2023Updated 2 years ago
- Implementation of an attack/decay model for piano transcription☆11Feb 1, 2018Updated 8 years ago
- Reproducible research code for the experiments presented in our article "Kara1k: a karaoke dataset for cover song identification and sing…☆10Jan 9, 2018Updated 8 years ago
- ☆14Nov 13, 2023Updated 2 years ago
- Impementation of corresponding paper using MATLAB/Simulink☆10Mar 30, 2019Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Text pair classification☆13Jun 24, 2017Updated 9 years ago
- PANiC - PAraphrasing Noun-Compounds☆15Apr 6, 2018Updated 8 years ago
- Repository for Weak Label Learning for Audio Events - A closer look. Uses Audioset subset data provided for reproducibility.☆32Sep 13, 2023Updated 2 years ago
- Google Speech Command Dataset Classification Neural Network, CNN, RNN☆26Aug 29, 2017Updated 8 years ago
- Mycroft Skills Kit☆27Jun 2, 2022Updated 4 years ago
- Learning to Separate Object Sounds by Watching Unlabeled Video (ECCV 2018)☆50Sep 24, 2019Updated 6 years ago
- The repo for "On-the-fly Modulation for Balanced Multimodal Learning", T-PAMI 2024☆19Sep 29, 2024Updated last year
- Bag-of-Features Acoustic Event Detection☆14Oct 5, 2016Updated 9 years ago
- Lightweight and flexible plugin framework for Python.☆14Nov 1, 2021Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A teen girl face generator based on GANs,中文译名“萌妹生成器”。☆20Jun 28, 2019Updated 7 years ago
- Codes for "Event localization in music auto-tagging"☆30Mar 15, 2017Updated 9 years ago
- keras project for audio deep learning☆40Apr 10, 2018Updated 8 years ago
- A streaming Speech to Text server using DeepSpeech☆16May 10, 2020Updated 6 years ago
- Deep Learning on Music Information Retrieval Tutorial☆50Oct 4, 2017Updated 8 years ago
- ☆15Jan 30, 2024Updated 2 years ago
- SoundNet, built in Keras with pre-trained 8-layer model.☆29Oct 15, 2019Updated 6 years ago