In this project, several approaches for training/finetuning an audio gender recognition is provided. The code can simply be used for any other audio classification task by simply changing the number of classes and the input dataset.
☆43Jan 11, 2025Updated last year
Alternatives and similar repositories for audio-classification-pytorch
Users that are interested in audio-classification-pytorch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆11May 9, 2022Updated 4 years ago
- ☆25Aug 2, 2022Updated 4 years ago
- pytorch lightening image classification☆17Dec 29, 2022Updated 3 years ago
- A simple OCR labeling tool built on Flask and PyQt5☆24Dec 6, 2022Updated 3 years ago
- ☆15Dec 10, 2021Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- This repo contains a series of notes for using Git and GitHub. It's aimed at making it easy to work with Git and GitHub practically. I al…☆30Dec 21, 2022Updated 3 years ago
- This is a simple dockerfile for running tflite without installing TensorFlow☆17Dec 26, 2021Updated 4 years ago
- Converting coco-like annotation json files to png masks☆26Jun 1, 2022Updated 4 years ago
- A complete instruction for training a Persian spell checker and a language model based on SymSpell and KenLM, respectively using Wikipedi…☆35Jul 20, 2022Updated 4 years ago
- Converting Vox files to Wav or other formats that are easy to workaround.☆23Dec 23, 2021Updated 4 years ago
- ☆17Nov 3, 2021Updated 4 years ago
- In this repository, I share codes of the introduction to python courses published on my YouTube channel☆42Dec 23, 2022Updated 3 years ago
- This is a minimal implementation of face-detection models using flask, gunicorn, nginx, docker, and docker-compose☆38Sep 12, 2022Updated 4 years ago
- added session 02☆20Feb 15, 2020Updated 6 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Generative AI based image editing/inpainting made super easy to work with.☆20Nov 5, 2023Updated 2 years ago
- ☆23Mar 25, 2020Updated 6 years ago
- Visualizing Yolov5's layers using GradCam☆297Nov 27, 2023Updated 2 years ago
- Sharif Emotional Speech Database☆39Jan 9, 2021Updated 5 years ago
- Will tidy your sass and scss☆12Jul 12, 2018Updated 8 years ago
- Tacotron 2 - Persian☆37Dec 28, 2021Updated 4 years ago
- 日本音響学会誌用BibTeXスタイルファイル☆11Jan 24, 2022Updated 4 years ago
- A reference .NET application implementing an eCommerce site☆10Sep 8, 2025Updated last year
- Implementation of Hybrid CTC/Attention Architecture for End-to-End Speech Recognition in pure python and PyTorch☆26Jul 25, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- An implementation of Speech Emotion Recognition, based on HuBERT model, training with PyTorch and HuggingFace framework, and fine-tuning …☆33May 18, 2022Updated 4 years ago
- ☆13Dec 2, 2024Updated last year
- A large-scale validated database for Persian speech emotion detection.☆25May 9, 2022Updated 4 years ago
- Awesome-4D-Radar☆12Feb 17, 2024Updated 2 years ago
- ☆13Nov 21, 2023Updated 2 years ago
- LI-FPN is an excellent model for depression recognition based on facial expression.☆15Apr 5, 2024Updated 2 years ago
- Human ID classification using mmwave radar point cloud☆14May 30, 2026Updated 3 months ago
- Deep learning-based audio spoofing attack detection experiments for speaker verification.☆14Aug 9, 2026Updated last month
- ☆15Aug 22, 2023Updated 3 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆15Updated this week
- ☆16May 3, 2025Updated last year
- official implementation of MGA-CLAP (ACM MM 2024)☆28Oct 25, 2024Updated last year
- Dataset/code for AudioMarkBench: Benchmarking Robustness of Audio Watermarking☆48Aug 23, 2024Updated 2 years ago
- ☆12Sep 12, 2024Updated 2 years ago
- EMO-SUPERB: a reproducible speech emotion recognition benchmark with leakage-free splits for 6 datasets and 15 speech SSL models (IEEE SL…☆52Jul 30, 2026Updated last month
- ☆27Mar 29, 2021Updated 5 years ago