Generating sound spectrograms using short-time Fourier transform that can be used for purposes such as sound classification by machine learning algorithms.
☆37May 9, 2021Updated 5 years ago
Alternatives and similar repositories for Audio-Spectrogram
Users that are interested in Audio-Spectrogram are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for our paper "Acoustic Features Fusion using Attentive Multi-channel Deep Architecture" in Keras and tensorflow☆26Nov 23, 2018Updated 7 years ago
- Speech Emotion Recognition☆27Jun 19, 2020Updated 6 years ago
- A CNN audio classifier via spectrogram images.☆10Jul 21, 2017Updated 9 years ago
- A PyTorch implementation of "Revisiting Multi-Task Learning with ROCK: a Deep Residual Auxiliary Block for Visual Detection"☆14Jun 29, 2020Updated 6 years ago
- Codebase and project page for EDMSound☆35Nov 20, 2023Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- The project tries to solve a speaker diarization problem using audio features, face recognition and video feature extraction from face im…☆16Feb 10, 2019Updated 7 years ago
- Speech command classification on Speech-Command v0.02 dataset using PyTorch and torchaudio. In this example, three models have been train…☆10Dec 5, 2022Updated 3 years ago
- We present a deep learning approach towards the large-scale prediction and analysis of bird acoustics from 100 different bird species☆21Jun 17, 2024Updated 2 years ago
- ☆25Mar 21, 2024Updated 2 years ago
- Dual-Adversarial Domain Adaptation for replay spoofing detection in automatic speaker verification.☆19Jul 17, 2026Updated last week
- Perform three types of feature extraction: STFT, MFCC and MelSpectrogram. Apply CNN/VGG with or without RNN architecture. Able to achieve…☆15Jun 28, 2020Updated 6 years ago
- 实现对视频进行简单的编辑,exe文件直接打开就能用,包括视频截取、视频高度裁剪(去字幕)、视频拼接、音频分离☆15Aug 10, 2018Updated 7 years ago
- Alzheimer's Dementia Recognition through Spontaneous Speech The ADReSSo Challenge☆15Aug 6, 2023Updated 2 years ago
- Supervised Speech Representation Learning for Parkinson's Disease Classification☆18Oct 26, 2021Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Save jpeg images in h5py☆13May 1, 2019Updated 7 years ago
- Depression-Detection represents a machine learning algorithm to classify audio using acoustic features in human speech, thus detecting de…☆14Jul 10, 2020Updated 6 years ago
- Here the code of EmoAudioNet is a deep neural network for speech classification (published in ICPR 2020)☆14Jul 13, 2020Updated 6 years ago
- ☆12Feb 23, 2021Updated 5 years ago
- It's an open source Altium database Libraries for Altium Designer.☆24Oct 9, 2023Updated 2 years ago
- Source code for paper Multi-Task Learning for Depression Detection in Dialogs (SIGDial 2022)☆12Jan 18, 2025Updated last year
- Detect Depression with AI Sub-challenge (DSS) of AVEC2019 experienment version via YZK☆15May 28, 2021Updated 5 years ago
- ☆13Nov 21, 2023Updated 2 years ago
- 基于Qt和Mysql的教务管理系统☆15Jun 10, 2019Updated 7 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Write a random novel using Markov chains☆12Nov 27, 2016Updated 9 years ago
- 采用三种方式 (1)利用keras库搭建seq2seq (2)利用keras_transformer库 (3)利用fastnlp框架 实现问答机器人、机器翻译、文本摘要等功能☆14Nov 16, 2020Updated 5 years ago
- ☆11Jun 20, 2023Updated 3 years ago
- Baseline scripts for AVEC 2019, Depression Detection Sub-challenge☆16Jul 11, 2019Updated 7 years ago
- The code and data used for the publication: Noise Reduction in X-ray Photon Correlation Spectroscopy with Convolutional Neural Networks E…☆13Nov 29, 2021Updated 4 years ago
- C++ header-only DSP library for Cortex-M☆19Mar 10, 2023Updated 3 years ago
- Depression detection and severity estimation☆13May 17, 2017Updated 9 years ago
- A Machine Learning Approach for the Diagnosis of Parkinson's Disease via Speech Analysis☆21Dec 27, 2020Updated 5 years ago
- Lista de modelos y aplicaciones basadas en diffusion☆11May 4, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- The final coursework for AI in Mental Health @ PKU.☆23Jan 5, 2024Updated 2 years ago
- The description of FMFCC-A (audio track of FMFCC) dataset and Challenge resluts.☆25Apr 14, 2022Updated 4 years ago
- Code & Data for the Paper "Time Masking for Temporal Language Models", WSDM 2022☆21Apr 23, 2023Updated 3 years ago
- Download and preprocess voxceleb datasets.☆41Jun 18, 2025Updated last year
- Multiform Ensemble Self-Supervised Learning for Few-Shot Remote Sensing Scene Classification☆13Mar 10, 2023Updated 3 years ago
- The aim of this project is to predict whether a person is depressed or not using different machine learning algorithms based on the tweet…☆18Jul 12, 2022Updated 4 years ago
- [AAAI 2024] DTF-AT: Decoupled Time-Frequency Audio Transformer for Event Classification☆12Mar 10, 2025Updated last year