Librosa equivalent Java library to process audio file adn extract features from it.
☆121May 14, 2024Updated 2 years ago
Alternatives and similar repositories for jlibrosa
Users that are interested in jlibrosa are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- TensorFlow on mobile with speech-to-text DL models.☆165Nov 21, 2017Updated 8 years ago
- [ICASSP 2020] Speech Emotion Recognition with Dual-Sequence LSTM Architecture☆12Jan 17, 2025Updated last year
- Experiment with JNI access to some Kaldi functions.☆12Dec 31, 2018Updated 7 years ago
- https://www.kaggle.com/c/tensorflow-speech-recognition-challenge/☆21Mar 1, 2018Updated 8 years ago
- Javascript library to convert between MIDI data and MusicXML☆20Oct 19, 2023Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- C code to extract mfcc or fbank features from wav files☆17Oct 25, 2019Updated 6 years ago
- CNTK implementation of Fully Convolutional Networks (FCN) with ResNet for semantic segmentation☆12Aug 18, 2017Updated 9 years ago
- C/C++实现Python音频处理库librosa中melspectrogram的计算过程☆31Jan 14, 2022Updated 4 years ago
- ☆47Sep 27, 2020Updated 5 years ago
- magicspeech competition recipe☆18Jun 29, 2020Updated 6 years ago
- Export an ONNX graph that performs ISTFT. Designed for TTS models.☆28Apr 23, 2024Updated 2 years ago
- A convenience Java wrapper around GloVe word vectors and converter to more space efficient binary files.☆25Apr 1, 2021Updated 5 years ago
- it's ASR decoder and make graph project☆33May 26, 2022Updated 4 years ago
- Convert WSJ sphere format to waveform and do data simulation.☆16Feb 20, 2020Updated 6 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆11May 27, 2026Updated 3 months ago
- A tool for reverse engineering Android apk files☆13Nov 19, 2012Updated 13 years ago
- LEAF is a learnable alternative to audio features such as mel-filterbanks, that can be initialized as an approximation of mel-filterbanks…☆531Mar 1, 2022Updated 4 years ago
- [INTERSPEECH'2022] Accurate Emotion Strength Assessment for Seen and Unseen Speech Based on Data-Driven Deep Learning☆83Nov 4, 2022Updated 3 years ago
- ☆45Jan 13, 2022Updated 4 years ago
- Code for our paper "Efficient Speech Emotion Recognition Using Multi-Scale CNN and Attention" (ICASSP 2021, co-first authorship)☆28Jun 8, 2021Updated 5 years ago
- Dual cross modality attention audio-visual speech recognition model based on vgg transformer with hybrid CTC/attention architecture using…☆14Jul 2, 2020Updated 6 years ago
- A tensorflow implementation of TasNet (ICASSP 2018)☆16Nov 27, 2018Updated 7 years ago
- Repository of code for Speech emotion recognition using voiced speech and attention model, submitted to ICSigSys 2019☆13Jan 6, 2020Updated 6 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Tensorflow Implementation for "Pre-trained Deep Convolution Neural Network Model With Attention for Speech Emotion Recognition"☆10Dec 19, 2021Updated 4 years ago
- This repository created for the NHN ASR hackathon competition.☆12Sep 20, 2023Updated 3 years ago
- Frontend filterbank learning module with HVQT initialization capabilities.☆21Feb 27, 2024Updated 2 years ago
- android rtp player☆13Feb 9, 2017Updated 9 years ago
- PyTorch implementation of the LEAF audio frontend☆79Mar 29, 2023Updated 3 years ago
- A short tutorial on Keras for the co-utilization of audio and text data (multi-modal analysis)☆16Nov 21, 2022Updated 3 years ago
- Presented by Seeed Studio, we offer our series open source hardware based on Raspberry Pi☆22Jan 8, 2025Updated last year
- ☆21Jan 13, 2020Updated 6 years ago
- Conv TaSNet follow work of KaiTuo Xu in TF-keras☆14Oct 19, 2020Updated 5 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Visualization toolbox for Sound Event Detection☆123Feb 26, 2024Updated 2 years ago
- Group Gated Fusion on Attention-based Bidirectional Alignment for Multimodal Emotion Recognition☆15May 10, 2022Updated 4 years ago
- MelGAN and Tacotron 2 in PyTorch☆11Oct 22, 2019Updated 6 years ago
- Python wrapper for phonetisaurus grapheme to phoneme tool☆12Mar 11, 2021Updated 5 years ago
- mxnet csharp Interface☆23Oct 8, 2017Updated 8 years ago
- Download and preperation tool for free speech corpora.☆16Apr 28, 2019Updated 7 years ago
- Android video semantic segmentation using DeeplabV3+ lite☆10Sep 20, 2019Updated 7 years ago