This repository documents Barry's journey in learning deep learning for speech processing. Here, you'll find scripts and code snippets related to environment setup, data preprocessing, speech frontend, speech recognition, voice conversion, speech synthesis, and more. Let's explore the fascinating world of speech processing together! 🚀🚀🚀
☆13Oct 8, 2025Updated 11 months ago
Alternatives and similar repositories for barry_speech_tools
Users that are interested in barry_speech_tools are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A Chinese Expressive Long-dialogue Speech Dataset with Scripts☆21Nov 11, 2024Updated last year
- ☆29Sep 14, 2024Updated 2 years ago
- 中国科学院大学2023-2024课程(更新中)☆13Jan 12, 2026Updated 8 months ago
- VOICOR: A Residual Iterative Voice Correction Framework for Monaural Speech Enhancement☆47Sep 12, 2024Updated 2 years ago
- ☆12Apr 26, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Hierarchical Vision Transformers for Disease Progression Detection in Chest X-Ray Images☆11Jan 11, 2024Updated 2 years ago
- ☆15Sep 16, 2024Updated 2 years ago
- ☆20Aug 23, 2024Updated 2 years ago
- This is the implementation of the manuscript "Learning General All-Neural Speech Enhancement based on Taylor's Approximation Theory", whi…☆14Nov 25, 2022Updated 3 years ago
- Greifswald Sleep Stage Classifier - a deep-learning based EEG sleep stage classifier☆17Aug 22, 2025Updated last year
- This repository contains code for an acoustic simulation framework that can be used for acoustic/ultrasonic indoor positioning and/or dat…☆15May 7, 2024Updated 2 years ago
- The implementation of TaylorBeamformer, which is in submission to Interspeech2022☆49Jun 10, 2022Updated 4 years ago
- Fairness-Aware Representation Learning by Suppressing Attribute-Class Associations☆13Mar 19, 2026Updated 6 months ago
- [ICLR'25] Official repository for "AVHBench: A Cross-Modal Hallucination Evaluation for Audio-Visual Large Language Models"☆26Mar 8, 2026Updated 6 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆16Jun 15, 2022Updated 4 years ago
- The baselines of ARC-Challenge-Interspeech2026☆62Dec 1, 2025Updated 9 months ago
- An interpreter in C for the language brainfuck.☆11Apr 12, 2023Updated 3 years ago
- MRSAudio: A Large-Scale Multimodal Recorded Spatial Audio Dataset with Refined Annotations☆43Aug 29, 2026Updated 3 weeks ago
- ☆16Nov 6, 2023Updated 2 years ago
- ☆24Feb 28, 2023Updated 3 years ago
- PyPI package for Number Token Loss (ICML 2025)☆25Aug 20, 2026Updated last month
- [ICLR 2025] Enhancing Self-Supervised Models with Audio Mixtures for Polyphonic Soundscapes☆81Oct 8, 2025Updated 11 months ago
- ☆17Dec 22, 2023Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- SLT 2024 Mandarin Stuttering Event Detection and Automatic Speech Recognition Challenge☆12Jun 11, 2024Updated 2 years ago
- ☆25Sep 30, 2019Updated 6 years ago
- [NeurIPS 2025] Benchmark data and code for MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix☆222Feb 25, 2026Updated 6 months ago
- speech enhancement\speech seperation\sound source localization☆15Apr 22, 2020Updated 6 years ago
- A collection of tools to improve TJUer's life experience☆22Feb 29, 2024Updated 2 years ago
- Codes and datasets for our ICASSP2023 paper, Evaluating parameter-efficient transfer learning approaches on SURE benchmark for speech und…☆43Mar 12, 2023Updated 3 years ago
- This is the official implement of Mamba-SEUNet: Mamba UNet for Monaural Speech Enhancement☆95May 26, 2025Updated last year
- Implementation of "Improving Whispered Speech Recognition Performance using Pseudo-whispered based Data Augmentation"☆15Sep 11, 2026Updated last week
- TDBRAIN EEG Database pre-processing code☆26May 8, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- 封装了百度、捷通华声和讯飞语音识别的库,以及捷通华声、民族语文翻译、小牛翻译的封装。☆15Sep 10, 2019Updated 7 years ago
- Some useful tools☆20Nov 28, 2019Updated 6 years ago
- A STFT/iSTFT written up in PyTorch using 1D Convolutions☆32Jul 9, 2024Updated 2 years ago
- Paper, Code and Resources for Speech Language Model and End2End Speech Dialogue System.☆204Jun 7, 2026Updated 3 months ago
- An unofficial implementation of Lite-RTSE, a cost-effective lite model for real-time speech enhancement☆14Nov 19, 2023Updated 2 years ago
- ☆13Jun 24, 2021Updated 5 years ago
- 基于Pytorch框架,使用CNN模型应用于Minist数据集上的分类任务☆23Mar 31, 2021Updated 5 years ago