AI Engineering Class (Fundamental)
☆21Aug 14, 2026Updated last month
Alternatives and similar repositories for AIE-F
Users that are interested in AIE-F are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Implementation of "Improving Whispered Speech Recognition Performance using Pseudo-whispered based Data Augmentation"☆15Sep 11, 2026Updated last week
- unofficial implementation of "CPTNN: CROSS-PARALLEL TRANSFORMER NEURAL NETWORK FOR TIME-DOMAIN SPEECH ENHANCEMENT"☆15Nov 14, 2023Updated 2 years ago
- Voice Conversion method based on speaker style☆14Aug 7, 2021Updated 5 years ago
- ☆19Jan 6, 2025Updated last year
- ☆24Sep 1, 2023Updated 3 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Tensorflow 1.x solution for chinese NER task, using ALBERT-LSTM-CRF model☆18Apr 19, 2020Updated 6 years ago
- A neural speech codec based on discrete WavLM representations☆27Aug 28, 2024Updated 2 years ago
- Blind Source Separation and Dereverberation☆21Mar 26, 2021Updated 5 years ago
- ☆30Nov 7, 2023Updated 2 years ago
- This repository includes training, inference, evaluation, and utility scripts developed for fine-tuning the Whisper medium.en model on Ai…☆30Oct 9, 2024Updated last year
- A multilingual phoneme recognizer capable of generalizing zero-shot to unseen phoneme inventories.☆30Mar 14, 2025Updated last year
- Source code and demo for INTERPSEECH 2023 paper: DuTa-VC: A Duration-aware Typical-to-atypical Voice Conversion Approach with Diffusion P…☆38Dec 5, 2023Updated 2 years ago
- RoBERTa + BiLSTM + CRF for Chinese NER Task☆36Jul 5, 2021Updated 5 years ago
- Document Q& A using RAG - GEMINI PRO☆44Feb 13, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official code for MUSE: Flexible Voiceprint Receptive Fields and Multi-Path Fusion Enhanced Taylor Transformer for U-Net-based Speech Enh…☆58Mar 5, 2025Updated last year
- Transcribing Speech with Multinomial Diffusion, training code and models.☆80Sep 27, 2023Updated 2 years ago
- demo page https://MingjieChen.github.io/dygan-vc☆66Apr 13, 2022Updated 4 years ago
- Mispronunciation Detection using a pretrained and finetuned wav2vec2 model for phoneme recognition and diagnosis and feedback using large…☆60May 6, 2024Updated 2 years ago
- ASR course at Chula 2018☆65Jun 15, 2018Updated 8 years ago
- Code for paper "Noise-aware Speech Enhancement using Diffusion Probabilistic Model"☆89Jun 10, 2024Updated 2 years ago
- A Corpus for Research on Robust Automatic Speech Recognition and Natural Language Understanding of Air Traffic Control Communications☆90Mar 24, 2023Updated 3 years ago
- This repo provides the processed samples of the manuscript "MossFormer: Pushing the Performance Limit of Monaural Speech Separation using…☆108Nov 28, 2024Updated last year
- Finetune VITS and MMS using HuggingFace's tools☆205Mar 31, 2024Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- This is the official implementation of the paper AGAIN-VC: A One-shot Voice Conversion using Activation Guidance and Adaptive Instance No…☆114Dec 7, 2020Updated 5 years ago
- This is the implementation for "ControlVC: Zero-Shot Voice Conversion with Time-Varying Controls on Pitch and Rhythm"☆132Nov 29, 2023Updated 2 years ago
- Implementation of Kaneko et al.'s MaskCycleGAN-VC model for non-parallel voice conversion.☆117Jun 6, 2021Updated 5 years ago
- Official Implementation of StyleTTS-VC☆200Jan 14, 2025Updated last year
- GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling☆177Feb 28, 2025Updated last year
- Phoneme Recognition using pre-trained models Wav2vec2, HuBERT and WavLM. Throughout this project, we compared specifically three differen…☆267May 9, 2022Updated 4 years ago
- Materials for the EMNLP 2020 Tutorial on "Interpreting Predictions of NLP Models"☆198Dec 2, 2020Updated 5 years ago
- Easy-to-Use Speech MOS predictors☆366Oct 24, 2023Updated 2 years ago
- Mathematics of Deep Learning, Courant Insititute, Spring 19☆283Mar 15, 2019Updated 7 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Speech Enhancement Generative Adversarial Network in PyTorch☆409Aug 16, 2023Updated 3 years ago
- This is the main repository of open-sourced speech technology by Huawei Noah's Ark Lab.☆601Sep 18, 2023Updated 3 years ago
- A Tutorial about Programming for Natural Language Processing☆440Oct 29, 2015Updated 10 years ago
- The code for the bark-voicecloning model. Training and inference.☆710Sep 13, 2023Updated 3 years ago
- Score-based Generative Models (Diffusion Models) for Speech Enhancement and Dereverberation☆771May 12, 2026Updated 4 months ago
- [WIP] Resources for AI engineers. Also contains supporting materials for the book AI Engineering (Chip Huyen, 2025)☆17,478Jul 3, 2026Updated 2 months ago
- https://huyenchip.com/ml-interviews-book/☆4,767Mar 21, 2025Updated last year