The official PyTorch implementation of VM-ASR, a model designed for high-fidelity audio super-resolution.
☆24Sep 8, 2025Updated 11 months ago
Alternatives and similar repositories for VM-ASR
Users that are interested in VM-ASR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICASSP 2025] "FLowHigh: Towards efficient and high-quality audio super-resolution with single-step flow matching"☆34May 12, 2025Updated last year
- Bandwidth Extension of Historical Recordings using Generative Adversarial Networks☆38May 25, 2023Updated 3 years ago
- FastWave is a lightweight diffusion model for general audio super-resolution (any -> 48 kHz). SOTA quality reconstruction metrics with ju…☆18May 16, 2026Updated 3 months ago
- ☆79Jan 25, 2025Updated last year
- ☆87May 21, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Towards High-Quality and Efficient Speech Bandwidth Extension with Parallel Amplitude and Phase Prediction☆198Apr 15, 2025Updated last year
- AI-powered audio denoising using a custom-trained 1D Operational GAN (Kiranyaz et al. 2022). FastAPI + React/TypeScript. 607 tests, uv, r…☆15Jul 31, 2026Updated 2 weeks ago
- fat_llama is a Python package for upscaling audio files to FLAC or WAV formats using advanced audio processing techniques. It utilizes CU…☆36Apr 28, 2026Updated 3 months ago
- Audio-to-Audio Schrodinger Bridges is a diffusion-based audio restoration model for bandwidth extension and inpainting.☆153Aug 13, 2025Updated last year
- Unofficial Pytorch Lightning Implementation of "Towards Robust Speech Super-Resolution"☆10May 8, 2023Updated 3 years ago
- Implementation of Learning Bandwidth Expansion Using Perceptually-Motivated Loss (ICASSP 2019)☆11May 18, 2022Updated 4 years ago
- DiffPhase: Generative Diffusion-based STFT Phase Retrieval☆16Sep 21, 2023Updated 2 years ago
- Synthesis of percussion sounds using sinusoidal modelling, DDSP noise synthesis, and a neural source filter approach.☆36Jan 7, 2025Updated last year
- ☆16Jul 23, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code release for "TinySpeech: Attention Condensers for Deep Speech Recognition Neural Networks on Edge Devices"☆23Jun 7, 2025Updated last year
- An initial foray into deploying an exported RNBO Max MSP object as an Android app.☆14Jan 25, 2023Updated 3 years ago
- [ICASSP 2025] "FLowHigh: Towards efficient and high-quality audio super-resolution with single-step flow matching"☆118Jan 17, 2025Updated last year
- NU-Wave: A Diffusion Probabilistic Model for Neural Audio Upsampling @ INTERSPEECH 2021☆283Jul 22, 2022Updated 4 years ago
- ASLP Summer Inter@NPU☆13Jul 30, 2024Updated 2 years ago
- DeepLearningで音楽をアップサンプリングします☆20Mar 24, 2018Updated 8 years ago
- ☆60Jun 14, 2024Updated 2 years ago
- ☆14Feb 3, 2026Updated 6 months ago
- Official implementation of "AEROMamba: An efficient architecture for audio super-resolution using generative adversarial networks and sta…☆50Nov 11, 2025Updated 9 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- An adversarial neural network that restores the high fidelity audio lost during lossy compression☆21Aug 10, 2026Updated last week
- ☆27Feb 8, 2025Updated last year
- ☆23Sep 16, 2025Updated 11 months ago
- Unofficial implementation of HiFi-GAN+ from the paper "Bandwidth Extension is All You Need" by Su, et al.☆225Oct 20, 2023Updated 2 years ago
- An invertible and differentiable implementation of the Constant-Q Transform (CQT).☆73Dec 9, 2022Updated 3 years ago
- Code for the paper "Comparative Analysis of CNN-based Spatiotemporal Reasoning in Videos"☆14May 3, 2024Updated 2 years ago
- Mp3 to wav super resolution model for audio restoration & enhancement. U-Net + Discrete Wavelet Transform (DWT) Architecture☆21Dec 1, 2025Updated 8 months ago
- Music repair method to convert lossy MP3 compressed music to lossless music.☆402Updated this week
- PyTorch implementation of MLP-Mixer architecture.☆12May 24, 2021Updated 5 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Code for the paper "Deep Attention Based Semi-Supervised 2D-Pose Estimation for Surgical Instruments"☆12Dec 14, 2019Updated 6 years ago
- This is the repo for the work "Where and What: Driver Attention-based Object Detection".☆10May 10, 2022Updated 4 years ago
- Romance Meter Extension for SillyTavern☆15Aug 4, 2025Updated last year
- ☆28Jun 1, 2023Updated 3 years ago
- Implementation of paper Generalised Image Outpainting with UTransformer☆22Mar 26, 2024Updated 2 years ago
- ☆26Dec 14, 2023Updated 2 years ago
- This is the repo for the work "Talking with your hands: Scaling hand gestures and recognition with cnns".☆13Apr 30, 2022Updated 4 years ago