☆18May 28, 2025Updated last year
Alternatives and similar repositories for BAST
Users that are interested in BAST are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for the paper "DSpAST: Disentangled Representations for Spatial Audio Reasoning with Large Language Models"☆17Oct 23, 2025Updated 10 months ago
- Implementation of the paper "Binaural Sound Source Distance Estimation and Localization for a Moving Listener"☆23Mar 2, 2025Updated last year
- ☆15Aug 13, 2023Updated 3 years ago
- A python implementation of “Self-Supervised Learning of Spatial Acoustic Representation with Cross-Channel Signal Reconstruction and Mult…☆41Oct 11, 2024Updated last year
- DeepEar: Sound Localization with Binaural Microphones☆16Nov 20, 2025Updated 9 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- A python algorithm to change the pitch of the voice in real time☆13Dec 13, 2020Updated 5 years ago
- Official page of "DeFTAN-II: Efficient multichannel speech enhancement with subgroup processing", IEEE/ACM Transactions on Audio, Speech,…☆34Nov 21, 2024Updated last year
- Mirror of the Auditory Modelling Toolbox http://amtoolbox.sourceforge.net/☆11Jan 28, 2019Updated 7 years ago
- The End-to-End Magnitude Least Squares Binaural Renderer for Spherical Microphone Array Signals☆41Feb 17, 2026Updated 6 months ago
- Model for selecting perceptually relevant early reflections for parametric spatial sound rendering☆13Oct 26, 2023Updated 2 years ago
- Dynamical Systems with JAX☆12Jun 3, 2026Updated 3 months ago
- ☆20Jun 29, 2025Updated last year
- This is the official implementation of our multi-channel multi-speaker multi-spatial neural audio codec architecture.☆55Mar 17, 2025Updated last year
- Lightweight streaming Voice Activity Detection (VAD) tool with ONNX runtime☆24Mar 18, 2026Updated 5 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- This sample includes simeple CNN classifier for music and audio-folder dataloader just like ImageFolder in torchvision.☆11Oct 30, 2018Updated 7 years ago
- ☆34Jun 10, 2025Updated last year
- Official code of SenSE.☆91Oct 30, 2025Updated 10 months ago
- This repository contains the dataset used to train the neural network model descried in the paper "Implicit HRTF Modeling Using Tempora…☆11Aug 4, 2023Updated 3 years ago
- Official Implementation of DMT: Dual Mean-Teacher in PyTorch.☆10Oct 27, 2023Updated 2 years ago
- 2.5D visual sound☆121Jul 25, 2023Updated 3 years ago
- Code to create networks that localize sounds sources in 3D environments☆53Jan 27, 2024Updated 2 years ago
- RAVEN: Recognition of Audio-Visual Emotional Nuances - a project on building multimodal emotion recognition system☆16Jun 24, 2025Updated last year
- ☆12Nov 1, 2024Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- A Python Library for Full Reference Binaural Fidelity Testing, Visualization & Feature Generation☆30Oct 30, 2025Updated 10 months ago
- arxiv daily for speech translation, legal. Ref: Vincentqyw/cv-arxiv-daily☆15Jan 6, 2025Updated last year
- AUCO ResNet: an end-to-end network for Covid-19 pre-screening from cough and breath☆13Mar 18, 2022Updated 4 years ago
- Classification of eeg signals using knn and svm based upon significant features☆14Jan 10, 2023Updated 3 years ago
- 支持多种图表类型的绘制工具,包括思维导图、流程图、数据可视化图表、数学函数图等;可根据用户需求生成 Mermaid、ECharts、Mindmap、DrawIO、GeoGebra 等格式的图表,并导出为 PNG、SVG、HTML 等格式☆17Jan 26, 2026Updated 7 months ago
- ☆17Jun 2, 2025Updated last year
- CleanUMamba: A Compact Mamba Network for Speech Denoising using Channel Pruning [Official PyTorch implementation]☆29Jun 12, 2025Updated last year
- Lyrics and Vocal Melody Generation conditioned on Accompaniment☆28Aug 27, 2022Updated 4 years ago
- Understanding and Tackling Hallucinations in Large Audio-Language Models | ICASSP 2025, Interspeech 2024☆34Mar 14, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- semantic tokenizer for speech and music☆20Jul 6, 2025Updated last year
- Code for the submitted 2021 DCASE Workshop paper: "Waveforms and Spectrograms: Enhancing Acoustic Scene Classification Using Multimodal F…☆16Aug 9, 2021Updated 5 years ago
- This is the CoNNear human auditory periphery model that simulates cochlear, IHC and AN processing across the human hearing range.☆26Feb 21, 2024Updated 2 years ago
- Material from paper "HRTF Individualization using Deep Learning", Miccini and Spagnol, 2020☆16Oct 25, 2021Updated 4 years ago
- ☆28Apr 17, 2023Updated 3 years ago
- Official repository for the paper "xLSTM-SENet: xLSTM for Single-Channel Speech Enhancement" (Accepted to INTERSPEECH 2025)☆60Aug 28, 2025Updated last year
- The official repo for Both Ears Wide Open: Towards Language-Driven Spatial Audio Generation☆65Jul 2, 2025Updated last year