3D Sound Source Localization using Masked Autoencoders
☆21Feb 12, 2025Updated last year
Alternatives and similar repositories for wav2pos
Users that are interested in wav2pos are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Graph Neural Networks for Sound Source Localization☆29Oct 31, 2023Updated 2 years ago
- Sound Event Localization and Detection using Neural Generalized Cross-Correlations☆36Feb 11, 2025Updated last year
- The Neural-SRP method for DOA estimation☆37May 24, 2024Updated 2 years ago
- DynamicSound Simulator is a modular Python library for generating virtual acoustic scenes with configurable microphones, sound sources, a…☆18Jul 15, 2026Updated last week
- Neural Generalized Cross Correlations https://arxiv.org/abs/2208.04654☆37Feb 11, 2025Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- A tutorial for Sound Source Localization researchers and practitioners. The purpose of this repo is to organize the world’s resources for…☆59Mar 17, 2023Updated 3 years ago
- A python implementation of “SRP-DNN: Learning Direct-Path Phase Difference for Multiple Moving Sound Source Localization” [ICASSP 2022]☆67Sep 28, 2024Updated last year
- ResNet-STFT Model for Sound Source Localization☆20Aug 25, 2022Updated 3 years ago
- This repo is for the paper "Uncertainty Estimation for Sound Source Localization".☆15Mar 13, 2025Updated last year
- Files for the paper: "Sound Source Localization using Deep Residual Learning"☆24Nov 13, 2017Updated 8 years ago
- Code for paper Audio Visual Speaker Localization from EgoCentric Views☆11Jul 3, 2024Updated 2 years ago
- PSELDNets: Pre-trained Neural Networks on Large-scale Synthetic Datasets for Sound Event Localization and Detection☆47Sep 17, 2025Updated 10 months ago
- The Official PyTorch Implementation of FN-SSL & IPDnet for Sound Source Localization [INTERSPEECH2023 & TASLP2024]☆159Mar 10, 2026Updated 4 months ago
- CST-former: Transformer with Channel-Spectro-Temporal Attention for Sound Event Localization and Detection (ICASSP 2024)☆39May 20, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Neural Network based Sound Source Localization Models☆51Aug 29, 2023Updated 2 years ago
- A Deep Convolutional Neural Network (DCNN) designed for the task of localizing human speech to 168 location classes using binaural microp…☆10Dec 16, 2017Updated 8 years ago
- ☆34Feb 19, 2025Updated last year
- Four neural network architectures to classify sound source direction☆11Oct 3, 2020Updated 5 years ago
- A Transformer-based Prediction Method for Depth of Anesthesia During Target-controlled Infusion of Propofol and Remifentanil.☆16Feb 17, 2025Updated last year
- Complex-valued neural networks for DOA estimation☆31Jan 25, 2023Updated 3 years ago
- Code repository for the paper Direction of Arrival Estimation of Sound Sources Using Icosahedral CNNs☆49May 19, 2022Updated 4 years ago
- Ichigo Whisper is a compact (22M parameters), open-source speech tokenizer for the Whisper-medium, designed to enhance performance on mul…☆16Jan 20, 2025Updated last year
- Repository of the IJCV'26 & WACV'24 paper☆34Apr 27, 2026Updated 2 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- This software is a demonstration of Audio Signal Processing and Machine Learning using Python and Tensorflow. The software contains a GU…☆12Dec 7, 2023Updated 2 years ago
- ☆13Jan 28, 2022Updated 4 years ago
- Pytorch implementation of "spectro-temporal attention-based voice activity detection"☆13Jun 4, 2024Updated 2 years ago
- ☆36Feb 14, 2025Updated last year
- End-to-End binaural sound localization☆17Feb 27, 2020Updated 6 years ago
- ☆32Apr 21, 2025Updated last year
- 🦇 Encoder of BAT (Learning to Reason about Spatial Sounds with Large Language Models)☆87Feb 13, 2025Updated last year
- In this repository, we deal with developing different estimators to localize Transvahan - the e-vehicle on IISc Campus using measurements…☆20Jul 2, 2020Updated 6 years ago
- Tr-VAD: An Efficient Transformer based Voice Activity Detection Model☆18Aug 1, 2024Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆82Dec 2, 2024Updated last year
- KETI Mobius platform resource managament tool☆11Jun 22, 2022Updated 4 years ago
- Sound field reconstruction using neural processes with dynamic kernels☆16Mar 25, 2025Updated last year
- Master repository for 3D Spatial Audio Reproduction Toolbox☆22Jul 25, 2016Updated 10 years ago
- Paper Review about Speech Recognition · NLP☆10Mar 25, 2021Updated 5 years ago
- Audio classification deep learning model using TensorFlow 2.0 to detect Gunshots. 97.5% test set accuracy and 99% training set accuracy w…☆23Feb 16, 2020Updated 6 years ago
- Silero VAD(ncnn): pre-trained enterprise-grade Voice Activity Detector.☆26Aug 21, 2024Updated last year