Audio-Visual Room Impulse Response Estimation
☆25Jul 22, 2024Updated 2 years ago
Alternatives and similar repositories for AV-RIR
Users that are interested in AV-RIR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Real Acoustic Fields An Audio-Visual Room Acoustics Dataset and Benchmark☆64Aug 29, 2024Updated 2 years ago
- Implementation of FiNS model for RIR estimation☆38Nov 1, 2023Updated 2 years ago
- Official PyTorch implementation of 'VINP: Variational Bayesian Inference with Neural Speech Prior for Joint ASR-Effective Speech Dereverb…☆36Feb 23, 2026Updated 6 months ago
- Towards Efficient Audio-Visual Learners via Empowering Pre-trained Vision Transformers with Cross-Modal Adaptation☆15Apr 13, 2024Updated 2 years ago
- Grouped Feedback Delay Networks for Coupled Room Modeling☆41May 27, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- This is the official implementation of reverberant speech to room impulse response estimator☆42Aug 7, 2024Updated 2 years ago
- [ICCV 2021] Image2Reverb: Cross-Modal Reverb Impulse Response Synthesis.☆91Oct 12, 2021Updated 4 years ago
- [CVPR 2024] Code and datasets for 'Learning Spatial Features from Audio-Visual Correspondence in Egocentric Videos'☆14Jun 16, 2024Updated 2 years ago
- Translating Synthetic RIRs to Real RIRs☆46Sep 15, 2023Updated 2 years ago
- Ego4DSounds: A diverse egocentric dataset with high action-audio correspondence☆21Jun 14, 2024Updated 2 years ago
- ☆12Apr 26, 2025Updated last year
- This repository contains code for the generation of binaural Room Impulse Responses using the Paraspax method and implementing a 6 DoF en…☆31Nov 20, 2024Updated last year
- This is the official implementation of our neural-network-based fast diffuse room impulse response generator (FAST-RIR) for generating r…☆183Mar 19, 2026Updated 5 months ago
- This is the official implementation of our mesh-based neural network (MESH2IR) to generate acoustic impulse responses (IRs) for indoor 3D…☆110Jul 24, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Repo for our research paper "Learning Acoustic Scattering Fields for Dynamic Interactive Sound Propagation"☆17Apr 6, 2021Updated 5 years ago
- This is the official implementation of our neural-network-based fast diffuse room impulse response generator (FAST-RIR) for generating r…☆12Nov 30, 2021Updated 4 years ago
- Repo for Visual Acoustic Matching, CVPR 2022☆71Feb 28, 2023Updated 3 years ago
- Code and datasets for 'Few-Shot Audio-Visual Learning of Environment Acoustics' (NeurIPS 2022)☆25Jun 16, 2026Updated 2 months ago
- [ICASSP 2026] The official pytorch implementation of ACVIS☆15Jan 19, 2026Updated 7 months ago
- Implementation of the paper: "Audio Mamba: Bidirectional State Space Model for Audio Representation Learning" in pytorch☆15Updated this week
- Github repository for the paper accepted in ICASSP 2024 : Blind estimation of audio effects using an auto-encoder approach and differenti…☆15Apr 11, 2024Updated 2 years ago
- to release the source code for reproducing the results reported in our paper: https://arxiv.org/abs/2409.17550☆14Nov 15, 2024Updated last year
- soundvista☆17Dec 31, 2025Updated 7 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Room Impulse Response reconstruction with Physics Informed Neural Networks☆57Feb 15, 2024Updated 2 years ago
- Enhanced sound event localization and detection in real 360-degree audio-visual soundscapes (DCASE task3 format)☆14Mar 21, 2025Updated last year
- Download scripts and tools for Replay dataset.☆39Jun 23, 2023Updated 3 years ago
- Time-varying, nonlinear, fun-loving FDNs☆40May 11, 2020Updated 6 years ago
- A PyTorch implementation of the paper: "AMSS-Net: Audio Manipulation on User-Specified Sources with Textual Queries" (ACM Multimedia 2021…☆21Jul 4, 2021Updated 5 years ago
- Code for LAVSS: Location-Guided Audio-Visual Spatial Audio Separation☆19Feb 25, 2025Updated last year
- An auralisation system that takes a head-worn microphone array recordings as input and renders the audio for binaural playback; taking in…☆37Oct 10, 2023Updated 2 years ago
- ☆26Jan 18, 2022Updated 4 years ago
- Companion code of DAFx23 "Differentiable Feedback Delay Network for Colorless Reverberation"☆57Apr 7, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆14Aug 17, 2024Updated 2 years ago
- Official Implementation of "Open-Vocabulary Audio-Visual Semantic Segmentation" [ACM MM 2024 Oral].☆37Nov 2, 2024Updated last year
- Feedback Delay Network in real-time☆26Dec 2, 2022Updated 3 years ago
- MeshRIR: Dataset of room impulse responses on meshed grid points☆43Jul 27, 2026Updated last month
- ☆19Jul 22, 2025Updated last year
- CS230 Final Project - Audio Super Resolution☆13Jun 18, 2018Updated 8 years ago
- Acoustic impulse response generation using diffusion models☆77Oct 3, 2023Updated 2 years ago