Localize to Binauralize: Audio Spatialization from Visual Sound Source Localization (ICCV 2021)
☆10Oct 11, 2021Updated 4 years ago
Alternatives and similar repositories for Localize-to-Binauralize
Users that are interested in Localize-to-Binauralize are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Project website for "Telling left from right: Learning spatial correspondence between sight and sound"☆29Jun 6, 2022Updated 4 years ago
- Official Implementation of our Interspeech 2021 paper "An Empirical Study on Channel Effects for Synthetic Voice Spoofing Countermeasure …☆20Feb 15, 2022Updated 4 years ago
- [ICCV'21] The Right to Talk: An Audio-Visual Transformer Approach☆20Aug 2, 2021Updated 5 years ago
- ☆15Jan 24, 2025Updated last year
- ☆24Jun 28, 2019Updated 7 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Code for "Recognizing Scenes from Novel Viewpoints"☆29Sep 16, 2022Updated 4 years ago
- Codebase for the paper "Visually Informed Binaural Audio Generation without Binaural Audios" (CVPR 2021)☆72Jul 8, 2021Updated 5 years ago
- Project MANAS Official Website☆13Jan 10, 2022Updated 4 years ago
- A voice spoofing detection system, based on paper presented at ICSPIS 2021☆11Feb 11, 2022Updated 4 years ago
- Codebase for the paper "Sep-Stereo: Visually Guided Stereophonic Audio Generation by Associating Source Separation" (ECCV2020)☆72Oct 20, 2020Updated 5 years ago
- ☆49Sep 11, 2026Updated last week
- Code for WACV24 work for multiview acoustic-visual detection☆13Mar 22, 2024Updated 2 years ago
- 🚀 海南大学编译原理 pl0 语言编译器扩充☆11Dec 19, 2020Updated 5 years ago
- MRSAudio: A Large-Scale Multimodal Recorded Spatial Audio Dataset with Refined Annotations☆43Aug 29, 2026Updated 3 weeks ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆31Jun 14, 2022Updated 4 years ago
- Towards Intelligibility-Oriented Audio-Visual Speech Enhancement☆15Sep 6, 2024Updated 2 years ago
- PyTorch implementation of Retriever: Learning Content-Style Representation☆12Jan 27, 2023Updated 3 years ago
- Code and data recipes for the paper: Heterogeneous Target Speech Separation☆44Dec 6, 2022Updated 3 years ago
- A Benchmark Corpus for Low-Resource Cantonese Punctuation Restoration from Speech Transcripts☆15Dec 3, 2024Updated last year
- PaperBot: Learning to Design Real-World Tools Using Paper☆13Mar 15, 2024Updated 2 years ago
- Official repository of the work "Low-complexity Unsupervised Audio Anomaly Detection exploiting Separable Convolutions and Angular Loss" …☆11Nov 6, 2024Updated last year
- [CVPR 2025] LumiNet: Latent Intrinsics Meets Diffusion Models for Indoor Scene Relighting☆47Sep 16, 2025Updated last year
- N/A☆190May 19, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆14Nov 22, 2022Updated 3 years ago
- Source code for paper "Breaking Security-Critical Voice Authentication".☆13Jul 10, 2023Updated 3 years ago
- ☆14Dec 8, 2025Updated 9 months ago
- This repository contains materials for the paper: Towards generating ambisonics using audio-visual cue for virtual reality☆13Jul 2, 2019Updated 7 years ago
- ☆10Apr 17, 2024Updated 2 years ago
- CANTE: Automatic transcription of flamenco singing.☆14Feb 13, 2018Updated 8 years ago
- Code for Deep Multimodal Clustering for Unsupervised Audiovisual Learning (CVPR2019)☆15May 27, 2020Updated 6 years ago
- Neural model for prediction of stress position in Russian words☆13Jun 22, 2025Updated last year
- Prabhupadavani: A Code-mixed Speech Translation Data for 25 languages☆14Oct 12, 2022Updated 3 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Light Field Super-Resolution with Zero-Shot Learning, CVPR 2021, Oral.☆30Oct 21, 2021Updated 4 years ago
- ☆14Aug 16, 2023Updated 3 years ago
- Code for the paper "Representing Spatial Trajectories as Distributions"☆13Jan 17, 2023Updated 3 years ago
- Dynamic vision-guided speaker embedding for audio-visual speaker diarization☆12Jul 5, 2022Updated 4 years ago
- ☆18Jun 2, 2026Updated 3 months ago
- Denoising autoencoders for speaker identification on MCE 2018 challenge☆12Nov 8, 2018Updated 7 years ago
- A chinese singing voice dataset, professional male singer, 105 songs, 132 minutes☆13Oct 19, 2023Updated 2 years ago