Project website for "Telling left from right: Learning spatial correspondence between sight and sound"
☆29Jun 6, 2022Updated 4 years ago
Alternatives and similar repositories for telling-left-from-right
Users that are interested in telling-left-from-right are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Localize to Binauralize: Audio Spatialization from Visual Sound Source Localization (ICCV 2021)☆10Oct 11, 2021Updated 4 years ago
- ☆31Jun 14, 2022Updated 4 years ago
- MRSAudio: A Large-Scale Multimodal Recorded Spatial Audio Dataset with Refined Annotations☆43Aug 29, 2026Updated 3 weeks ago
- Co-Separating Sounds of Visual Objects (ICCV 2019)☆98Jul 25, 2023Updated 3 years ago
- 2.5D visual sound dataset☆108Sep 21, 2021Updated 5 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Codebase and Dataset for the paper: Learning to Localize Sound Source in Visual Scenes☆102Dec 4, 2024Updated last year
- Code for the ISMIR 2021 tutorial "Programming MIR Baselines from Scratch: Three Cases Studies"☆31Nov 21, 2021Updated 4 years ago
- Multimodal Variational Auto-encoder based Audio-Visual Segmentation [ICCV2023].☆20Sep 19, 2024Updated 2 years ago
- Official implementation of the paper How to Listen? Rethinking Visual Sound Localization☆18Apr 25, 2022Updated 4 years ago
- ☆19Jan 10, 2024Updated 2 years ago
- Official Implementation of our Interspeech 2021 paper "An Empirical Study on Channel Effects for Synthetic Voice Spoofing Countermeasure …☆20Feb 15, 2022Updated 4 years ago
- Official implementation for MGN☆20Dec 22, 2022Updated 3 years ago
- [NeurIPS 2022] Your Transformer May Not be as Powerful as You Expect (official implementation)☆35Aug 6, 2023Updated 3 years ago
- [INTERSPEECH 2024] Official pytorch code for the paper "Disentangled Representation Learning for Environment-agnostic Speaker Recognition…☆20Jul 23, 2024Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Official repository for the paper "Audio xLSTMs: Learning Self-supervised audio representations with xLSTMs"☆22Sep 7, 2025Updated last year
- Evaluation kit for the HEAR Benchmark☆65Feb 12, 2026Updated 7 months ago
- Official repo for MMAU-Pro Benchmark☆22Sep 25, 2025Updated last year
- [ICCV'21] The Right to Talk: An Audio-Visual Transformer Approach☆20Aug 2, 2021Updated 5 years ago
- [ICLR'25] Official repository for "AVHBench: A Cross-Modal Hallucination Evaluation for Audio-Visual Large Language Models"☆26Mar 8, 2026Updated 6 months ago
- Spatial Audio Generation☆116Mar 24, 2023Updated 3 years ago
- ☆29Jun 22, 2022Updated 4 years ago
- Implementation for ECCV20 paper "Self-Supervised Learning of audio-visual objects from video"☆114Nov 16, 2020Updated 5 years ago
- Official Repository for "SingFake: Singing Voice Deepfake Detection"☆65Feb 26, 2024Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Containing SOTA methods that follows time-varying conditions for Text-to-Music☆24Jan 1, 2026Updated 8 months ago
- ☆12Oct 2, 2020Updated 5 years ago
- Baseline system for SVDD 2024 Challenge CtrSVDD track☆30Nov 16, 2024Updated last year
- ☆24Jun 28, 2019Updated 7 years ago
- A first-of-its-kind acoustic simulation platform for audio-visual embodied AI research. It supports training and evaluating multiple task…☆472Sep 29, 2023Updated 2 years ago
- Contrastive Language-Audio Pretraining☆15May 18, 2021Updated 5 years ago
- [CVPR 2024] "Towards Robust Audiovisual Segmentation in Complex Environments with Quantization-based Semantic Decomposition"☆11Feb 27, 2024Updated 2 years ago
- Open-source audio embedding models, submitted to the HEAR 2021 challenge☆11Feb 15, 2026Updated 7 months ago
- Official PyTorch implementation of "Conditional Generation of Audio from Video via Foley Analogies".☆93Dec 8, 2023Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Official code and pretrained models for Linear Consistency Autoencoders (Lin-CAE), a method to induce linearity in audio autoencoders via…☆18Feb 12, 2026Updated 7 months ago
- PyTorch code for "Self-Supervised Predictive Learning: A Negative-Free Method for Sound Source Localization in Visual Scenes" (CVPR, 2022…☆32Jul 8, 2024Updated 2 years ago
- ☆14Nov 22, 2022Updated 3 years ago
- ☆14Feb 26, 2024Updated 2 years ago
- Official implementation of the SPL paper "One-class Learning Towards Synthetic Voice Spoofing Detection"☆140Aug 30, 2024Updated 2 years ago
- ☆37Jun 30, 2022Updated 4 years ago
- ☆10Apr 22, 2016Updated 10 years ago