Project website for "Telling left from right: Learning spatial correspondence between sight and sound"
☆29Jun 6, 2022Updated 4 years ago
Alternatives and similar repositories for telling-left-from-right
Users that are interested in telling-left-from-right are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Localize to Binauralize: Audio Spatialization from Visual Sound Source Localization (ICCV 2021)☆10Oct 11, 2021Updated 4 years ago
- ☆31Jun 14, 2022Updated 4 years ago
- MRSAudio: A Large-Scale Multimodal Recorded Spatial Audio Dataset with Refined Annotations☆43Aug 29, 2026Updated last week
- Co-Separating Sounds of Visual Objects (ICCV 2019)☆98Jul 25, 2023Updated 3 years ago
- 2.5D visual sound dataset☆108Sep 21, 2021Updated 4 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Codebase and Dataset for the paper: Learning to Localize Sound Source in Visual Scenes☆102Dec 4, 2024Updated last year
- Code for the ISMIR 2021 tutorial "Programming MIR Baselines from Scratch: Three Cases Studies"☆31Nov 21, 2021Updated 4 years ago
- Code supporting the ISMIR 2020 Klio Tutorial☆20Oct 11, 2020Updated 5 years ago
- Codebase for the paper "Sep-Stereo: Visually Guided Stereophonic Audio Generation by Associating Source Separation" (ECCV2020)☆72Oct 20, 2020Updated 5 years ago
- Multimodal Variational Auto-encoder based Audio-Visual Segmentation [ICCV2023].☆20Sep 19, 2024Updated last year
- Official implementation of the paper How to Listen? Rethinking Visual Sound Localization☆18Apr 25, 2022Updated 4 years ago
- ☆19Jan 10, 2024Updated 2 years ago
- Official Implementation of our Interspeech 2021 paper "An Empirical Study on Channel Effects for Synthetic Voice Spoofing Countermeasure …☆19Feb 15, 2022Updated 4 years ago
- Official implementation for MGN☆20Dec 22, 2022Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [NeurIPS 2022] Your Transformer May Not be as Powerful as You Expect (official implementation)☆35Aug 6, 2023Updated 3 years ago
- [INTERSPEECH 2024] Official pytorch code for the paper "Disentangled Representation Learning for Environment-agnostic Speaker Recognition…☆19Jul 23, 2024Updated 2 years ago
- Evaluation kit for the HEAR Benchmark☆65Feb 12, 2026Updated 6 months ago
- Official repo for MMAU-Pro Benchmark☆22Sep 25, 2025Updated 11 months ago
- [ICCV'21] The Right to Talk: An Audio-Visual Transformer Approach☆20Aug 2, 2021Updated 5 years ago
- Baseline method for audio-visual sound event localization and detection task of DCASE 2023 challenge☆69Mar 19, 2025Updated last year
- [ICLR'25] Official repository for "AVHBench: A Cross-Modal Hallucination Evaluation for Audio-Visual Large Language Models"☆26Mar 8, 2026Updated 5 months ago
- Spatial Audio Generation☆116Mar 24, 2023Updated 3 years ago
- ☆29Jun 22, 2022Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Implementation for ECCV20 paper "Self-Supervised Learning of audio-visual objects from video"☆114Nov 16, 2020Updated 5 years ago
- Official Repository for "SingFake: Singing Voice Deepfake Detection"☆65Feb 26, 2024Updated 2 years ago
- Containing SOTA methods that follows time-varying conditions for Text-to-Music☆24Jan 1, 2026Updated 8 months ago
- ☆12Oct 2, 2020Updated 5 years ago
- Baseline system for SVDD 2024 Challenge CtrSVDD track☆30Nov 16, 2024Updated last year
- ☆24Jun 28, 2019Updated 7 years ago
- A first-of-its-kind acoustic simulation platform for audio-visual embodied AI research. It supports training and evaluating multiple task…☆470Sep 29, 2023Updated 2 years ago
- Contrastive Language-Audio Pretraining☆15May 18, 2021Updated 5 years ago
- [CVPR 2024] "Towards Robust Audiovisual Segmentation in Complex Environments with Quantization-based Semantic Decomposition"☆11Feb 27, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Open-source audio embedding models, submitted to the HEAR 2021 challenge☆11Feb 15, 2026Updated 6 months ago
- Official PyTorch implementation of "Conditional Generation of Audio from Video via Foley Analogies".☆93Dec 8, 2023Updated 2 years ago
- Official code and pretrained models for Linear Consistency Autoencoders (Lin-CAE), a method to induce linearity in audio autoencoders via…☆17Feb 12, 2026Updated 6 months ago
- ☆14Nov 22, 2022Updated 3 years ago
- ☆14Feb 26, 2024Updated 2 years ago
- Official implementation of the SPL paper "One-class Learning Towards Synthetic Voice Spoofing Detection"☆140Aug 30, 2024Updated 2 years ago
- ☆37Jun 30, 2022Updated 4 years ago