SAAVN Code release for paper "Sound Adversarial Audio-Visual Navigation,ICLR2022" (In PyTorch)
☆21Nov 9, 2022Updated 3 years ago
Alternatives and similar repositories for SAAVN
Users that are interested in SAAVN are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Starter code for SoundSpaces challenge at CVPR 21's Embodied AI workshop☆16Mar 2, 2023Updated 3 years ago
- Audio propagation engine - Meta Reality Labs Research.☆24Nov 1, 2022Updated 3 years ago
- Code and datasets for 'Move2Hear: Active Audio-Visual Source Separation' (ICCV 2021)☆16Jun 17, 2026Updated last month
- [ECCV 2020] Code release for "Resolution Switchable Networks for Runtime Efficient Image Recognition"☆40Aug 11, 2020Updated 5 years ago
- The official project website of "Ske2Grid: Skeleton-to-Grid Representation Learning for Action Recognition" (The paper of Ske2Grid is pub…☆19Sep 6, 2023Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- A first-of-its-kind acoustic simulation platform for audio-visual embodied AI research. It supports training and evaluating multiple task…☆468Sep 29, 2023Updated 2 years ago
- The official project website of "NORM: Knowledge Distillation via N-to-One Representation Matching" (The paper of NORM is published in IC…☆20Sep 18, 2023Updated 2 years ago
- [ISER 2023] The official implementation of Audio Visual Language Maps for Robot Navigation☆68May 11, 2024Updated 2 years ago
- [ACMMM 2021, Oral] Code release for "Elastic Tactile Simulation Towards Tactile-Visual Perception"☆50Jul 20, 2022Updated 4 years ago
- Code and datasets for 'Few-Shot Audio-Visual Learning of Environment Acoustics' (NeurIPS 2022)☆24Jun 16, 2026Updated last month
- A Closed-form Solution to Universal Style Transfer - ICCV 2019☆34Dec 2, 2019Updated 6 years ago
- The official project website of "SliderQuant: Accurate Post-Training Quantization for LLMs" (accepted to ICLR 2026).☆24Jun 15, 2026Updated last month
- Sound field reconstruction using neural processes with dynamic kernels☆16Mar 25, 2025Updated last year
- Code for paper Learning Audio-Visual Dereverberation☆32Aug 10, 2022Updated 3 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Repo for Visual Acoustic Matching, CVPR 2022☆71Feb 28, 2023Updated 3 years ago
- [ICLR 2024 Poster] SCHEMA: State CHangEs MAtter for Procedure Planning in Instructional Videos☆20Aug 21, 2025Updated 11 months ago
- Fastest CUDA RGB to grayscale: 5-30x faster than OpenCV. For image processing/computer vision.☆16Mar 23, 2021Updated 5 years ago
- We propose a novel approach for reconstructing human expressiveness in piano performance with a multi-layer bi-directional Transformer. (…☆21May 16, 2024Updated 2 years ago
- ☆14Nov 13, 2023Updated 2 years ago
- Deep learning for pedestrians: backpropagation in CNNs. Latex and PyTorch code to verify theoretical derivations.☆13Jun 21, 2022Updated 4 years ago
- [AAAI 2025] Towards Audio-visual Navigation in Noisy Environments: A Large-scale Benchmark Dataset and An Architecture Considering Multip…☆16May 21, 2026Updated 2 months ago
- ☆13Sep 4, 2023Updated 2 years ago
- Improving Recording Device Generalization using Impulse Response Augmentation☆21Apr 24, 2025Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- The repo for "On-the-fly Modulation for Balanced Multimodal Learning", T-PAMI 2024☆19Sep 29, 2024Updated last year
- VisualEchoes Dataset (ECCV 2020)☆37Aug 31, 2021Updated 4 years ago
- Official implementation for MGN☆20Dec 22, 2022Updated 3 years ago
- ☆12Sep 28, 2021Updated 4 years ago
- 🦇 Encoder of BAT (Learning to Reason about Spatial Sounds with Large Language Models)☆87Feb 13, 2025Updated last year
- ☆10Sep 7, 2021Updated 4 years ago
- The official project website of "ScaleKD: Strong Vision Transformers Could Be Excellent Teachers" (ScaleKD for short, accepted to NeurIPS…☆68Apr 15, 2026Updated 3 months ago
- Official Implementation of "Inference and Denoise: Causal Inference-based Neural Speech Enhancement"☆28Feb 26, 2023Updated 3 years ago
- CMU Masters Thesis Project: UAV Path Planning and Human Trajectory Prediction for Navigation through Work Sites.☆11May 4, 2021Updated 5 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆14Jul 20, 2021Updated 5 years ago
- Official implementation of NeurIPS 2022 paper "Learning Active Camera for Multi-Object Navigation"☆14Apr 23, 2023Updated 3 years ago
- Learning to Separate Object Sounds by Watching Unlabeled Video (ECCV 2018)☆50Sep 24, 2019Updated 6 years ago
- Unofficial PyTorch implementation of MapNet: An Allocentric Spatial Memory for Mapping Environments☆12Jun 4, 2020Updated 6 years ago
- the code for 'Global HRTF Personalization Using Anthropometric Measures'(AES 150th convention)☆36Jul 24, 2022Updated 4 years ago
- ☆10Oct 11, 2022Updated 3 years ago
- [CVPR 2023] Egocentric Audio-Visual Object Localization☆27Jan 6, 2024Updated 2 years ago