SAAVN Code release for paper "Sound Adversarial Audio-Visual Navigation,ICLR2022" (In PyTorch)
☆21Nov 9, 2022Updated 3 years ago
Alternatives and similar repositories for SAAVN
Users that are interested in SAAVN are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Starter code for SoundSpaces challenge at CVPR 21's Embodied AI workshop☆16Mar 2, 2023Updated 3 years ago
- Audio propagation engine - Meta Reality Labs Research.☆24Nov 1, 2022Updated 3 years ago
- [ICCV 2021] Code release for "Sub-bit Neural Networks: Learning to Compress and Accelerate Binary Neural Networks"☆34Jul 24, 2022Updated 4 years ago
- Code and datasets for 'Move2Hear: Active Audio-Visual Source Separation' (ICCV 2021)☆16Jun 17, 2026Updated 2 months ago
- [ECCV 2020] Code release for "Resolution Switchable Networks for Runtime Efficient Image Recognition"☆40Aug 11, 2020Updated 6 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- A first-of-its-kind acoustic simulation platform for audio-visual embodied AI research. It supports training and evaluating multiple task…☆470Sep 29, 2023Updated 2 years ago
- The official project website of "NORM: Knowledge Distillation via N-to-One Representation Matching" (The paper of NORM is published in IC…☆20Sep 18, 2023Updated 2 years ago
- [ISER 2023] The official implementation of Audio Visual Language Maps for Robot Navigation☆69May 11, 2024Updated 2 years ago
- [ACMMM 2021, Oral] Code release for "Elastic Tactile Simulation Towards Tactile-Visual Perception"☆50Jul 20, 2022Updated 4 years ago
- ☆23Mar 20, 2024Updated 2 years ago
- Code and datasets for 'Few-Shot Audio-Visual Learning of Environment Acoustics' (NeurIPS 2022)☆25Jun 16, 2026Updated 2 months ago
- The official project website of "SliderQuant: Accurate Post-Training Quantization for LLMs" (accepted to ICLR 2026).☆26Jun 15, 2026Updated 2 months ago
- Code for paper Learning Audio-Visual Dereverberation☆32Aug 10, 2022Updated 4 years ago
- Repo for Visual Acoustic Matching, CVPR 2022☆71Feb 28, 2023Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- The official project website of "3D Human Pose Lifting with Grid Convolution" (GridConv for short, oral in AAAI 2023)☆34Dec 19, 2023Updated 2 years ago
- A structural model for HRTF individualization combining measured, synthesized, and selected components.☆23Jan 15, 2023Updated 3 years ago
- [ICLR 2024 Poster] SCHEMA: State CHangEs MAtter for Procedure Planning in Instructional Videos☆20Aug 21, 2025Updated last year
- Fastest CUDA RGB to grayscale: 5-30x faster than OpenCV. For image processing/computer vision.☆16Mar 23, 2021Updated 5 years ago
- We propose a novel approach for reconstructing human expressiveness in piano performance with a multi-layer bi-directional Transformer. (…☆20May 16, 2024Updated 2 years ago
- Deep learning for pedestrians: backpropagation in CNNs. Latex and PyTorch code to verify theoretical derivations.☆13Jun 21, 2022Updated 4 years ago
- Indoor Navigation for a Differential-Drive Robot in an Unknown and Dynamic Environments☆12May 14, 2021Updated 5 years ago
- cordial-sync is a software package than can be used to reproduce the results from the paper "A Cordial Sync: Going Beyond Marginal Polici…☆41Jan 13, 2021Updated 5 years ago
- [AAAI 2025] Towards Audio-visual Navigation in Noisy Environments: A Large-scale Benchmark Dataset and An Architecture Considering Multip…☆17May 21, 2026Updated 3 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆13Sep 4, 2023Updated 3 years ago
- Improving Recording Device Generalization using Impulse Response Augmentation☆21Apr 24, 2025Updated last year
- Official implementation for MGN☆20Dec 22, 2022Updated 3 years ago
- ☆12Sep 28, 2021Updated 4 years ago
- 🦇 Encoder of BAT (Learning to Reason about Spatial Sounds with Large Language Models)☆90Feb 13, 2025Updated last year
- ☆10Sep 7, 2021Updated 4 years ago
- An alternative EQA paradigm and informative benchmark + models (BMVC 2019, ViGIL 2019 spotlight)☆25Jun 22, 2022Updated 4 years ago
- Official Implementation of "Inference and Denoise: Causal Inference-based Neural Speech Enhancement"☆28Feb 26, 2023Updated 3 years ago
- CMU Masters Thesis Project: UAV Path Planning and Human Trajectory Prediction for Navigation through Work Sites.☆11May 4, 2021Updated 5 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Official implementation of NeurIPS 2022 paper "Learning Active Camera for Multi-Object Navigation"☆14Apr 23, 2023Updated 3 years ago
- Learning to Separate Object Sounds by Watching Unlabeled Video (ECCV 2018)☆50Sep 24, 2019Updated 6 years ago
- Unofficial PyTorch implementation of MapNet: An Allocentric Spatial Memory for Mapping Environments☆12Jun 4, 2020Updated 6 years ago
- the code for 'Global HRTF Personalization Using Anthropometric Measures'(AES 150th convention)☆36Jul 24, 2022Updated 4 years ago
- Knowledge Transfer via Dense Cross-layer Mutual-distillation (ECCV'2020)☆30Aug 19, 2020Updated 6 years ago
- [CVPR 2023] Egocentric Audio-Visual Object Localization☆27Jan 6, 2024Updated 2 years ago
- [ECCV 2022] Joint-Modal Label Denoising for Weakly-Supervised Audio-Visual Video Parsing☆27Jul 15, 2022Updated 4 years ago