[CVPR'26] Semantic Audio-Visual Navigation in Continuous Environments
☆34Sep 5, 2026Updated last week
Alternatives and similar repositories for SAVN-CE
Users that are interested in SAVN-CE are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [AAAI 2025] Towards Audio-visual Navigation in Noisy Environments: A Large-scale Benchmark Dataset and An Architecture Considering Multip…☆17May 21, 2026Updated 3 months ago
- [ISER 2023] The official implementation of Audio Visual Language Maps for Robot Navigation☆69May 11, 2024Updated 2 years ago
- Audio propagation engine - Meta Reality Labs Research.☆24Nov 1, 2022Updated 3 years ago
- ☆15May 25, 2026Updated 3 months ago
- ☆25Jan 16, 2026Updated 7 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆15Dec 6, 2024Updated last year
- ☆12Mar 28, 2025Updated last year
- Awesome Audio-Visual Intelligence, Survey of Audio-Visual Intelligence☆89May 8, 2026Updated 4 months ago
- (CVPR 2023) HypLiLoc: Towards Effective LiDAR Pose Regression with Hyperbolic Fusion☆57Dec 17, 2023Updated 2 years ago
- 🎉 [ICLR 2026] All-Day Multi-Scenes Lifelong Vision-and-Language Navigation with Tucker Adaptation☆40Jun 29, 2026Updated 2 months ago
- Habitat ROS is a ROS 1 package for robot simulation in habitat-sim providing customizable robotic sensors (2D Laser, 3D Lidar, RGBD camer…☆43Jun 29, 2024Updated 2 years ago
- Code and datasets for 'Move2Hear: Active Audio-Visual Source Separation' (ICCV 2021)☆16Jun 17, 2026Updated 2 months ago
- Fastest CUDA RGB to grayscale: 5-30x faster than OpenCV. For image processing/computer vision.☆16Mar 23, 2021Updated 5 years ago
- ☆11Jul 16, 2024Updated 2 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Code for paper Audio Visual Speaker Localization from EgoCentric Views☆11Jul 3, 2024Updated 2 years ago
- Code repo for paper: InfiniBench: Infinite Benchmarking for Visual Spatial Reasoning with Customizable Scene Complexity☆19May 13, 2026Updated 4 months ago
- Official Implementation of paper: [Nav-R2:Dual‑Relation Reasoning for Generalizable Open‑Vocabulary Object‑Goal Navigation]☆21Dec 10, 2025Updated 9 months ago
- [TIV 2025] C2L-PR: Cross-modal Camera-to-LiDAR Place Recognition via Modality Alignment and Orientation Voting.☆20Mar 28, 2026Updated 5 months ago
- Implementation of algorithms for refinement of direction of arrival estimators by optimization☆15Jun 2, 2021Updated 5 years ago
- Official implementation of Why Only Text: Empowering Vision-and-Language Navigation with Multi-modal Prompts(IJCAI 2024)☆15Oct 16, 2024Updated last year
- A stacked self-attention network for two-dimensional direction-of-arrival estimation in hands-free speech communication☆12Sep 12, 2024Updated 2 years ago
- Official Repository for the ACM MM 2024 paper "Navigating Beyond Instructions: Vision-and-Language Navigation in Obstructed Environments"☆16May 16, 2025Updated last year
- 📚 2025 Scene Graph ArXiv Paper List — Updated Daily☆16Mar 18, 2026Updated 5 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- ☆45Mar 13, 2023Updated 3 years ago
- Official code for IROS 2025 paper "TextInPlace: Indoor Visual Place Recognition in Repetitive Structures with Scene Text Spotting and Ver…☆19Dec 27, 2025Updated 8 months ago
- Implementation (R2R part) for the paper "Iterative Vision-and-Language Navigation"☆18Apr 4, 2024Updated 2 years ago
- A tutorial for Sound Source Localization researchers and practitioners. The purpose of this repo is to organize the world’s resources for…☆60Mar 17, 2023Updated 3 years ago
- Tools to convert sigsep mus dataset from STEMS <-> WAV☆12Jul 15, 2020Updated 6 years ago
- ☆21Feb 12, 2025Updated last year
- ☆26Feb 4, 2026Updated 7 months ago
- ☆18Jan 26, 2021Updated 5 years ago
- Implementation of CVPR2024 paper "TransLoc4D: Transformer-based 4D Radar Place Recognition."☆47Aug 5, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A description of "RealMAN: A Real-Recorded and Annotated Microphone Array Dataset for Dynamic Speech Enhancement and Localization" [NeurI…☆178Apr 29, 2025Updated last year
- Code and datasets for 'Few-Shot Audio-Visual Learning of Environment Acoustics' (NeurIPS 2022)☆25Jun 16, 2026Updated 2 months ago
- Dynamic vision-guided speaker embedding for audio-visual speaker diarization☆12Jul 5, 2022Updated 4 years ago
- This repository is for the paper Incorporating External POS Tagger for Punctuation Restoration. Proc. Interspeech 2021, 1987-1991, doi: 1…☆11May 24, 2026Updated 3 months ago
- ☆15Sep 26, 2023Updated 2 years ago
- My implementation of a scene memory transformer module for reinforcement learning☆14Jun 19, 2019Updated 7 years ago
- SAAVN Code release for paper "Sound Adversarial Audio-Visual Navigation,ICLR2022" (In PyTorch)☆21Nov 9, 2022Updated 3 years ago