Data generator for stereo sound event localization and detection task of DCASE 2025 challenge
☆17Jul 17, 2025Updated last year
Alternatives and similar repositories for dcase2025_stereo_seld_data_generator
Users that are interested in dcase2025_stereo_seld_data_generator are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆29May 27, 2025Updated last year
- Enhanced sound event localization and detection in real 360-degree audio-visual soundscapes (DCASE task3 format)☆14Mar 21, 2025Updated last year
- Python implementation of the paper "Fusion of Audio and Visual Embeddings for Sound Event Localization and Detection"☆32Apr 26, 2024Updated 2 years ago
- The program ranked first in Audio-only track of DCASE2024 Challenge task3.☆23Mar 2, 2026Updated 6 months ago
- ☆78Aug 7, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- PSELDNets: Pre-trained Neural Networks on Large-scale Synthetic Datasets for Sound Event Localization and Detection☆49Sep 17, 2025Updated last year
- For more detailed information, please refer to the paper titled "MVANet: Multi-Stage Video Attention Network for Sound Event Localization…☆35May 20, 2025Updated last year
- ☆53Dec 13, 2025Updated 9 months ago
- Sound Event Localization and Detection using Neural Generalized Cross-Correlations☆38Feb 11, 2025Updated last year
- The Neural-SRP method for DOA estimation☆37May 24, 2024Updated 2 years ago
- Data generator for sound event localization and detection clips, including 4-ch microphone-array-format signals and first-order-ambisonic…☆22Nov 13, 2024Updated last year
- Data generator for creating synthetic audio mixtures suitable for DCASE Challenge 2022 Task 3☆47Apr 5, 2023Updated 3 years ago
- Implementation of the paper "Binaural Sound Source Distance Estimation and Localization for a Moving Listener"☆23Mar 2, 2025Updated last year
- A simple Python script to convert FOA audio to binaural.☆17Nov 29, 2022Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Onset-and-Offset-Aware Sound Event Detection☆22Feb 10, 2025Updated last year
- A controllable, end-to-end API for soundscape synthesis across ray-traced & real-world measured acoustics☆29Apr 1, 2026Updated 5 months ago
- ICASSP 2024: Robust DOA estimation from deep acoustic imaging☆25Apr 14, 2024Updated 2 years ago
- Official implementation of "sound distance estimation" WASPAA 23☆20Dec 31, 2023Updated 2 years ago
- The official repo for Both Ears Wide Open: Towards Language-Driven Spatial Audio Generation☆66Jul 2, 2025Updated last year
- A python implementation of “SRP-DNN: Learning Direct-Path Phase Difference for Multiple Moving Sound Source Localization” [ICASSP 2022]☆69Sep 28, 2024Updated last year
- CST-former: Transformer with Channel-Spectro-Temporal Attention for Sound Event Localization and Detection (ICASSP 2024)☆40May 20, 2025Updated last year
- Official codes for 'Preserving Full Degradation Details for Blind Image Super-Resolution'☆12Jul 2, 2024Updated 2 years ago
- Unofficial Pytorch Lightning Implementation of "Towards Robust Speech Super-Resolution"☆10May 8, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Balatro calculator☆18Jul 10, 2026Updated 2 months ago
- This repository contains the code of the CP JKU submission to DCASE23 Task 1 "Low-complexity Acoustic Scene Classification"☆33Sep 18, 2023Updated 3 years ago
- ☆17Oct 24, 2025Updated 11 months ago
- ☆15May 25, 2026Updated 3 months ago
- [AAAI 2025] Towards Audio-visual Navigation in Noisy Environments: A Large-scale Benchmark Dataset and An Architecture Considering Multip…☆17May 21, 2026Updated 4 months ago
- ☆12Apr 1, 2020Updated 6 years ago
- ☆13Sep 4, 2023Updated 3 years ago
- Modular implementation of the Steered Response Power method and its variants☆48Mar 25, 2026Updated 5 months ago
- The Official PyTorch Implementation of FN-SSL & IPDnet for Sound Source Localization [INTERSPEECH2023 & TASLP2024]☆165Mar 10, 2026Updated 6 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- codes for RFSR: Improving ISR Diffusion Models via Reward Feedback Learning☆18Dec 8, 2024Updated last year
- ☆12Mar 5, 2024Updated 2 years ago
- ☆17Apr 3, 2025Updated last year
- ☆50Jul 8, 2025Updated last year
- About Official PyTorch(MMCV) implementation of “SUMix: Mixup with Semantic and Uncertain Information” (ECCV 2024)☆12Sep 2, 2024Updated 2 years ago
- ☆19Oct 18, 2025Updated 11 months ago
- WildDESED: A LLM-Powered Dataset for Wild Domestic Environment Sound Event Detection☆19Nov 19, 2024Updated last year