download the vggsound dataset
☆22Feb 22, 2022Updated 4 years ago
Alternatives and similar repositories for vggsound_download
Users that are interested in vggsound_download are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official Codebase of "A Closer Look at Weakly-Supervised Audio-Visual Source Localization" (NeurIPS 2022)☆22Dec 6, 2022Updated 3 years ago
- ☆23Mar 20, 2024Updated 2 years ago
- This repository contains the code for our CVPR 2022 paper on "Audio-visual Generalised Zero-shot Learning with Cross-modal Attention and …☆43Nov 29, 2022Updated 3 years ago
- Localizing Visual Sounds the Hard Way☆84Jul 6, 2022Updated 4 years ago
- [CVPR 2023] Official implementation of our paper - Learning Audio-Visual Source Localization via False Negative Aware Contrastive Learnin…☆30Apr 10, 2023Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Scripts for download AudioSet☆89Nov 7, 2017Updated 8 years ago
- Official code for WACV 2024 paper, "Annotation-free Audio-Visual Segmentation"☆38Oct 11, 2024Updated last year
- [2024 ECCV] Label-anticipated Event Disentanglement for Audio-Visual Video Parsing☆14Nov 17, 2024Updated last year
- Cross-Modal Relation-Aware Networks for Audio-Visual Event Localization, ACM MM 2020☆33Nov 6, 2020Updated 5 years ago
- Official Codebase of "A Unified Audio-Visual Learning Framework for Localization, Separation, and Recognition" (ICML 2023)☆12Jun 1, 2023Updated 3 years ago
- Unified Multisensory Perception: Weakly-Supervised Audio-Visual Video Parsing, ECCV, 2020. (Spotlight)☆90Jul 25, 2024Updated 2 years ago
- Research code for NeurIPS 2023 paper "Modality-Independent Teachers Meet Weakly-Supervised Audio-Visual Event Parser"☆17Jul 13, 2025Updated last year
- A repo for publishing solution to 3DCoMPaT++ challenge on an improved large-scale 3D vision dataset for compositional recognition☆14Jun 22, 2023Updated 3 years ago
- Code for Discriminative Sounding Objects Localization (NeurIPS 2020)☆61Jan 19, 2022Updated 4 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- [NeurIPS 24] Implementation of "Advancing Video Anomaly Detection: A Concise Review and a New Dataset".☆23Apr 16, 2025Updated last year
- Solution for CarRacing-v0 environment from OpenAI Gym. It uses the Deep Deterministic Policy Gradient algorithm.☆12Nov 18, 2022Updated 3 years ago
- This repository contains code for AAAI2025 paper "Dense Audio-Visual Event Localization under Cross-Modal Consistency and Multi-Temporal …☆26Aug 18, 2025Updated last year
- Sapsucker Woods 60 Audiovisual Dataset☆19Oct 7, 2022Updated 3 years ago
- [AAAI 2024] AVSegFormer: Audio-Visual Segmentation with Transformer☆74Mar 6, 2025Updated last year
- repository for paper "Audio-Visual Speech Recognition in MISP2021 Challenge: Dataset Release and Deep Analysis"☆18Jun 17, 2022Updated 4 years ago
- 复旦研究生入学教育测试☆30Aug 28, 2025Updated last year
- label smoothing PyTorch implementation☆33Nov 2, 2020Updated 5 years ago
- A reading list for research topics in multimodal deception detection.☆46Aug 29, 2023Updated 3 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- [ICCV 2023] Audio-Visual Class-Incremental Learning☆38Sep 29, 2024Updated last year
- Pytorch implementation of our paper: Audio-Visual Speech Separation with Visual Features Enhanced by Adversarial Training.☆19Jul 11, 2022Updated 4 years ago
- A curated list of audio-visual learning methods and datasets.☆290Dec 3, 2024Updated last year
- Acoustic camera system for measuring ultrasound communication in rodents☆23Jun 24, 2026Updated 2 months ago
- This is a repo for the paper "Networking Systems for Video Anomaly Detection: A Tutorial and Survey". Paper: https://arxiv.org/abs/2405.1…☆33Mar 26, 2025Updated last year
- A dataset collected from synchronized ad-hoc microphone arrays☆19Apr 24, 2023Updated 3 years ago
- This repository contains the code for our ECCV 2022 paper "Temporal and cross-modal attention for audio-visual zero-shot learning"☆25Sep 12, 2025Updated 11 months ago
- Polyphonic Sound Detection Score (PSDS)☆20Jan 20, 2020Updated 6 years ago
- The code repo for ICASSP 2023 Paper "MMCosine: Multi-Modal Cosine Loss Towards Balanced Audio-Visual Fine-Grained Learning"☆26May 18, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [CVPR 2026] Implementation of HAMMER: Harnessing MLLMs via Cross-Modal Integration for Intention-Driven 3D Affordance Grounding☆23Aug 2, 2026Updated 3 weeks ago
- CVPR2022:Learning from Untrimmed Videos: Self-Supervised Video Representation Learning with Hierarchical Consistency☆18Aug 10, 2022Updated 4 years ago
- Official PyTorch Implementation of MC3D-AD: A Unified Geometry-aware Reconstruction Model for Multi-category 3D Anomaly Detection. Accept…☆15Jan 18, 2026Updated 7 months ago
- Collaborative Learning of Anomalies with Privacy (CLAP) for Unsupervised Video Anomaly Detection: A New Baseline☆23Sep 30, 2024Updated last year
- Prediction of sound event bounding boxes (SEBBs)☆35Aug 2, 2024Updated 2 years ago
- ☆18Oct 2, 2023Updated 2 years ago
- A python implementation of “Learning Deep Direct-Path Relative Transfer Function for Binaural Sound Source Localization” [TASLP 2021]☆27Feb 11, 2023Updated 3 years ago