[CVPR25] Official Implementation of CAV-MAE Sync
☆33Apr 5, 2026Updated 5 months ago
Alternatives and similar repositories for cav-mae-sync
Users that are interested in cav-mae-sync are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- VGGSounder, a multi-label audio-visual classification dataset with modality annotations.☆18Jun 30, 2026Updated 2 months ago
- ☆12Mar 12, 2023Updated 3 years ago
- ☆82Mar 14, 2025Updated last year
- Solos: A Dataset for Audio-Visual Music Analysis☆24Feb 17, 2023Updated 3 years ago
- Code for the C2KD paper (ICASSP 2023)☆20May 15, 2023Updated 3 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Unsupervised word segmentation and clustering of speech☆13Feb 17, 2017Updated 9 years ago
- [ICLR2025] Frechet Wavelet Distance: A metric to detect domain bias in Generative models.☆18Sep 2, 2025Updated last year
- Temperature Schedules for self-supervised contrastive methods on long-tail data (ICLR'23)☆18Apr 25, 2023Updated 3 years ago
- Unified Multisensory Perception: Weakly-Supervised Audio-Visual Video Parsing, ECCV, 2020. (Spotlight)☆90Jul 25, 2024Updated 2 years ago
- ☆23Dec 5, 2023Updated 2 years ago
- [CHIL 2024] Interpretation of Intracardiac Electrograms Through Textual Representations☆12Sep 4, 2024Updated 2 years ago
- [ICCV25] Official Implementation of LeGrad☆99Oct 14, 2024Updated last year
- ☆15Mar 27, 2025Updated last year
- Fast training of unitary deep network layers from low-rank updates☆30Dec 11, 2022Updated 3 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- This is an official pytorch implementation of Learning To Recognize Procedural Activities with Distant Supervision. In this repository, w…☆43Feb 21, 2023Updated 3 years ago
- 👁️🗨️ Scientists often do the same bad stuff. Automate giving deterministic feedback during peer review with determinstic (LLM-free)☆32May 9, 2025Updated last year
- Segmentation of prostatic zones (peripheral zone, central gland, AFS and distal prostatic urethra )☆19Apr 12, 2021Updated 5 years ago
- Efficient Training for Multilingual Visual Speech Recognition: Pre-training with Discretized Visual Speech Representation (ACM MM 2024)☆20Mar 17, 2025Updated last year
- 👆PyTorch Implementation of JEDi Metric described in "Beyond FVD: Enhanced Evaluation Metrics for Video Generation Quality"☆37Dec 8, 2024Updated last year
- [CVPR 2025] Pytorch implementation of the paper "Learning to Highlight Audio by Watching Movies"☆15Oct 1, 2025Updated 11 months ago
- ☆12Mar 24, 2024Updated 2 years ago
- Learnable Weight Initialization for Volumetric Medical Image Segmentation [Elsevier AIM2024]☆22Oct 27, 2024Updated last year
- awesome-audio-visual-robustness☆12Jan 27, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆36Jan 20, 2025Updated last year
- Code for the AVLnet (Interspeech 2021) and Cascaded Multilingual (Interspeech 2021) papers.☆54Mar 30, 2022Updated 4 years ago
- Video descriptions of research papers relating to foundation models and scaling☆29Mar 16, 2023Updated 3 years ago
- Visual Speech Recognition For Low-Resource Languages with Automatic Labels (ICASSP 2024)☆17Mar 17, 2025Updated last year
- Repository for the paper: dense and aligned captions (dac) promote compositional reasoning in vl models☆28Nov 29, 2023Updated 2 years ago
- Offical code for the CVPR 2024 Paper: Separating the "Chirp" from the "Chat": Self-supervised Visual Grounding of Sound and Language☆90Jun 12, 2024Updated 2 years ago
- Humans-in-Kitchens Dataset API ([NeurIPS 2023 Dataset and Benchmark Track])☆41Sep 29, 2024Updated last year
- WildVSR☆22Dec 13, 2023Updated 2 years ago
- Official code for the paper "Scaling Multilingual Visual Speech Recognition"☆20Aug 15, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Official implementation of "VideoZeroBench: Probing the Limits of Video MLLMs with Spatio-Temporal Evidence Verification"☆26May 7, 2026Updated 4 months ago
- Official repository for the MMFM challenge☆26Jun 18, 2024Updated 2 years ago
- Pitman-Yor processes in python☆26Apr 18, 2014Updated 12 years ago
- 哈工大2021秋计算机网络☆13Mar 30, 2023Updated 3 years ago
- Download audioset data super fastly with youtube-dl, ffmpeg and python multiprocessing☆48Aug 1, 2024Updated 2 years ago
- PyTorch DataSet and Jupyter demos for MusicNet☆75Mar 21, 2024Updated 2 years ago
- Continual Online Recalibration with Pseudo-labels☆15Jun 20, 2024Updated 2 years ago