PySlowFast: video understanding codebase from FAIR for reproducing state-of-the-art video models.
☆14Aug 11, 2020Updated 6 years ago
Alternatives and similar repositories for SlowFast
Users that are interested in SlowFast are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Codebase for VidHal: Benchmarking Hallucinations in Vision LLMs☆14Apr 23, 2026Updated 3 months ago
- Code for CVPR 2021 paper Exploring Heterogeneous Clues for Weakly-Supervised Audio-Visual Video Parsing☆24Dec 29, 2021Updated 4 years ago
- The repo for "Class-aware Sounding Objects Localization", TPAMI 2021.☆29Mar 4, 2022Updated 4 years ago
- Based on StackExchange.Redis that operates Tair For Redis Modules.☆11Feb 28, 2025Updated last year
- [NeurIPS 2025] PreFM: Online Audio-Visual Event Parsing via Predictive Future Modeling☆20Oct 26, 2025Updated 9 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆10Jun 28, 2023Updated 3 years ago
- ☆13Nov 28, 2021Updated 4 years ago
- A paper list of Weakly Supervised Object Detection (WSOD) resources.☆13May 6, 2021Updated 5 years ago
- ☆13Jul 10, 2024Updated 2 years ago
- ☆11Aug 27, 2018Updated 7 years ago
- Pytorch implemenation of structure from motion using Libviso2, SIFT, SuperPoint, SPyNet and Sfm Learner.☆21Oct 12, 2021Updated 4 years ago
- [Arxiv2022] Revitalize Region Feature for Democratizing Video-Language Pre-training☆22Mar 19, 2022Updated 4 years ago
- The repository contains the Pytorch Implementation of the paper Age invariant face recognition and retrieval by coupled auto-encoder netw…☆13Dec 17, 2022Updated 3 years ago
- Self-Supervised Learning by Cross-Modal Audio-Video Clustering (NeurIPS 2020)☆91Oct 24, 2022Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆15Jan 9, 2026Updated 7 months ago
- ☆10Jul 27, 2019Updated 7 years ago
- This is an official implementation of GRIT-VLP☆20Aug 8, 2022Updated 4 years ago
- Listen to Look: Action Recognition by Previewing Audio (CVPR 2020)☆130Aug 31, 2021Updated 4 years ago
- Official repository for "Self-Supervised Video Transformer" (CVPR'22)☆109Jun 26, 2024Updated 2 years ago
- Unofficial implementation of Variational Diffusion Models in PyTorch (Lightning)☆12Aug 31, 2023Updated 2 years ago
- Generalized cross-modal NNs; new audiovisual benchmark (IEEE TNNLS 2019)☆31Apr 13, 2020Updated 6 years ago
- Official Code of ICCV 2021 Paper: Learning to Cut by Watching Movies☆51Nov 9, 2022Updated 3 years ago
- HDU - 在期末的时候给老师评价的小脚本,需要在控制台打开☆14May 21, 2016Updated 10 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆12Oct 23, 2021Updated 4 years ago
- ☆14Nov 13, 2023Updated 2 years ago
- A MaskGIT port from JAX to PyTorch☆18Jun 18, 2022Updated 4 years ago
- Shaping Visual Representations with Language for Few-shot Classification, ACL 2020☆16May 9, 2021Updated 5 years ago
- ☆11Jan 29, 2023Updated 3 years ago
- PyTorch GPU distributed training code for MIL-NCE HowTo100M☆221Jul 5, 2022Updated 4 years ago
- Analyzing Airline data to predict delays☆19May 15, 2014Updated 12 years ago
- ☆13May 23, 2018Updated 8 years ago
- Implementation of "Slow-Fast Auditory Streams for Audio Recognition, ICASSP, 2021" in PyTorch☆73Sep 27, 2021Updated 4 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Models, data, and codes for the paper: MetaAligner: Towards Generalizable Multi-Objective Alignment of Language Models☆24Sep 26, 2024Updated last year
- [CVPR 2024] KEPP: Why Not Use Your Textbook? Knowledge-Enhanced Procedure Planning of Instructional Videos☆12Sep 24, 2024Updated last year
- ☆11Aug 11, 2023Updated 3 years ago
- Code for the Active Speakers in Context Paper (CVPR2020)☆58May 19, 2021Updated 5 years ago
- VGGSound: A Large-scale Audio-Visual Dataset☆358Sep 13, 2021Updated 4 years ago
- Transformers at any scale☆42Jan 18, 2024Updated 2 years ago
- Youku-mPLUG: A 10 Million Large-scale Chinese Video-Language Pre-training Dataset and Benchmarks☆307Jan 8, 2024Updated 2 years ago