A curated list of awesome self-supervised learning methods in videos
☆173Aug 17, 2026Updated this week
Alternatives and similar repositories for awesome-video-self-supervised-learning
Users that are interested in awesome-video-self-supervised-learning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official code repo for TCLR: Temporal Contrastive Learning for Video Representation [CVIU-2022]☆41Feb 28, 2024Updated 2 years ago
- [NeurIPS 2023 (Spotlight)] Uncovering the Hidden Dynamics of Video Self-supervised Learning under Distribution Shifts☆13Jan 30, 2024Updated 2 years ago
- Menagerie of video models trained on various video datasets☆10Oct 13, 2024Updated last year
- This is the official implementation of Global-local Motion Transformer for Unsupervised Skeleton-based Action Learning (ECCV 2022).☆23Nov 6, 2023Updated 2 years ago
- [CVPR'23] AdaMAE: Adaptive Masking for Efficient Spatiotemporal Learning with Masked Autoencoders☆84Feb 2, 2024Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆22Jul 3, 2025Updated last year
- Official implementation of the ICCV 2023 paper "Masked Motion Predictors are Strong 3D Action Representation Learners"☆53Sep 22, 2023Updated 2 years ago
- [NeurIPS 2022 Spotlight] VideoMAE: Masked Autoencoders are Data-Efficient Learners for Self-Supervised Video Pre-Training☆1,782Dec 8, 2023Updated 2 years ago
- Code repository for "Parameter Efficient Self-supervised Geospatial Domain Adaptation", CVPR 2024☆40Jul 29, 2024Updated 2 years ago
- Strefer: Empowering Video LLMs with Space-Time Referring and Reasoning via Synthetic Instruction Data☆19Jun 2, 2026Updated 2 months ago
- ☆24May 11, 2025Updated last year
- Time Does Tell: Self-Supervised Time-Tuning of Dense Image Representations ICCV23☆30Dec 30, 2024Updated last year
- Hierarchical Consistent Contrastive Learning for Skeleton-Based Action Recognition with Growing Augmentations, AAAI 2023☆29Dec 8, 2022Updated 3 years ago
- Github repo for referring atomic video action recognition☆21Oct 2, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆28Oct 8, 2023Updated 2 years ago
- official implementation of CVPR 23 paper "M3Video: Masked Motion Modeling for Self-Supervised Video Representation Learning"☆52Dec 8, 2023Updated 2 years ago
- ☆66Oct 27, 2023Updated 2 years ago
- The official project website of "Ske2Grid: Skeleton-to-Grid Representation Learning for Action Recognition" (The paper of Ske2Grid is pub…☆19Sep 6, 2023Updated 2 years ago
- Learning Debiased and Disentangled Representations for Semantic Segmentation (NeurIPS 2021)☆13Jan 23, 2022Updated 4 years ago
- [AAAI 2023 (Oral)] CrissCross: Self-Supervised Audio-Visual Representation Learning with Relaxed Cross-Modal Synchronicity☆26Jul 11, 2023Updated 3 years ago
- [Survey] Masked Modeling for Self-supervised Representation Learning on Vision and Beyond (https://arxiv.org/abs/2401.00897)☆354Apr 23, 2025Updated last year
- Official implementation of ECCV 2024 paper: Take A Step Back: Rethinking the Two Stages in Visual Reasoning☆13Jun 1, 2025Updated last year
- Official code repository for SPAct: Self-supervised Privacy Preservation for Action Recognition [CVPR-2022]☆22Jun 5, 2022Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Pytorch implementation of Swin MAE https://arxiv.org/abs/2212.13805☆107Jul 7, 2025Updated last year
- A curated list of awesome temporal action segmentation resources.☆254Apr 4, 2024Updated 2 years ago
- Official Open Source code for "Masked Autoencoders As Spatiotemporal Learners"☆372Updated this week
- Official repository for "Self-Supervised Video Transformer" (CVPR'22)☆109Jun 26, 2024Updated 2 years ago
- Official implementation of "A simple, efficient and scalable contrastive masked autoencoder for learning visual representations".☆37Apr 3, 2023Updated 3 years ago
- ☆31Sep 12, 2024Updated last year
- HieraRS☆16Jul 9, 2026Updated last month
- This is official Pytorch implementation of "Decouple and Weight Semi-supervised Semantic Segmentation of Remote Sensing Images, " ISPRS, …☆11May 16, 2024Updated 2 years ago
- Video Representation Learning by Recognizing Temporal Transformations. In ECCV, 2020.☆49Mar 18, 2021Updated 5 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Foundation Models for Video Understanding: A Survey☆145Jul 9, 2025Updated last year
- [Main Conference @ EACL'26] [Workshop @ NeurIPS'24] 🎞️ LVNet.☆45Feb 10, 2026Updated 6 months ago
- Video datasets☆1,659Mar 8, 2023Updated 3 years ago
- [CVPR 2023] VideoMAE V2: Scaling Video Masked Autoencoders with Dual Masking☆814Oct 8, 2024Updated last year
- 🔥🔥🔥 [IEEE TCSVT] Latest Papers, Codes and Datasets on Vid-LLMs.☆3,268Aug 12, 2026Updated last week
- [PR 2024] TFS-ViT: Token-Level Feature Stylization for Domain Generalization☆27Mar 29, 2023Updated 3 years ago
- Self-Supervised Learning by Cross-Modal Audio-Video Clustering (NeurIPS 2020)☆91Oct 24, 2022Updated 3 years ago