☆69Apr 26, 2021Updated 5 years ago
Alternatives and similar repositories for ViViT-pytorch
Users that are interested in ViViT-pytorch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Implementation of ViViT: A Video Vision Transformer - Zipping Coding Challenge☆32Jun 10, 2021Updated 5 years ago
- Implementation of ViViT: A Video Vision Transformer☆559Jun 21, 2021Updated 5 years ago
- Official code for CVPR2024 “VideoMAC: Video Masked Autoencoders Meet ConvNets”☆16May 12, 2026Updated 4 months ago
- PyTorch implementation of a collections of scalable Video Transformer Benchmarks.☆306May 4, 2022Updated 4 years ago
- A modular PyTorch library for vision transformer models☆164Oct 28, 2023Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ICME'19: Removing Rain in Videos: A Large-scale Database and A Two-stream ConvLSTM Approach☆11Jul 4, 2022Updated 4 years ago
- Unofficial PyTorch implementation of TokenLearner by Google AI☆68Jan 28, 2023Updated 3 years ago
- Official Implementation of Visual Transformer Pooling for Lip reading☆42Aug 8, 2022Updated 4 years ago
- ☆20Mar 6, 2023Updated 3 years ago
- Official PyTorch implementation of "Video Prediction Recalling Long-term Motion Context via Memory Alignment Learning" (CVPR 2021 Oral)☆89Aug 9, 2021Updated 5 years ago
- [ECCV2022] The official repository of Emotion-aware Multi-view Contrastive Learning for Facial Emotion Recognition☆25Aug 21, 2023Updated 3 years ago
- TCM: Temporal Correlation Module☆17Apr 24, 2021Updated 5 years ago
- Code for Self-supervised Spatiotemporal Feature Learning by Video Geometric Transformations☆16Sep 11, 2019Updated 7 years ago
- Object-Region Video Transformers☆24Mar 24, 2022Updated 4 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- PySlowFast: video understanding codebase from FAIR for reproducing state-of-the-art video models.☆87Sep 13, 2021Updated 5 years ago
- Emonet unofficial Implemented "Estimation of continuous valence and arousal levels from faces in naturalistic conditions" published in Na…☆21Dec 20, 2022Updated 3 years ago
- SGD/ADAM/Amsgrad/AdamW/RAdam/Lookahead☆10Nov 18, 2019Updated 6 years ago
- Video classification tools using 3D ResNet☆25Sep 25, 2017Updated 8 years ago
- Source codes for "Unsupervised Curriculum Domain Adaptation for No-Reference Video Quality Assessment"☆20Dec 19, 2021Updated 4 years ago
- Implementation of the paper Video Action Transformer Network☆138Apr 5, 2021Updated 5 years ago
- The official pytorch code for paper "Facial Emotion Recognition with Noisy Multi-task Annotations" (2021 WACV)☆25Aug 18, 2021Updated 5 years ago
- Corpus and code for Aligned Recipe Actions (ARA) corpus, EMNLP 2021☆10May 22, 2024Updated 2 years ago
- [IEEE FG 2021] Official implementation: Exploiting Emotional Dependencies with Graph Convolutional Networks for Facial Expression Recogni…☆34May 15, 2022Updated 4 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [IEEE TIP 2022] RealSRQ-KLTSRQA☆38Jul 23, 2023Updated 3 years ago
- Implementation of "Temporal Recurrent Networks for Online Action Detection"☆23May 6, 2019Updated 7 years ago
- This is an official implementation for "Video Swin Transformers".☆1,665Mar 8, 2023Updated 3 years ago
- Official PyTorch implementation of ACTION-Net: Multipath Excitation for Action Recognition (CVPR'21)☆208Apr 19, 2021Updated 5 years ago
- PyTorch implementation of "Multi-modality Associative Bridging through Memory: Speech Sound Recollected from Face Video" (ICCV2021)☆22Apr 11, 2022Updated 4 years ago
- [ICCV2023] Spatio-temporal Prompting Network for Robust Video Feature Extraction☆11Aug 17, 2023Updated 3 years ago
- Implementation of frequent pattern mining using Apriori in python. It can also mine closed and max patterns from frequent itemsets.☆13Jun 6, 2018Updated 8 years ago
- [CVPR 2023] Castling-ViT: Compressing Self-Attention via Switching Towards Linear-Angular Attention During Vision Transformer Inference☆31Mar 14, 2024Updated 2 years ago
- Harmonic-NAS: Hardware-Aware Multimodal Neural Architecture Search on Resource-constrained Devices (ACML 2023)☆16May 7, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Transformer for Action Recognition in PyTorch☆39Mar 14, 2020Updated 6 years ago
- A Spatial–Temporal Video Quality Assessment Method via Comprehensive HVS Simulation☆17Jan 13, 2024Updated 2 years ago
- The official PyTorch implementation of "Context Matters: Self-Attention for sign Language Recognition"☆24Oct 17, 2021Updated 4 years ago
- This is my attempt at the ActivityNet Challenge 2017. Thanks to the organizers for providing the boilerplate code and annotated datasets.…☆10Jul 19, 2017Updated 9 years ago
- ☆29Jul 1, 2021Updated 5 years ago
- ☆28Jun 6, 2023Updated 3 years ago
- Video Transformer Network☆41Jun 8, 2021Updated 5 years ago