☆69Apr 26, 2021Updated 5 years ago
Alternatives and similar repositories for ViViT-pytorch
Users that are interested in ViViT-pytorch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Implementation of ViViT: A Video Vision Transformer - Zipping Coding Challenge☆33Jun 10, 2021Updated 5 years ago
- Implementation of ViViT: A Video Vision Transformer☆559Jun 21, 2021Updated 5 years ago
- PyTorch implementation of a collections of scalable Video Transformer Benchmarks.☆306May 4, 2022Updated 4 years ago
- A modular PyTorch library for vision transformer models☆165Oct 28, 2023Updated 2 years ago
- ICME'19: Removing Rain in Videos: A Large-scale Database and A Two-stream ConvLSTM Approach☆12Jul 4, 2022Updated 4 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Pose driven attention mechanism☆44Mar 31, 2022Updated 4 years ago
- Unofficial PyTorch implementation of TokenLearner by Google AI☆67Jan 28, 2023Updated 3 years ago
- Source code related to the research paper entitled RVENet: A Large Echocardiographic Dataset for the Deep Learning-Based Assessment of Ri…☆12Mar 10, 2024Updated 2 years ago
- Official Implementation of Visual Transformer Pooling for Lip reading☆41Aug 8, 2022Updated 3 years ago
- Scenic: A Jax Library for Computer Vision Research and Beyond☆3,819Jul 9, 2026Updated last week
- ☆20Mar 6, 2023Updated 3 years ago
- [ECCV2022] The official repository of Emotion-aware Multi-view Contrastive Learning for Facial Emotion Recognition☆25Aug 21, 2023Updated 2 years ago
- TCM: Temporal Correlation Module☆17Apr 24, 2021Updated 5 years ago
- Efficient dual attention SlowFast networks for video action recognition☆24Jul 10, 2022Updated 4 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Code for Self-supervised Spatiotemporal Feature Learning by Video Geometric Transformations☆16Sep 11, 2019Updated 6 years ago
- Object-Region Video Transformers☆24Mar 24, 2022Updated 4 years ago
- PySlowFast: video understanding codebase from FAIR for reproducing state-of-the-art video models.☆87Sep 13, 2021Updated 4 years ago
- Emonet unofficial Implemented "Estimation of continuous valence and arousal levels from faces in naturalistic conditions" published in Na…☆21Dec 20, 2022Updated 3 years ago
- SGD/ADAM/Amsgrad/AdamW/RAdam/Lookahead☆10Nov 18, 2019Updated 6 years ago
- Convolutional Neural Network for Full-Reference color Image Quality Assessment☆16Jun 3, 2019Updated 7 years ago
- Source codes for "Unsupervised Curriculum Domain Adaptation for No-Reference Video Quality Assessment"☆20Dec 19, 2021Updated 4 years ago
- Implementation of the paper Video Action Transformer Network☆138Apr 5, 2021Updated 5 years ago
- Spatial Attention-based Non-reference Perceptual Quality Prediction Network for Omnidirectional Images (IEEE ICME'2021))☆20Jan 27, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Implementations of Transformers for Video☆24Mar 26, 2021Updated 5 years ago
- ☆13Jul 31, 2019Updated 6 years ago
- [IEEE FG 2021] Official implementation: Exploiting Emotional Dependencies with Graph Convolutional Networks for Facial Expression Recogni…☆34May 15, 2022Updated 4 years ago
- Deep Attentive Center Loss☆62Feb 4, 2025Updated last year
- code for our paper "Attention Distillation: self-supervised vision transformer students need more guidance" in BMVC 2022☆17Oct 4, 2022Updated 3 years ago
- Machine Translation Metrics Unit TesTing☆13Jun 4, 2016Updated 10 years ago
- ☆18Jul 6, 2023Updated 3 years ago
- Action Classification using CNN and LSTM☆12Jan 17, 2019Updated 7 years ago
- This is an official implementation for "Video Swin Transformers".☆1,667Mar 8, 2023Updated 3 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Official PyTorch implementation of ACTION-Net: Multipath Excitation for Action Recognition (CVPR'21)☆209Apr 19, 2021Updated 5 years ago
- Source code of the publication MICRA-Net: MICRoscopy Analysis Neural Network to solve detection, classification, and segmentation from a …☆10Feb 19, 2024Updated 2 years ago
- [CVPR 2021] FMO Deblurring Benchmark☆15Jan 12, 2022Updated 4 years ago
- Lazy python recipes.☆10Apr 17, 2026Updated 3 months ago
- [ICCV2023] Spatio-temporal Prompting Network for Robust Video Feature Extraction☆10Aug 17, 2023Updated 2 years ago
- Harmonic-NAS: Hardware-Aware Multimodal Neural Architecture Search on Resource-constrained Devices (ACML 2023)☆16May 7, 2024Updated 2 years ago
- Codebase for Linguistic Collapse: Neural Collapse in (Large) Language Models [NeurIPS 2024] [arXiv:2405.17767]☆19Apr 14, 2025Updated last year