☆69Apr 26, 2021Updated 5 years ago
Alternatives and similar repositories for ViViT-pytorch
Users that are interested in ViViT-pytorch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Implementation of ViViT: A Video Vision Transformer - Zipping Coding Challenge☆32Jun 10, 2021Updated 5 years ago
- Implementation of ViViT: A Video Vision Transformer☆560Jun 21, 2021Updated 5 years ago
- PyTorch implementation of a collections of scalable Video Transformer Benchmarks.☆307May 4, 2022Updated 4 years ago
- ICME'19: Removing Rain in Videos: A Large-scale Database and A Two-stream ConvLSTM Approach☆11Jul 4, 2022Updated 4 years ago
- Pose driven attention mechanism☆45Mar 31, 2022Updated 4 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Unofficial PyTorch implementation of TokenLearner by Google AI☆68Jan 28, 2023Updated 3 years ago
- Source code related to the research paper entitled RVENet: A Large Echocardiographic Dataset for the Deep Learning-Based Assessment of Ri…☆13Mar 10, 2024Updated 2 years ago
- Implementation of STAM (Space Time Attention Model), a pure and simple attention model that reaches SOTA for video classification☆134Apr 1, 2021Updated 5 years ago
- Official Implementation of Visual Transformer Pooling for Lip reading☆42Aug 8, 2022Updated 4 years ago
- Scenic: A Jax Library for Computer Vision Research and Beyond☆3,842Sep 28, 2026Updated last week
- ☆20Mar 6, 2023Updated 3 years ago
- Official PyTorch implementation of "Video Prediction Recalling Long-term Motion Context via Memory Alignment Learning" (CVPR 2021 Oral)☆89Aug 9, 2021Updated 5 years ago
- [ECCV2022] The official repository of Emotion-aware Multi-view Contrastive Learning for Facial Emotion Recognition☆25Aug 21, 2023Updated 3 years ago
- Efficient dual attention SlowFast networks for video action recognition☆24Jul 10, 2022Updated 4 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- This repo holds the code for: {Transformer-based Spatio-temporal Analysis for Automatic Classification of Aortic Stenosis Severity from B…☆11Nov 29, 2022Updated 3 years ago
- collection of skeleton-based human action recognition☆10Jun 28, 2020Updated 6 years ago
- Code for Self-supervised Spatiotemporal Feature Learning by Video Geometric Transformations☆16Sep 11, 2019Updated 7 years ago
- Object-Region Video Transformers☆24Mar 24, 2022Updated 4 years ago
- Convolutional Neural Network for Full-Reference color Image Quality Assessment☆16Jun 3, 2019Updated 7 years ago
- Implementation of the paper Video Action Transformer Network☆138Apr 5, 2021Updated 5 years ago
- Implementations of Transformers for Video☆25Mar 26, 2021Updated 5 years ago
- The official pytorch code for paper "Facial Emotion Recognition with Noisy Multi-task Annotations" (2021 WACV)☆25Aug 18, 2021Updated 5 years ago
- The official pytorch implementation of our paper "Is Space-Time Attention All You Need for Video Understanding?"☆1,862Apr 9, 2024Updated 2 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- [IEEE FG 2021] Official implementation: Exploiting Emotional Dependencies with Graph Convolutional Networks for Facial Expression Recogni…☆34May 15, 2022Updated 4 years ago
- Deep Attentive Center Loss☆62Feb 4, 2025Updated last year
- code for our paper "Attention Distillation: self-supervised vision transformer students need more guidance" in BMVC 2022☆17Oct 4, 2022Updated 4 years ago
- [IEEE TIP 2022] RealSRQ-KLTSRQA☆38Jul 23, 2023Updated 3 years ago
- Implementation of "Temporal Recurrent Networks for Online Action Detection"☆23May 6, 2019Updated 7 years ago
- ☆18Jul 6, 2023Updated 3 years ago
- Action Classification using CNN and LSTM☆12Jan 17, 2019Updated 7 years ago
- This is an official implementation for "Video Swin Transformers".☆1,666Mar 8, 2023Updated 3 years ago
- Official PyTorch implementation of ACTION-Net: Multipath Excitation for Action Recognition (CVPR'21)☆209Apr 19, 2021Updated 5 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Source code of the publication MICRA-Net: MICRoscopy Analysis Neural Network to solve detection, classification, and segmentation from a …☆10Feb 19, 2024Updated 2 years ago
- [CVPR 2021] FMO Deblurring Benchmark☆15Jan 12, 2022Updated 4 years ago
- [ICCV2023] Spatio-temporal Prompting Network for Robust Video Feature Extraction☆11Aug 17, 2023Updated 3 years ago
- Implementation of frequent pattern mining using Apriori in python. It can also mine closed and max patterns from frequent itemsets.☆13Jun 6, 2018Updated 8 years ago
- [CVPR 2023] Castling-ViT: Compressing Self-Attention via Switching Towards Linear-Angular Attention During Vision Transformer Inference☆31Mar 14, 2024Updated 2 years ago
- Harmonic-NAS: Hardware-Aware Multimodal Neural Architecture Search on Resource-constrained Devices (ACML 2023)☆16May 7, 2024Updated 2 years ago
- Codebase for Linguistic Collapse: Neural Collapse in (Large) Language Models [NeurIPS 2024] [arXiv:2405.17767]☆18Apr 14, 2025Updated last year