Official PyTorch implementation of the ECCV 2022 paper: Efficient Video Transformers with Spatial-Temporal Token Selection.
☆52Jul 13, 2022Updated 4 years ago
Alternatives and similar repositories for STTS
Users that are interested in STTS are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [CVPR 2022] Official repository of AdaFocusV2.☆91Dec 15, 2024Updated last year
- ☆12Dec 15, 2023Updated 2 years ago
- [ICCV2023 Oral] Implicit Temporal Modeling with Learnable Alignment for Video Recognition☆41Nov 29, 2023Updated 2 years ago
- ☆70Oct 6, 2023Updated 2 years ago
- ☆14May 16, 2021Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- This repository contains the official implementation of CoMix (NeurIPS 2021) https://arxiv.org/pdf/2110.15128.pdf.☆22Jan 12, 2022Updated 4 years ago
- Official PyTorch implementation of ResFormer: Scaling ViTs with Multi-Resolution Training, CVPR2023☆30Jun 22, 2023Updated 3 years ago
- [AAAI 2022] DCAN: Improving Temporal Action Detection via Dual Context Aggregation☆17Nov 13, 2022Updated 3 years ago
- Contrastive Video Question Answering via Video Graph Transformer (IEEE T-PAMI'23)☆20Mar 9, 2024Updated 2 years ago
- Can I Trust Your Answer? Visually Grounded Video Question Answering (CVPR'24, Highlight)☆89Jul 1, 2024Updated 2 years ago
- ☆21Jan 17, 2025Updated last year
- ☆36Nov 4, 2022Updated 3 years ago
- This repository is the official Pytorch implementation of Balanced Product of Calibrated Experts for Long-Tailed Recognition (CVPR 2023).☆17Mar 13, 2025Updated last year
- Code for Enhancing Self-supervised Video Representation Learning via Multi-level Feature Optimization.☆10Sep 28, 2021Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆58Dec 2, 2025Updated 8 months ago
- AFNet(NeurIPS 2022)☆20Nov 24, 2022Updated 3 years ago
- [CVPR2022] SVIP: Sequence VerIfication for Procedures in Videos☆24Feb 24, 2023Updated 3 years ago
- Accelerating T2t-ViT by 1.6-3.6x.☆260Nov 25, 2021Updated 4 years ago
- [CVPR 2022] An Empirical Study of End-to-end Temporal Action Detection☆86Feb 19, 2023Updated 3 years ago
- ☆108Dec 23, 2022Updated 3 years ago
- Code for our paper "Eventful Transformers: Leveraging Temporal Redundancy in Vision Transformers"☆39Jan 27, 2026Updated 6 months ago
- ☆123May 12, 2022Updated 4 years ago
- The code is for the CVPR 2019 paper 'Dance with Flow: Two-in-One Stream for Action Detection '☆32Nov 21, 2022Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Risky Object Localization (ROL) in a Driving Scene Dataset☆15Dec 24, 2023Updated 2 years ago
- ☆37Dec 20, 2023Updated 2 years ago
- [IEEE T-IP 2022] TCGL: Temporal Contrastive Graph for Self-supervised Video Representation Learning☆24Dec 19, 2023Updated 2 years ago
- Hyperspectral Imagery One Class Classification (ISPRS 2022 & TGRS 2023)☆13Jan 28, 2026Updated 6 months ago
- ☆34Jul 28, 2022Updated 4 years ago
- [CVPR 2022] End-to-End Semi-Supervised Learning for Video Action Detection☆35May 3, 2023Updated 3 years ago
- [ICML 2026] VideoLoom: A Video Large Language Model for Joint Spatial-Temporal Understanding☆28Jul 3, 2026Updated last month
- Code and data for the ACM CIKM 2024 paper "Adversarial Text Rewriting for Text-aware Recommender Systems"☆12Aug 1, 2024Updated 2 years ago
- Repository for ECCV 2022 paper "Source-free Video Domain Adaptation by Learning Temporal Consistency for Action Recognition"☆24Mar 9, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [ECCV 2022] A pytorch implementation for TS2-Net: Token Shift and Selection Transformer for Text-Video Retrieval☆80Nov 29, 2022Updated 3 years ago
- Implementation of ViViT: A Video Vision Transformer☆558Jun 21, 2021Updated 5 years ago
- Reducing spatial redundancy in video recognition. SOTA computational efficiency.☆128Dec 15, 2024Updated last year
- Python package for ML developers and researchers to change certain variables while their code is executing to make the task of training a…☆11May 12, 2026Updated 2 months ago
- Official repo for CVPR 2022 (Oral) paper: Revisiting the "Video" in Video-Language Understanding. Contains code for the Atemporal Probe (…☆51May 29, 2024Updated 2 years ago
- [ICCV2023] UniFormerV2: Spatiotemporal Learning by Arming Image ViTs with Video UniFormer☆352Apr 2, 2024Updated 2 years ago
- PyTorch implementation of a collections of scalable Video Transformer Benchmarks.☆306May 4, 2022Updated 4 years ago