Official PyTorch implementation of the ECCV 2022 paper: Efficient Video Transformers with Spatial-Temporal Token Selection.
☆52Jul 13, 2022Updated 4 years ago
Alternatives and similar repositories for STTS
Users that are interested in STTS are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [CVPR 2022] Official repository of AdaFocusV2.☆91Dec 15, 2024Updated last year
- [ICCV2023 Oral] Implicit Temporal Modeling with Learnable Alignment for Video Recognition☆41Nov 29, 2023Updated 2 years ago
- ☆70Oct 6, 2023Updated 2 years ago
- ☆14May 16, 2021Updated 5 years ago
- This repository contains the official implementation of CoMix (NeurIPS 2021) https://arxiv.org/pdf/2110.15128.pdf.☆22Jan 12, 2022Updated 4 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Official PyTorch implementation of ResFormer: Scaling ViTs with Multi-Resolution Training, CVPR2023☆30Jun 22, 2023Updated 3 years ago
- [AAAI 2022] DCAN: Improving Temporal Action Detection via Dual Context Aggregation☆17Nov 13, 2022Updated 3 years ago
- Contrastive Video Question Answering via Video Graph Transformer (IEEE T-PAMI'23)☆20Mar 9, 2024Updated 2 years ago
- Can I Trust Your Answer? Visually Grounded Video Question Answering (CVPR'24, Highlight)☆89Jul 1, 2024Updated 2 years ago
- ☆21Jan 17, 2025Updated last year
- ☆36Nov 4, 2022Updated 3 years ago
- Learning An Effective Transformer for Remote Sensing Satellite Image Dehazing☆12Sep 25, 2023Updated 2 years ago
- Code for Enhancing Self-supervised Video Representation Learning via Multi-level Feature Optimization.☆10Sep 28, 2021Updated 4 years ago
- ☆25Dec 26, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Risky Object Localization (ROL) in a Driving Scene Dataset☆15Dec 24, 2023Updated 2 years ago
- ☆58Dec 2, 2025Updated 7 months ago
- [ECCV 22] LocVTP: Video-Text Pre-training for Temporal Localization☆39Jul 29, 2022Updated 3 years ago
- PyDeepFakeDet is an integrated and scalable tool for Deepfake detection.☆114Nov 6, 2022Updated 3 years ago
- Official Pytorch implementation of MuCo: Multi-turn Contrastive Learning for Multimodal Embedding Model (CVPR 2026)☆15Apr 16, 2026Updated 3 months ago
- AFNet(NeurIPS 2022)☆20Nov 24, 2022Updated 3 years ago
- [CVPR2022] SVIP: Sequence VerIfication for Procedures in Videos☆24Feb 24, 2023Updated 3 years ago
- Accelerating T2t-ViT by 1.6-3.6x.☆260Nov 25, 2021Updated 4 years ago
- [CVPR 2022] An Empirical Study of End-to-end Temporal Action Detection☆87Feb 19, 2023Updated 3 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- ☆109Dec 23, 2022Updated 3 years ago
- Code for our paper "Eventful Transformers: Leveraging Temporal Redundancy in Vision Transformers"☆39Jan 27, 2026Updated 5 months ago
- The code is for the CVPR 2019 paper 'Dance with Flow: Two-in-One Stream for Action Detection '☆32Nov 21, 2022Updated 3 years ago
- ☆37Dec 20, 2023Updated 2 years ago
- [ICLR2021] AdaFuse: Adaptive Temporal Fusion Network for Efficient Action Recognition☆35Apr 8, 2021Updated 5 years ago
- [IEEE T-IP 2022] TCGL: Temporal Contrastive Graph for Self-supervised Video Representation Learning☆24Dec 19, 2023Updated 2 years ago
- ☆33Jul 28, 2022Updated 3 years ago
- [ICML 2026] VideoLoom: A Video Large Language Model for Joint Spatial-Temporal Understanding☆27Jul 3, 2026Updated 2 weeks ago
- Deep Multi-layer Fusion Dense Network for Hyperspectral Image Classification.☆11Apr 25, 2021Updated 5 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Repository for ECCV 2022 paper "Source-free Video Domain Adaptation by Learning Temporal Consistency for Action Recognition"☆24Mar 9, 2023Updated 3 years ago
- [ECCV 2022] A pytorch implementation for TS2-Net: Token Shift and Selection Transformer for Text-Video Retrieval☆80Nov 29, 2022Updated 3 years ago
- ☆184Aug 20, 2022Updated 3 years ago
- Implementation of ViViT: A Video Vision Transformer☆559Jun 21, 2021Updated 5 years ago
- Reducing spatial redundancy in video recognition. SOTA computational efficiency.☆128Dec 15, 2024Updated last year
- Official repo for CVPR 2022 (Oral) paper: Revisiting the "Video" in Video-Language Understanding. Contains code for the Atemporal Probe (…☆51May 29, 2024Updated 2 years ago
- [ICCV2023] UniFormerV2: Spatiotemporal Learning by Arming Image ViTs with Video UniFormer☆350Apr 2, 2024Updated 2 years ago