Code for the paper: Anticipative Feature Fusion Transformer for Multi-Modal Action Anticipation.
☆32Aug 15, 2023Updated 3 years ago
Alternatives and similar repositories for AFFT
Users that are interested in AFFT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Future Transformer for Long-term Action Anticipation (CVPR 2022)☆48Dec 22, 2022Updated 3 years ago
- Code for the Paper: Antonino Furnari and Giovanni Maria Farinella. What Would You Expect? Anticipating Egocentric Actions with Rolling-Un…☆137Aug 23, 2023Updated 3 years ago
- ☆19Sep 10, 2021Updated 4 years ago
- Code release for ICCV 2021 paper "Anticipative Video Transformer"☆154Feb 11, 2022Updated 4 years ago
- Code Release for MeMViT Memory-Augmented Multiscale Vision Transformer for Efficient Long-Term Video Recognition, CVPR 2022☆155Nov 30, 2022Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ICCV 2023] Official implementation of Memory-and-Anticipation Transformer for Online Action Understanding☆49Oct 7, 2023Updated 2 years ago
- Video + CLIP Baseline for Ego4D Long Term Action Anticipation Challenge (CVPR 2022)☆15Jul 4, 2022Updated 4 years ago
- This project uses RGB and Depth images as input into two different convolutional network of same architecture (namely VGGNet, RESNet, Ale…☆13Jul 10, 2019Updated 7 years ago
- 利用kafka+storm+mysql/redis构建日志监控系统☆13May 6, 2018Updated 8 years ago
- Official PyTorch implementation of "IntegralAction: Pose-driven Feature Integration for Robust Human Action Recognition in Videos", CVPRW…☆36Jul 10, 2024Updated 2 years ago
- Digit classification with Convolutional Neural Networks using Keras☆20May 12, 2018Updated 8 years ago
- Annotations for the public release of the EPIC-KITCHENS-100 dataset☆172Aug 1, 2022Updated 4 years ago
- ☆13Aug 24, 2018Updated 8 years ago
- ☆78Aug 16, 2021Updated 5 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- CoaT: Co-Scale Conv-Attentional Image Transformers☆15Apr 20, 2021Updated 5 years ago
- [PRCV-2023, IEEE TMM-2025] Learning Bottleneck Transformer for Event Image-Voxel Feature Fusion based Classification☆12Dec 20, 2025Updated 8 months ago
- Training for multi-modal image fusion with PyTorch.☆37Nov 30, 2023Updated 2 years ago
- The official implementation of "Multi-Glimpse Network: A Robust and Efficient Classification Architecture based on Recurrent Downsampled …☆13Nov 4, 2021Updated 4 years ago
- This is the repository to the article "NEWBEE: A Multi-Modal Gait Database of Natural Everyday-Walk in an Urban Environment", 2022☆13Aug 2, 2022Updated 4 years ago
- [ICMR'21, Best Poster Paper Award] Medical Visual Question Answering with Multi-task Pre-training and Cross-modal Self-attention☆34Dec 15, 2022Updated 3 years ago
- Unofficial PyTorch implementation of DALL-E 2 by OpenAI☆10Apr 6, 2022Updated 4 years ago
- [IEEE TMI'22] VQAMix: Conditional Triplet Mixup for Medical Visual Question Answering☆16Oct 9, 2022Updated 3 years ago
- This repository contains the official code for "Flexible Biometrics Recognition: Bridging the Multimodality Gap through Attention, Alignm…☆12Oct 9, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- GaitParsing: Human Semantic Parsing for Gait Recognition (IEEE TMM)☆13May 20, 2024Updated 2 years ago
- ☆82Jan 5, 2024Updated 2 years ago
- 2021 腾讯广告赛算法大赛 赛道二 决赛第六名☆42Oct 7, 2022Updated 3 years ago
- ☆15Mar 11, 2023Updated 3 years ago
- Code accompanying EGO-TOPO: Environment Affordances from Egocentric Video (CVPR 2020)☆31Aug 3, 2022Updated 4 years ago
- The repo for "MMPareto: Boosting Multimodal Learning with Innocent Unimodal Assistance", ICML 2024☆56Jun 28, 2024Updated 2 years ago
- Various object detection testing using YOLO and other algorithms, Raspberry pi based integration experiments.☆13Dec 9, 2024Updated last year
- ☆13Jul 6, 2022Updated 4 years ago
- This repository shows the implementation of the Trained Born Iterative Method (TBIM) applied for electromagnetic imaging.☆12Nov 9, 2022Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆13Mar 8, 2024Updated 2 years ago
- The Land-Diffuser is a novel application of the Denoising Diffusion Probabilistic Model (DDPM) in the realm of 3D Talking Head generation…☆13Dec 23, 2023Updated 2 years ago
- Official repository for "Boosting Audio Visual Question Answering via Key Semantic-Aware Cues" in ACM MM 2024.☆17Oct 25, 2024Updated last year
- PMR: Prototypical Modal Rebalance for Multimodal Learning☆47Mar 10, 2023Updated 3 years ago
- ☆14Nov 13, 2023Updated 2 years ago
- Spatial-Temporal Knowledge-Embedded Transformer for Video Scene Graph Generation (TIP 2024, ACM MM 2023)☆19Mar 13, 2024Updated 2 years ago
- Download AudioSet for Vision-Audio-Text Pre-training☆13May 16, 2022Updated 4 years ago