[CVPR 2024] Code and models for pi-ViT, a video transformer for understanding activities of daily living
☆32Nov 12, 2025Updated 10 months ago
Alternatives and similar repositories for pi-vit
Users that are interested in pi-vit are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for the paper Seeing the Pose in the Pixels: Learning Pose-Aware Representations in Vision Transformers☆22Aug 2, 2024Updated 2 years ago
- Official Repository of "Fibottention: Inceptive Visual Representation Learning with Diverse Attention Across Heads"☆18Aug 24, 2026Updated last month
- This is the offical repository of LLAVIDAL☆25Oct 4, 2025Updated 11 months ago
- [AAAI 2025] Official Repository of 'SKI Models: Skeleton Induced Vision-Language Embeddings for Understanding Activities of Daily Living'☆26Sep 17, 2025Updated last year
- ☆50Nov 24, 2023Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- IJCAI 2024 Shap-Mix: Shapley Value Guided Mixing for Long-Tailed Skeleton Based Action Recognition☆16Nov 25, 2024Updated last year
- SI-MIL☆36Jan 3, 2025Updated last year
- Pytorch I3D implmentation on Toyota Smarthome Dataset☆20Apr 23, 2022Updated 4 years ago
- ☆15Nov 15, 2023Updated 2 years ago
- UVA-Human-Skeleton-Preprocessing☆10May 4, 2023Updated 3 years ago
- Custom Iterable Dataset Class for Large-Scale Data Loading☆14Dec 8, 2021Updated 4 years ago
- Video + CLIP Baseline for Ego4D Long Term Action Anticipation Challenge (CVPR 2022)☆15Jul 4, 2022Updated 4 years ago
- Official PyTorch implementation of "DeGCN : Deformable Graph Convolutional Networks for Skeleton-Based Action Recognition"☆52Jun 18, 2024Updated 2 years ago
- [CVPR 2026] Official Repository of 'MS-Temba: Multi-Scale Temporal Mamba for Understanding Long Untrimmed Videos'☆49Jun 22, 2026Updated 3 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- [ICCV 2023] How Much Temporal Long-Term Context is Needed for Action Segmentation?☆50Jun 21, 2024Updated 2 years ago
- [ACMMM 2023] Skeleton-MixFormer: Multivariate Topology Representation for Skeleton-based Action Recognition☆25Sep 24, 2023Updated 3 years ago
- Paper reading: Jamba — Hybrid Transformer-Mamba LM (SSM → S4 → S6 → Jamba)☆16May 22, 2024Updated 2 years ago
- We have implemented Track # 1 for ICME 2024: Spatial Action Localization on Chaotic World dataset. Our mAP on the validation set reaches …☆14Nov 11, 2024Updated last year
- Official Implementation of the paper "Unified Fully and Timestamp Supervised Temporal Action Segmentation via Sequence to Sequence Transl…☆40Nov 30, 2022Updated 3 years ago
- CNN+LSTM type "classical" visuomotor behavior cloning framework☆17Mar 3, 2025Updated last year
- The Official PyTorch implementation of "Part Aware Contrastive Learning for Self-Supervised Action Recognition" in IJCAI 2023☆13Nov 9, 2023Updated 2 years ago
- [ICCV 2023] Hierarchically Decomposed Graph Convolutional Networks for Skeleton-Based Action Recognition☆171Oct 5, 2023Updated 2 years ago
- This is a repo of extension of VPN for Recognition of Activities of Daily Living☆16May 17, 2021Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Placeholder☆10Jul 17, 2023Updated 3 years ago
- ☆21Jan 29, 2023Updated 3 years ago
- Implementation of the paper: VG4D: Vision-Language Model Goes 4D Video Recognition(ICRA 2024)☆15Apr 23, 2024Updated 2 years ago
- The official codebase of FineAction dataset. We will update the data and code of our FineAction.☆24Apr 10, 2025Updated last year
- This code is provided for reproducibility of results in the paper: Multiview Aerial Visual Recognition (MAVREC): Can Multi-view Improve A…☆24Feb 6, 2025Updated last year
- [CAC2023] Bilateral Network with Residual U-blocks and Dual-Guided Attention for Real-time Semantic Segmentation☆11Nov 28, 2024Updated last year
- [ICCV2023] Chaotic World: A Large and Challenging Benchmark for Human Behavior Understanding in Chaotic Events☆10Dec 7, 2024Updated last year
- ☆12Apr 28, 2018Updated 8 years ago
- TMI 2023: FoPro-KD: Fourier Prompted Effective Knowledge Distillation for Long-Tailed Medical Image Recognition☆12Mar 19, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Deformable Graph Convolutional Networks (Author's PyTorch implementation for the AAAI 2022 paper)☆27Sep 22, 2022Updated 4 years ago
- Tools for Toyota Smarthome datasets☆17Nov 16, 2022Updated 3 years ago
- A project about deploying a yolo server to support inferring image sent by different clients.☆10Mar 23, 2024Updated 2 years ago
- [Codes of paper]: Busy-Quiet Video Disentangling for Video Classification☆14Jan 17, 2022Updated 4 years ago
- [NeurIPS 2024] CHASE: Learning Convex Hull Adaptive Shift for Skeleton-based Multi-Entity Action Recognition & [IJCV'26] De-biasing Skele…☆16Updated this week
- [ECAI-2024] OmniCLIP: Adapting CLIP for Video Recognition with Spatial-Temporal Omni-Scale Feature Learning☆16Jan 7, 2025Updated last year
- [ICCV 2023] Latent Action Composition for Skeleton-based Action Segmentation☆22Oct 25, 2023Updated 2 years ago