The AVA dataset densely annotates 80 atomic visual actions in 351k movie clips with actions localized in space and time, resulting in 1.65M action labels with multiple labels per human occurring frequently.
☆350Feb 9, 2022Updated 4 years ago
Alternatives and similar repositories for ava-dataset
Users that are interested in ava-dataset are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- STEP: Spatio-Temporal Progressive Learning for Video Action Detection. CVPR'19 (Oral)☆251Oct 19, 2019Updated 6 years ago
- 国内下载google AVA dataset;download google AVA dataset in China☆70Jun 6, 2022Updated 4 years ago
- This repository is intended to host tools and demos for ActivityNet☆977Mar 21, 2024Updated 2 years ago
- Scripts for downloading the AVA (Atomic Visual Actions) dataset https://research.google.com/ava/ and do postprocessing of it.☆29May 2, 2019Updated 7 years ago
- Preprocessing tools for Google AVA Dataset☆49Apr 27, 2018Updated 8 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Code for the Active Speakers in Context Paper (CVPR2020)☆58May 19, 2021Updated 5 years ago
- ☆24Jul 23, 2026Updated 2 months ago
- Caffe: a fast open framework for deep learning.☆106Feb 27, 2018Updated 8 years ago
- Spatio-Temporal Action Localization System☆426May 21, 2022Updated 4 years ago
- GPU implementation of improved dense trajectory☆10Apr 14, 2015Updated 11 years ago
- You Only Watch Once: A Unified CNN Architecture for Real-Time Spatiotemporal Action Localization☆913Oct 28, 2024Updated last year
- An open-source toolbox for action understanding based on PyTorch☆1,874Apr 8, 2022Updated 4 years ago
- Diagnostic tools and additional visualizations from "What Actions are Needed for Understanding Human Actions in Videos?" ICCV 2017☆87Dec 19, 2017Updated 8 years ago
- [CVPR 2021] Actor-Context-Actor Relation Network for Spatio-temporal Action Localization☆215Oct 8, 2021Updated 4 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- PySlowFast: video understanding codebase from FAIR for reproducing state-of-the-art video models.☆7,423Mar 16, 2026Updated 6 months ago
- Convolutional neural network model for video classification trained on the Kinetics dataset.☆1,837Sep 12, 2019Updated 7 years ago
- This repository host the code for real-time action detection paper☆321Feb 23, 2021Updated 5 years ago
- [ICCV 2019] TSM: Temporal Shift Module for Efficient Video Understanding☆2,224Jul 11, 2024Updated 2 years ago
- HACS: Human Action Clips and Segments Dataset☆199Apr 23, 2020Updated 6 years ago
- Real-time Action detection demo for the work Actor Conditioned Attention Maps. This repo includes a complete pipeline for person detectio…☆154Dec 8, 2022Updated 3 years ago
- Code for Oops! Predicting Unintentional Action in Video☆80Apr 13, 2020Updated 6 years ago
- ☆31Jan 6, 2019Updated 7 years ago
- Accepted by TMM 2022☆20Aug 18, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- VMZ: Model Zoo for Video Modeling☆1,052Jun 17, 2025Updated last year
- Non-local Neural Networks for Video Classification☆1,987Sep 15, 2021Updated 5 years ago
- Code & Models for Temporal Segment Networks (TSN) in ECCV 2016☆1,576Oct 27, 2020Updated 5 years ago
- Temporal Relation Networks☆789May 6, 2021Updated 5 years ago
- A curated list of action recognition and related area resources☆4,033May 13, 2023Updated 3 years ago
- ☆996May 15, 2024Updated 2 years ago
- Long-Term Feature Banks for Detailed Video Understanding☆383Aug 30, 2021Updated 5 years ago
- Temporal Segment Networks (TSN) in PyTorch☆1,074Jun 21, 2019Updated 7 years ago
- ☆30Apr 10, 2018Updated 8 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Code for Temporal Relation Networks☆24Dec 30, 2017Updated 8 years ago
- ☆83Feb 20, 2021Updated 5 years ago
- Mini-Kinetics-200 data splits used in paper "Rethinking Spatiotemporal Feature Learning For Video Understanding"☆80Dec 24, 2017Updated 8 years ago
- ACM MM 2021: 'Is Someone Speaking? Exploring Long-term Temporal Features for Audio-visual Active Speaker Detection'☆505Oct 23, 2023Updated 2 years ago
- Inflated i3d network with inception backbone, weights transfered from tensorflow☆547May 23, 2024Updated 2 years ago
- Context-aware RCNN: a Baseline for Action Detection in Videos☆51Oct 13, 2020Updated 5 years ago
- PyTorch implementation of "SlowFast Networks for Video Recognition".☆350Mar 13, 2019Updated 7 years ago