This repository contains video datasets that can be used for training coarse to fine-grained (phase, step and action) temporal classification tasks.
☆16Oct 26, 2021Updated 4 years ago
Alternatives and similar repositories for video-action-recognition-datasets
Users that are interested in video-action-recognition-datasets are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Indexity is a web-based tool designed for medical video annotation in surgical data science projects.☆11Jun 27, 2023Updated 3 years ago
- ☆22Sep 19, 2025Updated 11 months ago
- ☆39Apr 5, 2025Updated last year
- Towards context-aware head-mounted display-based augmented reality for surgical guidance.☆25Jun 17, 2022Updated 4 years ago
- CholecTriplet 2022 challenge on surgical action triplet detection☆14Sep 17, 2025Updated 11 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Laparoscopic video dataset for surgical action triplet recognition☆44Sep 17, 2025Updated 11 months ago
- Finalist entry for the M2CAI Workflow Challenge 2016☆10Nov 25, 2016Updated 9 years ago
- ☆73Feb 1, 2024Updated 2 years ago
- Reading list and publicly available datasets for surgical vision☆41Dec 2, 2021Updated 4 years ago
- [CVPR2022] Bridge-Prompt: Towards Ordinal Action Understanding in Instructional Videos☆102Oct 30, 2022Updated 3 years ago
- Algorithms notes learning from ZuoShen.☆10Jun 30, 2022Updated 4 years ago
- Papers of ComputerVision x Surgery☆116Jan 7, 2024Updated 2 years ago
- Official repository for "Dissecting Self-Supervised Learning Methods for Surgical Computer Vision"☆47May 23, 2025Updated last year
- ☆24May 31, 2026Updated 3 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- How to build Text-to-Image app using stable diffusion via hugging face☆10May 28, 2023Updated 3 years ago
- Proposed splits for the LREC Wikipron paper☆15Apr 7, 2020Updated 6 years ago
- This repository contains code for our paper titled "A semi-supervised teacher-student framework for surgical tool detection and localizat…☆10Nov 16, 2023Updated 2 years ago
- Official implementation of the paper "No frame left behind: Full Video Action Recognition" (CVPR 2021)☆17Aug 29, 2021Updated 5 years ago
- The Code for M2CAI19 Paper: Hard Frame Detection and Online Mapping for Surgical Phase Recognition☆14Oct 31, 2019Updated 6 years ago
- 双目图像拼接,双目相机立体匹配,得到深度图像等。☆11Apr 27, 2020Updated 6 years ago
- ☆17Nov 28, 2024Updated last year
- training a vision transformer based model to detect violence in real life videos☆11Dec 7, 2023Updated 2 years ago
- The software used in NeurIPS 2022 Paper Don't Pour Cereal into Coffee: Differentiable Temporal Logic for Temporal Action Segmentation.☆20Aug 22, 2023Updated 3 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆16May 19, 2023Updated 3 years ago
- Official repository of the GraSP dataset and implemention of TAPIS☆60Dec 31, 2024Updated last year
- List of surgical tool datasets organised by task.☆180Aug 30, 2024Updated 2 years ago
- It can detect a Cat😺 Faces from image or video using OpenCV☆19Sep 8, 2020Updated 5 years ago
- [ECCV 2024] Official Implementation of "OphNet: A Large-Scale Video Benchmark for Ophthalmic Surgical Workflow Understanding"☆63Jul 5, 2025Updated last year
- Code for Diffusion Action Segmentation (ICCV 2023)☆81Aug 16, 2023Updated 3 years ago
- This repository contains the code associated with our 2023 TMI paper "Latent Graph Representations for Critical View of Safety Assessment…☆40Sep 17, 2025Updated 11 months ago
- [CVPR 2022] Understanding 3D Object Articulation in Internet Videos☆33Mar 7, 2024Updated 2 years ago
- CV codes from CIAM Group at SUSTech, Shenzhen, China☆11Aug 26, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Official PyTorch implementation of: "Cannot See the Forest for the Trees: Aggregating Multiple Viewpoints to Better Classify Objects in V…☆14Aug 29, 2022Updated 4 years ago
- ☆14Jul 11, 2022Updated 4 years ago
- [ECCV 2024] Official Implementation of "Appearance-Based Refinement for Object-Centric Motion Segmentation" Junyu Xie, Weidi Xie, Andrew …☆13Oct 23, 2024Updated last year
- ☆19Sep 17, 2025Updated 11 months ago
- Perform RAG (Retrieval-Augmented Generation) from your PDFs using this Colab notebook! Powered by Llama 2☆17Mar 24, 2024Updated 2 years ago
- [BMVC 2021]OMAD: Object Model with Articulated Deformations for Pose Estimation and Retrieval☆12Dec 17, 2021Updated 4 years ago
- Long Surgical Phase Recognition☆26Nov 7, 2024Updated last year