Video feature extraction pipeline that supports diverse models including I3D, SlowFast, EgoVLP, and CLIP.
☆13Apr 20, 2024Updated 2 years ago
Alternatives and similar repositories for video_features
Users that are interested in video_features are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This repo holds the implementation of PAVE: Patching and Adapting Video Large Language Models (CVPR2025)☆28Sep 6, 2025Updated last year
- Official Repository of RefChartQA: Grounding Visual Answer on Chart Images through Instruction Tuning☆15Jul 9, 2025Updated last year
- [CVPR 2024] Code and datasets for 'Learning Spatial Features from Audio-Visual Correspondence in Egocentric Videos'☆14Jun 16, 2024Updated 2 years ago
- Official Implementation for "ESCAPE: Encoding Super-keypoints for Category-Agnostic Pose Estimation", CVPR 2024.☆10Jun 17, 2024Updated 2 years ago
- Extract video features. Currently, the models includes I3D, will be continuously updated.☆12Jun 4, 2020Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- official codes for FRD-UVAD(10 crop version)☆16Nov 2, 2024Updated last year
- ☆13Jul 6, 2022Updated 4 years ago
- Video Anomaly Detection with Motion and Appearance Guided Patch Diffusion Model☆17Apr 14, 2025Updated last year
- official repository for the Instance Prototype Contrastive Learning (IPCL)☆18Jun 20, 2022Updated 4 years ago
- Remove the "block" element on Y-Tub☆16Aug 23, 2024Updated 2 years ago
- Support library for the MaskRCNN masks extracted on EPIC-KITCHENS-100☆14Dec 1, 2020Updated 5 years ago
- Visual Concept Connectome☆15Jun 23, 2024Updated 2 years ago
- Official code implementation for the paper "Do Vision & Language Decoders use Images and Text equally? How Self-consistent are their Expl…☆12Jul 14, 2026Updated 2 months ago
- This project hosts the code for implementing the ISAL algorithm for object detection and image classification (ICCV2021)☆29Nov 19, 2021Updated 4 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Dataset for Indoor Fire Detection☆16May 9, 2025Updated last year
- Extract video features from raw videos using multiple GPUs. We support RAFT flow frames as well as S3D, I3D, R(2+1)D, VGGish, CLIP, and T…☆658Feb 1, 2026Updated 7 months ago
- Replace the MS-TCN with ASFormer in asrf☆23Oct 28, 2021Updated 4 years ago
- Overlooked Video Classification in Video Anomaly Detection☆19Oct 17, 2022Updated 3 years ago
- ☆24Mar 24, 2023Updated 3 years ago
- A charuco based calibrator for multi-camera setups (intrinsic and extrinsic coordinates).☆25Aug 6, 2026Updated last month
- ☆31Mar 1, 2024Updated 2 years ago
- [CVPR 2024] Official code for paper: Prompt-Enhanced Multiple Instance Learning for Weakly Supervised Video Anomaly Detection.☆27Aug 19, 2024Updated 2 years ago
- ☆19Sep 10, 2021Updated 5 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Simple PyTorch Dataset for the EPIC-Kitchens-55 and EPIC-Kitchens-100 that handles frames and features (rgb, optical flow, and objects) f…☆24Jan 22, 2023Updated 3 years ago
- Official code for "Federated Weakly Supervised Video Anomaly Detection with Multimodal Prompt" (AAAI2025)☆27May 27, 2025Updated last year
- Dense Optical Flow extraction rewritten from https://github.com/wanglimin/dense_flow☆31Feb 16, 2019Updated 7 years ago
- In this study, we propose a deep-learning-based method to image through dynamic scattering media in a non-invasive manner under incoheren…☆28Sep 11, 2024Updated 2 years ago
- Code release for ActionFormer (ECCV 2022)☆576Apr 11, 2024Updated 2 years ago
- PyTorch implementation of "Wasserstein Iterative Networks for Barycenter Estimation" (NeurIPS 2022)☆20Jul 3, 2023Updated 3 years ago
- Implementation of "Temporal Recurrent Networks for Online Action Detection"☆23May 6, 2019Updated 7 years ago
- 👾 E.T. Bench: Towards Open-Ended Event-Level Video-Language Understanding (NeurIPS 2024)☆74Jan 20, 2025Updated last year
- TF 2 implementation Learning to Resize Images for Computer Vision Tasks (https://arxiv.org/abs/2103.09950v1).☆54Oct 12, 2021Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆57Jul 21, 2024Updated 2 years ago
- audio-extract is a Python library that allows you to extract audio from video files and trim the audio according to your needs.☆30Jul 28, 2025Updated last year
- What and How Well You Performed? A Multitask Learning Approach to Action Quality Assessment [CVPR 2019]☆77May 5, 2025Updated last year
- EPIC-Kitchens-100 Action Recognition baselines: TSN, TRN, TSM☆33Mar 15, 2022Updated 4 years ago
- [ICCV 2025] Official code for "AIM: Adaptive Inference of Multi-Modal LLMs via Token Merging and Pruning"☆65Oct 9, 2025Updated 11 months ago
- ☆37Jan 30, 2026Updated 7 months ago
- [CVPR 2023] Official code for "Learning Procedure-aware Video Representation from Instructional Videos and Their Narrations"☆56Aug 8, 2023Updated 3 years ago