facebookresearch/ActivityNet-Entities

Readme badge preview -

If you own this repo, copy the snippet below and add it to your README.md

[![RelatedRepos](https://img.shields.io/badge/related-repos-yellow)](https://relatedrepos.com/gh/facebookresearch/ActivityNet-Entities)

facebookresearch / ActivityNet-Entities

A Dataset for Grounded Video Description

☆165

Alternatives and similar repositories for ActivityNet-Entities

Users that are interested in ActivityNet-Entities are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.

Sorting:

facebookresearch / grounded-video-description
View on GitHub
Video Grounding and Captioning
☆332Oct 12, 2021Updated 4 years ago
zfchenUnique / WSSTG
View on GitHub
This repository contains the main baselines introduced in WSSTG (ACL 2019).
☆57Jul 8, 2024Updated 2 years ago
TheShadow29 / vognet-pytorch
View on GitHub
[CVPR20] Video Object Grounding using Semantic Roles in Language Description (https://arxiv.org/abs/2003.10606)
☆69Jun 10, 2020Updated 6 years ago
LuoweiZhou / detectron-vlp
View on GitHub
Detectron for image/video region feature extraction, inspired by Xinlei's repo
☆22Nov 21, 2020Updated 5 years ago
jayleicn / TVQAplus
View on GitHub
[ACL 2020] PyTorch code for TVQA+: Spatio-Temporal Grounding for Video Question Answering
☆132Oct 25, 2022Updated 3 years ago
1-Click AI Models by DigitalOcean Gradient • Ad
Deploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
ranjaykrishna / densevid_eval
View on GitHub
Evaluation code for Dense-Captioning Events in Videos
☆130Jun 11, 2019Updated 7 years ago
chihyaoma / cyclical-visual-captioning
View on GitHub
PyTorch code for: Learning to Generate Grounded Visual Captions without Localization Supervision
☆46Jul 29, 2020Updated 5 years ago
Guaranteer / VidSTG-Dataset
View on GitHub
This repository provides the dataset introduced by the paper "Where Does It Exist: Spatio-Temporal Video Grounding for Multi-Form Sentenc…
☆70May 1, 2020Updated 6 years ago
salesforce / densecap
View on GitHub
☆191Jun 16, 2025Updated last year
XgDuan / WSDEC
View on GitHub
Weakly Supervised Dense Event Captioning in Videos, i.e. generating multiple sentence descriptions for a video in a weakly-supervised man…
☆104Mar 21, 2020Updated 6 years ago
MichiganCOG / Video-Grounding-from-Text
View on GitHub
Source code for "Weakly-Supervised Video Object Grounding from Text by Loss Weighting and Object Interaction"
☆47Jun 22, 2024Updated 2 years ago
JaywongWang / CBP
View on GitHub
Official Tensorflow Implementation of the AAAI-2020 paper "Temporally Grounding Language Queries in Videos by Contextual Boundary-aware P…
☆59Mar 24, 2023Updated 3 years ago
LuoweiZhou / anet2016-cuhk-feature
View on GitHub
Feature Extraction Toolbox from CUHK&ETHZ&SIAT submission to ActivityNet 2016
☆32Mar 31, 2019Updated 7 years ago
JaywongWang / DenseVideoCaptioning
View on GitHub
Official Tensorflow Implementation of the paper "Bidirectional Attentive Fusion with Context Gating for Dense Video Captioning" in CVPR 2…
☆151Jul 8, 2019Updated 7 years ago
Managed Kubernetes at scale on DigitalOcean • Ad
DigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
vsislab / Controllable_XGating
View on GitHub
ICCV2019: Controllable Video Captioning with POS Sequence Guidance Based on Gated Fusion Network
☆68Nov 19, 2019Updated 6 years ago
LuoweiZhou / densecap
View on GitHub
Dense video captioning in PyTorch
☆41Aug 30, 2019Updated 6 years ago
zfchenUnique / VID-Sentence
View on GitHub
This repository provides the dataset introduced by our WSSTG paper
☆13Jul 21, 2019Updated 7 years ago
simon-ging / coot-videotext
View on GitHub
COOT: Cooperative Hierarchical Transformer for Video-Text Representation Learning
☆291Sep 6, 2022Updated 3 years ago
YuanEZhou / Grounded-Image-Captioning
View on GitHub
☆64Jan 5, 2022Updated 4 years ago
shikorab / SceneGraph
View on GitHub
Scene Graphs with Permutation-Invariant Structured Prediction
☆72Nov 21, 2022Updated 3 years ago
waybarrios / Anet_tools2.0
View on GitHub
Download Activity Net Videos
☆10Nov 3, 2015Updated 10 years ago
yaohungt / Gated-Spatio-Temporal-Energy-Graph
View on GitHub
[CVPR'19] [PyTorch] Gated Spatio Temporal Energy Graph
☆153Feb 20, 2020Updated 6 years ago
Sundrops / video-caption.pytorch
View on GitHub
☆33Apr 20, 2018Updated 8 years ago
Deploy to Railway using AI coding agents - Free Credits Offer • Ad
Use Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
facebookresearch / video-long-term-feature-banks
View on GitHub
Long-Term Feature Banks for Detailed Video Understanding
☆383Aug 30, 2021Updated 4 years ago
cshizhe / hgr_v2t
View on GitHub
Code accompanying the paper "Fine-grained Video-Text Retrieval with Hierarchical Graph Reasoning".
☆211Jun 12, 2020Updated 6 years ago
ikuinen / CMIN_moment_retrieval
View on GitHub
Cross-Modal Interaction Networks for Query-Based Moment Retrieval in Videos
☆87Nov 22, 2020Updated 5 years ago
activitynet / ActivityNet
View on GitHub
This repository is intended to host tools and demos for ActivityNet
☆976Mar 21, 2024Updated 2 years ago
doc-doc / vRGV
View on GitHub
Visual Relation Grounding in Videos (ECCV'20, Spotlight)
☆57Dec 8, 2022Updated 3 years ago
fanchenyou / HME-VideoQA
View on GitHub
Heterogeneous Memory Enhanced Multimodal Attention Model for VideoQA
☆55Sep 13, 2021Updated 4 years ago
jayleicn / TVCaption
View on GitHub
[ECCV 2020] PyTorch code of MMT (a multimodal transformer captioning model) on TVCaption dataset
☆91Sep 6, 2023Updated 2 years ago
ttengwang / dense-video-captioning-pytorch
View on GitHub
Second-place solution to dense video captioning task in ActivityNet Challenge (CVPR 2020 workshop)
☆75Aug 25, 2021Updated 4 years ago
jayleicn / recurrent-transformer
View on GitHub
[ACL 2020] PyTorch code for MART: Memory-Augmented Recurrent Transformer for Coherent Video Paragraph Captioning
☆170Dec 4, 2020Updated 5 years ago
GPUs on demand by Runpod - Special Offer Available • Ad
Run AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
facebookresearch / binary-image-selection
View on GitHub
BISON: Binary Image SelectiON
☆50Sep 15, 2021Updated 4 years ago
hyounghk / VideoQADenseCapFrameGate-ACL2020
View on GitHub
Code for ACL 2020 paper "Dense-Caption Matching and Frame-Selection Gating for Temporal Localization in VideoQA." Hyounghun Kim, Zineng T…
☆34May 14, 2020Updated 6 years ago
TheShadow29 / awesome-grounding
View on GitHub
awesome grounding: A curated list of research papers in visual grounding
☆1,127Sep 21, 2025Updated 10 months ago
niluthpol / multimodal_vtt
View on GitHub
Joint Embedding with Multimodal Cues for Cross-Modal Video-Text Retrieval
☆68Apr 10, 2020Updated 6 years ago
tgc1997 / RMN
View on GitHub
IJCAI2020: Learning to Discretely Compose Reasoning Module Networks for Video Captioning
☆79Nov 23, 2020Updated 5 years ago
wzmsltw / BSN-boundary-sensitive-network
View on GitHub
Codes of our paper: "BSN: Boundary Sensitive Network for Temporal Action Proposal Generation"
☆411Mar 1, 2019Updated 7 years ago
linjieli222 / HERO
View on GitHub
Research code for EMNLP 2020 paper "HERO: Hierarchical Encoder for Video+Language Omni-representation Pre-training"
☆235Sep 16, 2021Updated 4 years ago