视频的文本摘要(标注),输入一段视频,通过深度学习网络和人工智能程序识别视频主要表达的意思(Input a video output a txt decribing the video)。
☆189Mar 20, 2018Updated 8 years ago
Alternatives and similar repositories for VideoCaption
Users that are interested in VideoCaption are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- pytorch implementation of video captioning☆400Aug 19, 2019Updated 6 years ago
- This repository contains the code for a video captioning system inspired by Sequence to Sequence -- Video to Text. This system takes as i…☆170Oct 12, 2019Updated 6 years ago
- Official Tensorflow Implementation of the paper "Bidirectional Attentive Fusion with Context Gating for Dense Video Captioning" in CVPR 2…☆152Jul 8, 2019Updated 7 years ago
- Video Grounding and Captioning☆331Oct 12, 2021Updated 4 years ago
- Video to Text: Natural language description generator for some given video. [Video Captioning]☆363May 3, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Pytorch implementation of audio-visual fusion video captioning model☆27Jul 26, 2018Updated 8 years ago
- 🎬 Video Captioning: ICCV '15 paper implementation☆47May 30, 2018Updated 8 years ago
- A video captioning tool using S2VT method and attention mechanism (TensorFlow)☆15Oct 14, 2018Updated 7 years ago
- video synopsis, video enrichment, 视频浓缩,视频摘要☆37Aug 20, 2020Updated 5 years ago
- Implemention of Baidu's DenseBox used for multi-task learning of object detection and landmark(key-point) localization 用PyTorch实现了百度的Den…☆94Oct 10, 2020Updated 5 years ago
- Image Caption workout with NIC and NBT☆16Apr 5, 2019Updated 7 years ago
- Using Semantic Compositional Networks for Video Captioning☆96Nov 27, 2018Updated 7 years ago
- Video Captioning is an encoder decoder mode based on sequence to sequence learning☆138Apr 9, 2024Updated 2 years ago
- Face recognition using triplet loss, implementing FaceNet with pytorch.人脸识别项目,提供一个小型数据集用作验证,使用三元组损失函数(Triplet loss)提升准确率和泛化能力,对FaceNet进行了…☆133Apr 26, 2018Updated 8 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Implementing RepNet(a two-stream multitask learning network) to do vehicle Re-identification, vehicle search(or vehicle match) with PyTo…☆252Oct 10, 2020Updated 5 years ago
- Soft attention mechanism for video caption generation☆154Jul 17, 2017Updated 9 years ago
- [ACM MM 2017 & IEEE TMM 2020] This is the Theano code for the paper "Video Description with Spatial Temporal Attention"☆61Oct 20, 2020Updated 5 years ago
- Word2VisualVec : Predicting Visual Features from Text for Image and Video Caption Retrieval☆70Jan 27, 2020Updated 6 years ago
- A curated list of research papers in Video Captioning☆122Jan 5, 2021Updated 5 years ago
- Joint Embedding with Multimodal Cues for Cross-Modal Video-Text Retrieval☆68Apr 10, 2020Updated 6 years ago
- mumu-spark是一个学习项目,主要通过这个项目来了解和学习spark的基本使用方式和工作原理。mumu-spark主要包括弹性数据集rdd、spark sql、机器学习语言mlib、实时工作流streaming、图形数据库graphx。通过这些模块的学习,初步掌握sp…☆14Sep 8, 2022Updated 3 years ago
- Optimized code based on M2 for faster image captioning training☆21Nov 18, 2022Updated 3 years ago
- Video classification using convGRU☆13Feb 15, 2018Updated 8 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- A fork of FairMOT used to do vehicle MOT.用于跟踪车辆的多目标跟踪, 自定义数据进行单类别多目标实时跟踪☆204Oct 3, 2023Updated 2 years ago
- ☆21Nov 6, 2020Updated 5 years ago
- 图像中文描述+视觉注意力☆192Jan 9, 2020Updated 6 years ago
- Experimenting with different Summarizing techniques on SumMe Dataset☆143Jul 7, 2020Updated 6 years ago
- 中文文本摘要生成模型☆21Jul 29, 2022Updated 4 years ago
- Detectron for image/video region feature extraction, inspired by Xinlei's repo☆22Nov 21, 2020Updated 5 years ago
- ☆13Dec 25, 2018Updated 7 years ago
- Papers, codes collection of video summarization / video highlight detection / video key frame selection☆37Jul 16, 2021Updated 5 years ago
- 这是一个利用Spring Cloud,Dubbo,Thrift三个微服务框架整合开发的IM社交系统,并用到了Netty即时通讯技术,Tensorflow深度学习框架与Haar+Adaboost人脸识别技术,每个模块都可以被完整的被拿来直接使用,适合对微服务,即时通信感兴趣的…☆11Nov 16, 2022Updated 3 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Implement en-fr translation task by implenting seq2seq, encoder-decoder in RNN layers with Attention mechanism and Beamsearch inference d…☆20Feb 14, 2018Updated 8 years ago
- Python implementation of extraction of several visual features representations from videos☆23Jul 19, 2021Updated 5 years ago
- 目标检测,关键点检测。A pure version of CenterNet, convenient for secondary development and easy to understand.☆21Dec 9, 2020Updated 5 years ago
- A curated list of the Video Summarization subject which is a computer science using machine learning and deep learning☆42May 29, 2020Updated 6 years ago
- Detect the face in each key frame which extracts from the movie☆25Feb 26, 2021Updated 5 years ago
- End-to-End Dense Video Captioning with Parallel Decoding (ICCV 2021)☆230Jan 3, 2024Updated 2 years ago
- Code for Unsupervised Image Captioning☆223Mar 24, 2023Updated 3 years ago