Implementation of "GL-RG: Global-Local Representation Granularity for Video Captioning".
☆27Dec 16, 2021Updated 4 years ago
Alternatives and similar repositories for GL-RG
Users that are interested in GL-RG are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆60Mar 30, 2022Updated 4 years ago
- ☆15Sep 6, 2021Updated 4 years ago
- ☆26Oct 20, 2021Updated 4 years ago
- Official pytorch implementation of the AAAI 2021 paper "Semantic Grouping Network for Video Captioning"☆54Jul 9, 2021Updated 5 years ago
- ☆14Mar 7, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code for the paper "Controllable Video Captioning with an Exemplar Sentence"☆12Apr 14, 2021Updated 5 years ago
- pytorch implementation of Semantics-AssistedVideoCaptioning☆11Feb 16, 2023Updated 3 years ago
- [ECCV 2022] Multimodal Transformer with Variable-length Memory for Vision-and-Language Navigation☆19Jul 18, 2022Updated 4 years ago
- JAX tutorials for PyTorch users☆14Feb 18, 2023Updated 3 years ago
- Houses deployable code for the SCORCH scoring function and docking pipeline from the related publication: https://doi.org/10.1016/j.jare.…☆18Nov 29, 2022Updated 3 years ago
- ☆35Mar 22, 2019Updated 7 years ago
- ☆18May 15, 2026Updated 3 months ago
- [NeurIPS 2022 Spotlight] Learning Equivariant Segmentation with Instance-Unique Querying☆22Dec 17, 2022Updated 3 years ago
- The PyTorch code of the AAAI2021 paper "Non-Autoregressive Coarse-to-Fine Video Captioning".☆57Oct 22, 2023Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆23Aug 21, 2021Updated 5 years ago
- Official code of *Towards Event-oriented Long Video Understanding*☆12Jul 26, 2024Updated 2 years ago
- Reimplemention of "Mask-Guided Attention Network for Occluded Pedestrian Detection" based on mmdetection toolbox☆10Aug 20, 2020Updated 6 years ago
- 短信验证码模块☆10Jul 25, 2021Updated 5 years ago
- ☆62May 11, 2021Updated 5 years ago
- BINANA (BINding ANAlyzer) analyzes the geometries of predicted ligand poses to identify molecular interactions that contribute to binding…☆25Jun 30, 2025Updated last year
- S2VT pytorch implementation☆20Jun 28, 2019Updated 7 years ago
- Video to Language Challenge (MSR-VTT Challenge 2016)☆32Dec 28, 2017Updated 8 years ago
- ☆18Sep 8, 2021Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆12Mar 8, 2021Updated 5 years ago
- [ACM MM 2021 Oral] Exploiting BERT For Multimodal Target Sentiment Classification Through Input Space Translation"☆39Aug 8, 2021Updated 5 years ago
- Data release for Step Differences in Instructional Video (CVPR24)☆15Jun 19, 2024Updated 2 years ago
- A PyTorch implementation of Uni-Mol3.☆24Mar 24, 2026Updated 5 months ago
- A curated list of Multimodal Captioning related research(including image captioning, video captioning, and text captioning)☆114Jun 6, 2022Updated 4 years ago
- ☆12Dec 19, 2016Updated 9 years ago
- Integration of Clinical Embeddings with Neural ODEs☆11Jan 6, 2025Updated last year
- The code of 《HAM: Hidden Anchor Mechanism for Scene Text Detection》☆11Sep 22, 2020Updated 5 years ago
- ☆10Jan 9, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A Unified GPU-Accelerated Platform for High-Throughput Structure-Based, Ligand-Based, and Synergistic Hybrid Virtual Screening☆31Feb 24, 2026Updated 6 months ago
- PedHunter: Occlusion Robust Pedestrian Detector in Crowded Scenes☆11Nov 21, 2019Updated 6 years ago
- Implementation of Boundary Attributions for Normal (Vector) Explanations☆11Aug 13, 2021Updated 5 years ago
- python codes for CIDEr - Consensus-based Image Caption Evaluation☆32Jun 25, 2019Updated 7 years ago
- [ICLR 2025] Causal Graphical Models for Vision-Language Compositional Understanding☆10Apr 15, 2025Updated last year
- Source code of the paper: Video Inpainting Localization with Contrastive Learning, IEEE SPL 2025.☆12Aug 9, 2025Updated last year
- This dataset contains about 110k images annotated with the depth and occlusion relationships between arbitrary objects. It enables resear…☆16Apr 28, 2021Updated 5 years ago