Feature extraction and visualization scripts for nocaps baselines.
☆18Jan 22, 2021Updated 5 years ago
Alternatives and similar repositories for image-feature-extractors
Users that are interested in image-feature-extractors are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Baseline model for nocaps benchmark, ICCV 2019 paper "nocaps: novel object captioning at scale".☆77Oct 3, 2023Updated 2 years ago
- Website for TextVQA dataset.☆29Apr 30, 2023Updated 2 years ago
- Implementation of CVPR 2016 paper☆74Jan 31, 2021Updated 5 years ago
- ☆22Oct 9, 2021Updated 4 years ago
- ☆20May 12, 2022Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [ICLR 2022 Spotlight] Multi-Stage Episodic Control for Strategic Exploration in Text Games☆15Feb 8, 2026Updated 2 months ago
- ☆11Nov 13, 2020Updated 5 years ago
- Code for Decoupled Novel Object Captioner☆28Feb 26, 2020Updated 6 years ago
- This repository provides the dataset introduced by our WSSTG paper☆13Jul 21, 2019Updated 6 years ago
- Official code for the paper "Contrast and Classify: Training Robust VQA Models" published at ICCV, 2021☆19Jul 27, 2021Updated 4 years ago
- a parody of the ever-increasing amount of papers that appear on arXiv☆38Jan 6, 2025Updated last year
- Multitask Multilingual Multimodal Pre-training☆73Nov 27, 2022Updated 3 years ago
- Pytorch version of VidLanKD: Improving Language Understanding viaVideo-Distilled Knowledge Transfer (NeurIPS 2021))☆56Feb 6, 2023Updated 3 years ago
- ☆16Apr 27, 2021Updated 4 years ago
- Deploy open-source AI quickly and easily - Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- PyTorch implementation of Chinese image captioning on AI_challenger dataset☆13Sep 24, 2017Updated 8 years ago
- The Student Cluster Guide, tutorial 2 of ETH's Digital Humans 2024 course.☆10Jul 26, 2025Updated 8 months ago
- GLAC Net: GLocal Attention Cascading Network for the Visual Storytelling Challenge☆45Aug 26, 2020Updated 5 years ago
- ☆54Dec 13, 2019Updated 6 years ago
- ☆24Jul 8, 2020Updated 5 years ago
- Code for WACV 2021 Paper "Meta Module Network for Compositional Visual Reasoning"☆43May 13, 2021Updated 4 years ago
- This is a python library. Install with "python3 -m pip install rp" then run with "python3 -m rp" or just "rp". Requires python≥3.5☆13Mar 17, 2026Updated last month
- Show, Control and Tell: A Framework for Generating Controllable and Grounded Captions. CVPR 2019☆282Dec 21, 2022Updated 3 years ago
- Pytorch implementation of Multimodal Neural Machine Translation(MNMT).☆12Jan 21, 2021Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Show, Edit and Tell: A Framework for Editing Image Captions, CVPR 2020☆82Jul 17, 2020Updated 5 years ago
- Bottom-up attention model for image captioning and VQA, based on Faster R-CNN and Visual Genome☆23Aug 22, 2019Updated 6 years ago
- Code for ICML 2019 paper "Probabilistic Neural-symbolic Models for Interpretable Visual Question Answering" [long-oral]☆67Aug 3, 2023Updated 2 years ago
- (ECCV2024) Within the Dynamic Context: Inertia-aware 3D Human Modeling with Pose Sequence☆19Jun 27, 2025Updated 9 months ago
- ☆28Dec 8, 2022Updated 3 years ago
- 针对 markdown 文件的命令行翻译☆14Feb 2, 2023Updated 3 years ago
- Stack-Captioning: Coarse-to-Fine Learning for Image Captioning☆63Apr 18, 2018Updated 8 years ago
- The source code and the data for ACL 2022 paper "Show Me More Details: Discovering Hierarchies of Procedures from Semi-structured Web Dat…☆14Apr 21, 2023Updated 2 years ago
- Code for the CVPR 2020 paper 'Action Modifiers: Learning from Adverbs in Instructional Videos'☆23May 17, 2021Updated 4 years ago
- Serverless GPU API endpoints on Runpod - Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆12Dec 13, 2022Updated 3 years ago
- ☆39May 28, 2018Updated 7 years ago
- Unpaired Image Captioning☆36Mar 25, 2021Updated 5 years ago
- Code base for the paper "Latent variable model for multi-modal translation".☆17Jul 25, 2024Updated last year
- ☆26Mar 26, 2025Updated last year
- Pytorch implementation of https://arxiv.org/pdf/1909.10470.pdf☆32Aug 23, 2021Updated 4 years ago
- ☆45Oct 11, 2021Updated 4 years ago