A neural network architecture(CNN+LSTM) that automatically generates captions from the images. The model uses ResNet architecture to train the Encoder while DecoderRNN has to be trained with our choice of trainable parameters. I have trained the model on the Microsoft Common Objects in COntext (MS COCO) dataset and have tested the network on fic…
☆25Jan 13, 2020Updated 6 years ago
Alternatives and similar repositories for Automatic-Image-Captioning
Users that are interested in Automatic-Image-Captioning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- CNN-Encoder and RNN-Decoder (Bahdanau Attention) for image caption or image to text on MS-COCO dataset. 图片描述☆36Jun 30, 2019Updated 7 years ago
- ☆10Apr 20, 2018Updated 8 years ago
- An implementation of bidirectional LSTM-CRF for Named Entity Relationship on custom corpus with custom word embeddings☆14Apr 9, 2019Updated 7 years ago
- Image Caption workout with NIC and NBT☆16Apr 5, 2019Updated 7 years ago
- Reimplementation of ECCV paper "NeRF: Representing Scenes as Neural Radiance Fields for View Synthesis" with PyTorch Library.☆37Apr 7, 2022Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Neural Reflectance Field from Shading and Shadow under a Fixed Viewpoint☆16Aug 8, 2022Updated 4 years ago
- A simplified implementation of RetinaNet from https://arxiv.org/pdf/1708.02002.pdf using TF2.0☆13Aug 5, 2020Updated 6 years ago
- Public repository for paper "Finding Your (3D) Center: 3D object detection using a learned loss"☆17Nov 21, 2022Updated 3 years ago
- SODEN: A Scalable Continuous-Time Survival Model through Ordinary Differential Equation Networks☆16Mar 2, 2023Updated 3 years ago
- Implementation of 3D reconstruction from accidental motion, CVPR 2014☆12Dec 8, 2022Updated 3 years ago
- This repository contains the source code, models and data files for the work titled: "Unsupervised Image Style Embeddings for Retrieval a…☆13May 29, 2021Updated 5 years ago
- ☆14Oct 24, 2023Updated 2 years ago
- https://github.com/mitsuba-renderer/mitsuba2 in docker☆10Jun 13, 2020Updated 6 years ago
- Pytorch implementation of audio-visual fusion video captioning model☆27Jul 26, 2018Updated 8 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Sentiment analysis with variable length sequences in pytorch☆34Jul 13, 2019Updated 7 years ago
- Experiments dashboard for LabML☆17Dec 11, 2022Updated 3 years ago
- ☆10Apr 11, 2023Updated 3 years ago
- Code for the DataPipes article☆15Jun 14, 2022Updated 4 years ago
- ☆21Apr 12, 2022Updated 4 years ago
- Reading list for research topics in multimodal machine learning☆11Mar 14, 2023Updated 3 years ago
- This sample includes simeple CNN classifier for music and audio-folder dataloader just like ImageFolder in torchvision.☆11Oct 30, 2018Updated 7 years ago
- Open Source Deep Learning Computer Vision (DLCV) Library☆16Nov 26, 2020Updated 5 years ago
- Image Caption using keras, VGG16 pretrained model, CNN and RNN☆44Sep 24, 2019Updated 6 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Imagenet Pretraining for Covid-19 Xray Identification☆10Apr 5, 2020Updated 6 years ago
- Examples of Generative Adversarial Networks built using torchgan☆12Jun 11, 2019Updated 7 years ago
- Llama causal LM fully recreated in LibTorch. Designed to be used in Unreal Engine 5☆16Sep 19, 2024Updated 2 years ago
- PyTorch implementation of Chinese image captioning on AI_challenger dataset☆33Dec 25, 2019Updated 6 years ago
- ☆17Mar 23, 2023Updated 3 years ago
- fish eye correct☆19Mar 20, 2015Updated 11 years ago
- [ICML'2022] Estimating Instance-dependent Bayes-label Transition Matrix using a Deep Neural Network☆21Jul 19, 2022Updated 4 years ago
- Training Deep Convolutional Networks on visual classification tasks☆11Mar 7, 2016Updated 10 years ago
- Pytorch implement Show, Attend and Tell: Neural Image Caption Generation with Visual Attention☆95Dec 25, 2018Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Collection of (unfinished) notebooks☆14Sep 16, 2020Updated 6 years ago
- Summarization of Multimodal articles☆10Oct 14, 2022Updated 3 years ago
- Self Project on making Anime Face using DCGAN☆15Apr 13, 2019Updated 7 years ago
- Text-Independent Speaker Recognition Using Gaussian Mixture Models☆12Jul 1, 2015Updated 11 years ago
- ☆12Aug 27, 2020Updated 6 years ago
- SSR: An Efficient and Robust Framework for Learning with Unknown Label Noise (BMVC2022)☆32Mar 22, 2025Updated last year
- ☆15Mar 5, 2019Updated 7 years ago