A multi-task model which does image captioning, sentence paraphrasing and cross-modal retrieval.
☆19Nov 21, 2019Updated 6 years ago
Alternatives and similar repositories for STT
Users that are interested in STT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆10Apr 20, 2018Updated 8 years ago
- ☆23Aug 18, 2018Updated 7 years ago
- Learning Fragment Self-Attention Embeddings for Image-Text Matching, in ACM MM 2019☆41Sep 24, 2019Updated 6 years ago
- In this work, we implement different cross-modal learning schemes such as Siamese Network, Correlational Network and Deep Cross-Modal Pro…☆11Aug 23, 2021Updated 4 years ago
- Cross-modal Coherence Modeling for Caption Generation☆11Jul 24, 2020Updated 6 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Scraping Program for Pascal Sentence Dataset☆17Sep 9, 2015Updated 10 years ago
- For ICDAR 2019 Paper on End-to-end License Plate and Scene Text Recognition with multi-head attention models☆25Aug 14, 2021Updated 4 years ago
- web3js + infura + android + solidity☆10Feb 2, 2019Updated 7 years ago
- I implemented a detection algorithm with a classification data set that does not have annotation information for the bounding box. Based …☆31Jan 29, 2018Updated 8 years ago
- Extract the key frame from the tested video, and then search the most similar Images from the database, which consists over 1,4000 pictur…☆10Mar 13, 2014Updated 12 years ago
- Measure the diversity of image descriptions, repository for our COLING 2018 paper.☆13Dec 29, 2019Updated 6 years ago
- My assignments for CN course [CSE232] [IIIT-Delhi].☆13May 24, 2018Updated 8 years ago
- kubernetes setup to bootstrap distributed on google container engine☆13Mar 20, 2018Updated 8 years ago
- Partially Non-Autoregressive Image Captioning☆10Sep 30, 2021Updated 4 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Code for paper "Adaptively Aligned Image Captioning via Adaptive Attention Time". NeurIPS 2019☆50Dec 18, 2019Updated 6 years ago
- Code for "simNet: Stepwise Image-Topic Merging Network for Generating Detailed and Comprehensive Image Captions" (EMNLP 2018)☆36Sep 5, 2018Updated 7 years ago
- Self-Supervised Domain Adaptation with Consistency Training☆20Oct 28, 2020Updated 5 years ago
- an implementation of Deformation Graph compatible with CUDA C++ and used in warping defamations in real-time non-rigid registration☆10Jan 22, 2020Updated 6 years ago
- NeurIPS22 "RankFeat: Rank-1 Feature Removal for Out-of-distribution Detection" and T-PAMI Extension☆20Feb 21, 2025Updated last year
- Person Keypoint Detection in PyTorch☆13Mar 20, 2020Updated 6 years ago
- ☆16Jan 30, 2022Updated 4 years ago
- Code for ComEx [CVPR 2022]☆12Dec 5, 2022Updated 3 years ago
- Task-Adaptive Feature Sub-Space Learning for few-shot classification☆12Sep 26, 2020Updated 5 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Matlab demo code for "MTFH: A Matrix Tri-Factorization Hashing Framework for Efficient Cross-Modal Retrieval"☆17Sep 13, 2019Updated 6 years ago
- ☆12Jan 8, 2025Updated last year
- The Pytorch code of "Asymmetric Distribution Measure for Few-shot Learning", IJCAI 2020.☆15Oct 9, 2020Updated 5 years ago
- (AAAI'20) The source code for the paper "Joint Parsing and Generation for Abstractive Summarization".☆15Apr 3, 2020Updated 6 years ago
- [AAAI'20] Code release for "HAL: Improved Text-Image Matching by Mitigating Visual Semantic Hubs".☆38Oct 4, 2023Updated 2 years ago
- Simple tool to change the INPUT and OUTPUT shape of ONNX.☆15Apr 1, 2025Updated last year
- Code for Unsupervised Image Captioning☆223Mar 24, 2023Updated 3 years ago
- pytorch implementation of Semantics-AssistedVideoCaptioning☆11Feb 16, 2023Updated 3 years ago
- Polysemous Visual-Semantic Embedding for Cross-Modal Retrieval (CVPR 2019)☆135Mar 15, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Coco datasets Visualization.☆10Aug 9, 2021Updated 5 years ago
- Will store links to known evaluation datasets alongside stats to characterize them☆24Mar 9, 2016Updated 10 years ago
- COVID-19 data from the repo CSSEGISandData/COVID-19> https://github.com/CSSEGISandData/COVID-19☆16Mar 10, 2023Updated 3 years ago
- ☆15Nov 26, 2023Updated 2 years ago
- Lightweight Transformer for Multi-modal Tasks☆16Dec 9, 2022Updated 3 years ago
- Unofficial implementation of the paper I2V-Adapter: A General Image-to-Video Adapter for Video Diffusion Models.☆19Mar 13, 2024Updated 2 years ago
- The code for “Attention and Language Ensemble for Scene Text Recognition with Convolutional Sequence Modeling”☆43Nov 29, 2018Updated 7 years ago