A Light weight deep learning model with with a web application to answer image-based questions with a non-generative approach for the VizWiz grand challenge 2023 by carefully curating the answer vocabulary and adding linear layer on top of Open AI's CLIP model as image and text encoder
☆16Jun 27, 2023Updated 3 years ago
Alternatives and similar repositories for Visual-Question-Answering
Users that are interested in Visual-Question-Answering are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A self-evident application of the VQA task is to design systems that aid blind people with sight reliant queries. The VizWiz VQA dataset …☆15Dec 12, 2023Updated 2 years ago
- Local self-attention in Transformer for visual question answering☆13Mar 17, 2024Updated 2 years ago
- Variational Information Bottleneck☆16Nov 26, 2018Updated 7 years ago
- Reproducible code for Augmentation paper☆17Jan 23, 2019Updated 7 years ago
- ☆12Jan 5, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆12Updated this week
- The codes for ACM Multimedia 2023 paper 'DAOT: Domain-Agnostically Aligned Optimal Transport for Domain-Adaptive Crowd Counting. '☆13Jan 12, 2024Updated 2 years ago
- The active learning algorithm, mismatch-first farthest-traversal. Implementation and visualization.☆12Dec 25, 2021Updated 4 years ago
- 科大讯飞线下销量挑战赛top7方案☆13Aug 21, 2021Updated 5 years ago
- Neural Fuzzy Repair (NFR) is a data augmentation pipeline, which integrates fuzzy matches (i.e. similar translations) into neural machine…☆12Aug 14, 2024Updated 2 years ago
- Multicultural Proverbs and Sayings☆13Jan 11, 2025Updated last year
- This is a code repository of Graphhopper: Multi-Hop Scene GraphReasoning for Visual Question Answering☆19Oct 30, 2021Updated 4 years ago
- A library for automatically extracting color palettes from images☆10Mar 20, 2016Updated 10 years ago
- USTC网络安全实验室网站源码☆11Sep 1, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆12Jan 4, 2022Updated 4 years ago
- A Persian Image Captioning model based on Vision Encoder Decoder Models of the transformers🤗.☆20Feb 27, 2022Updated 4 years ago
- ☆15Jan 28, 2021Updated 5 years ago
- Machine learning bootcamp https://ztlevi.gitbook.io/ml-101/☆14Jul 22, 2023Updated 3 years ago
- Matlab code for fast Hausdorff distance for binary images or segmentation maps☆10Mar 10, 2019Updated 7 years ago
- 人人都能看懂的轻量级解决方案☆15Jul 10, 2020Updated 6 years ago
- The open source implementation of "NeVA: NeMo Vision and Language Assistant"☆17Aug 26, 2023Updated 3 years ago
- Alpha version of our data-centric visual benchmark for training data selection☆16Aug 28, 2023Updated 3 years ago
- ☆17Nov 1, 2023Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Code and dataset release for "PACS: A Dataset for Physical Audiovisual CommonSense Reasoning" (ECCV 2022)☆18Dec 20, 2022Updated 3 years ago
- Hyperpatameter Bayesian Optimization for Image Classification in PyTorch☆11Aug 20, 2019Updated 7 years ago
- PySpark Tutorial for Beginners on Google Colab: Hands-On Guide☆17Sep 13, 2020Updated 5 years ago
- C++11多线程入门☆12Apr 24, 2019Updated 7 years ago
- 基于多模态检索的互联网图文匹配☆15Mar 17, 2024Updated 2 years ago
- 清华大学软件学院-数据结构C++/qt大作业☆12Sep 2, 2020Updated 6 years ago
- Image similarity estimation using a Siamese Network with a triplet loss☆11Jul 27, 2023Updated 3 years ago
- CMU 15-441 项目一 Liso Web服务器☆17Nov 4, 2018Updated 7 years ago
- 微信大数据挑战赛2021☆18Sep 6, 2021Updated 5 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- AMS实时推荐系统☆17Nov 4, 2022Updated 3 years ago
- ☆13Dec 16, 2022Updated 3 years ago
- Variational Attention: Propagating Domain-Specific Knowledge for Multi-Domain Learning in Crowd Counting https://arxiv.org/abs/2108.08023…☆22Sep 9, 2021Updated 4 years ago
- 1D-CNN models for NAFLD diagnosis and liver fat fraction quantification using radiofrequency ultrasound signals☆13Jun 10, 2020Updated 6 years ago
- "Fair Federated AI" Summer School, July 19-21, 2024; 16:00 — 19:30 CET time☆13Aug 19, 2024Updated 2 years ago
- 电商广告推荐系统☆14Jun 3, 2022Updated 4 years ago
- Code for the 15th place submission at Trading at the Close competition☆19Jun 22, 2024Updated 2 years ago