aurooj / WeakGroundedVQA_CapsulesView external linksLinks
☆18Apr 10, 2023Updated 2 years ago
Alternatives and similar repositories for WeakGroundedVQA_Capsules
Users that are interested in WeakGroundedVQA_Capsules are comparing it to the libraries listed below
Sorting:
- Weakly Supervised Grounding for VQA in Vision-Language Transformers☆16May 6, 2023Updated 2 years ago
- Repo for ICCV 2021 paper: Beyond Question-Based Biases: Assessing Multimodal Shortcut Learning in Visual Question Answering☆28Jul 1, 2024Updated last year
- BottomUpTopDown VQA model with question-type debiasing☆22Oct 6, 2019Updated 6 years ago
- ☆27Oct 7, 2021Updated 4 years ago
- Code for Greedy Gradient Ensemble for Visual Question Answering (ICCV 2021, Oral)☆27Mar 28, 2022Updated 3 years ago
- Local self-attention in Transformer for visual question answering☆13Mar 17, 2024Updated last year
- ☆32Apr 24, 2024Updated last year
- [NAACL 2021] Designing a Minimal Retrieve-and-Read System for Open-Domain Question Answering☆36Apr 20, 2021Updated 4 years ago
- Actlist Plugin library to development and debugging.☆14Feb 17, 2022Updated 3 years ago
- 11th Solution of Kaggle TalkingData AdTracking Fraud Detection Challenge☆10May 10, 2018Updated 7 years ago
- sklearn implementation of gap-statistic☆10May 25, 2019Updated 6 years ago
- Deep Learning library for Python. Convnets, recurrent neural networks, and more. Runs on Theano or TensorFlow.☆12Dec 24, 2016Updated 9 years ago
- Code release for Park et al. Multimodal Multimodal Explanations: Justifying Decisions and Pointing to the Evidence. in CVPR, 2018☆48Jul 27, 2018Updated 7 years ago
- Evaluation codes of "From Images to Textual Prompts: Zero-shot VQA with Frozen Large Language Models".☆16May 15, 2023Updated 2 years ago
- Code for WACV 2021 Paper "Meta Module Network for Compositional Visual Reasoning"☆43May 13, 2021Updated 4 years ago
- Web Interface for gaze recording: CVPR 2018☆10Jul 10, 2018Updated 7 years ago
- Code for "RADCoT: Retrieval-Augmented Distillation to Specialization Models for Generating Chain-of-Thoughts in Query Expansion", LREC-CO…☆11May 25, 2024Updated last year
- A unified network structure specialized for smoke detection and concentration evaluation in the wild.☆15Apr 19, 2019Updated 6 years ago
- Deep Learning for Video Retrieval by Natural Language☆11Oct 20, 2019Updated 6 years ago
- AAAI-22 paper: Synthetic Disinformation Attacks on Automated Fact Verification Systems☆12Feb 23, 2022Updated 3 years ago
- [EMNLP 2024 Industry track] MERLIN : Multimodal Embedding Refinement via LLM-based Iterative Navigation for Text-Video Retrieval-Rerank P…☆14Mar 4, 2025Updated 11 months ago
- Official Implementation for CVPR 2023 paper "Divide and Conquer: Answering Questions with Object Factorization and Compositional Reasonin…☆10Jun 16, 2024Updated last year
- The official implement of "Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings"☆18Dec 5, 2024Updated last year
- Enhancing Legal Case Retrieval via Scaling High-quality Synthetic Query-Candidate Pairs (EMNLP 2024)☆16Nov 17, 2024Updated last year
- Adaptive Passage Encoder for Open-domain Question Answering☆15Jun 1, 2021Updated 4 years ago
- Research Code for ICCV 2019 paper "Relation-aware Graph Attention Network for Visual Question Answering"☆187Apr 15, 2021Updated 4 years ago
- Joint Neural Model for Entity & Relation Extraction☆15Oct 18, 2021Updated 4 years ago
- EMNLP 2020: Filtering before Iteratively Referring for Knowledge-Grounded Response Selection in Retrieval-Based Chatbots☆12Dec 15, 2020Updated 5 years ago
- Demo codes in our presentation about MXNet in AWS Seoul Summit 2017☆12Apr 24, 2017Updated 8 years ago
- Identifying the language of input text using character-level n-grams, with support for 45 languages☆11Dec 26, 2022Updated 3 years ago
- For early fire detection, smoke must be detect first. This project create smoke videos to feed deep laerning dataset☆11Apr 25, 2018Updated 7 years ago
- FreebaseAPI is a library to use the Freebase API (data mapper + low level API)☆43Nov 21, 2014Updated 11 years ago
- Implementation for the paper "Dynamic Language Binding in Relational Visual Reasoning" (Le et al., IJCAI 2020)☆13Jul 25, 2024Updated last year
- code for the table-based open domain question answering project, with paper title: "Reasoning over Hybrid Chain for Table-and-Text Open D…☆12Sep 16, 2022Updated 3 years ago
- Routines for implementing various statistical and machine learning techniques.☆19Nov 28, 2022Updated 3 years ago
- ☆15Jan 16, 2024Updated 2 years ago
- A pytorch-version implementation codes of paper: "BSN++: Complementary Boundary Regressor with Scale-Balanced Relation Modeling for Tempo…☆14Oct 10, 2021Updated 4 years ago
- ☆13May 26, 2022Updated 3 years ago
- This code was used to collect, process, and validate the REFLACX (Reports and Eye-Tracking Data for Localization of Abnormalities in Ches…☆18Apr 6, 2022Updated 3 years ago