aurooj/WeakGroundedVQA_Capsules

Readme badge preview -

If you own this repo, copy the snippet below and add it to your README.md

[![RelatedRepos](https://img.shields.io/badge/related-repos-yellow)](https://relatedrepos.com/gh/aurooj/WeakGroundedVQA_Capsules)

aurooj / WeakGroundedVQA_Capsules

☆18

Alternatives and similar repositories for WeakGroundedVQA_Capsules

Users that are interested in WeakGroundedVQA_Capsules are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.

Sorting:

chrisc36 / bottom-up-attention-vqa
View on GitHub
BottomUpTopDown VQA model with question-type debiasing
☆22Oct 6, 2019Updated 6 years ago
GeraldHan / GGE
View on GitHub
Code for Greedy Gradient Ensemble for Visual Question Answering （ICCV 2021, Oral）
☆27Mar 28, 2022Updated 4 years ago
SpencerWhitehead / novelvqa
View on GitHub
☆27Oct 7, 2021Updated 4 years ago
DoubtedSteam / DyVTE
View on GitHub
The official implement of "Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings"
☆18Dec 5, 2024Updated last year
shenxiang-vqa / LSAT
View on GitHub
Local self-attention in Transformer for visual question answering
☆13Mar 17, 2024Updated 2 years ago
Wordpress hosting with auto-scaling - Free Trial Offer • Ad
Fully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
liyidi / soundnet_localize_sound_source
View on GitHub
soundnet and localize sound source
☆12Dec 7, 2020Updated 5 years ago
guoyang9 / UnifER
View on GitHub
Official implementation for the MM'22 paper.
☆14Jun 30, 2022Updated 4 years ago
szzexpoi / POEM
View on GitHub
Official Implementation for CVPR 2023 paper "Divide and Conquer: Answering Questions with Object Factorization and Compositional Reasonin…
☆10Jun 16, 2024Updated 2 years ago
CR-Gjx / Img2Prompt
View on GitHub
Evaluation codes of "From Images to Textual Prompts: Zero-shot VQA with Frozen Large Language Models".
☆18May 15, 2023Updated 3 years ago
JasonForJoy / FIRE
View on GitHub
EMNLP 2020: Filtering before Iteratively Referring for Knowledge-Grounded Response Selection in Retrieval-Based Chatbots
☆12Dec 15, 2020Updated 5 years ago
linjieli222 / VQA_ReGAT
View on GitHub
Research Code for ICCV 2019 paper "Relation-aware Graph Attention Network for Visual Question Answering"
☆187Apr 15, 2021Updated 5 years ago
li-xirong / video-retrieval
View on GitHub
Deep Learning for Video Retrieval by Natural Language
☆11Oct 20, 2019Updated 6 years ago
praveena2j / JointCrossAttentional-AV-Fusion
View on GitHub
ABAW3 (CVPRW): A Joint Cross-Attention Model for Audio-Visual Fusion in Dimensional Emotion Recognition
☆50Jan 15, 2024Updated 2 years ago
arunbalajeev / gaze-interface
View on GitHub
Web Interface for gaze recording: CVPR 2018
☆10Jul 10, 2018Updated 8 years ago
Deploy open-source AI quickly and easily - Special Bonus Offer • Ad
Runpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
AlonMendelson / SGVL
View on GitHub
☆17Dec 13, 2023Updated 2 years ago
thunlp / LEAD
View on GitHub
Enhancing Legal Case Retrieval via Scaling High-quality Synthetic Query-Candidate Pairs (EMNLP 2024)
☆17Nov 17, 2024Updated last year
ronghanghu / snmn
View on GitHub
Code release for Hu et al., Explainable Neural Computation via Stack Neural Module Networks. in ECCV, 2018
☆71Nov 17, 2019Updated 6 years ago
thaolmk54 / LOGNet-VQA
View on GitHub
Implementation for the paper "Dynamic Language Binding in Relational Visual Reasoning" (Le et al., IJCAI 2020)
☆13Jul 25, 2024Updated 2 years ago
ronghanghu / gqa_single_hop_baseline
View on GitHub
A simple but well-performing "single-hop" visual attention model for the GQA dataset
☆20Aug 8, 2019Updated 6 years ago
tjdevWorks / TEASEL
View on GitHub
☆26May 8, 2022Updated 4 years ago
stys / kaggle-talkingdata-adtracking-fraud-detection
View on GitHub
TalkingData AdTracking Fraud Detection Challenge
☆10May 8, 2018Updated 8 years ago
sxjscience / aws-summit-2017-seoul
View on GitHub
Demo codes in our presentation about MXNet in AWS Seoul Summit 2017
☆12Apr 24, 2017Updated 9 years ago
val-iisc / RMLVQA
View on GitHub
☆19May 31, 2023Updated 3 years ago
Deploy to Railway using AI coding agents - Free Credits Offer • Ad
Use Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
facebookresearch / selective-vqa_ood
View on GitHub
Implementation for the CVPR 2023 paper "Improving Selective Visual Question Answering by Learning from Your Peers" (https://arxiv.org/abs…
☆26Jul 20, 2023Updated 3 years ago
wenz116 / DRFT
View on GitHub
End-to-end Multi-modal Video Temporal Grounding, NeurIPS 2021
☆18Oct 24, 2021Updated 4 years ago
ceyzaguirre4 / NSM
View on GitHub
Neural State Machine implemented in PyTorch
☆71Oct 10, 2019Updated 6 years ago
namemzy / Sniffer-Net
View on GitHub
A unified network structure specialized for smoke detection and concentration evaluation in the wild.
☆15Apr 19, 2019Updated 7 years ago
xmu-xiaoma666 / SDATR
View on GitHub
Official Code for "Knowing what it is: Semantic-enhanced Dual Attention Transformer" (TMM2022)
☆19Oct 15, 2022Updated 3 years ago
dogusyuksel / CreatingSyntheticSmokeVideos
View on GitHub
For early fire detection, smoke must be detect first. This project create smoke videos to feed deep laerning dataset
☆11Apr 25, 2018Updated 8 years ago
adymaharana / DeepPath_PyTorch
View on GitHub
☆15Jan 15, 2021Updated 5 years ago
ZiyueWu59 / CCA
View on GitHub
☆15Jan 16, 2024Updated 2 years ago
DoubtedSteam / Flash_Attn_with_Score
View on GitHub
Flash Attention implementation that returns both output and attention scores. High-performance, memory-efficient attention with score ext…
☆16Feb 6, 2026Updated 5 months ago
Bare Metal GPUs on DigitalOcean Gradient AI • Ad
Purpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
EIHW / MuSe2022
View on GitHub
☆28May 13, 2022Updated 4 years ago
lupantech / rvqa
View on GitHub
R-VQA: Visual Question Answering with Relation Facts
☆19May 11, 2021Updated 5 years ago
xxcheng0708 / BSNPlusPlus-boundary-sensitive-network
View on GitHub
A pytorch-version implementation codes of paper: "BSN++: Complementary Boundary Regressor with Scale-Balanced Relation Modeling for Tempo…
☆14Oct 10, 2021Updated 4 years ago
ForJadeForest / LIVE-Learnable-In-Context-Vector
View on GitHub
【NeurIPS 2024】The implementation of LIVE: Learnable In-Context Vector for Visual Question Answering https://arxiv.org/abs/2406.13185
☆23May 31, 2025Updated last year
ovguyo / captions-in-VQA
View on GitHub
Using image captions with LLM for zero-shot VQA
☆19Mar 14, 2024Updated 2 years ago
brandonckelly / bck_stats
View on GitHub
Routines for implementing various statistical and machine learning techniques.
☆19Nov 28, 2022Updated 3 years ago
MILVLG / mcan-vqa
View on GitHub
Deep Modular Co-Attention Networks for Visual Question Answering
☆459Dec 16, 2020Updated 5 years ago