[ICCV 2021] Official implementation of the paper "TRAR: Routing the Attention Spans in Transformers for Visual Question Answering"
☆68Oct 11, 2021Updated 4 years ago
Alternatives and similar repositories for TRAR-VQA
Users that are interested in TRAR-VQA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Local self-attention in Transformer for visual question answering☆13Mar 17, 2024Updated 2 years ago
- Official Repository for CVPR 2022 paper "REX: Reasoning-aware and Grounded Explanation"☆22Nov 21, 2023Updated 2 years ago
- Official implementation of Dynamic Routing Transformer Network for Multimodal Sarcasm Detection (ACL'23)☆35Jul 9, 2023Updated 3 years ago
- ☆19May 31, 2023Updated 3 years ago
- Official Code for 'RSTNet: Captioning with Adaptive Attention on Visual and Non-Visual Words' (CVPR 2021)☆123Dec 17, 2022Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Optimized code based on M2 for faster image captioning training☆21Nov 18, 2022Updated 3 years ago
- VQACL: A Novel Visual Question Answering Continual Learning Setting (CVPR'23)☆46Mar 28, 2024Updated 2 years ago
- Official code for paper "Spatially Aware Multimodal Transformers for TextVQA" published at ECCV, 2020.☆64Sep 15, 2021Updated 4 years ago
- Repository for the paper "Data Efficient Masked Language Modeling for Vision and Language".☆18Sep 17, 2021Updated 4 years ago
- Coarse-to-Fine Reasoning for Visual Question Answering (CVPRW'22)☆48Apr 22, 2026Updated 4 months ago
- Official implementation for the MM'22 paper.☆13Jun 30, 2022Updated 4 years ago
- A pytorch implemetation of data augmentation method for visual question answering☆21May 25, 2023Updated 3 years ago
- Official pytorch implementation of paper "Dual-Level Collaborative Transformer for Image Captioning" (AAAI 2021).☆203Jun 8, 2022Updated 4 years ago
- ☆19Jan 7, 2026Updated 8 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Code release for Hu et al., Language-Conditioned Graph Networks for Relational Reasoning. in ICCV, 2019☆92Aug 9, 2019Updated 7 years ago
- Deep Modular Co-Attention Networks for Visual Question Answering☆458Dec 16, 2020Updated 5 years ago
- Offical PyTorch implementation of Clover: Towards A Unified Video-Language Alignment and Fusion Model (CVPR2023)☆38Feb 15, 2023Updated 3 years ago
- The code of IJCAI2022 paper, Declaration-based Prompt Tuning for Visual Question Answering☆20May 10, 2022Updated 4 years ago
- Visual Question Answering through transformers.☆13Sep 21, 2018Updated 7 years ago
- SeqTR: A Simple yet Universal Network for Visual Grounding☆144Oct 30, 2024Updated last year
- code for downloading videos from HowTo100M dataset☆18May 13, 2021Updated 5 years ago
- ☆12Mar 8, 2021Updated 5 years ago
- [AAAI2023] Symbolic Replay: Scene Graph as Prompt for Continual Learning on VQA Task (Oral)☆43Mar 23, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆79Oct 8, 2022Updated 3 years ago
- The official implementation of the paper "DIP: Dual Incongruity Perceiving Network for Sarcasm Detection"☆36Dec 6, 2024Updated last year
- SOIT: Segmenting Objects with Instance-Aware Transformers☆14Jun 6, 2022Updated 4 years ago
- Learning Situation Hyper-Graphs for Video Question Answering☆23Feb 16, 2024Updated 2 years ago
- [CVPR 2022] This repository is for the paper ``DIFNet: Boosting Visual Information Flow for Image Captioning'' .☆21Nov 28, 2022Updated 3 years ago
- Deep Learning ❤️ OneFlow☆19Aug 26, 2021Updated 5 years ago
- A lightweight, scalable, and general framework for visual question answering research☆335Sep 3, 2021Updated 5 years ago
- ☆15May 10, 2021Updated 5 years ago
- A curated list of Visual Question Answering(VQA)(Image/Video Question Answering),Visual Question Generation ,Visual Dialog ,Visual Common…☆673Jul 6, 2023Updated 3 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Simple tutorials on Pytorch DDP training☆277Aug 19, 2022Updated 4 years ago
- [NeurIPS 2022] Make Sharpness-Aware Minimization Stronger: A Sparsified Perturbation Approach -- Official Implementation☆48Jun 29, 2023Updated 3 years ago
- PixelFolder: An Efficient Progressive Pixel Synthesis Network for Image Generation (ECCV 2022)☆33Jul 21, 2022Updated 4 years ago
- ☆12Dec 20, 2024Updated last year
- [NeurIPS 2021] Introspective Distillation for Robust Question Answering☆13Dec 7, 2021Updated 4 years ago
- ☆39Nov 29, 2022Updated 3 years ago
- Code for Greedy Gradient Ensemble for Visual Question Answering (ICCV 2021, Oral)☆27Mar 28, 2022Updated 4 years ago