[CVPR 2021] SUTD-TrafficQA: A Question Answering Benchmark and an Efficient Network for Video Reasoning over Traffic Events
☆66Aug 31, 2026Updated 3 weeks ago
Alternatives and similar repositories for SUTD-TrafficQA
Users that are interested in SUTD-TrafficQA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Implementation for the paper "Hierarchical Conditional Relation Networks for Video Question Answering" (Le et al., CVPR 2020, Oral)☆135Jul 25, 2024Updated 2 years ago
- [ICCV2023] Tem-adapter: Adapting Image-Text Pretraining for Video Question Answer☆37Oct 18, 2023Updated 2 years ago
- [IEEE T-PAMI 2023] Cross-Modal Causal Relational Reasoning for Event-Level Visual Question Answering☆21Jul 6, 2023Updated 3 years ago
- Heterogeneous Memory Enhanced Multimodal Attention Model for VideoQA☆55Sep 13, 2021Updated 5 years ago
- [CVPRW 2024] TrafficVLM: A Controllable Visual Language Model for Traffic Video Captioning. Official code for the 3rd place solution of t…☆59Feb 11, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [ICCV2023] Chaotic World: A Large and Challenging Benchmark for Human Behavior Understanding in Chaotic Events☆10Dec 7, 2024Updated last year
- A reading list of papers about Visual Question Answering.☆36Aug 17, 2022Updated 4 years ago
- Video Graph Transformer for Video Question Answering (ECCV'22)☆49Jun 8, 2023Updated 3 years ago
- [ICCV 2021 Oral + TPAMI] Just Ask: Learning to Answer Questions from Millions of Narrated Videos☆128Sep 29, 2023Updated 2 years ago
- videoqa,天池江之杯视频问答比赛☆13Dec 19, 2018Updated 7 years ago
- Video Question Answering via Gradually Refined Attention over Appearance and Motion☆178Dec 5, 2017Updated 8 years ago
- Spatial-Temporal Graph Learning with Self-supervised Spatial State Module for Traffic Accident Anticipation☆16Dec 11, 2022Updated 3 years ago
- Implementation of the paper: "BRAVE : Broadening the visual encoding of vision-language models"☆26Jun 22, 2026Updated 3 months ago
- [ACM MM 2020] CCD dataset for traffic accident anticipation.☆158Sep 2, 2023Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [CVPR2024 Highlight] The official repo for paper "Abductive Ego-View Accident Video Understanding for Safe Driving Perception"☆78Mar 24, 2025Updated last year
- Load and visualize different datasets in video question answering☆10May 11, 2021Updated 5 years ago
- [CVPR 2021] Pytorch implementation for Probabilistic Modeling of Semantic Ambiguity for Scene Graph Generation☆20May 7, 2021Updated 5 years ago
- [NeurIPS 2023] LMC: Large Model Collaboration with Cross-assessment for Training-Free Open-Set Object Recognition☆20May 26, 2024Updated 2 years ago
- Repository for our CVPR 2017 and IJCV: TGIF-QA☆181Sep 6, 2021Updated 5 years ago
- Connective Cognition Network for Directional Visual Commonsense Reasoning☆15May 6, 2021Updated 5 years ago
- [EMNLP 2018] PyTorch code for TVQA: Localized, Compositional Video Question Answering☆182Oct 25, 2022Updated 3 years ago
- Repository for Traffic Accident Benchmark for Causality Recognition (ECCV 2020)☆34Jun 30, 2021Updated 5 years ago
- ☆10Mar 30, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆15Dec 12, 2024Updated last year
- ☆15Aug 12, 2022Updated 4 years ago
- YOLOv9 implement with mmyolo☆12Jul 8, 2024Updated 2 years ago
- ☆11Jul 11, 2023Updated 3 years ago
- In-the-wild Question Answering☆15May 10, 2023Updated 3 years ago
- Dataset accompanying the paper "Adaptive Methods for Real-World Domain Generalization"☆16Aug 17, 2023Updated 3 years ago
- Code for the paper BiST: Bi-directional Spatio-Temporal Reasoning for Video-Grounded Dialogues (EMNLP20)☆11Jun 16, 2025Updated last year
- Activity Grammars for Temporal Action Segmentation (NeurIPS 2023)☆14Jun 14, 2024Updated 2 years ago
- ☆36Apr 18, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆30Feb 18, 2022Updated 4 years ago
- Uni-ViGU: Towards Unified Video Generation and Understanding via A Diffusion-Based Video Generator☆35Apr 15, 2026Updated 5 months ago
- Code of paper "A Video Dataset for Falling Object Detection around Buildings" https://arxiv.org/abs/2408.05750☆22Jul 10, 2025Updated last year
- Implementations of autoencoders (VAE, AAE, and others)☆11Oct 1, 2018Updated 7 years ago
- generate ROS/ROS2 node based on DBC files☆22Jun 28, 2025Updated last year
- The efficient tuning method for VLMs☆84Mar 10, 2024Updated 2 years ago
- Code for the paper "Crash To Not Crash: Learn to Identify Dangerous Vehicles Using a Simulator" presented in AAAI 2019 (Oral)☆17Mar 15, 2019Updated 7 years ago