The implementations of various baselines in our CIKM 2022 paper: ChiQA: A Large Scale Image-based Real-World Question Answering Dataset for Multi-Modal Understanding.
☆34May 13, 2024Updated 2 years ago
Alternatives and similar repositories for ChiQA
Users that are interested in ChiQA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A reimplementation of KOSMOS-1 from "Language Is Not All You Need: Aligning Perception with Language Models"☆27Mar 3, 2023Updated 3 years ago
- CoADNet: Collaborative Aggregation-and-Distribution Networks for Co-Salient Object Detection☆19Jan 8, 2021Updated 5 years ago
- ☆10Apr 4, 2018Updated 8 years ago
- [CVPR 2024] Code and datasets for 'Learning Spatial Features from Audio-Visual Correspondence in Egocentric Videos'☆14Jun 16, 2024Updated 2 years ago
- ☆13Nov 28, 2021Updated 4 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Human-centric environment representations from egocentric video☆15Feb 5, 2026Updated 5 months ago
- Code and data for paper "Context-faithful Prompting for Large Language Models".☆41Mar 23, 2023Updated 3 years ago
- The demo for "Discretization and Re-synthesis: an alternative method to solve the Cocktail Party Problem".☆12Oct 25, 2021Updated 4 years ago
- Audio Generation model working with GPT-2 and VQVAE compressed representation of MelSpectrograms☆18Oct 8, 2023Updated 2 years ago
- [ECCV2024] The official implementation of "Listen to Look into the Future: Audio-Visual Egocentric Gaze Anticipation".☆16Feb 24, 2025Updated last year
- Official Implementation of DMT: Dual Mean-Teacher in PyTorch.☆10Oct 27, 2023Updated 2 years ago
- The repository for papaer "Distance between Relevant Information Pieces Causes Bias in Long-Context LLMs"☆14Dec 16, 2024Updated last year
- ACM Multimedia 2023 (Oral) - RTQ: Rethinking Video-language Understanding Based on Image-text Model☆15Apr 7, 2026Updated 3 months ago
- ☆15Oct 19, 2020Updated 5 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [ICML2022] "Identity-Disentangled Adversarial Augmentation for Self-Supervised Learning"☆10Jul 24, 2022Updated 4 years ago
- Source code for the Paper "Mind the Gap: Benchmarking Spatial Reasoning in Vision-Language Models"☆19Feb 1, 2026Updated 5 months ago
- The substitution of qsub.☆12Jan 25, 2019Updated 7 years ago
- Code for our ACL2021 paper: "Check It Again: Progressive Visual Question Answering via Visual Entailment"☆31Nov 24, 2021Updated 4 years ago
- RUArt: A Novel Text-Centered Solution for Text-Based Visual Question Answering☆10Nov 27, 2022Updated 3 years ago
- Boundaries and Region Representation Fusion☆12Mar 24, 2023Updated 3 years ago
- 毕业设计: 基于深度学习的视觉问答☆13Jun 20, 2018Updated 8 years ago
- A three-column graphical LaTeX2e resume☆12Dec 12, 2019Updated 6 years ago
- An FL algorithm inspired by FedGMA☆11Oct 21, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆21Feb 15, 2022Updated 4 years ago
- Tensorflow implementation of CARN☆10Oct 3, 2018Updated 7 years ago
- TaiSu(太素)--a large-scale Chinese multimodal dataset(亿级大规模中文视觉语言预训练数据集)☆192Nov 17, 2023Updated 2 years ago
- ☆10Jul 21, 2021Updated 5 years ago
- ☆18Apr 16, 2024Updated 2 years ago
- This is the official repo of "QuickLLaMA: Query-aware Inference Acceleration for Large Language Models"☆54Jul 16, 2024Updated 2 years ago
- Action2Sound: Ambient-Aware Generation of Action Sounds from Egocentric Videos☆26Oct 1, 2024Updated last year
- Official implementation of HawkEye: Training Video-Text LLMs for Grounding Text in Videos☆47Apr 29, 2024Updated 2 years ago
- 基于树形条件随机场的高阶句法分析☆16Apr 28, 2022Updated 4 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [COLM 2025] Official code for "When To Solve, When To Verify: Compute-Optimal Problem Solving and Generative Verification for LLM Reasoni…☆15Oct 31, 2025Updated 8 months ago
- Code and data release for the paper "Learning Fine-grained View-Invariant Representations from Unpaired Ego-Exo Videos via Temporal Align…☆19Apr 5, 2024Updated 2 years ago
- NLPCC-KBQA Dataset☆15Dec 7, 2021Updated 4 years ago
- Implement SSD using Gluon in only 300 lines of codes!☆10Nov 12, 2017Updated 8 years ago
- ☆15Jan 9, 2026Updated 6 months ago
- A Systematic Study and Comprehensive Evaluation of ChatGPT on Benchmark Datasets.☆15Jul 10, 2023Updated 3 years ago
- Implementation of BapFL: You can Backdoor Attack Personalized Federated Learning☆15Sep 18, 2023Updated 2 years ago