[NeurIPS 2024] CharXiv: Charting Gaps in Realistic Chart Understanding in Multimodal LLMs
☆159Apr 22, 2025Updated last year
Alternatives and similar repositories for CharXiv
Users that are interested in CharXiv are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The codebase for our EMNLP24 paper: Multimodal Self-Instruct: Synthetic Abstract Image and Visual Reasoning Instruction Using Language Mo…☆85Jan 27, 2025Updated last year
- The proposed simulated dataset consisting of 9,536 charts and associated data annotations in CSV format.☆26Feb 22, 2024Updated 2 years ago
- Learning from Negative samples for Biomedical Generative Entity Linking☆18May 25, 2025Updated last year
- ☆46May 21, 2024Updated 2 years ago
- ☆259Apr 18, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆44May 29, 2025Updated last year
- A curated list of recent and past chart understanding work based on our IEEE TKDE survey paper: From Pixels to Insights: A Survey on Auto…☆241Dec 17, 2025Updated 7 months ago
- [ACL'25 Main] ChartCoder: Advancing Multimodal Large Language Model for Chart-to-Code Generation☆79Dec 8, 2025Updated 7 months ago
- MathVista: data, code, and evaluation for Mathematical Reasoning in Visual Contexts☆367Sep 29, 2025Updated 9 months ago
- Dataset introduced in PlotQA: Reasoning over Scientific Plots☆83Jun 20, 2023Updated 3 years ago
- Official Repo for the paper: VCR: Visual Caption Restoration. Check arxiv.org/pdf/2406.06462 for details.☆32Feb 26, 2025Updated last year
- Dataset and evaluation suite enabling LLM instruction-following for scientific literature understanding.☆48Mar 17, 2025Updated last year
- [ACL 2024] ChartAssistant is a chart-based vision-language model for universal chart comprehension and reasoning.☆135Sep 7, 2024Updated last year
- This repository contains the code and data for the paper "VisOnlyQA: Large Vision Language Models Still Struggle with Visual Perception o…☆29Jul 9, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [ICLR 2025] ChartMimic: Evaluating LMM’s Cross-Modal Reasoning Capability via Chart-to-Code Generation☆132Dec 19, 2025Updated 7 months ago
- [NeurIPS 2024] MATH-Vision dataset and code to measure multimodal mathematical reasoning capabilities.☆139May 16, 2025Updated last year
- Supporting code for ReCEval paper☆32Sep 14, 2024Updated last year
- ICLR 2023: Learning to Extrapolate: A Transductive Approach☆11Aug 15, 2023Updated 2 years ago
- ☆18Oct 22, 2022Updated 3 years ago
- The official dataset of the flowvqa project.☆24Mar 26, 2024Updated 2 years ago
- [ICLR2025 Oral] ChartMoE: Mixture of Diversely Aligned Expert Connector for Chart Understanding☆100Apr 1, 2025Updated last year
- A huge dataset for Document Visual Question Answering☆24Jul 29, 2024Updated last year
- [NAACL 2024] MMC: Advancing Multimodal Chart Understanding with LLM Instruction Tuning☆95Jan 7, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Continual Memorization of Factoids in Large Language Models☆12Nov 20, 2024Updated last year
- [ICML 2024] | MMT-Bench: A Comprehensive Multimodal Benchmark for Evaluating Large Vision-Language Models Towards Multitask AGI☆119Apr 6, 2026Updated 3 months ago
- Code for paper "Unraveling Cross-Modality Knowledge Conflicts in Large Vision-Language Models."☆54Oct 19, 2024Updated last year
- Evaluating Multimodal Generative AI with Korean Educational Standards, NAACL 2025.☆27May 15, 2025Updated last year
- WikiVideo: Article Generation from Multiple Videos☆15Nov 14, 2025Updated 8 months ago
- Code release for "SPIQA: A Dataset for Multimodal Question Answering on Scientific Papers" [NeurIPS D&B, 2024]☆76Jan 13, 2025Updated last year
- ☆76Jul 14, 2024Updated 2 years ago
- ☆14Dec 25, 2024Updated last year
- ☆19Oct 12, 2025Updated 9 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆92Mar 12, 2026Updated 4 months ago
- [SCIS 2024] The official implementation of the paper "MMInstruct: A High-Quality Multi-Modal Instruction Tuning Dataset with Extensive Di…☆64Nov 7, 2024Updated last year
- Code for Paper: Harnessing Webpage Uis For Text Rich Visual Understanding☆54Dec 12, 2024Updated last year
- [IEEE VIS 2024] LLaVA-Chart: Advancing Multimodal Large Language Models in Chart Question Answering with Visualization-Referenced Instruc…☆75Jan 22, 2025Updated last year
- Official PyTorch implementation of Extract Free Dense Misalignment from CLIP (AAAI'25)☆24Apr 20, 2025Updated last year
- A Python library for processing and filtering TabLib☆14Aug 24, 2024Updated last year
- A Universal Platform for Training and Evaluation of Mobile Interaction☆63Sep 24, 2025Updated 9 months ago