[ICML 2026] GenExam: A Multidisciplinary Text-to-Image Exam
☆70May 26, 2026Updated 2 months ago
Alternatives and similar repositories for GenExam
Users that are interested in GenExam are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ECCV'26] GRADE: Grounded Reasoning Assessment for Discipline-informed Editing☆29Apr 23, 2026Updated 3 months ago
- ⚽️🤖 Benchmarking LLMs and deep-research agents on real-world football prediction — from the tactical "who scores in minute 67" to the st…☆22Jul 22, 2026Updated 2 weeks ago
- [CVPR 2025] Mono-InternVL: Pushing the Boundaries of Monolithic Multimodal Large Language Models with Endogenous Visual Pre-training☆109Jul 18, 2025Updated last year
- We provide TextEdit, a high-quality, multi-scenario text editing benchmark for generation models.☆21Mar 16, 2026Updated 4 months ago
- ☆16Oct 11, 2025Updated 9 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- [ICLR 2026] SpaCE-10: A Comprehensive Benchmark for Multimodal Large Language Models in Compositional Spatial Intelligence☆20Jan 26, 2026Updated 6 months ago
- Bridging the gap between image generation and real-world design: a benchmark for structured, multi-constraint commercial visual content g…☆21Apr 24, 2026Updated 3 months ago
- [NIPS 2025 DB Oral] Official Repository of paper: Envisioning Beyond the Pixels: Benchmarking Reasoning-Informed Visual Editing☆155May 18, 2026Updated 2 months ago
- ☆15Nov 13, 2025Updated 8 months ago
- Implementation code for the paper "Draw2Think: Harnessing Geometry Reasoning through Constraint Engine Interaction"☆17May 28, 2026Updated 2 months ago
- [ACL 2026] A benchmark for evaluating the reliability of text-to-infographic generation with curated test cases and automated question-ba…☆15Jun 8, 2026Updated 2 months ago
- Official respository for ReasonGen-R1☆75Jun 23, 2025Updated last year
- The first unified, efficient, and extensible evaluation toolkit for evaluating image generation and editing models across multiple benchm…☆50Apr 12, 2026Updated 3 months ago
- Sequential Diffusion Language Model (SDLM) enhances pre-trained autoregressive language models by adaptively determining generation lengt…☆98Dec 27, 2025Updated 7 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ECCV 2026] Official repository of "Reliable Reasoning in SVG-LLMs via Multi-Task Multi-Reward Reinforcement Learning".☆24Jul 17, 2026Updated 3 weeks ago
- ☆21Jul 3, 2025Updated last year
- ☆44Jul 9, 2025Updated last year
- ☆16Aug 19, 2023Updated 2 years ago
- [ICLR 2026 Oral] DiffusionNFT: Online Diffusion Reinforcement with Forward Process☆1,007Feb 10, 2026Updated 6 months ago
- TELL: Test-time Experiential Lifelong Learning – a single LLM agent that learns from experience at test time, achieving 43.9% on ARC-AGI-…☆27Apr 29, 2026Updated 3 months ago
- [NeurIPS 2025 DB] OneIG-Bench is a meticulously designed comprehensive benchmark framework for fine-grained evaluation of T2I models acro…☆122Feb 10, 2026Updated 6 months ago
- [ECCV 2026] VKnowU: Evaluating Visual Knowledge Understanding in Multimodal LLMs☆17Feb 3, 2026Updated 6 months ago
- [ICLR 2026] Official repository of "InternSVG: Towards Unified SVG Tasks with Multimodal Large Language Models".☆122Feb 6, 2026Updated 6 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆40May 20, 2025Updated last year
- ☆53Jun 13, 2025Updated last year
- Official implementation of HPSv3: Towards Wide-Spectrum Human Preference Score (ICCV2025)☆333Dec 5, 2025Updated 8 months ago
- Doodling our way to AGI ✏️ 🖼️ 🧠☆128May 29, 2025Updated last year
- ☆44May 29, 2025Updated last year
- Official repo of "MMBench-GUI: Hierarchical Multi-Platform Evaluation Framework for GUI Agents". It can be used to evaluate a GUI agent w…☆113Sep 8, 2025Updated 11 months ago
- (ICCV2025) EEdit⚡: Rethinking the Spatial and Temporal Redundancy for Efficient Image Editing☆63Sep 17, 2025Updated 10 months ago
- AAAI2026 X2Edit: Revisiting Arbitrary-Instruction Image Editing through Self-Constructed Data and Task-Aware Representation Learning☆97Nov 21, 2025Updated 8 months ago
- Evaluation codes and data for GenEval2☆84Jan 8, 2026Updated 7 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official Repository of paper MM-HELIX: Boosting Multimodal Long-Chain Reflective Reasoning with Holistic Platform and Adaptive Hybrid Pol…☆69Jan 26, 2026Updated 6 months ago
- [ACM MM25] LongWriter-V: Enabling Ultra-Long and High-Fidelity Generation in Vision-Language Models☆24Mar 29, 2025Updated last year
- Official inference code and LongText-Bench benchmark for our paper X-Omni (https://arxiv.org/pdf/2507.22058).☆428Aug 26, 2025Updated 11 months ago
- 【ICML2026】Reasoning to Edit: Hypothetical Instruction-Based Image Editing with Visual Reasoning☆27May 18, 2026Updated 2 months ago
- ☆16Mar 8, 2026Updated 5 months ago
- Official repository for the FIRM Reward series☆41Updated this week
- Official repository for Interleave-VLA☆22Apr 5, 2026Updated 4 months ago