[ICML 2026] GenExam: A Multidisciplinary Text-to-Image Exam
☆69May 26, 2026Updated last month
Alternatives and similar repositories for GenExam
Users that are interested in GenExam are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ECCV'26] GRADE: Grounded Reasoning Assessment for Discipline-informed Editing☆28Apr 23, 2026Updated 2 months ago
- ⚽️🤖 Benchmarking LLMs and deep-research agents on real-world football prediction — from the tactical "who scores in minute 67" to the st…☆18Updated this week
- [CVPR 2025] Mono-InternVL: Pushing the Boundaries of Monolithic Multimodal Large Language Models with Endogenous Visual Pre-training☆109Jul 18, 2025Updated last year
- We provide TextEdit, a high-quality, multi-scenario text editing benchmark for generation models.☆20Mar 16, 2026Updated 4 months ago
- ☆16Oct 11, 2025Updated 9 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ICLR 2026] SpaCE-10: A Comprehensive Benchmark for Multimodal Large Language Models in Compositional Spatial Intelligence☆20Jan 26, 2026Updated 5 months ago
- Bridging the gap between image generation and real-world design: a benchmark for structured, multi-constraint commercial visual content g…☆20Apr 24, 2026Updated 2 months ago
- [NIPS 2025 DB Oral] Official Repository of paper: Envisioning Beyond the Pixels: Benchmarking Reasoning-Informed Visual Editing☆154May 18, 2026Updated 2 months ago
- ☆15Nov 13, 2025Updated 8 months ago
- Implementation code for the paper "Draw2Think: Harnessing Geometry Reasoning through Constraint Engine Interaction"☆17May 28, 2026Updated last month
- [ACL 2026] A benchmark for evaluating the reliability of text-to-infographic generation with curated test cases and automated question-ba…☆15Jun 8, 2026Updated last month
- Official respository for ReasonGen-R1☆75Jun 23, 2025Updated last year
- The first unified, efficient, and extensible evaluation toolkit for evaluating image generation and editing models across multiple benchm…☆50Apr 12, 2026Updated 3 months ago
- Sequential Diffusion Language Model (SDLM) enhances pre-trained autoregressive language models by adaptively determining generation lengt…☆98Dec 27, 2025Updated 6 months ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- [ECCV 2026] Official repository of "Reliable Reasoning in SVG-LLMs via Multi-Task Multi-Reward Reinforcement Learning".☆22Updated this week
- ☆21Jul 3, 2025Updated last year
- ☆44Jul 9, 2025Updated last year
- ☆16Aug 19, 2023Updated 2 years ago
- [ICLR 2026 Oral] DiffusionNFT: Online Diffusion Reinforcement with Forward Process☆974Feb 10, 2026Updated 5 months ago
- TELL: Test-time Experiential Lifelong Learning – a single LLM agent that learns from experience at test time, achieving 43.9% on ARC-AGI-…☆26Apr 29, 2026Updated 2 months ago
- [NeurIPS 2025 DB] OneIG-Bench is a meticulously designed comprehensive benchmark framework for fine-grained evaluation of T2I models acro…☆120Feb 10, 2026Updated 5 months ago
- [ECCV 2026] VKnowU: Evaluating Visual Knowledge Understanding in Multimodal LLMs☆15Feb 3, 2026Updated 5 months ago
- [ICLR 2026] Official repository of "InternSVG: Towards Unified SVG Tasks with Multimodal Large Language Models".☆120Feb 6, 2026Updated 5 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆39May 20, 2025Updated last year
- ☆53Jun 13, 2025Updated last year
- Official implementation of HPSv3: Towards Wide-Spectrum Human Preference Score (ICCV2025)☆325Dec 5, 2025Updated 7 months ago
- Doodling our way to AGI ✏️ 🖼️ 🧠☆128May 29, 2025Updated last year
- ☆44May 29, 2025Updated last year
- Official repo of "MMBench-GUI: Hierarchical Multi-Platform Evaluation Framework for GUI Agents". It can be used to evaluate a GUI agent w…☆112Sep 8, 2025Updated 10 months ago
- (ICCV2025) EEdit⚡: Rethinking the Spatial and Temporal Redundancy for Efficient Image Editing☆62Sep 17, 2025Updated 10 months ago
- AAAI2026 X2Edit: Revisiting Arbitrary-Instruction Image Editing through Self-Constructed Data and Task-Aware Representation Learning☆97Nov 21, 2025Updated 8 months ago
- Evaluation codes and data for GenEval2☆80Jan 8, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official inference code and LongText-Bench benchmark for our paper X-Omni (https://arxiv.org/pdf/2507.22058).☆426Aug 26, 2025Updated 10 months ago
- [ACM MM25] LongWriter-V: Enabling Ultra-Long and High-Fidelity Generation in Vision-Language Models☆24Mar 29, 2025Updated last year
- ☆16Mar 8, 2026Updated 4 months ago
- 【ICML2026】Reasoning to Edit: Hypothetical Instruction-Based Image Editing with Visual Reasoning☆27May 18, 2026Updated 2 months ago
- Trust Your Critic: Robust Reward Modeling and Reinforcement Learning for Faithful Image Editing and Generation☆40Mar 13, 2026Updated 4 months ago
- This is the official repository for the paper "FLUX-Reason-6M & PRISM-Bench: A Million-Scale Text-to-Image Reasoning Dataset and Comprehe…☆131Jan 29, 2026Updated 5 months ago
- Official PyTorch implementation of ProCLIP: Progressive Vision-Language Alignment via LLM-based Embedder☆25Dec 4, 2025Updated 7 months ago