PICABench: How Far Are We from Physically Realistic Image Editing?
☆39Nov 5, 2025Updated 9 months ago
Alternatives and similar repositories for PICABench
Users that are interested in PICABench are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICLR2026] Factuality Matters: When Image Generation and Editing Meet Structured Visuals☆37Nov 13, 2025Updated 9 months ago
- Official implementation of StableI2I (ICML 2026)☆20Updated this week
- 【ICML2026】Reasoning to Edit: Hypothetical Instruction-Based Image Editing with Visual Reasoning☆27May 18, 2026Updated 3 months ago
- [ICLR 2026] This is an early exploration to introduce Interleaving Reasoning to Text-to-image Generation field and achieve the SoTA bench…☆101Jan 26, 2026Updated 7 months ago
- ☆30Jun 30, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This is the official repository for the paper "FLUX-Reason-6M & PRISM-Bench: A Million-Scale Text-to-Image Reasoning Dataset and Comprehe…☆132Jan 29, 2026Updated 7 months ago
- [NIPS 25'] Evaluation code of paper "KRIS-Bench: Benchmarking Next-Level Intelligent Image Editing Models"☆47Oct 19, 2025Updated 10 months ago
- Open source community's implementation of the model from "LANGUAGE MODEL BEATS DIFFUSION — TOKENIZER IS KEY TO VISUAL GENERATION"☆15Nov 11, 2024Updated last year
- Learning from Next-Frame Prediction: Autoregressive Video Modeling Encodes Effective Representations☆22Dec 24, 2025Updated 8 months ago
- This is a simple toolkit to view and crop image patches for image/video super-resolution tasks.☆11Jan 6, 2023Updated 3 years ago
- Visual Instruction-guided Explainable Metric. Code for "Towards Explainable Metrics for Conditional Image Synthesis Evaluation" (ACL 2024…☆68Nov 19, 2024Updated last year
- Official repository of "GoT: Unleashing Reasoning Capability of Multimodal Large Language Model for Visual Generation and Editing"☆317Sep 28, 2025Updated 11 months ago
- [NeurIPS 2025 D&B🔥] ImgEdit: A Unified Image Editing Dataset and Benchmark☆335Nov 5, 2025Updated 9 months ago
- Lumina-DiMOO - An Open-Sourced Multi-Modal Large Diffusion Language Model☆1,016May 19, 2026Updated 3 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- A question-conditioned, reasoning-aware image editor designed to serve as a decoupled visual reasoning assistant for Multimodal Large Lan…☆23May 25, 2026Updated 3 months ago
- SOLACE: Improving Text-to-Image Generation with Intrinsic Self-Confidence Rewards (CVPR 2026)☆17Jun 2, 2026Updated 2 months ago
- ThinkGen: Generalized Thinking for Visual Generation☆61Dec 30, 2025Updated 8 months ago
- Unifying Image Processing as Visual Prompting Question Answering☆23Jun 17, 2024Updated 2 years ago
- [ICML2025] The code and data of Paper: Towards World Simulator: Crafting Physical Commonsense-Based Benchmark for Video Generation☆165Oct 25, 2024Updated last year
- ECCV2024:A Comparative Study of Image Restoration Networks for General Backbone Network Design☆155Jan 7, 2025Updated last year
- Fine-tune of Florence-2 for shot categorization.☆26Mar 6, 2025Updated last year
- [AAAI 2024] Decoupled Textual Embeddings for Customized Image Generation☆30Feb 29, 2024Updated 2 years ago
- GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset☆243Aug 15, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [NTIRE2024] official code for "Towards Real-world Video Face Restoration: A New Benchmark"☆31Jul 29, 2024Updated 2 years ago
- Evaluation codes and data for GenEval2☆87Jan 8, 2026Updated 7 months ago
- UICrit is a dataset containing human-generated natural language design critiques, corresponding bounding boxes for each critique, and des…☆27Nov 19, 2024Updated last year
- [ECCV 2026] Offline implementation of UniREditBench: A Unified Reasoning-based Image Editing Benchmark.☆58Aug 14, 2026Updated 2 weeks ago
- ☆189Jun 27, 2025Updated last year
- Official eval code for ROVER: Benchmarking Reciprocal Cross-Modal Reasoning for Omnimodal Generation☆28Dec 12, 2025Updated 8 months ago
- [ICLR2022] CoordX: Accelerating Implicit Neural Representation with a Split MLP Architecture☆19Feb 21, 2022Updated 4 years ago
- ReNeg: Learning Negative Embedding with Reward Guidance☆35Dec 22, 2025Updated 8 months ago
- Official Repository for "LLMs as Visual Explainers: Advancing Image Classification with Evolving Visual Descriptions"☆15Apr 20, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [ACM Multimedia 2025 Datasets Track] EditWorld: Simulating World Dynamics for Instruction-Following Image Editing☆141Aug 2, 2025Updated last year
- Official Implementation of Paper Transfer between Modalities with MetaQueries☆326Oct 12, 2025Updated 10 months ago
- [NeurIPS2025] The official implementation of MindOmni: Unleashing Reasoning Generation in Vision Language Models with RGPO☆140Oct 15, 2025Updated 10 months ago
- [NIPS 2025 DB Oral] Official Repository of paper: Envisioning Beyond the Pixels: Benchmarking Reasoning-Informed Visual Editing☆154May 18, 2026Updated 3 months ago
- MIPS OS on R3000 (course assignment for BUAA-Operating-System)☆23Mar 16, 2022Updated 4 years ago
- Motion-sensing game control system based on bone point recognition☆11Dec 1, 2023Updated 2 years ago
- ☆54Apr 11, 2025Updated last year