[ICLR2026] Factuality Matters: When Image Generation and Editing Meet Structured Visuals
☆37Nov 13, 2025Updated 8 months ago
Alternatives and similar repositories for Structured-Visuals
Users that are interested in Structured-Visuals are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- PICABench: How Far Are We from Physically Realistic Image Editing?☆39Nov 5, 2025Updated 9 months ago
- Official Repository for "LLMs as Visual Explainers: Advancing Image Classification with Evolving Visual Descriptions"☆15Apr 20, 2025Updated last year
- [ECCV'26] GRADE: Grounded Reasoning Assessment for Discipline-informed Editing☆29Apr 23, 2026Updated 3 months ago
- [NIPS 25'] Evaluation code of paper "KRIS-Bench: Benchmarking Next-Level Intelligent Image Editing Models"☆46Oct 19, 2025Updated 9 months ago
- RISE-Video: Can Video Generators Decode Implicit World Rules?☆28Mar 26, 2026Updated 4 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [NIPS 2025 DB Oral] Official Repository of paper: Envisioning Beyond the Pixels: Benchmarking Reasoning-Informed Visual Editing☆155May 18, 2026Updated 2 months ago
- Bridging the gap between image generation and real-world design: a benchmark for structured, multi-constraint commercial visual content g…☆21Apr 24, 2026Updated 3 months ago
- The official pytorch implementation of our paper "Motion meets Attention: Video Motion Prompts" (The 16th Asian Conference on Machine Lea…☆15Oct 30, 2024Updated last year
- Code release for "Generative Modeling of Weights: Generalization or Memorization?"☆23Apr 9, 2026Updated 4 months ago
- [ICCV2025] VEGGIE: Instructional Editing and Reasoning Video Concepts with Grounded Generation☆34Aug 18, 2025Updated 11 months ago
- CoCo: Code as CoT for Text-to-Image Preview and Rare Concept Generation☆55Updated this week
- [CVPR 2025 Oral] VideoEspresso: A Large-Scale Chain-of-Thought Dataset for Fine-Grained Video Reasoning via Core Frame Selection☆141Jul 28, 2025Updated last year
- Math-VR Benchmark & CodePlot-CoT: Mathematical Visual Reasoning by Thinking with Code-Driven Images☆63Nov 4, 2025Updated 9 months ago
- [ICCV 2025] Official Implementation of RefEdit: A Benchmark and Method for Improving Instruction-based Image Editing Model for Referring …☆20Jun 27, 2025Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆56Nov 26, 2024Updated last year
- Offical Repository for Paper: DraCo: Draft as CoT for Text-to-Image Preview and Rare Concept Generation☆19Dec 7, 2025Updated 8 months ago
- [ACL 2026 Findings, ICCV 2025 Workshop Outstanding Paper Award] VChain: Chain-of-Visual-Thought for Reasoning in Video Generation☆120Apr 8, 2026Updated 4 months ago
- Uni-ViGU: Towards Unified Video Generation and Understanding via A Diffusion-Based Video Generator☆34Apr 15, 2026Updated 3 months ago
- [ECCV 2026] Offline implementation of UniREditBench: A Unified Reasoning-based Image Editing Benchmark.☆58Jun 21, 2026Updated last month
- Code for "Weakly-supervised Fingerspelling Recognition in British Sign Language Videos", BMVC 2022.☆12Jun 22, 2023Updated 3 years ago
- A curated list of papers and resources for text-to-image evaluation.☆30Sep 6, 2023Updated 2 years ago
- EditReward: A Human-Aligned Reward Model for Instruction-Guided Image Editing [ICLR 2026]☆158Jul 26, 2026Updated 2 weeks ago
- Montreal Forced Aligner for Vietnamese☆15Oct 23, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Interpreting Chest X-rays Like a Radiologist: A Benchmark with Clinical Reasoning, release the dataset and the model weight☆13May 26, 2025Updated last year
- [NeurIPS 2025 D&B🔥] ImgEdit: A Unified Image Editing Dataset and Benchmark☆330Nov 5, 2025Updated 9 months ago
- ☆30Jun 30, 2025Updated last year
- ☆10Nov 27, 2024Updated last year
- A library for multilingual word, phrase and sentence segmentation.☆16Jul 24, 2026Updated 2 weeks ago
- ☆19May 6, 2024Updated 2 years ago
- [ICCV 2025] Scaling Inference-Time Optimization for Text-to-Image Diffusion Models via Reflection Tuning☆220Nov 5, 2025Updated 9 months ago
- Minimal Differentiable Image Reward Functions☆119Mar 30, 2026Updated 4 months ago
- ☆14Jul 17, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [Remote Sensing 2026] Co-Training Vision Language Models for Remote Sensing Multi-task Learning☆38Jul 8, 2026Updated last month
- [ACM Multimedia 2025 Datasets Track] EditWorld: Simulating World Dynamics for Instruction-Following Image Editing☆142Aug 2, 2025Updated last year
- [EMNLP 2024] IFCap: Image-like Retrieval and Frequency-based Entity Filtering for Zero-shot Captioning☆15May 13, 2025Updated last year
- Pixel Parsing. A reproduction of OCR-free end-to-end document understanding models with open data☆24Jul 30, 2024Updated 2 years ago
- Official repository for the FIRM Reward series☆41Updated this week
- Structured Decomposition and Conditional Skill Orchestration for Complex Image Generation☆32Jul 30, 2026Updated last week
- [NeurIPS 2025] T2I-R1: Reinforcing Image Generation with Collaborative Semantic-level and Token-level CoT☆433Sep 18, 2025Updated 10 months ago