Bridging the gap between image generation and real-world design: a benchmark for structured, multi-constraint commercial visual content generation.
☆24Apr 24, 2026Updated 5 months ago
Alternatives and similar repositories for BizGenEval
Users that are interested in BizGenEval are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Build coherent and visually polished multimodal webpages with hierarchical planning, AIGC tools, and iterative reflection.☆18May 17, 2026Updated 4 months ago
- [ECCV'26] GRADE: Grounded Reasoning Assessment for Discipline-informed Editing☆29Sep 24, 2026Updated last week
- ☆19Jun 2, 2026Updated 4 months ago
- [ECCV'26] Code repo for "EvoTok: A Unified Image Tokenizer via Residual Latent Evolution for Visual Understanding and Generation"☆23Jun 18, 2026Updated 3 months ago
- [ICML26] AVGen-Bench is a task-driven benchmark for multi-granular evaluation of Text-to-Audio-Video (T2AV) generation.☆31Jul 2, 2026Updated 3 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Official repository for the FIRM Reward series☆49Aug 25, 2026Updated last month
- RISE-Video: Can Video Generators Decode Implicit World Rules?☆28Mar 26, 2026Updated 6 months ago
- Official repo for Directional Self-supervised Learning for Heavy Image Augmentations [CVPR2022]☆12Jun 29, 2022Updated 4 years ago
- [CVPR'26] AdapTok: Learning Adaptive and Temporally Causal Video Tokenization in a 1D Latent Space☆31Mar 15, 2026Updated 6 months ago
- CodexOpt: Optize your Agents.MD and Skills for Codex with GEPA☆21May 26, 2026Updated 4 months ago
- [ICLR 2026] SpaCE-10: A Comprehensive Benchmark for Multimodal Large Language Models in Compositional Spatial Intelligence☆21Jan 26, 2026Updated 8 months ago
- 本项目将《动手学 深度学习》(Dive into Deep Learning)原书中的MXNet实现的部分代码改成tensorflow2.0的实现☆11Jul 29, 2020Updated 6 years ago
- This repository catalogs cutting-edge research papers, practical tools, datasets, and learning materials for AI-powered SVG generation, p…☆19Dec 19, 2025Updated 9 months ago
- [ICLR2026] Factuality Matters: When Image Generation and Editing Meet Structured Visuals☆37Nov 13, 2025Updated 10 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Interpreting Chest X-rays Like a Radiologist: A Benchmark with Clinical Reasoning, release the dataset and the model weight☆13May 26, 2025Updated last year
- InternVL-U is a 4B-parameter unified multimodal model (UMM) that brings multimodal understanding, reasoning, image generation, image edit…☆296Mar 21, 2026Updated 6 months ago
- [ICML 2026] GenExam: A Multidisciplinary Text-to-Image Exam☆72Sep 26, 2026Updated last week
- Official Repository of paper MM-HELIX: Boosting Multimodal Long-Chain Reflective Reasoning with Holistic Platform and Adaptive Hybrid Pol…☆68Jan 26, 2026Updated 8 months ago
- ☆40May 20, 2025Updated last year
- ☆18Oct 5, 2024Updated 2 years ago
- [ECCV 2026] Official repository of "Reliable Reasoning in SVG-LLMs via Multi-Task Multi-Reward Reinforcement Learning".☆30Jul 17, 2026Updated 2 months ago
- This is a repository for awesome any2any work collection.☆32Updated this week
- OmniRefiner: Reinforcement-Guided Local Diffusion Refinement☆18Nov 26, 2025Updated 10 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- A segmentation project based on aniseg, trained on yolov8-seg☆12Jul 15, 2023Updated 3 years ago
- ☆10Jul 12, 2022Updated 4 years ago
- [CVPR'25] Official repo of "Point2RBox-v2:Rethinking Point-supervised Oriented Object Detection with Spatial Layout Among Instances"☆46Aug 3, 2026Updated 2 months ago
- ☆17Oct 24, 2024Updated last year
- ☆12Jul 4, 2024Updated 2 years ago
- Stepping VLMs onto the Court: Benchmarking Spatial Intelligence in Sports☆70Mar 15, 2026Updated 6 months ago
- The official GitHub page for the survey paper "From Models to Systems: A Comprehensive Survey of Efficient Multimodal Learning". And this…☆23Sep 18, 2026Updated 2 weeks ago
- ☆11Oct 19, 2022Updated 3 years ago
- ☆12Feb 2, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- The official repo of CrossEarth-SAR, a sar-centric and billion-scale geospatial foundation model for cross-domain semantic segmentation☆48Mar 18, 2026Updated 6 months ago
- Implementation of Em_Garde: a proposal-retrieval framework for streaming video understanding☆34Jun 24, 2026Updated 3 months ago
- Official repository for "Structure-Enhanced Pop Music Generation via Harmony-Aware Learning", ACM MM 2022.☆14Mar 22, 2023Updated 3 years ago
- Aerial Detection Toolbox☆11Jan 18, 2023Updated 3 years ago
- JoVA: Unified Multimodal Learning for Joint Video-Audio Generation☆34Dec 22, 2025Updated 9 months ago
- Semantic-decoupled Spatial Partition Guided Point-supervised Oriented Object Detection☆13Jul 7, 2026Updated 2 months ago
- ☆12Apr 30, 2025Updated last year