(NeurIPS 2025 D&B Track) OverLayBench: A Benchmark for Layout-to-Image Generation with Dense Overlaps
☆27May 4, 2026Updated 3 months ago
Alternatives and similar repositories for OverLayBench
Users that are interested in OverLayBench are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- (ICCV 2025) DepR: Depth Guided Single-view Scene Reconstruction with Instance-level Diffusion☆24Mar 7, 2026Updated 5 months ago
- The MIG benchmark of CVPR2024 MIGC☆15Mar 3, 2024Updated 2 years ago
- Code release for our paper "Divide and Conquer: Language Models can Plan and Self-Correct for Compositional Text-to-Image Generation".☆18Jan 30, 2024Updated 2 years ago
- [ICCV 2025] CreatiLayout: Siamese Multimodal Diffusion Transformer for Creative Layout-to-Image Generation☆135Aug 6, 2025Updated last year
- VideoNSA: Native Sparse Attention Scales Video Understanding☆88Nov 16, 2025Updated 8 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Latest Advances on Autoregressive Visual Models.📖☆28Mar 15, 2025Updated last year
- RoboVIP: Multi-View Video Generation with Visual Identity Prompting Augments Robot Manipulation☆30Apr 3, 2026Updated 4 months ago
- 旋转倒立摆matlab物理模型仿真☆11Jun 5, 2022Updated 4 years ago
- [ICML 2026] Official PyTorch implementation of paper “CoCoEdit: Content-Consistent Image Editing via Region Regularized Reinforcement Lea…☆26Jun 14, 2026Updated last month
- Unified layout planning and image generation, ICCV2025☆46Jan 19, 2026Updated 6 months ago
- [ICCV 2025] DreamRenderer: Taming Multi-Instance Attribute Control in Large-Scale Text-to-Image Models (official implement)☆156May 21, 2025Updated last year
- implement the HOG(histogram of Gradient) feature extraction in matlab.☆10Nov 5, 2015Updated 10 years ago
- [ICLR 2026] ContextGen: Contextual Layout Anchoring for Identity-Consistent Multi-Instance Generation☆87Apr 19, 2026Updated 3 months ago
- Code for "Memory Forcing: Spatio-Temporal Memory for Consistent Scene Generation on Minecraft"☆19Oct 11, 2025Updated 9 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Human Detection in images using HoG + SVM and 3D shape extraction using fringe projection☆11Apr 24, 2019Updated 7 years ago
- [CVPR 2025] Science-T2I: Addressing Scientific Illusions in Image Synthesis☆62Mar 31, 2026Updated 4 months ago
- [ICCV 2025] UniVerse: Unleashing the Scene Prior of Video Diffusion Models for Robust Radiance Field Reconstruction☆27Oct 3, 2025Updated 10 months ago
- [KDD 2023] code for "Test accuracy vs. generalization gap: model selection in NLP without accessing training or testing data" https://arx…☆12Oct 17, 2022Updated 3 years ago
- ☆13Apr 5, 2020Updated 6 years ago
- (2024) The Official Repository of Paper "SISP: A Benchmark Dataset for Fine-grained Ship Instance Segmentation in Panchromatic Satellite …☆15Feb 7, 2024Updated 2 years ago
- [NeurIPS 2024] Official Implementation of GrounDiT☆58Dec 12, 2024Updated last year
- The official implementation of the paper "Self-Updatable Large Language Models by Integrating Context into Model Parameters"☆15May 18, 2025Updated last year
- 安全检测系统-多目标识别(YOLOv5)和人脸识别(Facenet)快速部署系统。功能上:本项目使用YOLOv5实现多目标识别,使用Facenet实现人脸识别,最终需要人脸和此人应具备的多目标同时满足才能通过安全检测,部署上:使用pyqt5实现前端可视化,在前端页面运行YO…☆17Oct 28, 2025Updated 9 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code for "Generalize or Detect? Towards Robust Semantic Segmentation Under Multiple Distribution Shift". (NeurIPS 24)☆19Apr 21, 2025Updated last year
- ☆16Sep 6, 2024Updated last year
- The official code of paper WristWorld.☆31Nov 8, 2025Updated 9 months ago
- A framework for camera-controllable image editing using unified geometric guidance and video models.☆65Jun 25, 2026Updated last month
- 基于HOG和SVM的人脸口罩识别算法-使用Matlab☆16Oct 29, 2021Updated 4 years ago
- High-Resolution Visual Reasoning via Multi-Turn Grounding-Based Reinforcement Learning☆55Jul 23, 2025Updated last year
- ☆18Jan 17, 2024Updated 2 years ago
- A collection of diffusion models based on FLUX/DiT for image/video generation, editing, reconstruction, inpainting .etc.☆86Jun 20, 2025Updated last year
- ☆17Jul 26, 2023Updated 3 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Official repository of "GoT: Unleashing Reasoning Capability of Multimodal Large Language Model for Visual Generation and Editing"☆317Sep 28, 2025Updated 10 months ago
- Public release of the code for "Accelerating Vision Transformers with Adaptive Patches"☆117May 6, 2026Updated 3 months ago
- Unsupervised domain adaptation for cross-modality liver segmentation via joint adversarial learning and self-learning☆16Feb 10, 2023Updated 3 years ago
- LogiCity@NeurIPS'24, D&B track. A multi-agent inductive learning environment for "abstractions".☆27Jun 10, 2025Updated last year
- 基于YOLO11-pose 的运动计数系统,支持俯卧撑、深蹲、引体向上、仰卧起坐计数。界面用pyside6开发☆22Feb 25, 2025Updated last year
- EMNLP 2024 "Re-reading improves reasoning in large language models". Simply repeating the question to get bidirectional understanding for…☆30Dec 10, 2024Updated last year
- Codebase for the paper Aerial Diffusion: Text Guided Ground-to-Aerial View Translation from a Single Image using Diffusion Models☆13Oct 3, 2023Updated 2 years ago