[EMNLP2026] Does Understanding Inform Generation in Unified Multimodal Models? From Analysis to Path Forward
β61Nov 27, 2025Updated 10 months ago
Alternatives and similar repositories for UniSandBox
Users that are interested in UniSandBox are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICML 2026π₯] WISE: A World Knowledge-Informed Semantic Evaluation for Text-to-Image Generationβ220Aug 3, 2026Updated last month
- [ECCV 2026π₯] SRUM: Fine-Grained Self-Rewarding for Unified Multimodal Modelsβ100Nov 26, 2025Updated 10 months ago
- β68Aug 7, 2026Updated last month
- Evaluation codes and data for GenEval2β91Jan 8, 2026Updated 8 months ago
- Official eval code for ROVER: Benchmarking Reciprocal Cross-Modal Reasoning for Omnimodal Generationβ29Dec 12, 2025Updated 9 months ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- OSP-Nextβ68Jun 22, 2026Updated 3 months ago
- γCOLING 2025π₯γCode for the paper "Is Parameter Collision Hindering Continual Learning in LLMs?".β40Dec 5, 2024Updated last year
- [NeurIPS 2025 D&Bπ₯] ImgEdit: A Unified Image Editing Dataset and Benchmarkβ338Nov 5, 2025Updated 10 months ago
- [CVPR 2026π₯] Enhancing Spatial Understanding in Image Generation via Reward Modelingβ86Mar 2, 2026Updated 6 months ago
- Official repository for the UAE paper, unified-GRPO, and unified-Benchβ166Sep 12, 2025Updated last year
- Code for the paper "AsFT: Anchoring Safety During LLM Fune-Tuning Within Narrow Safety Basin".β38Jul 10, 2025Updated last year
- [ACM MM 2025] HoloTime: Taming Video Diffusion Models for Panoramic 4D Scene Generationβ160Sep 4, 2025Updated last year
- π This is a repository for organizing papers, codes, and other resources related to unified multimodal models.β366Jan 8, 2026Updated 8 months ago
- [π₯ACMMM 2025] mplemetation of "E-4DGS: High-Fidelity Dynamic Reconstruction from the Multi-view Event Cameras"β19Aug 14, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- This repository is the official implementation of "Look-Back: Implicit Visual Re-focusing in MLLM Reasoning".β98Jul 10, 2025Updated last year
- [ICLR 2026] This is an early exploration to introduce Interleaving Reasoning to Text-to-image Generation field and achieve the SoTA benchβ¦β100Jan 26, 2026Updated 8 months ago
- β43May 9, 2026Updated 4 months ago
- LLM Reasoning Benchmark & Chain-of-Thoughts Dataset for Chemistryβ57Oct 9, 2025Updated 11 months ago
- UniWorld: High-Resolution Semantic Encoders for Unified Visual Understanding and Generationβ891Dec 23, 2025Updated 9 months ago
- Code for "Understanding-in-Generation:Reinforcing Generative Capability of Unified Model via Infusing Understanding into Generation"β16Nov 11, 2025Updated 10 months ago
- [ICLR 2026 π₯ ] Official implementation of "UniLiP: Adapting CLIP for Unified Multimodal Understanding, Generation and Editing"β152Jan 26, 2026Updated 8 months ago
- Offical Repository for Paper: DraCo: Draft as CoT for Text-to-Image Preview and Rare Concept Generationβ19Dec 7, 2025Updated 9 months ago
- [ICLR 2026] Uni-CoT: Towards Unified Chain-of-Thought Reasoning Across Text and Visionβ237May 31, 2026Updated 3 months ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Echo: "Constantly Improving Image Models Need Constantly Improving Benchmarks" (ICLR 2026)β20Updated this week
- Edit-R1: Reinforce Image Editing with Diffusion Negative-Aware Finetuning and MLLM Implicit Feedbackβ299Jan 24, 2026Updated 8 months ago
- [NeurIPS 2025 D&Bπ₯] OpenS2V-Nexus: A Detailed Benchmark and Million-Scale Dataset for Subject-to-Video Generationβ231May 19, 2026Updated 4 months ago
- β18Mar 14, 2026Updated 6 months ago
- LLMBind: A Unified Modality-Task Integration Frameworkβ19Jun 16, 2024Updated 2 years ago
- [NeurIPS 2024] Artemis: Towards Referential Understanding in Complex Videosβ27Apr 8, 2025Updated last year
- [ICLR 2026] RecA: visual understanding help generation through self-supervised learningβ417Sep 11, 2026Updated 2 weeks ago
- γNature Computational Science 2025π₯γDeep peak property learning for efficient chiral molecules ECD spectra predictionβ52Jan 12, 2025Updated last year
- https://huggingface.co/datasets/multimodal-reasoning-lab/Zebra-CoTβ140Jan 30, 2026Updated 7 months ago
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- iFSQ & LlamaGen-REPAβ106Jan 27, 2026Updated 8 months ago
- [AAAI26] Next Patch Predictionβ129Jan 2, 2025Updated last year
- [ACL2026 oral] Uni-MMMU : A Massive Multi-discipline Multimodal Unified Benchmarkβ27Apr 13, 2026Updated 5 months ago
- [NeurIPS 2025] Official repository of the paper "Unlocking Aha Moments via Reinforcement Learning: Advancing Collaborative Visual Comprehβ¦β22Sep 27, 2025Updated last year
- CoCo: Code as CoT for Text-to-Image Preview and Rare Concept Generationβ56Aug 7, 2026Updated last month
- [ICCV2025] TokenBridge: Bridging Continuous and Discrete Tokens for Autoregressive Visual Generation. https://yuqingwang1029.github.io/Toβ¦β161Jul 24, 2025Updated last year
- GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Datasetβ245Aug 15, 2025Updated last year