Official repository for the UAE paper, unified-GRPO, and unified-Bench
β166Sep 12, 2025Updated 10 months ago
Alternatives and similar repositories for UAE
Users that are interested in UAE are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICML 2026π₯] WISE: A World Knowledge-Informed Semantic Evaluation for Text-to-Image Generationβ212Jun 26, 2026Updated last month
- UniWorld: High-Resolution Semantic Encoders for Unified Visual Understanding and Generationβ886Dec 23, 2025Updated 7 months ago
- Does Understanding Inform Generation in Unified Multimodal Models? From Analysis to Path Forwardβ60Nov 27, 2025Updated 8 months ago
- Code for the paper "AsFT: Anchoring Safety During LLM Fune-Tuning Within Narrow Safety Basin".β37Jul 10, 2025Updated last year
- γCOLING 2025π₯γCode for the paper "Is Parameter Collision Hindering Continual Learning in LLMs?".β38Dec 5, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [NeurIPS 2025 D&Bπ₯] ImgEdit: A Unified Image Editing Dataset and Benchmarkβ330Nov 5, 2025Updated 8 months ago
- [ICLR 2026] Official repo of paper "Reconstruction Alignment Improves Unified Multimodal Models". Unlocking the Massive Zero-shot Potentiβ¦β411May 23, 2026Updated 2 months ago
- Here is the official code for Nature Communications "Navigating Chemical-Linguistic Sharing Space with Heterogeneous Molecular Encoding".β23May 23, 2026Updated 2 months ago
- Edit-R1: Reinforce Image Editing with Diffusion Negative-Aware Finetuning and MLLM Implicit Feedbackβ295Jan 24, 2026Updated 6 months ago
- GPT as a Monte Carlo Language Tree: A Probabilistic Perspectiveβ46Jan 18, 2025Updated last year
- β28Oct 10, 2025Updated 9 months ago
- Official Implementation of Paper Transfer between Modalities with MetaQueriesβ325Oct 12, 2025Updated 9 months ago
- β65Jul 1, 2026Updated 3 weeks ago
- γNature Computational Science 2025π₯γDeep peak property learning for efficient chiral molecules ECD spectra predictionβ51Jan 12, 2025Updated last year
- End-to-end encrypted email - Proton Mail β’ AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Official inference code and LongText-Bench benchmark for our paper X-Omni (https://arxiv.org/pdf/2507.22058).β426Aug 26, 2025Updated 11 months ago
- [AAAI26] Next Patch Predictionβ129Jan 2, 2025Updated last year
- This repository is the official implementation of "Look-Back: Implicit Visual Re-focusing in MLLM Reasoning".β100Jul 10, 2025Updated last year
- OSP-Nextβ68Jun 22, 2026Updated last month
- [CVPR 2025] Exploring the Deep Fusion of Large Language Models and Diffusion Transformers for Text-to-Image Synthesisβ140May 16, 2025Updated last year
- Official Implementation of "UniFlow: A Unified Pixel Flow Tokenizer for Visual Understanding and Generation"β143Oct 17, 2025Updated 9 months ago
- [CVPR 2025π₯] Enhancing Video VAE by Wavelet-Driven Energy Flow for Latent Video Diffusion Modelβ205May 11, 2025Updated last year
- Code release for Ming-UniVision: Joint Image Understanding and Geneation with a Continuous Unified Tokenizerβ143Oct 14, 2025Updated 9 months ago
- Official implementation of BLIP3o-Seriesβ1,663Nov 29, 2025Updated 7 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits β’ AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Native Multimodal Models are World Learnersβ1,538Dec 30, 2025Updated 6 months ago
- https://huggingface.co/datasets/multimodal-reasoning-lab/Zebra-CoTβ137Jan 30, 2026Updated 5 months ago
- [NeurIPS 2025 D&Bπ₯] OpenS2V-Nexus: A Detailed Benchmark and Million-Scale Dataset for Subject-to-Video Generationβ223May 19, 2026Updated 2 months ago
- Co-Reinforcement Learning for Unified Multimodal Understanding and Generationβ48Jul 22, 2025Updated last year
- [ICML 2026] a unified reinforcement learning toolbox for joint RL on language models and diffusion modelsβ91May 26, 2026Updated 2 months ago
- β189Jun 27, 2025Updated last year
- Awesome Unified Multimodal Modelsβ1,306Mar 24, 2026Updated 4 months ago
- GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Datasetβ243Aug 15, 2025Updated 11 months ago
- π This is a repository for organizing papers, codes and other resources related to unified multimodal models.β830Oct 10, 2025Updated 9 months ago
- Simple, predictable pricing with DigitalOcean hosting β’ AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- iFSQ & LlamaGen-REPAβ103Jan 27, 2026Updated 6 months ago
- [ICLR 2026] Uni-CoT: Towards Unified Chain-of-Thought Reasoning Across Text and Visionβ234May 31, 2026Updated last month
- Unified Multi-modal IAA Baseline and Benchmarkβ94Sep 27, 2024Updated last year
- [ICLR'25] PiCO: Peer Review in LLMs based on the Consistency Optimization, https://arxiv.org/pdf/2402.01830β36Feb 16, 2025Updated last year
- The official code for "TaxDiff: Taxonomic-Guided Diffusion Model for Protein Sequence Generation"β75Aug 23, 2024Updated last year
- Official PyTorch Implementation of "Diffusion Transformers with Representation Autoencoders"β1,978Feb 25, 2026Updated 5 months ago
- [π ICLR 2026 Oral] NextStep-1: SOTA Autogressive Image Generation with Continuous Tokens. A research project developed by the StepFunβs β¦β692Feb 27, 2026Updated 5 months ago