[ICLR 2026 Oral & ICML 2026] Generative Universal Verifier as Multimodal Meta-Reasoner
☆69May 29, 2026Updated 3 months ago
Alternatives and similar repositories for OmniVerifier
Users that are interested in OmniVerifier are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A unified framework for controllable caption generation across images, videos, and audio. Supports multi-modal inputs and customizable ca…☆54Jul 24, 2025Updated last year
- Description for MV-MATH☆15Jul 20, 2025Updated last year
- [CVPR 2026] See Less, See Right: Bi-directional Perceptual Shaping For Multimodal Reasoning☆23Jun 28, 2026Updated 2 months ago
- Official eval code for ROVER: Benchmarking Reciprocal Cross-Modal Reasoning for Omnimodal Generation☆29Dec 12, 2025Updated 9 months ago
- [NeurIPS'25] The official code of "PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning"☆30Mar 30, 2026Updated 5 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ICML 2026] The official implementation of paper "Unified Multimodal Autoregressive Modeling with Shared Context—Visual Tokenizer is Key …☆58Jul 13, 2026Updated 2 months ago
- ☆46Jan 4, 2026Updated 8 months ago
- [ICML 2026] a unified reinforcement learning toolbox for joint RL on language models and diffusion models☆97May 26, 2026Updated 4 months ago
- Official impl. of "MagicMirror: A Large-Scale Dataset and Benchmark for Fine-Grained Artifacts Assessment in Text-to-Image Generation"☆24Sep 15, 2025Updated last year
- ☆43May 9, 2026Updated 4 months ago
- Code of StyleCrafter on SDXL☆20Jun 25, 2024Updated 2 years ago
- ☆35Feb 12, 2026Updated 7 months ago
- DreamVideo-Omni: Omni-Motion Controlled Multi-Subject Video Customization with Latent Identity Reinforcement Learning☆20May 27, 2026Updated 4 months ago
- Code for "Understanding-in-Generation:Reinforcing Generative Capability of Unified Model via Infusing Understanding into Generation"☆16Nov 11, 2025Updated 10 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Codebase of 'From Denoising to Refining: A Corrective Framework for Vision-Language Diffusion Model'☆45Jun 27, 2026Updated 3 months ago
- [CVPR 2026] Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens☆300Aug 2, 2025Updated last year
- T2I-ReasonBench: Benchmarking Reasoning-Informed Text-to-Image Generation☆38Sep 16, 2025Updated last year
- The first Interleaved framework for textual reasoning within the visual generation process☆167Mar 16, 2026Updated 6 months ago
- [EMNLP2026] Does Understanding Inform Generation in Unified Multimodal Models? From Analysis to Path Forward☆61Nov 27, 2025Updated 10 months ago
- Codes for our paper "AgentMonitor: A Plug-and-Play Framework for Predictive and Secure Multi-Agent Systems"☆14Dec 13, 2024Updated last year
- [MICCAI 2024] Implicit Representation Embraces Challenging Attributes of Pulmonary Airway Tree Structures☆14Nov 13, 2024Updated last year
- [NeurIPS2025] The official implementation of MindOmni: Unleashing Reasoning Generation in Vision Language Models with RGPO☆140Oct 15, 2025Updated 11 months ago
- The official pytorch implementation of “Diffusion Model as a Noise-Aware Latent Reward Model for Step-Level Preference Optimization”.☆19May 22, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [CVPR 2025] DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval☆22Jun 23, 2025Updated last year
- [ICLR 2026] Uni-CoT: Towards Unified Chain-of-Thought Reasoning Across Text and Vision☆237May 31, 2026Updated 3 months ago
- The code repository of UniRL☆55May 30, 2025Updated last year
- 🐧 Unify-Agent: An end-to-end unified multimodal agent for faithful, knowledge-grounded image generation.☆92May 2, 2026Updated 4 months ago
- [NeurIPS 2025 DB] OneIG-Bench is a meticulously designed comprehensive benchmark framework for fine-grained evaluation of T2I models acro…☆122Feb 10, 2026Updated 7 months ago
- Official Repo for the VideoVerse☆15Mar 29, 2026Updated 5 months ago
- [ICLR 2026] RecA: visual understanding help generation through self-supervised learning☆417Sep 11, 2026Updated 2 weeks ago
- [ECCV 2026🔥] SRUM: Fine-Grained Self-Rewarding for Unified Multimodal Models☆100Nov 26, 2025Updated 10 months ago
- [CPAL 2026 oral] Offical implementation of "ROSE: Reordered SparseGPT for More Accurate One-Shot Large Language Models Pruning”☆17Jul 31, 2026Updated last month
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [CVPR 2026] Official codes of "Monet: Reasoning in Latent Visual Space Beyond Image and Language"☆225Mar 19, 2026Updated 6 months ago
- [2025-TMLR] A Survey on the Honesty of Large Language Models☆67Dec 8, 2024Updated last year
- [ICLR 2026] High-Fidelity Visual Reasoning on Structured Images☆36Jul 17, 2026Updated 2 months ago
- 📖 This is a repository for organizing papers, codes and other resources related to visual tokenizers.☆18Updated this week
- [ICLR 2026] This is an early exploration to introduce Interleaving Reasoning to Text-to-image Generation field and achieve the SoTA bench…☆100Jan 26, 2026Updated 8 months ago
- LayoutDiT: Exploring Content-Graphic Balance in Layout Generation with Diffusion Transformer☆50Jan 6, 2026Updated 8 months ago
- An automated workflow for composing, rendering, and retargeting MMD assets.☆16Feb 23, 2026Updated 7 months ago