[CVPR'26] UniGame code implementation
☆20Apr 21, 2026Updated 3 months ago
Alternatives and similar repositories for UniGame
Users that are interested in UniGame are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆18Mar 14, 2026Updated 4 months ago
- Co-Reinforcement Learning for Unified Multimodal Understanding and Generation☆48Jul 22, 2025Updated last year
- Echo: "Constantly Improving Image Models Need Constantly Improving Benchmarks" (ICLR 2026)☆20Jan 29, 2026Updated 6 months ago
- ☆43May 9, 2026Updated 3 months ago
- A general large multimodal model for 4D scene understanding☆16Jul 31, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [ICML 2026] The official implementation of paper "Generation Enhances Understanding in Unified Multimodal Models via Multi-Representation…☆91May 25, 2026Updated 2 months ago
- ☆45Jan 4, 2026Updated 7 months ago
- [ICLR 2026] Official repo of paper "Reconstruction Alignment Improves Unified Multimodal Models". Unlocking the Massive Zero-shot Potenti…☆411Updated this week
- [ICCV2025]Code Release of Harmonizing Visual Representations for Unified Multimodal Understanding and Generation☆192May 21, 2025Updated last year
- Matlab code for "Joint Projection Learning and Tensor Decomposition Based Incomplete Multi-view Clustering".☆10Jun 5, 2023Updated 3 years ago
- [IMWUT/UbiComp 2024] Optimization-Free Test-Time Adaptation for Cross-Person Activity Recognition☆23Oct 28, 2023Updated 2 years ago
- [ICLR26] Understanding VS. Generation: Navigating Optimization Dilemma in Multimodal Models☆27May 6, 2026Updated 3 months ago
- Official eval code for ROVER: Benchmarking Reciprocal Cross-Modal Reasoning for Omnimodal Generation☆27Dec 12, 2025Updated 7 months ago
- Notion LifeOS PARA system — agent skill for Claude Code, OpenClaw, Codex and more☆23Mar 24, 2026Updated 4 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- The official implementation of COOPER: A Unified Model for Cooperative Perception and Reasoning in Spatial Intelligence.☆38Jul 1, 2026Updated last month
- FaceShield: Explainable Face Anti-Spoofing with Multimodal Large Language Models☆14Jul 14, 2026Updated 3 weeks ago
- ☆189Jun 27, 2025Updated last year
- [ICLR 2026] This is an early exploration to introduce Interleaving Reasoning to Text-to-image Generation field and achieve the SoTA bench…☆101Jan 26, 2026Updated 6 months ago
- ☆16Jun 1, 2026Updated 2 months ago
- my attempt at implementing the DiffEdit paper (WIP)☆16Oct 30, 2022Updated 3 years ago
- Code for the experiments and websites of the paper "Same Task, Different Circuits"☆37Jul 21, 2026Updated 2 weeks ago
- End-to-end Deep Linear Discriminant Analysis in Pytorch.☆18Mar 22, 2021Updated 5 years ago
- Official pytorch implementation of "RITUAL: Random Image Transformations as a Universal Anti-hallucination Lever in Large Vision Language…☆14Dec 16, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆93Apr 29, 2026Updated 3 months ago
- Code for experiments on transformers using Markovian data.☆22Nov 22, 2024Updated last year
- The MIG benchmark of CVPR2024 MIGC☆15Mar 3, 2024Updated 2 years ago
- A unified multimodal model toolkit☆560Jul 29, 2026Updated last week
- [ICLR 2026] ContextGen: Contextual Layout Anchoring for Identity-Consistent Multi-Instance Generation☆87Apr 19, 2026Updated 3 months ago
- ThinkJEPA: Empowering Latent World Models with Large Vision-Language Reasoning Model☆49Jul 19, 2026Updated 3 weeks ago
- Does Understanding Inform Generation in Unified Multimodal Models? From Analysis to Path Forward☆60Nov 27, 2025Updated 8 months ago
- UniDoc-RL: Unified Document Understanding with Reinforcement Learning☆17May 21, 2026Updated 2 months ago
- [AAAI 2026] PerTouch: VLM-Driven Agent for Personalized and Semantic Image Retouching☆27Feb 4, 2026Updated 6 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- A controllable and interactive simulation framework for vision research.☆16May 25, 2026Updated 2 months ago
- StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation☆43Jun 6, 2025Updated last year
- [NeurIPS2025] The official implementation of MindOmni: Unleashing Reasoning Generation in Vision Language Models with RGPO☆139Oct 15, 2025Updated 9 months ago
- https://huggingface.co/datasets/multimodal-reasoning-lab/Zebra-CoT☆137Jan 30, 2026Updated 6 months ago
- Visual Generation Tuning☆101Apr 16, 2026Updated 3 months ago
- Officail Implementation for "Unified Diffusion-Based Rigid and Non-Rigid Editing with Text and Image Guidance"☆19Jan 25, 2024Updated 2 years ago
- The code repository of UniRL☆54May 30, 2025Updated last year