[CVPR 2026π₯] Enhancing Spatial Understanding in Image Generation via Reward Modeling
β86Mar 2, 2026Updated 5 months ago
Alternatives and similar repositories for SpatialT2I
Users that are interested in SpatialT2I are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- γCOLING 2025π₯γCode for the paper "Is Parameter Collision Hindering Continual Learning in LLMs?".β39Dec 5, 2024Updated last year
- OSP-Nextβ68Jun 22, 2026Updated last month
- [AAAI 2026 π₯] Official implementation of "NeuralGS: Bridging Neural Fields and 3D Gaussian Splatting for Compact 3D Representation"β184Aug 14, 2025Updated last year
- [CVPR 2026] Adaptive Spectral Feature Forecasting for Diffusion Sampling Accelerationβ132Apr 30, 2026Updated 3 months ago
- Does Understanding Inform Generation in Unified Multimodal Models? From Analysis to Path Forwardβ60Nov 27, 2025Updated 8 months ago
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- β20Sep 17, 2024Updated last year
- MADAv2: Advanced Multi-Anchor Based Active Domain Adaptation Segmentationβ25Jul 8, 2023Updated 3 years ago
- γNature Computational Science 2025π₯γDeep peak property learning for efficient chiral molecules ECD spectra predictionβ51Jan 12, 2025Updated last year
- LLM Reasoning Benchmark & Chain-of-Thoughts Dataset for Chemistryβ57Oct 9, 2025Updated 10 months ago
- iFSQ & LlamaGen-REPAβ105Jan 27, 2026Updated 6 months ago
- β60Mar 16, 2025Updated last year
- [ICML 2026π₯] WISE: A World Knowledge-Informed Semantic Evaluation for Text-to-Image Generationβ216Aug 3, 2026Updated last week
- Code repo for EffectMaker: Unifying Reasoning and Generation for Customized Visual Effect Creationβ42Mar 6, 2026Updated 5 months ago
- [π₯ACMMM 2025] mplemetation of "E-4DGS: High-Fidelity Dynamic Reconstruction from the Multi-view Event Cameras"β19Aug 14, 2025Updated last year
- Open source password manager - Proton Pass β’ AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- [CVPR 2026] Offical implementation of the paper "HiFi-Inpaint: Towards High-Fidelity Reference-Based Inpainting for Generating Detail-Preβ¦β109Jun 7, 2026Updated 2 months ago
- [ICML 2026π₯]Rethinking Video Generation Model for the Embodied Worldβ93Jun 1, 2026Updated 2 months ago
- [ECCV 2026] CustomX: Unified Character, Action, and Scene Customization in Video World Modelsβ96Jun 25, 2026Updated last month
- GPT as a Monte Carlo Language Tree: A Probabilistic Perspectiveβ46Jan 18, 2025Updated last year
- The official code for "TaxDiff: Taxonomic-Guided Diffusion Model for Protein Sequence Generation"β75Aug 23, 2024Updated last year
- [ICML 2026] a unified reinforcement learning toolbox for joint RL on language models and diffusion modelsβ95May 26, 2026Updated 2 months ago
- Code for the paper "AsFT: Anchoring Safety During LLM Fune-Tuning Within Narrow Safety Basin".β37Jul 10, 2025Updated last year
- A Unified Visual Generator with Interleaved OmniModal Contextβ233Mar 5, 2026Updated 5 months ago
- A unified and fully open-source framework for instruction-guided and reference-guided video editing using natural language.β315May 13, 2026Updated 3 months ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Helios: Real Real-Time Long Video Generation Modelβ2,050Jul 28, 2026Updated 2 weeks ago
- [CVPR 2025π₯] Enhancing Video VAE by Wavelet-Driven Energy Flow for Latent Video Diffusion Modelβ206May 11, 2025Updated last year
- β104Mar 13, 2026Updated 5 months ago
- [NeurIPS 2025 D&Bπ₯] Implementation of "GS2E: Gaussian Splatting is an Effective Data Generator for Event Stream Generation"β20Jun 1, 2025Updated last year
- Implementation of <Streaming Autoregressive Video Generation via Diagonal Distillation> in ICLR 2026β131Mar 18, 2026Updated 4 months ago
- HY-WU (Part I): An Extensible Functional Neural Memory Framework and An Instantiation in Text-Guided Image Editingβ297Mar 18, 2026Updated 4 months ago
- EditReward: A Human-Aligned Reward Model for Instruction-Guided Image Editing [ICLR 2026]β158Jul 26, 2026Updated 2 weeks ago
- [ICLR'25] PiCO: Peer Review in LLMs based on the Consistency Optimization, https://arxiv.org/pdf/2402.01830β36Feb 16, 2025Updated last year
- β95Mar 14, 2026Updated 5 months ago
- Managed Database hosting by DigitalOcean β’ AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [ACM MM 2025] HoloTime: Taming Video Diffusion Models for Panoramic 4D Scene Generationβ159Sep 4, 2025Updated 11 months ago
- RationalRewards: a reasoning reward model for diffusion RL and test-time prompt tuningβ58Jun 4, 2026Updated 2 months ago
- GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Datasetβ243Aug 15, 2025Updated 11 months ago
- This repository is the official implementation of "Look-Back: Implicit Visual Re-focusing in MLLM Reasoning".β99Jul 10, 2025Updated last year
- Code for paper "Rethinking Text-based Protein Understanding: Retrieval or LLM?"β20Oct 7, 2025Updated 10 months ago
- [CVPR 2026] Spatio-Temporal Autoregressive 4K 360Β° Video Generation from Perspective Videoβ129Mar 24, 2026Updated 4 months ago
- [Official Code] PermaVid: Consistent Video Generation Across Edits via Disentangled Context Memoryβ43Jun 17, 2026Updated last month