[CVPR 2026π₯] Enhancing Spatial Understanding in Image Generation via Reward Modeling
β85Mar 2, 2026Updated 4 months ago
Alternatives and similar repositories for SpatialT2I
Users that are interested in SpatialT2I are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- γCOLING 2025π₯γCode for the paper "Is Parameter Collision Hindering Continual Learning in LLMs?".β38Dec 5, 2024Updated last year
- OSP-Nextβ66Jun 22, 2026Updated last month
- [AAAI 2026 π₯] Official implementation of "NeuralGS: Bridging Neural Fields and 3D Gaussian Splatting for Compact 3D Representation"β183Aug 14, 2025Updated 11 months ago
- [CVPR 2026] Adaptive Spectral Feature Forecasting for Diffusion Sampling Accelerationβ126Apr 30, 2026Updated 2 months ago
- Does Understanding Inform Generation in Unified Multimodal Models? From Analysis to Path Forwardβ60Nov 27, 2025Updated 7 months ago
- End-to-end encrypted cloud storage - Proton Drive β’ AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- β20Sep 17, 2024Updated last year
- MADAv2: Advanced Multi-Anchor Based Active Domain Adaptation Segmentationβ25Jul 8, 2023Updated 3 years ago
- γNature Computational Science 2025π₯γDeep peak property learning for efficient chiral molecules ECD spectra predictionβ51Jan 12, 2025Updated last year
- LLM Reasoning Benchmark & Chain-of-Thoughts Dataset for Chemistryβ55Oct 9, 2025Updated 9 months ago
- iFSQ & LlamaGen-REPAβ102Jan 27, 2026Updated 5 months ago
- β60Mar 16, 2025Updated last year
- [ICML 2026π₯] WISE: A World Knowledge-Informed Semantic Evaluation for Text-to-Image Generationβ212Jun 26, 2026Updated 3 weeks ago
- Code repo for EffectMaker: Unifying Reasoning and Generation for Customized Visual Effect Creationβ42Mar 6, 2026Updated 4 months ago
- [π₯ACMMM 2025] mplemetation of "E-4DGS: High-Fidelity Dynamic Reconstruction from the Multi-view Event Cameras"β19Aug 14, 2025Updated 11 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [CVPR 2026] Offical implementation of the paper "HiFi-Inpaint: Towards High-Fidelity Reference-Based Inpainting for Generating Detail-Preβ¦β105Jun 7, 2026Updated last month
- [ICML 2026π₯]Rethinking Video Generation Model for the Embodied Worldβ87Jun 1, 2026Updated last month
- [ECCV 2026] CustomX: Unified Character, Action, and Scene Customization in Video World Modelsβ96Jun 25, 2026Updated last month
- GPT as a Monte Carlo Language Tree: A Probabilistic Perspectiveβ46Jan 18, 2025Updated last year
- The official code for "TaxDiff: Taxonomic-Guided Diffusion Model for Protein Sequence Generation"β75Aug 23, 2024Updated last year
- [ICML 2026] a unified reinforcement learning toolbox for joint RL on language models and diffusion modelsβ91May 26, 2026Updated last month
- Code for the paper "AsFT: Anchoring Safety During LLM Fune-Tuning Within Narrow Safety Basin".β37Jul 10, 2025Updated last year
- A Unified Visual Generator with Interleaved OmniModal Contextβ232Mar 5, 2026Updated 4 months ago
- A unified and fully open-source framework for instruction-guided and reference-guided video editing using natural language.β306May 13, 2026Updated 2 months ago
- Managed Kubernetes at scale on DigitalOcean β’ AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Helios: Real Real-Time Long Video Generation Modelβ1,999Jun 10, 2026Updated last month
- [CVPR 2025π₯] Enhancing Video VAE by Wavelet-Driven Energy Flow for Latent Video Diffusion Modelβ205May 11, 2025Updated last year
- β103Mar 13, 2026Updated 4 months ago
- [NeurIPS 2025 D&Bπ₯] Implementation of "GS2E: Gaussian Splatting is an Effective Data Generator for Event Stream Generation"β20Jun 1, 2025Updated last year
- Implementation of <Streaming Autoregressive Video Generation via Diagonal Distillation> in ICLR 2026β130Mar 18, 2026Updated 4 months ago
- HY-WU (Part I): An Extensible Functional Neural Memory Framework and An Instantiation in Text-Guided Image Editingβ295Mar 18, 2026Updated 4 months ago
- EditReward: A Human-Aligned Reward Model for Instruction-Guided Image Editing [ICLR 2026]β155Apr 11, 2026Updated 3 months ago
- [ICLR'25] PiCO: Peer Review in LLMs based on the Consistency Optimization, https://arxiv.org/pdf/2402.01830β36Feb 16, 2025Updated last year
- β98Mar 14, 2026Updated 4 months ago
- Proton VPN Special Offer - Get 70% off β’ AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- [ACM MM 2025] HoloTime: Taming Video Diffusion Models for Panoramic 4D Scene Generationβ159Sep 4, 2025Updated 10 months ago
- RationalRewards: a reasoning reward model for diffusion RL and test-time prompt tuningβ56Jun 4, 2026Updated last month
- GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Datasetβ243Aug 15, 2025Updated 11 months ago
- This repository is the official implementation of "Look-Back: Implicit Visual Re-focusing in MLLM Reasoning".β100Jul 10, 2025Updated last year
- Code for paper "Rethinking Text-based Protein Understanding: Retrieval or LLM?"β20Oct 7, 2025Updated 9 months ago
- [CVPR 2026] Spatio-Temporal Autoregressive 4K 360Β° Video Generation from Perspective Videoβ129Mar 24, 2026Updated 4 months ago
- [Official Code] PermaVid: Consistent Video Generation Across Edits via Disentangled Context Memoryβ43Jun 17, 2026Updated last month