CoCo: Code as CoT for Text-to-Image Preview and Rare Concept Generation
☆56Aug 7, 2026Updated 3 weeks ago
Alternatives and similar repositories for CoCo
Users that are interested in CoCo are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆53Feb 25, 2026Updated 6 months ago
- Offical Repository for Paper: DraCo: Draft as CoT for Text-to-Image Preview and Rare Concept Generation☆19Dec 7, 2025Updated 8 months ago
- Learning Brain Representation with Hierarchical Visual Embeddings☆25Apr 18, 2026Updated 4 months ago
- Official repository for “Reasoning in the Dark: Interleaved Vision-Text Reasoning in Latent Space”☆18Jan 27, 2026Updated 7 months ago
- [EMNLP2026] Does Understanding Inform Generation in Unified Multimodal Models? From Analysis to Path Forward☆60Nov 27, 2025Updated 9 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- 🐧 Unify-Agent: An end-to-end unified multimodal agent for faithful, knowledge-grounded image generation.☆92May 2, 2026Updated 3 months ago
- GEditBench v2: A Human-Aligned Benchmark for General Image Editing☆62Jun 18, 2026Updated 2 months ago
- SpatialEvo: Self-Evolving Spatial Intelligence via Deterministic Geometric Environments☆83Apr 16, 2026Updated 4 months ago
- ☆28Feb 3, 2026Updated 6 months ago
- Step3-VL-10B: A compact yet frontier multimodal model achieving SOTA performance at the 10B scale, matching open-source models 10-20x its…☆413Jan 21, 2026Updated 7 months ago
- [NeurIPS 2025] The official repository for our paper, "Open Vision Reasoner: Transferring Linguistic Cognitive Behavior for Visual Reason…☆156Sep 12, 2025Updated 11 months ago
- [ECCV 2026] Official repository of "Reliable Reasoning in SVG-LLMs via Multi-Task Multi-Reward Reinforcement Learning".