[CVPR2026] Official codebase for the paper "Reasoning Within the Mind: Dynamic Multimodal Interleaving in Latent Space"
☆88May 12, 2026Updated 3 months ago
Alternatives and similar repositories for DMLR
Users that are interested in DMLR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [CVPR 2026] Official codes of "Monet: Reasoning in Latent Visual Space Beyond Image and Language"☆216Mar 19, 2026Updated 4 months ago
- Official codebase for the paper Latent Visual Reasoning☆172Oct 22, 2025Updated 9 months ago
- [ICLR 2026] Thinking on the Fly: Test-Time Reasoning Enhancement via Latent Thought Policy Optimization☆32Mar 6, 2026Updated 5 months ago
- Official Repository of LatentSeek☆86Jun 6, 2025Updated last year
- Official codes of "Sketch-in-Latents: Eliciting Unified Reasoning in MLLMs"☆17Feb 15, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ECCV 2026] Official repo of "Chain-of-Visual-Thought: Teaching VLMs to See and Think Better with Continuous Visual Tokens"☆392Jul 30, 2026Updated 2 weeks ago
- [CVPR 2026] Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens☆296Aug 2, 2025Updated last year
- [ACL'26 Oral] Interleaved Latent Visual Reasoning with Selective Perceptual Modeling☆66Aug 5, 2026Updated last week
- This repository is the official implementation of "Look-Back: Implicit Visual Re-focusing in MLLM Reasoning".☆99Jul 10, 2025Updated last year
- MCOUT: Multimodal Chain of Continuous Thought for Latent Reasoning☆21Oct 4, 2025Updated 10 months ago
- Official repository for “Reasoning in the Dark: Interleaved Vision-Text Reasoning in Latent Space”☆18Jan 27, 2026Updated 6 months ago
- [NeurIPS 2025] Official code for paper: Latent Chain-of-Thought for Visual Reasoning☆36Oct 16, 2025Updated 10 months ago
- [ICML2026] Imagination Helps Visual Reasoning, But Not Yet in Latent Space☆29May 4, 2026Updated 3 months ago
- [ICML 2026] Official implementation for paper: Learning Self-Correction in Vision–Language Models via Rollout Augmentation☆16Jun 4, 2026Updated 2 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Code for Reducing Hallucinations in Vision-Language Models via Latent Space Steering☆118Nov 23, 2024Updated last year
- ☆72Feb 27, 2026Updated 5 months ago
- ☆92Feb 5, 2026Updated 6 months ago
- A paper list of Awesome Latent Space.☆955Jul 13, 2026Updated last month
- Official codebase for the paper "WorldMemArena: Evaluating Multimodal Agent Memory Through Action–World Interaction"☆25May 29, 2026Updated 2 months ago
- [NeurIPS 2025] More Thinking, Less Seeing? Assessing Amplified Hallucination in Multimodal Reasoning Models☆82May 31, 2025Updated last year
- When and How Much to Imagine: Adaptive Test-Time Scaling with World Models for Visual Spatial Reasoning☆19Jun 2, 2026Updated 2 months ago
- ☆32Mar 17, 2026Updated 5 months ago
- ☆15Apr 15, 2026Updated 4 months ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Official implementation of the NeurIPS 2025 paper "Soft Thinking: Unlocking the Reasoning Potential of LLMs in Continuous Concept Space"☆349Jun 12, 2026Updated 2 months ago
- Rui Qian, Xin Yin, Dejing Dou†: Reasoning to Attend: Try to Understand How <SEG> Token Works (CVPR 2025)☆55Feb 4, 2026Updated 6 months ago
- [ACL 2026] Render-of-Thought: Rendering Textual Chain-of-Thought as Images for Visual Latent Reasoning☆93Jan 22, 2026Updated 6 months ago
- This is the official repository for paper: cross-modal information flow in multimodal large language models☆44May 21, 2025Updated last year
- HyperEyes is a parallel multimodal search agent that fuses visual grounding and retrieval into a single atomic action, enabling concurren…☆73May 23, 2026Updated 2 months ago
- ☆73Feb 1, 2026Updated 6 months ago
- Code for ICLR 2025 Paper: Visual Description Grounding Reduces Hallucinations and Boosts Reasoning in LVLMs☆25May 7, 2025Updated last year
- [CVPR2026] VideoAuto-R1: Video Auto Reasoning via Thinking Once, Answering Twice☆88Feb 27, 2026Updated 5 months ago
- [ICLR 26] Visual Multi-Agent System: Mitigating Hallucination Snowballing via Visual Flow☆46Oct 3, 2025Updated 10 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Source code For AAAI 2026 paper: "RaLiFlow: Scene Flow Estimation with 4D Radar and LiDAR Point Clouds"☆15Jun 5, 2026Updated 2 months ago
- [ICLR 2026] SwiReasoning: Switch-Thinking in Latent and Explicit for Pareto-Superior Reasoning LLMs☆139Aug 2, 2026Updated 2 weeks ago
- Generating Structured Pseudo Labels for Noise-resistant Zero-shot Video Sentence Localization☆16Jul 20, 2023Updated 3 years ago
- Code for EMNLP25 paper "Video-RTS: Rethinking Reinforcement Learning and Test-Time Scaling for Efficient and Enhanced Video Reasoning"☆24Feb 18, 2026Updated 6 months ago
- Reasoning Guided Embeddings: Leveraging MLLM Reasoning for Improved Multimodal Retrieval☆16Nov 29, 2025Updated 8 months ago
- ☆26Oct 9, 2025Updated 10 months ago
- [ICCV 2025] Enhancing Partially Relevant Video Retrieval with Hyperbolic Learning☆63May 11, 2026Updated 3 months ago