[CVPR2026] Official codebase for the paper "Reasoning Within the Mind: Dynamic Multimodal Interleaving in Latent Space"
☆85May 12, 2026Updated 2 months ago
Alternatives and similar repositories for DMLR
Users that are interested in DMLR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [CVPR 2026] Official codes of "Monet: Reasoning in Latent Visual Space Beyond Image and Language"☆215Mar 19, 2026Updated 4 months ago
- Official codebase for the paper Latent Visual Reasoning☆171Oct 22, 2025Updated 9 months ago
- Official codebase for the paper LaViT☆34Feb 15, 2026Updated 5 months ago
- [ICLR 2026] Thinking on the Fly: Test-Time Reasoning Enhancement via Latent Thought Policy Optimization☆32Mar 6, 2026Updated 4 months ago
- Official Repository of LatentSeek☆85Jun 6, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official codes of "Sketch-in-Latents: Eliciting Unified Reasoning in MLLMs"☆17Feb 15, 2026Updated 5 months ago
- [ECCV 2026] Official repo of "Chain-of-Visual-Thought: Teaching VLMs to See and Think Better with Continuous Visual Tokens"☆379Apr 17, 2026Updated 3 months ago
- [ACL'26 Oral] Interleaved Latent Visual Reasoning with Selective Perceptual Modeling☆66May 29, 2026Updated 2 months ago
- This repository is the official implementation of "Look-Back: Implicit Visual Re-focusing in MLLM Reasoning".☆100Jul 10, 2025Updated last year
- [CVPR 2026 Highlight] Thinking in Uncertainty: Mitigating Hallucinations in MLRMs with Latent Entropy-Aware Decoding☆95Apr 9, 2026Updated 3 months ago
- MCOUT: Multimodal Chain of Continuous Thought for Latent Reasoning☆21Oct 4, 2025Updated 9 months ago
- Official repository for “Reasoning in the Dark: Interleaved Vision-Text Reasoning in Latent Space”☆18Jan 27, 2026Updated 6 months ago
- [NeurIPS 2025] Official code for paper: Latent Chain-of-Thought for Visual Reasoning☆36Oct 16, 2025Updated 9 months ago
- [ICML2026] Imagination Helps Visual Reasoning, But Not Yet in Latent Space☆28May 4, 2026Updated 2 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [ICML 2026] Official implementation for paper: Learning Self-Correction in Vision–Language Models via Rollout Augmentation☆16Jun 4, 2026Updated last month
- Code for Reducing Hallucinations in Vision-Language Models via Latent Space Steering☆117Nov 23, 2024Updated last year
- ☆72Feb 27, 2026Updated 5 months ago
- ☆91Feb 5, 2026Updated 5 months ago
- ☆35Apr 22, 2026Updated 3 months ago
- Official codebase for the paper "WorldMemArena: Evaluating Multimodal Agent Memory Through Action–World Interaction"☆23May 29, 2026Updated 2 months ago
- [NeurIPS 2025] More Thinking, Less Seeing? Assessing Amplified Hallucination in Multimodal Reasoning Models☆82May 31, 2025Updated last year
- When and How Much to Imagine: Adaptive Test-Time Scaling with World Models for Visual Spatial Reasoning☆18Jun 2, 2026Updated last month
- ☆32Mar 17, 2026Updated 4 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Official implementation of the NeurIPS 2025 paper "Soft Thinking: Unlocking the Reasoning Potential of LLMs in Continuous Concept Space"☆347Jun 12, 2026Updated last month
- Rui Qian, Xin Yin, Dejing Dou†: Reasoning to Attend: Try to Understand How <SEG> Token Works (CVPR 2025)☆54Feb 4, 2026Updated 5 months ago
- This is the official repository for paper: cross-modal information flow in multimodal large language models☆44May 21, 2025Updated last year
- HyperEyes is a parallel multimodal search agent that fuses visual grounding and retrieval into a single atomic action, enabling concurren…☆70May 23, 2026Updated 2 months ago
- ☆73Feb 1, 2026Updated 5 months ago
- Code for ICLR 2025 Paper: Visual Description Grounding Reduces Hallucinations and Boosts Reasoning in LVLMs☆25May 7, 2025Updated last year
- [CVPR2026] VideoAuto-R1: Video Auto Reasoning via Thinking Once, Answering Twice☆88Feb 27, 2026Updated 5 months ago
- [ICLR 26] Visual Multi-Agent System: Mitigating Hallucination Snowballing via Visual Flow☆44Oct 3, 2025Updated 9 months ago
- Source code For AAAI 2026 paper: "RaLiFlow: Scene Flow Estimation with 4D Radar and LiDAR Point Clouds"☆15Jun 5, 2026Updated last month
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ICLR 2026] SwiReasoning: Switch-Thinking in Latent and Explicit for Pareto-Superior Reasoning LLMs☆137May 20, 2026Updated 2 months ago
- Generating Structured Pseudo Labels for Noise-resistant Zero-shot Video Sentence Localization☆16Jul 20, 2023Updated 3 years ago
- ☆27Oct 9, 2025Updated 9 months ago
- [ICCV 2025] Enhancing Partially Relevant Video Retrieval with Hyperbolic Learning☆62May 11, 2026Updated 2 months ago
- [ICLR'25] Official code for the paper 'MLLMs Know Where to Look: Training-free Perception of Small Visual Details with Multimodal LLMs'☆381Apr 20, 2025Updated last year
- ☆76May 8, 2026Updated 2 months ago
- Official implementation of "Reasoning by Superposition: A Theoretical Perspective on Chain of Continuous Thought" (NeurIPS 2025)☆44Oct 8, 2025Updated 9 months ago