☆34Mar 17, 2026Updated 6 months ago
Alternatives and similar repositories for CodeDance
Users that are interested in CodeDance are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICML 2026] What Does Vision Tool-Use Reinforcement Learning Really Learn? Disentangling Tool-Induced and Intrinsic Effects for Crop-and-…☆23May 15, 2026Updated 4 months ago
- [CVPR 2026] Thinking with Programming Vision: Towards a Unified View for Thinking with Images☆74Jan 23, 2026Updated 8 months ago
- Official repository for “Reasoning in the Dark: Interleaved Vision-Text Reasoning in Latent Space”☆18Jan 27, 2026Updated 7 months ago
- ☆28Mar 17, 2026Updated 6 months ago
- The repository of VG-Refiner paper☆20Dec 9, 2025Updated 9 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official data and code for the paper "VisBrowse-Bench: Benchmarking Visual-Native Search for Multimodal Browsing Agents".☆14Mar 18, 2026Updated 6 months ago
- This repository is the official implementation of "Look-Back: Implicit Visual Re-focusing in MLLM Reasoning".☆98Jul 10, 2025Updated last year
- [CVPR 2026 Oral] Code with Image☆34Dec 5, 2025Updated 9 months ago
- Official Code for paper "Active Video Perception: Iterative Evidence Seeking for Agentic Long Video Understanding""☆20Jun 2, 2026Updated 3 months ago
- ☆19Apr 7, 2026Updated 5 months ago
- Official Repository for NeurIPS'25 Paper "Tool-Augmented Spatiotemporal Reasoning for Streamlining Video Question Answering Task"☆23Sep 4, 2026Updated 2 weeks ago
- [CVPR2026] Official codebase for the paper "Reasoning Within the Mind: Dynamic Multimodal Interleaving in Latent Space"☆88May 12, 2026Updated 4 months ago
- This repository is the official implementation for VISD.☆25May 17, 2026Updated 4 months ago
- [CVPR 2026] Official release of "Spatial-SSRL: Enhancing Spatial Understanding via Self-Supervised Reinforcement Learning"☆141Apr 7, 2026Updated 5 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Use Hermes Agent as the control plane for local coding agents like Codex, Kimi Code, Claude Code, OpenCode, and Gemini CLI.☆37Jul 22, 2026Updated 2 months ago
- Code for RACE.☆15Nov 12, 2025Updated 10 months ago
- [ICLR 26] Visual Multi-Agent System: Mitigating Hallucination Snowballing via Visual Flow☆47Oct 3, 2025Updated 11 months ago
- ThinkGen: Generalized Thinking for Visual Generation☆61Dec 30, 2025Updated 8 months ago
- [CVPR 2026] Official codes of "Monet: Reasoning in Latent Visual Space Beyond Image and Language"☆224Mar 19, 2026Updated 6 months ago
- HyperEyes is a parallel multimodal search agent that fuses visual grounding and retrieval into a single atomic action, enabling concurren…☆76May 23, 2026Updated 3 months ago
- \infty-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation☆22Feb 14, 2025Updated last year
- ☆12Mar 31, 2025Updated last year
- [ICLR 2026] "VTool-R1: VLMs Learn to Think with Images via Reinforcement Learning on Multimodal Tool Use"☆209Mar 20, 2026Updated 6 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- The official code of "Thinking With Videos: Multimodal Tool-Augmented Reinforcement Learning for Long Video Reasoning"☆103Oct 15, 2025Updated 11 months ago
- [EMNLP 2026 Main] Think, Act, Build: An Agentic Framework with Vision Language Models for Zero-Shot 3D Visual Grounding☆28Aug 21, 2026Updated last month
- ParaVT: Taming the Tool Prior Paradox for Parallel Tool Use in Agentic Video Reinforcement Learning☆58Jun 2, 2026Updated 3 months ago
- Just a copy of https://github.com/RobynE23/CodeHS-Java-APCSA, but I added folders and some extra files that didn't exist. Another option …☆27Jan 23, 2024Updated 2 years ago
- ☆13Jul 25, 2023Updated 3 years ago
- [EMNLP 2024] Wrong-of-Thought: An Integrated Reasoning Framework with Multi-Perspective Verification and Wrong Information☆13Oct 1, 2024Updated last year
- A benchmark for evaluating future-event forecasting from audio and video context in multimodal language models☆29Jan 22, 2026Updated 8 months ago
- 📖Curated list about reasoning abilitiy of MLLM, including OpenAI o1, OpenAI o3-mini, and Slow-Thinking.☆14Feb 7, 2025Updated last year
- Project page and release repository for PhysEditWorld, a dataset toward physics-editable world models.☆18Jul 2, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆37Apr 13, 2026Updated 5 months ago
- [ACM MM 2022] Patch-based Knowledge Distillation for Lifelong Person Re-Identification☆11Apr 20, 2023Updated 3 years ago
- ☆23Dec 4, 2025Updated 9 months ago
- ☆71Feb 27, 2026Updated 6 months ago
- The official code of "CaLa: Complementary Association Learning for Augmenting Composed Image Retrieval"☆15Sep 19, 2024Updated 2 years ago
- [ACL 2025 Findings] Official pytorch implementation of "Don't Miss the Forest for the Trees: Attentional Vision Calibration for Large Vis…☆25Jul 21, 2024Updated 2 years ago
- ☆75Feb 12, 2026Updated 7 months ago