A curated guide to reasoning-enhancement methods for Multimodal Large Language Models (MLLMs), including dataset construction, training strategies, architectural designs, and evaluation benchmarks.
☆30Apr 20, 2025Updated last year
Alternatives and similar repositories for MLLM-Reasoning-Enhancement-Guide
Users that are interested in MLLM-Reasoning-Enhancement-Guide are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Data Science, Visualization, and Predictive Analytics: Kaggle Dataset - 1.6M accidents & traffic flow over 16 years☆17Feb 10, 2018Updated 8 years ago
- Official Code and data for ACL 2024 finding, "An Empirical Study on Parameter-Efficient Fine-Tuning for MultiModal Large Language Models"☆25Nov 10, 2024Updated last year
- Use Hermes Agent as the control plane for local coding agents like Codex, Kimi Code, Claude Code, OpenCode, and Gemini CLI.☆23May 28, 2026Updated last month
- 🪐 Agent2World: Learning to Generate Symbolic World Models via Adaptive Multi-Agent Feedback☆23Jan 29, 2026Updated 5 months ago
- Reverse Engineering Imperceptible Backdoor Attacks on Deep Neural Networks for Detection and Training Set Cleansing☆15Feb 18, 2021Updated 5 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- DART-GUI: Efficient Multi-turn RL for GUI Agents via Decoupled Training and Adaptive Data Curation☆94Feb 26, 2026Updated 4 months ago
- A python script to calculate radar cross section.☆12Dec 26, 2023Updated 2 years ago
- Official data and code for the paper "VisBrowse-Bench: Benchmarking Visual-Native Search for Multimodal Browsing Agents".☆15Mar 18, 2026Updated 4 months ago
- MLLM, DeepResearch, Agentic AI☆18Jun 1, 2026Updated last month
- CPU with 31 MIPS Instruction / 支持 31 条 MIPS 指令的 Minisys-1 单周期 CPU☆18May 18, 2021Updated 5 years ago
- Code for paper: PoisonPrompt: Backdoor Attack on Prompt-based Large Language Models, IEEE ICASSP 2024. Demo//124.220.228.133:11107☆21Aug 10, 2024Updated last year
- ReplayCode — first open-source rebuild of Claude Code that actually runs. Built from decompiled source with Node.js/esbuild☆20Apr 1, 2026Updated 3 months ago
- [ICML 2026] What Does Vision Tool-Use Reinforcement Learning Really Learn? Disentangling Tool-Induced and Intrinsic Effects for Crop-and-…☆21May 15, 2026Updated 2 months ago
- This repository is associated with the research paper titled ImageChain: Advancing Sequential Image-to-Text Reasoning in Multimodal Large…☆15Jun 4, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Estimating musical surprisal/information content in Audio☆34Apr 9, 2026Updated 3 months ago
- [ECCV 2026] Official implementation of "TIR-Bench: A Comprehensive Benchmark for Agentic Thinking-with-Images Reasoning"☆25Feb 8, 2026Updated 5 months ago
- A flexible & scalable MLLM-based AIGC detection pipeline☆40Jun 16, 2026Updated last month
- This is the code implementation of the paper titled "UAV Path Planning based on Road Extraction"☆13Feb 23, 2023Updated 3 years ago
- Direction Finding in Airborne Electronic Warfare Systems☆13Apr 18, 2022Updated 4 years ago
- Official Repository: A Comprehensive Benchmark for Logical Reasoning in MLLMs☆45Jun 17, 2025Updated last year
- Trace origins, shared sources, and contamination risk☆25May 27, 2026Updated last month
- ☆11May 24, 2022Updated 4 years ago
- ☆10Dec 19, 2019Updated 6 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- 具身智能学习站 · VLA 模型发展深度调研 + 88 篇论文细读(VitePress + Mermaid)☆29Updated this week
- Depth-Camera Calibration Toolbox (RGB-D Calibration ToolBox)☆15Dec 25, 2024Updated last year
- This repository contains the implementations of the iterated posterior linearisation filter (IPLF) for single-target tracking with direct…☆11Jul 26, 2024Updated last year
- ☆11Jan 30, 2023Updated 3 years ago
- Code for the paper "Transformer based Online Continuous Multi-Target Tracking with State Regression"☆14Mar 20, 2024Updated 2 years ago
- ☆32Mar 17, 2026Updated 4 months ago
- Some correction on IR video labels for Anti-UAV Dataset☆12Apr 13, 2021Updated 5 years ago
- A python package for Guidance, navigation, and control (GNC) of Autonomous Swarms Using Random finite sets (RFS) developed by the Laborat…☆17Mar 7, 2022Updated 4 years ago
- Code accompanying paper "Models as Agents: Optimizing Multi-Step Predictions of Interactive Local Models in Model-Based Multi-Agent Reinf…☆15Dec 2, 2023Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Flock and swarm multi-agent RL training environments implemented in JAX☆14Nov 19, 2025Updated 8 months ago
- This is the official repository for paper: cross-modal information flow in multimodal large language models☆44May 21, 2025Updated last year
- Multi Agent Reinforcement Learning Environment For Aerial Unmanned Vehicles☆13Apr 13, 2023Updated 3 years ago
- Implementation of Evo-Memory style learning for LLM agents. Agents learn from outcomes, refine strategies, and get smarter with every tas…☆48Dec 3, 2025Updated 7 months ago
- The model, data and code for OpenMobile☆49Jul 9, 2026Updated last week
- Simulation of massive MIMO radar with compressive sensing☆18Jun 21, 2021Updated 5 years ago
- Pytorch code for Tracklet Association Unsupervised Deep Learning (TAUDL)☆16Jan 5, 2021Updated 5 years ago