[NeurIPS 2025] Reasoning MLLM, Share-GRPO, advantage vanishing, sparse reward
☆38Sep 19, 2025Updated 11 months ago
Alternatives and similar repositories for R1-ShareVL
Users that are interested in R1-ShareVL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- MLLM, DeepResearch, Agentic AI☆19Jun 1, 2026Updated 2 months ago
- [EMNLP Main 2026]VTC-R1: Vision-Text Compression for Efficient Long-Context Reasoning.☆26Jul 20, 2026Updated last month
- [ICCV 2025] MMReason, MLLMs, step by step, reasoning benchmark, AGI☆15Apr 25, 2026Updated 4 months ago
- Agentic MLLMs☆215Oct 24, 2025Updated 10 months ago
- [NeurIPS 2025@FoRLM] R1-Compress: Long Chain-of-Thought Compression via Chunk Compression and Search☆17Jan 24, 2026Updated 7 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- This repository is aim to reproduce the R1-Zero on medical domain.☆32Jun 11, 2025Updated last year
- [ICML 2026] What Does Vision Tool-Use Reinforcement Learning Really Learn? Disentangling Tool-Induced and Intrinsic Effects for Crop-and-…☆23May 15, 2026Updated 3 months ago
- The official code of "VL-Rethinker: Incentivizing Self-Reflection of Vision-Language Models with Reinforcement Learning" [NeurIPS25]☆192Jun 5, 2025Updated last year
- Uni-OPD: Unifying On-Policy Distillation with a Dual-Perspective Recipe☆57Jun 10, 2026Updated 2 months ago
- [EMNLP'26] Code and data for VTCBench, a VLM benchmark for long-context understanding capabilities under vision-text compression paradigm…☆27Updated this week
- MEDREASON-R1: Learning to Reason for CT Diagnosis with Reinforcement Learning and Local Zoom☆16Oct 10, 2025Updated 10 months ago
- The repository of the ACCV 2024 paper "FG-CXR: A Radiologist-Aligned Gaze Dataset for Enhancing Interpretability in Chest X-Ray Report Ge…☆12Aug 15, 2026Updated 2 weeks ago
- Official code for NeurIPS 2025 paper "GRIT: Teaching MLLMs to Think with Images"☆192Jan 16, 2026Updated 7 months ago
- This is the code repo for the paper AceSearcher: Bootstrapping Reasoning and Search for LLMs via Reinforced Self-Play (NeurIPS 2025 Spotl…☆25Sep 29, 2025Updated 11 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [CVPR'26] VisPlay: Self-Evolving Vision-Language Models☆75Feb 25, 2026Updated 6 months ago
- ☆42Apr 7, 2024Updated 2 years ago
- A unified framework for vision-language environments with Gymnasium-compatible interface☆38Mar 17, 2026Updated 5 months ago
- Citrus-V: Advancing Medical Foundation Models with Unified Medical Image Grounding for Clinical Reasoning☆25Sep 26, 2025Updated 11 months ago
- Mitigating Shortcuts in Visual Reasoning with Reinforcement Learning☆44Jul 2, 2025Updated last year
- ☆18Sep 13, 2023Updated 2 years ago
- ☆33Oct 6, 2024Updated last year
- Official repo for "TiMo: Spatiotemporal Foundation Model for Satellite Image Time Series"☆29Jun 30, 2026Updated 2 months ago
- MCPL: Multi-modal Collaborative Prompt Learning for Medical Vision-Language Model (Initial Version)☆13Apr 17, 2024Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Official Implementation of Trajectory-Refined Distillation☆35Jun 9, 2026Updated 2 months ago
- Interpreting Chest X-rays Like a Radiologist: A Benchmark with Clinical Reasoning, release the dataset and the model weight☆13May 26, 2025Updated last year
- Implementation of the Paper Scene-Graph ViT☆10Dec 20, 2024Updated last year
- [ICCV 2023] This is the Pytorch code for our paper "Self-Supervised Cross-View Representation Reconstruction for Change Captioning".☆20Sep 25, 2025Updated 11 months ago
- ☆27Aug 31, 2025Updated last year
- DBPM is a simple algorithm designed as a lightweight plug-in without learnable parameters to enhance the performance of time series contr…☆15Mar 8, 2024Updated 2 years ago
- [ICLR 2026] Official repo for "Spotlight on Token Perception for Multimodal Reinforcement Learning"☆84Apr 3, 2026Updated 4 months ago
- MaXM is a suite of test-only benchmarks for multilingual visual question answering in 7 languages: English (en), French (fr), Hindi (hi),…☆13Jan 16, 2024Updated 2 years ago
- [ICCV 2025] The official pytorch implement of "LLaVA-SP: Enhancing Visual Representation with Visual Spatial Tokens for MLLMs".☆24Oct 28, 2025Updated 10 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆137Jul 22, 2025Updated last year
- [ICLR2025] Swiss Army Knife: Synergizing Biases in Knowledge from Vision Foundation Models for Multi-Task Learning☆15Apr 8, 2025Updated last year
- This repository is the official implementation of "Look-Back: Implicit Visual Re-focusing in MLLM Reasoning".☆98Jul 10, 2025Updated last year
- [ICML 2026] Official implementation of "PyVision-RL: Forging Open Agentic Vision Models via RL."☆69Feb 25, 2026Updated 6 months ago
- DeepResearchEval: An Automated Framework for Deep Research Task Construction and Agentic Evaluation.☆143Feb 10, 2026Updated 6 months ago
- MemOCR: an OCR-driven visual memory agent.☆34May 17, 2026Updated 3 months ago
- SCOPE: Signal-Calibrated On-Policy Distillation Enhancement with Dual-Path Adaptive Weighting☆29Jun 22, 2026Updated 2 months ago