☆24Nov 4, 2025Updated 9 months ago
Alternatives and similar repositories for VL-Cogito
Users that are interested in VL-Cogito are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ACL 2025] Analyzing LLMs' Multilingual Knowledge Boundary Cognition Across Languages Through the Lens of Internal Representations☆19Oct 18, 2025Updated 10 months ago
- ☆15Jan 14, 2026Updated 7 months ago
- [NeurIPS 2025] Scaling Language-centric Omnimodal Representation Learning☆48Apr 13, 2026Updated 4 months ago
- [ICCV 2025] MMReason, MLLMs, step by step, reasoning benchmark, AGI☆15Apr 25, 2026Updated 4 months ago
- codes for Efficient Test-Time Scaling via Self-Calibration☆21Sep 13, 2025Updated 11 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- MedEvalKit: A Unified Medical Evaluation Framework☆253Feb 24, 2026Updated 6 months ago
- M2-Reasoning: Empowering MLLMs with Unified General and Spatial Reasoning☆47Jul 17, 2025Updated last year
- Drivel-ology: Challenging LLMs with Interpreting Nonsense with Depth☆15Nov 26, 2025Updated 9 months ago
- [CVPR 2026] MMR1: Enhancing Multimodal Reasoning with Variance-Aware Sampling and Open Resources☆217Sep 26, 2025Updated 11 months ago
- Textual Localization: Decomposing Multi-concept Images for Subject-Driven Text-to-Image Generation☆16Mar 10, 2024Updated 2 years ago
- Pixels, Patterns, but no Poetry: To See the World like Humans☆18Aug 11, 2025Updated last year
- ☆15Mar 8, 2024Updated 2 years ago
- [ICML 2026 Spotlight] Critique-GRPO: Advancing LLM Reasoning with Natural Language and Numerical Feedback☆74Jun 3, 2026Updated 2 months ago
- SophiaVL-R1: Reinforcing MLLMs Reasoning with Thinking Reward☆94Aug 8, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Official code repository of Shuffle-R1☆26Feb 23, 2026Updated 6 months ago
- [CVPR' 25] Interleaved-Modal Chain-of-Thought☆113Dec 30, 2025Updated 8 months ago
- Implementation of the paper "CXR-IRGen: An Integrated Vision and Language Model for the Generation of Clinically Accurate Chest X-Ray Ima…☆21Jul 2, 2024Updated 2 years ago
- [EMNLP 2025 Findings] MEXA: Towards General Multimodal Reasoning with Dynamic Multi-Expert Aggregation☆15Aug 22, 2025Updated last year
- ☆25Aug 20, 2025Updated last year
- [ACL '26] source code for the paper: "Long-Chain Reasoning Distillation via Adaptive Prefix Alignment"☆16Jan 21, 2026Updated 7 months ago
- ICLR 2025: We propose Atomas, a hierarchical molecular representation learning framework that jointly learns representations from SMILES …☆20Feb 24, 2025Updated last year
- VideoEval-Pro: Robust and Realistic Long Video Understanding Evaluation [TMLR26]☆15Jun 1, 2026Updated 2 months ago
- AutoThink is a reinforcement learning framework designed to equip R1-style language models with adaptive reasoning capabilities. Instead …☆52Oct 14, 2025Updated 10 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This repository contains the code for our CVPR 2024 paper,☆16Aug 27, 2024Updated 2 years ago
- This code implements the algorithm of FIPO, a value-free RL recipe for eliciting deeper reasoning from a clean base model.☆130Apr 7, 2026Updated 4 months ago
- Code for the multi-agent computer use project.☆21Jul 3, 2026Updated last month
- code for the paper "CoReS: Orchestrating the Dance of Reasoning and Segmentation"☆23Nov 24, 2025Updated 9 months ago
- A comprehensive toolkit for streamlining data editing, search, and inspection for large-scale language model training and interpretabilit…☆21Oct 30, 2025Updated 10 months ago
- (ICML 2025) Rethinking Chain-of-Thought from the Perspective of Self-Training☆13Feb 15, 2025Updated last year
- [ICLR 2026] Official code for paper: TimeSearch-R: Adaptive Temporal Search for Long-Form Video Understanding via Self-Verification Reinf…☆39Jan 29, 2026Updated 7 months ago
- Source Code for our ICLR'26 paper☆17Feb 22, 2026Updated 6 months ago
- An agentic RL framework to enhance retreival-augmented reasoning in Diagnostic Policy☆106Feb 27, 2026Updated 6 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆18Feb 2, 2026Updated 6 months ago
- [ACM MM 2025] ViTCoT: Video-Text Interleaved Chain-of-Thought for Boosting Video Understanding in Large Language Models☆18Jul 15, 2025Updated last year
- [ICLR'26] Traceable Evidence Enhanced Visual Grounded Reasoning: Evaluation and Methodology☆93Jan 26, 2026Updated 7 months ago
- Unifying Specialized Visual Encoders for Video Language Models☆25Nov 22, 2025Updated 9 months ago
- OpenVLThinker [NeurIPS 2025] & OpenVLThinkerV2 [COLM 2026]☆156May 25, 2026Updated 3 months ago
- [ICML 2026] InftyThink+: Effective and Efficient Infinite-Horizon Reasoning via Reinforcement Learning☆34May 25, 2026Updated 3 months ago
- OpenVLA Lightweight Version(0.5B). It uses qwen2-0.5B and fine-tunes using mllm format, without occupying LLM's inherent tokens. It repre…☆19Jan 7, 2026Updated 7 months ago