[CVPR 2026] GThinker, Reasoning MLLM, Visual Cues, Visual Rethinking
☆19Mar 9, 2026Updated 5 months ago
Alternatives and similar repositories for GThinker
Users that are interested in GThinker are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Chart-R1: Chain-of-Thought Supervision and Reinforcement for Advanced Chart Reasoner☆24Aug 7, 2025Updated last year
- [ICLR 2026] M2-Miner: Multi-Agent Enhanced MCTS for Mobile GUI Agent Data Mining☆55Apr 22, 2026Updated 3 months ago
- ☆34Sep 19, 2025Updated 10 months ago
- Official repo of Griffon series including v1(ECCV 2024), v2(ICCV 2025), G, and R, and also the RL tool Vision-R1(CVPR 2026).☆250Apr 17, 2026Updated 3 months ago
- Grounded Visual Token Sampling (GroundVTS), a Vid-LLM architecture designed to enhance VTG performance through adaptive and efficient vis…☆16Jun 12, 2026Updated last month
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Training codebase for K2-V2☆22Dec 17, 2025Updated 7 months ago
- Official Repo for CVPR 2025 Paper -- DeCafNet: Delegate and Conquer for Efficient Temporal Grounding in Long Videos☆17Mar 16, 2026Updated 4 months ago
- Secure and Scalable Federated Learning using Serverless Computing☆13Jan 31, 2024Updated 2 years ago
- NeurIPS 2025 Poster☆22Oct 17, 2025Updated 9 months ago
- Our 2nd-gen LMM☆34May 22, 2024Updated 2 years ago
- (CVPR 26 Findings) Official implementation of the paper "Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-…☆34Apr 7, 2026Updated 4 months ago
- We introduce DreamPRM-1.5, an instance-reweighted framework that adaptively adjusts the importance of each training example via bi-level …☆16Nov 13, 2025Updated 8 months ago
- The official repo for "Unified Domain Adaptive Semantic Segmentation" (IEEE TPAMI 2025)☆34Aug 14, 2025Updated 11 months ago
- An interactive thinking and deep reasoning model. It provides a cognitive reasoning paradigm for complex multi-hop problems.☆87Nov 14, 2025Updated 8 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [NeurIPS 2025] Reinforcing Spatial Reasoning in Vision-Language Models with Interwoven Thinking and Visual Drawing☆97Jul 27, 2025Updated last year
- Self-Teaching Notes on Gradient Leakage Attacks against GPT-2 models.☆14Mar 18, 2024Updated 2 years ago
- Multimodal Federated Learning on IoT Data☆11Dec 17, 2023Updated 2 years ago
- EVA: Efficient Reinforcement Learning for End-to-End Video Agent☆26May 6, 2026Updated 3 months ago
- ☆16Sep 16, 2025Updated 10 months ago
- object tracker for VOT☆10Jun 22, 2016Updated 10 years ago
- [WACV 2024] Enhancing Multimodal Compositional Reasoning of Visual Language Models with Generative Negative Mining, WACV 2024☆13Jan 3, 2024Updated 2 years ago
- ☆21Jan 22, 2026Updated 6 months ago
- A C++ code that detects the heart-rate of an individual from a input video. It is inspired by reviewing recent work on Eulerian Video Mag…☆12Aug 25, 2018Updated 7 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Qwen-WisdomVast is a large model trained on 1 million high-quality Chinese multi-turn SFT data, 200,000 English multi-turn SFT data, and …☆17Apr 12, 2024Updated 2 years ago
- ArcherCodeR is an open-source initiative enhancing code reasoning in large language models through scalable, rule-governed reinforcement …☆44Aug 6, 2025Updated last year
- Advanced Embodied Intelligence Brain Model☆38Nov 5, 2025Updated 9 months ago
- Methods and evaluation for aligning language models temporally☆31Mar 2, 2024Updated 2 years ago
- [ICML 2026] VideoLoom: A Video Large Language Model for Joint Spatial-Temporal Understanding☆28Jul 3, 2026Updated last month
- [CVPR2026] VideoAuto-R1: Video Auto Reasoning via Thinking Once, Answering Twice☆89Feb 27, 2026Updated 5 months ago
- IPO: Interpretable Prompt Optimization for Vision-Language Models(NeurIPS 2024)☆15Jun 12, 2026Updated last month
- [ICML 2026] Revisiting Uncertainty: On Evidential Learning for Partially Relevant Video Retrieval☆25Jul 10, 2026Updated last month
- ☆16Jun 4, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ECCV 2026] Official implementation of CEED-VLA: Consistency Vision-Language-Action Model with Early-Exit Decoding.☆52Sep 15, 2025Updated 10 months ago
- ☆62Jul 21, 2025Updated last year
- SpatialDreamer: Incentivizing Spatial Reasoning via Active Mental Imagery☆15Feb 1, 2026Updated 6 months ago
- Official repo of Promoting Efficient Reasoning with Verifiable Stepwise Reward☆16Sep 9, 2025Updated 11 months ago
- The official repository of "R-4B: Incentivizing General-Purpose Auto-Thinking Capability in MLLMs via Bi-Mode Integration"☆141Sep 4, 2025Updated 11 months ago
- There are some python scripts processing dataset, inferencing etc. wrote when I am using OpenMMLab.☆18Feb 13, 2023Updated 3 years ago
- 🤗 LeRobot: Making AI for Robotics more accessible with end-to-end learning☆15Jun 2, 2026Updated 2 months ago