[CVPR 2026] GThinker, Reasoning MLLM, Visual Cues, Visual Rethinking
☆18Mar 9, 2026Updated 4 months ago
Alternatives and similar repositories for GThinker
Users that are interested in GThinker are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆33Jun 30, 2026Updated 2 weeks ago
- Chart-R1: Chain-of-Thought Supervision and Reinforcement for Advanced Chart Reasoner☆24Aug 7, 2025Updated 11 months ago
- [ICLR 2026] M2-Miner: Multi-Agent Enhanced MCTS for Mobile GUI Agent Data Mining☆55Apr 22, 2026Updated 2 months ago
- ☆34Sep 19, 2025Updated 10 months ago
- Official repo of Griffon series including v1(ECCV 2024), v2(ICCV 2025), G, and R, and also the RL tool Vision-R1(CVPR 2026).☆250Apr 17, 2026Updated 3 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆21May 14, 2026Updated 2 months ago
- Grounded Visual Token Sampling (GroundVTS), a Vid-LLM architecture designed to enhance VTG performance through adaptive and efficient vis…☆16Jun 12, 2026Updated last month
- Official Repo for CVPR 2025 Paper -- DeCafNet: Delegate and Conquer for Efficient Temporal Grounding in Long Videos☆17Mar 16, 2026Updated 4 months ago
- [ACM MM 2025 🔥🔥 ] MIRA: A first-of-its-kind medical RAG framework that fuses image features and retrieved knowledge with dynamic contex…☆23Aug 28, 2025Updated 10 months ago
- ☆78May 6, 2024Updated 2 years ago
- Secure and Scalable Federated Learning using Serverless Computing☆13Jan 31, 2024Updated 2 years ago
- KDD25: The source code of our paper "UoMo: A Universal Model of Mobile Traffic Forecasting for Wireless Network Optimization"☆16Aug 14, 2025Updated 11 months ago
- (CVPR 26 Findings) Official implementation of the paper "Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-…☆34Apr 7, 2026Updated 3 months ago
- We introduce DreamPRM-1.5, an instance-reweighted framework that adaptively adjusts the importance of each training example via bi-level …☆16Nov 13, 2025Updated 8 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- An interactive thinking and deep reasoning model. It provides a cognitive reasoning paradigm for complex multi-hop problems.☆85Nov 14, 2025Updated 8 months ago
- The official repo for "Unified Domain Adaptive Semantic Segmentation" (IEEE TPAMI 2025)☆34Aug 14, 2025Updated 11 months ago
- [NeurIPS 2025] Reinforcing Spatial Reasoning in Vision-Language Models with Interwoven Thinking and Visual Drawing☆98Jul 27, 2025Updated 11 months ago
- Multimodal Federated Learning on IoT Data☆11Dec 17, 2023Updated 2 years ago
- [WACV 2024] Enhancing Multimodal Compositional Reasoning of Visual Language Models with Generative Negative Mining, WACV 2024☆13Jan 3, 2024Updated 2 years ago
- ☆16Sep 16, 2025Updated 10 months ago
- Artifact evaluation of MobiSys25 SynCheck☆20Mar 24, 2025Updated last year
- ArcherCodeR is an open-source initiative enhancing code reasoning in large language models through scalable, rule-governed reinforcement …☆44Aug 6, 2025Updated 11 months ago
- Methods and evaluation for aligning language models temporally☆31Mar 2, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ICML 2026] VideoLoom: A Video Large Language Model for Joint Spatial-Temporal Understanding☆27Jul 3, 2026Updated 2 weeks ago
- [CVPR2026] VideoAuto-R1: Video Auto Reasoning via Thinking Once, Answering Twice☆88Feb 27, 2026Updated 4 months ago
- [ICML 2026] Revisiting Uncertainty: On Evidential Learning for Partially Relevant Video Retrieval☆24Jul 10, 2026Updated last week
- Code and data for paper "Exploring Hallucination of Large Multimodal Models in Video Understanding: Benchmark, Analysis and Mitigation".☆25Oct 22, 2025Updated 8 months ago
- ☆16Jun 4, 2025Updated last year
- [ICLR 2025] TRACE: Temporal Grounding Video LLM via Casual Event Modeling☆156Aug 22, 2025Updated 10 months ago
- ☆62Jul 21, 2025Updated last year
- SpatialDreamer: Incentivizing Spatial Reasoning via Active Mental Imagery☆15Feb 1, 2026Updated 5 months ago
- Official repo of Promoting Efficient Reasoning with Verifiable Stepwise Reward☆16Sep 9, 2025Updated 10 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- The official repository of "R-4B: Incentivizing General-Purpose Auto-Thinking Capability in MLLMs via Bi-Mode Integration"☆141Sep 4, 2025Updated 10 months ago
- Official code repo of Video-Browser: Towards Agentic Open-web Video Browsing☆28Jan 19, 2026Updated 6 months ago
- ☆20Jul 3, 2026Updated 2 weeks ago
- Cockatiel: Ensembling Synthetic and Human Preferenced Training for Detailed Video Caption☆38May 21, 2025Updated last year
- ☆34Feb 12, 2026Updated 5 months ago
- Official code of *Towards Event-oriented Long Video Understanding*☆12Jul 26, 2024Updated last year
- ☆15Sep 14, 2023Updated 2 years ago