[CVPR'26] Dr. Seg: Revisiting GRPO Training for Visual Large Language Models through Perception-Oriented Design
☆82Mar 7, 2026Updated 6 months ago
Alternatives and similar repositories for Dr-Seg
Users that are interested in Dr-Seg are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ECCV 2026] StAR: Segment Anything Reasoner☆25Apr 2, 2026Updated 5 months ago
- implementation of paper submitted to ISPRS Journal of Photogrammetry and Remote Sensing☆29Jul 3, 2025Updated last year
- ALTo: Adaptive-Length Tokenizer for Autoregressive Mask Generation☆29May 27, 2025Updated last year
- This is an official PyTorch implementation for "EAGLE: Expert-Augmented Attention Guidance for Tuning-Free Industrial Anomaly Detection i…☆20Feb 24, 2026Updated 6 months ago
- ☆27Dec 4, 2025Updated 9 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [NeurIPS-W 2025] Official Implementation of "Seg-R1: Segmentation Can Be Surprisingly Simple with Reinforcement Learning"☆73Jul 1, 2025Updated last year
- This is the repository for JUDO (ICLR 26)☆25Mar 9, 2026Updated 6 months ago
- [ICLR'26] MedVR: Annotation-Free Medical Visual Reasoning via Agentic Reinforcement Learning☆54Apr 3, 2026Updated 5 months ago
- [ICLR 2026] VisionReasoner: Unified Reasoning-Integrated Visual Perception via Reinforcement Learning☆353Feb 9, 2026Updated 7 months ago
- Rui Qian, Xin Yin, Chuanhang Deng, et al.: UGround: Towards Unified Visual Grounding with Unrolled Transformers (ICML 2026)☆29Jun 18, 2026Updated 2 months ago
- The official implementation of "PixelThink: Towards Efficient Chain-of-Pixel Reasoning" (ICML 2026)☆43Jul 4, 2026Updated 2 months ago
- Native GStreamer plugins that integrate SAHI (Slicing Aided Hyper Inference) into NVIDIA DeepStream for real-time small object detection …☆33Jun 8, 2026Updated 3 months ago
- [AAAI 2026 Oral] LENS: Learning to Segment Anything with Unified Reinforced Reasoning☆143Dec 3, 2025Updated 9 months ago
- This is the implementation of the paper "Test-Time Adaptive Object Detection with Foundation Model" (Neurips 2025)☆23Jan 30, 2026Updated 7 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [CVPR 2026] STAMP: Better, Stronger, Faster: Tackling the Trilemma in MLLM-based Segmentation with Simultaneous Textual Mask Prediction☆43Feb 21, 2026Updated 6 months ago
- EMIT: Enhancing MLLMs for Industrial Anomaly Detection via Difficulty-Aware GRPO☆30Jan 24, 2026Updated 7 months ago
- ☆54Oct 1, 2025Updated 11 months ago
- ☆48Jan 1, 2026Updated 8 months ago
- This repository is the official data collection of MMFundus (Multimodal Fundus) dataset.☆14Feb 2, 2026Updated 7 months ago
- [CVPR2024] Parameter Efficient Fine-tuning via Cross Block Orchestration for Segment Anything Model☆12Jul 31, 2024Updated 2 years ago
- Zone Evaluation: Revealing Spatial Bias in Object Detection (TPAMI 2024)☆46Dec 6, 2024Updated last year
- VKnowU: Evaluating Visual Knowledge Understanding in Multimodal LLMs☆18Feb 3, 2026Updated 7 months ago
- Self-supervised adversarial masking for point clouds☆11Jul 12, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ECCV 2024] SAM4MLLM: Enhance Multi-Modal Large Language Model for Referring Expression Segmentation☆52Mar 20, 2025Updated last year
- CLIPCleaner: Cleaning Noisy Labels with CLIP (ACM MM2024)☆16Apr 28, 2025Updated last year
- This repository is the official implementation of "Look-Back: Implicit Visual Re-focusing in MLLM Reasoning".☆98Jul 10, 2025Updated last year
- [ICLR25] Official Implementation of "Decoupling Angles and Strength in Low-rank Adaptation"☆15Dec 12, 2025Updated 9 months ago
- [ECCV 2024] Make Your ViT-based Multi-view 3D Detectors Faster via Token Compression☆53Sep 21, 2024Updated last year
- Rui Qian, Xin Yin, Dejing Dou†: Reasoning to Attend: Try to Understand How <SEG> Token Works (CVPR 2025)☆55Feb 4, 2026Updated 7 months ago
- Official PyTorch implementation of our AAAI 2026 paper, "YOLO-IOD: Towards Real Time Incremental Object Detection"☆46Apr 14, 2026Updated 5 months ago
- [CVPR 2025 Highlight] Official Pytorch codebase for paper: "Assessing and Learning Alignment of Unimodal Vision and Language Models"☆61Aug 15, 2025Updated last year
- This repo is the official pytorch implementation of the paper: CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-V…☆43Sep 10, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- U版yolov5 2.0的tensorrt加速☆37Aug 3, 2020Updated 6 years ago
- [NeurIPS 2024 Oral] RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation☆20Dec 22, 2024Updated last year
- 图书《视觉自监督模型DINOv3:原理、训练到部署》学习导航☆65Jul 6, 2026Updated 2 months ago
- Jax, Flax, examples (ImageClassification, SemanticSegmentation, and more...)☆10May 10, 2025Updated last year
- ☆14Mar 1, 2023Updated 3 years ago
- ☆15May 30, 2026Updated 3 months ago
- [ICLR 2026] Empowering Small VLMs to Think with Dynamic Memorization and Exploration☆19Mar 18, 2026Updated 5 months ago