[ICLR 2025] Official Pytorch Implementation of MMR: A Large-scale Benchmark Dataset for Multi-target and Multi-granularity Reasoning Segmentation
☆28Apr 3, 2025Updated last year
Alternatives and similar repositories for MMR
Users that are interested in MMR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICCV 2025] Official implementation of "InstructSeg: Unifying Instructed Visual Segmentation with Multi-modal Large Language Models"☆56Feb 10, 2025Updated last year
- Paper list for LLM/MLLM-based image segmentation☆48Dec 24, 2025Updated 8 months ago
- [ICLR2025] Text4Seg: Reimagining Image Segmentation as Text Generation☆176Nov 8, 2025Updated 9 months ago
- Official Pytorch Implementation of Unsupervised Image Denoising With Frequency Domain Knowledge (BMVC2021 Oral Accepted Paper)☆24Mar 15, 2022Updated 4 years ago
- Revisiting Multi-Task Visual Representation Learning☆22Jan 21, 2026Updated 7 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- A curated list of awesome remote sensing visual generative models, papers, datasets, and resources. This repository focuses exclusively o…☆25Aug 17, 2026Updated 2 weeks ago
- ☆23Jan 24, 2024Updated 2 years ago
- ☆30Sep 2, 2025Updated 11 months ago
- Sa2VA-i is an improved version of the popular Sa2VA model☆17Nov 25, 2025Updated 9 months ago
- ALTo: Adaptive-Length Tokenizer for Autoregressive Mask Generation☆30May 27, 2025Updated last year
- Code release for "SegLLM: Multi-round Reasoning Segmentation"☆129Feb 20, 2025Updated last year
- Paper List on Earth Observation in the Foundation Model Era☆32Aug 4, 2026Updated 3 weeks ago
- Official PyTorch implementation of “MaskRIS: Semantic Distortion-aware Data Augmentation for Referring Image Segmentation”☆18Dec 5, 2024Updated last year
- ☆16Dec 15, 2025Updated 8 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ECCV2026] X2SAM: Any Segmentation in Images and Videos☆102Jul 13, 2026Updated last month
- Project Page For "Seg-Zero: Reasoning-Chain Guided Segmentation via Cognitive Reinforcement"☆639Jan 17, 2026Updated 7 months ago
- A curated list of publications on image and video segmentation leveraging Multimodal Large Language Models (MLLMs), highlighting state-of…☆233Updated this week
- [NeurIPS 2024 Oral] RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation☆20Dec 22, 2024Updated last year
- An up-to-date & curated list of awesome layout to image papers, methods & resources.☆13Jun 28, 2024Updated 2 years ago
- Code for the paper Open-Vocabulary Attention Maps with Token Optimization for Semantic Segmentation in Diffusion Models @ CVPR 2024☆70Jun 14, 2024Updated 2 years ago
- PathMR: Multimodal Visual Reasoning for Interpretable Pathology Analysis☆15Aug 27, 2025Updated last year
- ☆17Nov 15, 2025Updated 9 months ago
- Referring Change Detection in Remote Sensing Imagery☆15Jan 17, 2026Updated 7 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- InstructSAM: A Training-Free Framework for Instruction-Oriented Remote Sensing Object Recognition (NeurIPS 2025)☆117Jul 30, 2026Updated last month
- ☆37Jul 1, 2024Updated 2 years ago
- ☆47Oct 3, 2023Updated 2 years ago
- Rui Qian, Xin Yin, Dejing Dou†: Reasoning to Attend: Try to Understand How <SEG> Token Works (CVPR 2025)☆55Feb 4, 2026Updated 6 months ago
- [CVPR2026]PixDLM: A Dual-Path Multimodal Language Model for UAV Reasoning Segmentation☆30Aug 24, 2026Updated last week
- [NeurIPS 2024] COVE: Unleashing the Diffusion Feature Correspondence for Consistent Video Editing☆25Jul 31, 2026Updated last month
- [ISPRS2026] DVGBench: Implicit-to-Explicit Visual Grounding Benchmark in UAV Imagery with Large Vision-Language Models☆33Mar 24, 2026Updated 5 months ago
- SOLACE: Improving Text-to-Image Generation with Intrinsic Self-Confidence Rewards (CVPR 2026)☆17Jun 2, 2026Updated 2 months ago
- ☆16Sep 25, 2025Updated 11 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Landsat-Bench: Datasets and Benchmarks for Landsat Foundation Models☆20Jun 18, 2025Updated last year
- ☆47Apr 16, 2026Updated 4 months ago
- Official PyTorch implementation of CorrespondentDream: Enhancing 3D Fidelity of Text-to-3D using Cross-View Correspondences (CVPR 2024 Po…☆19Apr 29, 2024Updated 2 years ago
- This repo holds the official code and data for "Unveiling Parts Beyond Objects: Towards Finer-Granularity Referring Expression Segmentati…☆74Jun 3, 2024Updated 2 years ago
- ☆27Jun 25, 2026Updated 2 months ago
- Open Source Road Datasets☆19Aug 30, 2024Updated 2 years ago
- [TIP 2025] Self-Calibrated CLIP for Training-Free Open-Vocabulary Segmentation☆73Aug 10, 2026Updated 3 weeks ago