Ref-Diff: Zero-shot Referring Image Segmentation with Generative Models
☆21May 29, 2025Updated last year
Alternatives and similar repositories for Ref-Diff
Users that are interested in Ref-Diff are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [EMNLP 2024] SURf: Teaching Large Vision-Language Models to Selectively Utilize Retrieved Information☆11Oct 11, 2024Updated last year
- Related papers about Referring Image Segmentation (RIS)☆16Dec 26, 2023Updated 2 years ago
- Responsible Visual Editing☆15Jul 10, 2024Updated 2 years ago
- Robust Referring Video Object Segmentation with Cyclic Structural Consistency [ICCV 2023]☆30Mar 13, 2024Updated 2 years ago
- CLIP-based simple image-text matching baseline for COCO and F30K☆15Sep 16, 2021Updated 5 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Handy Utilities for Computer Vision☆12Sep 16, 2026Updated 3 weeks ago
- Referring Image Segmentation Benchmarking with Segment Anything Model (SAM)☆39Apr 7, 2023Updated 3 years ago
- Code for "CARIS: Context-Aware Referring Image Segmentation" [ACM MM2023]☆30Nov 28, 2024Updated last year
- ☆21Jul 6, 2022Updated 4 years ago
- This is an official PyTorch code for our accepted paper "When All We Need is a Piece of the Pie: A Generic Framework for Optimizing Two-w…☆15Jul 7, 2022Updated 4 years ago
- Code for paper "W-RAG: Weakly Supervised Dense Retrieval in RAG for Open-domain Question Answering"☆16Oct 2, 2025Updated last year
- ☆31Nov 7, 2023Updated 2 years ago
- ORES: Open-vocabulary Responsible Visual Synthesis☆14Dec 12, 2023Updated 2 years ago
- [ECCV-24] This is the official implementation of the paper "SEGIC: Unleashing the Emergent Correspondence for In-Context Segmentation".☆29Oct 13, 2024Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- This is for EMNLP 2024 Paper: AppBench: Planning of Multiple APIs from Various APPs for Complex User Instruction☆16Nov 4, 2024Updated last year
- 🏠🔍 Auto check for new apartments in Hamburg from various real estate provides☆16Apr 15, 2026Updated 5 months ago
- ☆14Oct 17, 2024Updated last year
- ☆10Nov 29, 2022Updated 3 years ago
- This is the repository for our paper: Untying the Reversal Curse via Bidirectional Language Model Editing☆11May 25, 2025Updated last year
- Referring Video Object Segmentation / Multi-Object Tracking Repo☆91Jul 27, 2023Updated 3 years ago
- DALL-E for Detection: Language-driven Compositional Image Synthesis for Object Detection☆21Oct 5, 2023Updated 3 years ago
- simple and efficient baselines for practical semantic segmentation with plain ViTs☆19Mar 9, 2024Updated 2 years ago
- Repository for AAAI 2024 paper "Manifold-based Verbalizer Space Re-embedding for Tuning-free Prompt-based Classification"☆10Feb 6, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ICIP 2025] Scribble-Guided Diffusion for Training-free Text-to-Image Generation☆27Oct 2, 2024Updated 2 years ago
- Official PyTorch Implementation of FREE (ICCV'21)☆34Sep 7, 2022Updated 4 years ago
- Numpy/Python implementation of DenseCRF☆24Oct 14, 2022Updated 3 years ago
- ☆34Sep 1, 2025Updated last year
- A minimal implementation of Agentic RAG using GRPO☆17Jun 11, 2025Updated last year
- Code for the ECCV22 paper Demystifying Unsupervised Semantic Correspondence Estimation☆14Oct 18, 2022Updated 3 years ago
- ☆10Dec 23, 2020Updated 5 years ago
- [TMLR] Official PyTorch implementation of "λ-ECLIPSE: Multi-Concept Personalized Text-to-Image Diffusion Models by Leveraging CLIP Latent…☆53Nov 29, 2024Updated last year
- ☆165Jul 19, 2023Updated 3 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- T2I-ReasonBench: Benchmarking Reasoning-Informed Text-to-Image Generation☆39Sep 16, 2025Updated last year
- Code for paper: Unified Text-to-Image Generation and Retrieval☆15Jul 19, 2026Updated 2 months ago
- [ACM MM24] MotionMaster: Training-free Camera Motion Transfer For Video Generation☆103Oct 15, 2024Updated last year
- [CVPR 2026] Ego2Web: A Web Agent Benchmark Grounded in Egocentric Videos☆30Mar 25, 2026Updated 6 months ago
- ALTo: Adaptive-Length Tokenizer for Autoregressive Mask Generation☆29May 27, 2025Updated last year
- This is the official released code for our paper, The Emergence of Objectness: Learning Zero-Shot Segmentation from Videos, which has bee…☆53Apr 14, 2023Updated 3 years ago
- DMAOT ranked 1st in the VOTS 2023 challenge.☆17Dec 21, 2023Updated 2 years ago