🍑 relsim: Relational Visual Similarity | pip install relsim 🌍 (CVPR 2026)
☆90Jul 22, 2026Updated 2 months ago
Alternatives and similar repositories for relsim
Users that are interested in relsim are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆20Dec 22, 2025Updated 9 months ago
- 🌋👵🏻 Yo'LLaVA: Your Personalized Language and Vision Assistant (NeurIPS 2024)☆125Mar 26, 2025Updated last year
- [ICLR 2026 Oral] TRACE: Your Diffusion Model Is Secretly an Instance Edge Detector☆18Mar 2, 2026Updated 7 months ago
- A curated list of Awesome Personalized Large Multimodal Models resources☆60Aug 10, 2026Updated last month
- ☆17Oct 24, 2024Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Official Repository of Personalized Visual Instruct Tuning☆34Mar 6, 2025Updated last year
- [WACV 2026 Oral] LASER: Lip Landmark Assisted Speaker Detection for Robustness official implemntation☆30Feb 26, 2026Updated 7 months ago
- [CVPR'26] UniGame code implementation☆20Apr 21, 2026Updated 5 months ago
- The source code of the paper: Image vectorization and editing via linear gradient layer decomposition.☆38Dec 14, 2023Updated 2 years ago
- [CVPR 2026] Video-as-Answer: Predict and Generate Next Video Event with Joint-GRPO☆119Feb 28, 2026Updated 7 months ago
- Strefer: Empowering Video LLMs with Space-Time Referring and Reasoning via Synthetic Instruction Data☆19Jun 2, 2026Updated 4 months ago
- ✏️ Edit One for All: Interactive Batch Image Editing (CVPR 2024)☆68Aug 8, 2024Updated 2 years ago
- [ACM MM25] LongWriter-V: Enabling Ultra-Long and High-Fidelity Generation in Vision-Language Models☆25Mar 29, 2025Updated last year
- [ICLR 2026] Follow-Your-Shape: This repo is the official implementation of "Follow-Your-Shape: Shape-Aware Image Editing via Trajectory-…☆71Apr 10, 2026Updated 5 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Complex-Edit: CoT-Like Instruction Generation for Complexity-Controllable Image Editing Benchmark☆30Apr 22, 2025Updated last year
- [CVPR 2026 Highlight] Official implementation of BiCo: Composing Concepts from Images and Videos via Concept-prompt Binding☆87May 31, 2026Updated 4 months ago
- ☆148Dec 19, 2025Updated 9 months ago
- ☆17Jun 9, 2025Updated last year
- A Unified Framework for Stylized Text Editing and Generation in Graphic Design Images☆16Jan 6, 2026Updated 9 months ago
- 🦎 Yo'Chameleon: Your Personalized Chameleon (CVPR 2025)☆151May 13, 2025Updated last year
- Code for the paper: "StoryReasoning Dataset: Using Chain-of-Thought for Scene Understanding and Grounded Story Generation"☆42May 16, 2025Updated last year
- Stress-Testing Image Generation Models with Explainable Human Evaluation on Open-ended Real-World Tasks [ICLR 2026]☆32Apr 2, 2026Updated 6 months ago
- Official repo for paper "EMMA: Efficient Multimodal Understanding, Generation, and Editing with a Unified Architecture."☆62Dec 16, 2025Updated 9 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆18Oct 15, 2025Updated 11 months ago
- https://little-misfit.github.io/GRAG-Image-Editing/☆119Nov 27, 2025Updated 10 months ago
- [CVPR 2026 main] Imagine Before Concentration: Diffusion-Guided Registers Enhance Partially Relevant Video Retrieval☆39Jun 1, 2026Updated 4 months ago
- [ICCV 2025] Object-centric Video Question Answering with Visual Grounding and Referring☆24Aug 8, 2025Updated last year
- [CVPR 2026] Thinking with Programming Vision: Towards a Unified View for Thinking with Images☆74Jan 23, 2026Updated 8 months ago
- [ICLR 2026] The official repository for paper "ThinkMorph: Emergent Properties in Multimodal Interleaved Chain-of-Thought Reasoning"☆200May 1, 2026Updated 5 months ago
- Accepted to IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2026.☆21Jun 2, 2026Updated 4 months ago
- [NeurIPS 2025] Official implementation of "Instance-Level Composed Image Retrieval".☆55Updated this week
- [ICLR 2026 Oral] Through the Lens of Contrast: Self-Improving Visual Reasoning in VLMs☆22Apr 29, 2026Updated 5 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [CVPR 2026] DiP: Taming Diffusion Models in Pixel Space☆81Jun 15, 2026Updated 3 months ago
- Repo for paper: https://arxiv.org/abs/2404.06479☆30Oct 3, 2024Updated 2 years ago
- VFXMaster: Unlocking Dynamic Visual Effect Generation via In-Context Learning☆67Apr 7, 2026Updated 6 months ago
- ☆30Apr 28, 2026Updated 5 months ago
- ☆27Apr 25, 2025Updated last year
- [CVPR 2026] Reinforcing Text-Rich Video Reasoning with Visual Rumination☆29Jun 5, 2026Updated 4 months ago
- ☆43Sep 1, 2025Updated last year