[ICCV 2025] What Changed? Detecting and Evaluating Instruction-Guided Image Edits with Multimodal Large Language Models
☆16Nov 3, 2025Updated 10 months ago
Alternatives and similar repositories for DICE
Users that are interested in DICE are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [CVPR2026 Findings] VHS: Verifier on Hidden States, an efficient inference-time scaling verification framework for DiT-based image genera…☆16Mar 25, 2026Updated 5 months ago
- ☆17May 19, 2026Updated 4 months ago
- [ICCV 2025] Official repository of the paper "Talking to DINO: Bridging Self-Supervised Vision Backbones with Language for Open-Vocabular…☆203Jul 28, 2026Updated last month
- Recurrence Meets Transformers for Universal Multimodal Retrieval☆15Dec 15, 2025Updated 9 months ago
- [ICLR 2026] "Inverse Virtual Try-On: Generating Multi-Category Product-Style Images from Clothed Individuals"☆50Mar 6, 2026Updated 6 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ECCV-W] Official repo for the paper "ComiCap: A VLMs pipeline for dense captioning of Comic Panels"☆15Nov 20, 2024Updated last year
- [CVPR 2025] Augmenting Multimodal LLMs with Self-Reflective Tokens for Knowledge-based Visual Question Answering☆57Jul 14, 2025Updated last year
- [CVPR 2024 Highlight] Official repository of the paper "The devil is in the fine-grained details: Evaluating open-vocabulary object detec…☆68Apr 4, 2025Updated last year
- Model merging, task-vector rebasin, and fine-tuning for vision and LLM models.☆35Updated this week
- Learning to Count without Annotations☆24May 24, 2024Updated 2 years ago
- [IJCAI 2025] Image Captioning Evaluation in the Age of Multimodal LLMs: Challenges and Future Perspectives☆37Nov 25, 2025Updated 9 months ago
- [WACV 2026] Official implementation of the paper: “CountingDINO: A Training-free Pipeline for Exemplar-based Class-Agnostic Counting”☆65Jun 22, 2026Updated 2 months ago
- Django App per l'autocompilazione dei moduli missione☆32Jan 7, 2026Updated 8 months ago
- Evolutionary F1 Race Strategy [GECCO 2023] - Genetic Algorithms applied to F1 race strategy using the F1 2021 game data and real data fro…☆10Aug 22, 2023Updated 3 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- [ICPR 2024] Exemplar-free continual deepfake detector that leverages CLIP and domain-specific multi-modal prompts☆15Aug 1, 2024Updated 2 years ago
- This repository contains a curated list of research papers and resources focusing on saliency and scanpath prediction, human attention, h…☆66May 9, 2025Updated last year
- [ICLR 2025] - Cross the Gap: Exposing the Intra-modal Misalignment in CLIP via Modality Inversion☆70Nov 30, 2025Updated 9 months ago
- [ICCV 2025] MissRAG: Addressing the Missing Modality Challenge in Multimodal Large Language Models☆26May 12, 2026Updated 4 months ago
- Effective caching in differentially-private databases (SOSP '23)☆13Nov 1, 2023Updated 2 years ago
- FreeDA: Training-Free Open-Vocabulary Segmentation with Offline Diffusion-Augmented Prototype Generation (CVPR 2024)☆51Aug 28, 2024Updated 2 years ago
- Animation of an SMPLX character in an augmented reality application☆19Aug 22, 2024Updated 2 years ago
- ☆43Oct 6, 2025Updated 11 months ago
- ☆13Dec 12, 2022Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Official PyTorch implementation for "Merging and Splitting Diffusion Paths for Semantically Coherent Panoramas", presenting the Merge-Att…☆16Jul 9, 2025Updated last year
- Official PyTorch implementation of the WACV 2025 Oral paper "Composed Image Retrieval for Training-FREE DOMain Conversion".☆46Aug 31, 2025Updated last year
- Official PyTorch Implementation of "Rethinking HTG Evaluation: Bridging Generation and Recognition" (Oral) - 1st Workshop on Critical Eva…☆17Sep 23, 2024Updated last year
- The Land-Diffuser is a novel application of the Denoising Diffusion Probabilistic Model (DDPM) in the realm of 3D Talking Head generation…☆13Dec 23, 2023Updated 2 years ago
- Open source code for Mouth Haptics in VR using a Heaset Ultrasound Phased Array☆10Feb 21, 2024Updated 2 years ago
- PyTorch code for BMVC 2019 paper: Embodied Vision-and-Language Navigation with Dynamic Convolutional Filters☆20Jan 4, 2023Updated 3 years ago
- ☆18Oct 1, 2021Updated 4 years ago
- The code of paper "High-order Correlation Preserved Incomplete Multi-view Subspace Clustering" accepted by IEEE TIP2022☆12Jan 19, 2022Updated 4 years ago
- Repository for "CoMix: Comprehensive Benchmark for Multi-Task Comic Understanding"☆18Nov 20, 2024Updated last year
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- [CBMI 2024 Best Paper] Official repository of the paper "Is CLIP the main roadblock for fine-grained open-world perception?".☆31May 12, 2025Updated last year
- [COG24] - Official repository of "OfflineMania: A Benchmark Environment for Offline Reinforcement Learning in Racing Games"☆12Jul 15, 2024Updated 2 years ago
- ☆12Updated this week
- [ECCV'24] Official Implementation of Autoregressive Visual Entity Recognizer.☆14Mar 2, 2024Updated 2 years ago
- Implementation of the paper: "BRAVE : Broadening the visual encoding of vision-language models"☆26Jun 22, 2026Updated 2 months ago
- C++, OpenMP and CUDA implementation of Mean Shift clustering algorithm☆14Apr 27, 2020Updated 6 years ago
- [ICCV 2023] Official implementation of "Keep It SimPool: Who Said Supervised Transformers Suffer from Attention Deficit?".☆102Dec 5, 2023Updated 2 years ago