π₯ [ICLR 2025] Official PyTorch Model "Visual Haystacks: A Vision-Centric Needle-In-A-Haystack Benchmark"
β26Feb 9, 2025Updated last year
Alternatives and similar repositories for mirage
Users that are interested in mirage are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- π₯ [ICLR 2025] Official Benchmark Toolkits for "Visual Haystacks: A Vision-Centric Needle-In-A-Haystack Benchmark"β44Nov 21, 2025Updated 10 months ago
- Echo: "Constantly Improving Image Models Need Constantly Improving Benchmarks" (ICLR 2026)β20Jan 29, 2026Updated 7 months ago
- π₯ [NeurIPS 2025] Official implementation of "Generate, but Verify: Reducing Visual Hallucination in Vision-Language Models with Retrospeβ¦β59Jan 22, 2026Updated 8 months ago
- β12Dec 20, 2024Updated last year
- π₯ [CVPR 2024] Official implementation of "See, Say, and Segment: Teaching LMMs to Overcome False Premises (SESAME)"β47Jun 16, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean β’ AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- π Official pytorch implementation of "D2ADA: Dynamic Density-aware Active Domain Adaptation for Semantic Segmentation. Wu et al. ECCV 20β¦β25Feb 2, 2023Updated 3 years ago
- Fast, free, easy, and object-agnostic video anonymizationβ12Dec 12, 2020Updated 5 years ago
- VHTestβ16Oct 31, 2024Updated last year
- Through MATLAB Computer Vision and Image Processing | Mathworksβ10Aug 18, 2023Updated 3 years ago
- YesBut - Multimodal Satire Comprehension Datasetβ20Oct 23, 2024Updated last year
- β66Jun 27, 2024Updated 2 years ago
- β10Sep 25, 2019Updated 6 years ago
- Official Repository of VisGym: Diverse, Customizable, Scalable Environments for Multimodal Agentsβ115May 3, 2026Updated 4 months ago
- Official This-Is-My Dataset published in CVPR 2023β16Jul 18, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform β’ AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- β16Oct 13, 2025Updated 11 months ago
- β21Nov 13, 2023Updated 2 years ago
- Single-pass Adaptive Image Tokenization for Minimum Program Search | What's the Kolmogorov Complexity of an Image?β45Jul 26, 2025Updated last year
- π· Python package and CLI utility to create photo mosaics - now with GPU supportβ20Mar 6, 2026Updated 6 months ago
- Code release for "MORE: Multi-mOdal REtrieval Augmented Generative Commonsense Reasoning"β11Oct 11, 2024Updated last year
- [ICLR 2025] Distilled Decoding 1: One-step Sampling of Image Auto-regressive Models with Flow Matchingβ19Apr 21, 2025Updated last year
- β17Jan 10, 2024Updated 2 years ago
- Improved Implementation for Training GLIGEN: Open-Set Grounded Text-to-Image Generationβ46Jun 1, 2024Updated 2 years ago
- β27Feb 8, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform β’ AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- β28Jul 22, 2026Updated 2 months ago
- Repository for paper Visual-RAG: Benchmarking Text-to-Image Retrieval Augmented Generation for Visual Knowledge Intensive Queriesβ13Jul 16, 2026Updated 2 months ago
- Seeing from Another Perspective: Evaluating Multi-View Understanding in MLLMsβ71Mar 22, 2026Updated 6 months ago
- a set of tools for computer vision processingβ18Jul 9, 2016Updated 10 years ago
- [BMVC 2022] Information Theoretic Representation Distillationβ19Oct 6, 2023Updated 2 years ago
- Code release for "UnSAMv2: Self-Supervised Learning Enables Segment Anything at Any Granularity"β86Sep 15, 2026Updated last week
- domain adaptation with CF distance for medical image segmentationβ27Apr 12, 2022Updated 4 years ago
- [NeurIPS 2024 D&B] DetectRL: Benchmarking LLM-Generated Text Detection in Real-World Scenariosβ16Nov 19, 2024Updated last year
- β25Aug 2, 2024Updated 2 years ago
- End-to-end encrypted email - Proton Mail β’ AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- π₯ [CVPR2024] Official implementation of "Self-correcting LLM-controlled Diffusion Models (SLD)β187Apr 9, 2024Updated 2 years ago
- ζι³ζε -ζι³ζ―δΈͺζθΆ£ηδΈθ₯Ώβ16Oct 16, 2024Updated last year
- Official implementation of "Video-Foley: Two-Stage Video-To-Sound Generation via Temporal Event Condition For Foley Sound". IEEE TASLP 20β¦β19Feb 27, 2026Updated 6 months ago
- [CVPR'25] Attention IoU: Examining Biases in CelebA using Attention Mapsβ13Mar 26, 2025Updated last year
- β17Jun 10, 2025Updated last year
- iLLaVA: An Image is Worth Fewer Than 1/3 Input Tokens in Large Multimodal Models (ICLR2026)β23Jun 24, 2026Updated 3 months ago
- (ICCV2025) Official repository of paper "ViSpeak: Visual Instruction Feedback in Streaming Videos"β54Jul 1, 2025Updated last year