π₯ [ICLR 2025] Official PyTorch Model "Visual Haystacks: A Vision-Centric Needle-In-A-Haystack Benchmark"
β27Feb 9, 2025Updated last year
Alternatives and similar repositories for mirage
Users that are interested in mirage are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- π₯ [ICLR 2025] Official Benchmark Toolkits for "Visual Haystacks: A Vision-Centric Needle-In-A-Haystack Benchmark"β45Nov 21, 2025Updated 8 months ago
- π₯ [ICML 2026] Official implementation of "Are LRMs Interruptible?"β19Jun 18, 2026Updated last month
- Echo: "Constantly Improving Image Models Need Constantly Improving Benchmarks" (ICLR 2026)β20Jan 29, 2026Updated 6 months ago
- π₯ [NeurIPS 2025] Official implementation of "Generate, but Verify: Reducing Visual Hallucination in Vision-Language Models with Retrospeβ¦β58Jan 22, 2026Updated 6 months ago
- β14Jun 11, 2024Updated 2 years ago
- End-to-end encrypted cloud storage - Proton Drive β’ AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- [ICLR 2025] Vision-Centric Evaluation for Retrieval-Augmented Multimodal Modelsβ63Jan 22, 2025Updated last year
- β11Dec 20, 2024Updated last year
- π₯ [CVPR 2024] Official implementation of "See, Say, and Segment: Teaching LMMs to Overcome False Premises (SESAME)"β47Jun 16, 2024Updated 2 years ago
- π Official pytorch implementation of "D2ADA: Dynamic Density-aware Active Domain Adaptation for Semantic Segmentation. Wu et al. ECCV 20β¦β25Feb 2, 2023Updated 3 years ago
- VHTestβ16Oct 31, 2024Updated last year
- YesBut - Multimodal Satire Comprehension Datasetβ20Oct 23, 2024Updated last year
- β66Jun 27, 2024Updated 2 years ago
- Official Repository of VisGym: Diverse, Customizable, Scalable Environments for Multimodal Agentsβ114May 3, 2026Updated 3 months ago
- Official This-Is-My Dataset published in CVPR 2023β16Jul 18, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform β’ AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [CVPR 2026] EasyOmnimatte: Taming Pretrained Inpainting Diffusion Models for End-to-End Video Layered Decompositionβ25Dec 29, 2025Updated 7 months ago
- Tensorflow implementation of Shearlab, including a python wrapper of the Julia Shearlab APiβ12Apr 22, 2021Updated 5 years ago
- Baseline to denoise + learn descriptors in N-HPatchesβ17Mar 14, 2019Updated 7 years ago
- Single-pass Adaptive Image Tokenization for Minimum Program Search | What's the Kolmogorov Complexity of an Image?β45Jul 26, 2025Updated last year
- π· Python package and CLI utility to create photo mosaics - now with GPU supportβ18Mar 6, 2026Updated 5 months ago
- POMA-3D: The Point Map Way to 3D Scene Understanding.β16Nov 9, 2025Updated 9 months ago
- Code release for "MORE: Multi-mOdal REtrieval Augmented Generative Commonsense Reasoning"β11Oct 11, 2024Updated last year
- [ICLR 2025] Distilled Decoding 1: One-step Sampling of Image Auto-regressive Models with Flow Matchingβ19Apr 21, 2025Updated last year
- β27Feb 8, 2025Updated last year
- Simple, predictable pricing with DigitalOcean hosting β’ AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Implementation of Baseline for Scene Text-to-Scene Text Translationβ19Mar 30, 2025Updated last year
- Seeing from Another Perspective: Evaluating Multi-View Understanding in MLLMsβ71Mar 22, 2026Updated 4 months ago
- a set of tools for computer vision processingβ18Jul 9, 2016Updated 10 years ago
- β19Jun 14, 2024Updated 2 years ago
- Iterate on LLM-based structured generation forward and backwardβ23Mar 20, 2025Updated last year
- β17Jul 25, 2023Updated 3 years ago
- Evaluation code and datasets for the ACL 2024 paper, VISTA: Visualized Text Embedding for Universal Multi-Modal Retrieval. The original cβ¦β48Nov 16, 2024Updated last year
- We introduce new approach, Token Reduction using CLIP Metric (TRIM), aimed at improving the efficiency of MLLMs without sacrificing theirβ¦