Document Haystacks: Vision-Language Reasoning Over Piles of 1000+ Documents, CVPR 2025
☆26Jan 25, 2025Updated last year
Alternatives and similar repositories for dochaystacks
Users that are interested in dochaystacks are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Third place of 2021 IEEE GRSS Data Fusion Contest: Track MSD☆10Mar 31, 2021Updated 5 years ago
- Teeth Mold Point Cloud Completion Via Data Augmentation and Hybrid RL-GAN (Paper Code)☆13May 23, 2023Updated 3 years ago
- 【TNNLS 2021】DONet: Dual-Octave Network for Fast MR Image Reconstruction☆11Jun 4, 2021Updated 5 years ago
- ☆11Jun 2, 2019Updated 7 years ago
- [CVPR 2023] Code for "Improving Visual Grounding by Encouraging Consistent Gradient-based Explanations"☆19Oct 10, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Creative AI for Visual Art and Music slides and demos.☆11May 2, 2023Updated 3 years ago
- ☆15Apr 23, 2026Updated 3 months ago
- ☆14Jul 13, 2021Updated 5 years ago
- Official InfiniBench: A Benchmark for Large Multi-Modal Models in Long-Form Movies and TV Shows☆20Nov 4, 2025Updated 8 months ago
- ☆22Mar 3, 2023Updated 3 years ago
- 📝 Source code for "ECNU-SenseMaker at SemEval-2020 Task 4: Leveraging Heterogeneous Knowledge Resources for Commonsense Validation and E…☆23Jun 17, 2023Updated 3 years ago
- ☆48Feb 9, 2025Updated last year
- [ACL 2019] Context-aware Embedding for Targeted Aspect-based Sentiment Analysis☆17Nov 12, 2020Updated 5 years ago
- Code for the paper CVPR‘17 “Zero Shot Learning from Noisy Text Description at Part Precision”☆16Nov 22, 2019Updated 6 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Awesome Papers at the 2026 3D Vision Conference in Vancouver, BC☆19Apr 21, 2026Updated 3 months ago
- ☆42Apr 6, 2026Updated 3 months ago
- 【MICCAI 2024, Early Accept】Enhancing Label-efficient Medical Image Segmentation with Text-guided Diffusion Models☆32Sep 10, 2024Updated last year
- Math24o: 高中奥林匹克数学竞赛测评集 High School Olympiad Mathematics Chinese Benchmark☆14Mar 27, 2025Updated last year
- ☆15Jun 11, 2021Updated 5 years ago
- Some examples of image processing based on Opencv☆17Feb 22, 2019Updated 7 years ago
- Code accompanying paper "Fine-Grained Visual Entailment" [ECCV 2022].☆11Oct 31, 2022Updated 3 years ago
- Exploring Hierarchical Graph Representation for Large-Scale Zero-Shot Image Classification. ECCV 2022.☆18Jul 12, 2022Updated 4 years ago
- ☆11Jun 2, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This project is my attempt at automating work in Notion.☆17Aug 28, 2025Updated 10 months ago
- 【AAAI 2021】Dual-Octave Convolution for Accelerated Parallel MR Image Reconstruction☆26Oct 16, 2022Updated 3 years ago
- Artemis Speaker Tools B☆24Apr 4, 2021Updated 5 years ago
- [WACV2023] Intention-Conditioned Long-Term Human Egocentric Action Forecasting @ EGO4D Challenge 2022☆14Sep 3, 2023Updated 2 years ago
- It analyze facial micro-expressions, provides emotional states indicative of PTSD. The tool supports real-time analysis of live video str…☆13May 7, 2024Updated 2 years ago
- AAAI 2025: Adapting to Non-Stationary Environments: Multi-Armed Bandit Enhanced Retrieval-Augmented Generation on Knowledge Graphs☆18Nov 9, 2024Updated last year
- ☆15Jan 14, 2026Updated 6 months ago
- This is the official implementation for our paper;"LAR:Look Around and Refer".☆30Dec 1, 2022Updated 3 years ago
- Refactor your code with local LLM in VSCode☆13Mar 14, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code for Paper: Harnessing Webpage Uis For Text Rich Visual Understanding☆54Dec 12, 2024Updated last year
- [CVPR 2024 Highlight] ImageNet-D☆47Updated this week
- NeurIPS 2024: SciFIBench: Benchmarking Large Multimodal Models for Scientific Figure Interpretation☆13May 24, 2025Updated last year
- Repo for ICCV 2021 paper: Beyond Question-Based Biases: Assessing Multimodal Shortcut Learning in Visual Question Answering☆29Jul 1, 2024Updated 2 years ago
- [CVPR2021] Look before you leap: learning landmark features for one-stage visual grounding.☆50Aug 31, 2021Updated 4 years ago
- [IEEE TMM 2025 & ACL 2024 Findings] LLMs as Bridges: Reformulating Grounded Multimodal Named Entity Recognition☆41Jul 19, 2025Updated last year
- ☆23Apr 12, 2022Updated 4 years ago