Document Haystacks: Vision-Language Reasoning Over Piles of 1000+ Documents, CVPR 2025
☆26Jan 25, 2025Updated last year
Alternatives and similar repositories for dochaystacks
Users that are interested in dochaystacks are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Third place of 2021 IEEE GRSS Data Fusion Contest: Track MSD☆10Mar 31, 2021Updated 5 years ago
- Teeth Mold Point Cloud Completion Via Data Augmentation and Hybrid RL-GAN (Paper Code)☆13May 23, 2023Updated 3 years ago
- [CVPR 2023] Code for "Improving Visual Grounding by Encouraging Consistent Gradient-based Explanations"☆19Oct 10, 2023Updated 2 years ago
- [ACM MM 2021] A causal perspective for compositional action recognition, providing a counterfactual debiasing inference implementation to…☆20May 5, 2022Updated 4 years ago
- ☆15Apr 23, 2026Updated 3 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- ☆14Jul 13, 2021Updated 5 years ago
- Official InfiniBench: A Benchmark for Large Multi-Modal Models in Long-Form Movies and TV Shows☆20Nov 4, 2025Updated 9 months ago
- ☆22Mar 3, 2023Updated 3 years ago
- ☆48Feb 9, 2025Updated last year
- ☆16Jul 10, 2023Updated 3 years ago
- Code and resources for EMNLP 2022 paper on 'Robustness of Fusion-based Multimodal Classifiers to Cross-Modal Content Dilutions'☆10Mar 11, 2024Updated 2 years ago
- [ACL 2019] Context-aware Embedding for Targeted Aspect-based Sentiment Analysis☆17Nov 12, 2020Updated 5 years ago
- Code for the paper CVPR‘17 “Zero Shot Learning from Noisy Text Description at Part Precision”☆16Nov 22, 2019Updated 6 years ago
- Awesome Papers at the 2026 3D Vision Conference in Vancouver, BC☆19Apr 21, 2026Updated 3 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- IEEE/CVF International Conference on Computer Vision Workshop (2023)☆17Feb 7, 2024Updated 2 years ago
- ☆43Apr 6, 2026Updated 4 months ago
- The official implementation of CVPR 2021 Paper: Improving Weakly Supervised Visual Grounding by Contrastive Knowledge Distillation.☆12Oct 15, 2021Updated 4 years ago
- 【MICCAI 2024, Early Accept】Enhancing Label-efficient Medical Image Segmentation with Text-guided Diffusion Models☆32Sep 10, 2024Updated last year
- Math24o: 高中奥林匹克数学竞赛测评集 High School Olympiad Mathematics Chinese Benchmark☆14Mar 27, 2025Updated last year
- ☆14Dec 9, 2023Updated 2 years ago
- ☆15Jun 11, 2021Updated 5 years ago
- Some examples of image processing based on Opencv☆17Feb 22, 2019Updated 7 years ago
- Code accompanying paper "Fine-Grained Visual Entailment" [ECCV 2022].☆11Oct 31, 2022Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- aigc evals☆10Dec 2, 2023Updated 2 years ago
- Exploring Hierarchical Graph Representation for Large-Scale Zero-Shot Image Classification. ECCV 2022.☆18Jul 12, 2022Updated 4 years ago
- ☆12Jun 2, 2024Updated 2 years ago
- Codebase for RecSys 2024 paper, The Elephant in the Room: Rethinking the Usage of Pre-trained Language Model in Sequential Recommendation☆19Aug 7, 2024Updated 2 years ago
- Artemis Speaker Tools B☆24Apr 4, 2021Updated 5 years ago
- ☆35Oct 21, 2023Updated 2 years ago
- [WACV2023] Intention-Conditioned Long-Term Human Egocentric Action Forecasting @ EGO4D Challenge 2022☆14Sep 3, 2023Updated 2 years ago
- It analyze facial micro-expressions, provides emotional states indicative of PTSD. The tool supports real-time analysis of live video str…☆13May 7, 2024Updated 2 years ago
- ☆15Jan 14, 2026Updated 7 months ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- This is the official implementation for our paper;"LAR:Look Around and Refer".☆30Dec 1, 2022Updated 3 years ago
- Code for Paper: Harnessing Webpage Uis For Text Rich Visual Understanding☆54Dec 12, 2024Updated last year
- Source code and dataset for the CCKS2021 paper "Text-guided Legal Knowledge Graph Reasoning".☆23Jan 5, 2022Updated 4 years ago
- [CVPR 2024 Highlight] ImageNet-D☆47Jul 24, 2026Updated 3 weeks ago
- Repo for ICCV 2021 paper: Beyond Question-Based Biases: Assessing Multimodal Shortcut Learning in Visual Question Answering☆29Jul 1, 2024Updated 2 years ago
- Official repository of ECCV 2024 paper - "HAT: History-Augmented Anchor Transformer for Online Temporal Action Localization"☆20Aug 23, 2024Updated last year
- [CVPR2021] Look before you leap: learning landmark features for one-stage visual grounding.☆50Aug 31, 2021Updated 4 years ago