☆32Feb 8, 2024Updated 2 years ago
Alternatives and similar repositories for Mementos
Users that are interested in Mementos are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [EACL'23] COVID-VTS: Fact Extraction and Verification on Short Video Platforms☆12Sep 26, 2023Updated 2 years ago
- ☆24Jun 18, 2025Updated last year
- PyTorch implementation of DreamerV3, Mastering Diverse Domains through World Models.☆14Feb 16, 2024Updated 2 years ago
- VisualGPTScore for visio-linguistic reasoning☆27Oct 7, 2023Updated 2 years ago
- [ECCV 2024] "REVISION: Rendering Tools Enable Spatial Fidelity in Vision-Language Models"☆14Aug 6, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Identifying Visible Actions in Lifestyle Vlogs☆15Aug 3, 2023Updated 3 years ago
- Official resource for paper Investigating and Mitigating the Multimodal Hallucination Snowballing in Large Vision-Language Models (ACL 20…☆18Aug 12, 2024Updated 2 years ago
- Code and data for ACL 2024 paper on 'Cross-Modal Projection in Multimodal LLMs Doesn't Really Project Visual Attributes to Textual Space'☆18Jul 21, 2024Updated 2 years ago
- Code for paper "Open-Domain Hierarchical Event Schema Induction by Incremental Prompting and Verification"☆17Jul 4, 2023Updated 3 years ago
- An implementation of (Chambers and Jurafsky, 2008), using updated machine learning models, and different training data domains for an ind…☆14Dec 8, 2022Updated 3 years ago
- Papers on fairness☆12Oct 20, 2020Updated 5 years ago
- [ICPRAI 2024] DocumentCLIP: Linking Figures and Main Body Text in Reflowed Documents☆16Apr 4, 2024Updated 2 years ago
- [CVPR'24] HallusionBench: You See What You Think? Or You Think What You See? An Image-Context Reasoning Benchmark Challenging for GPT-4V(…☆342Oct 14, 2025Updated 10 months ago
- ☆23Apr 2, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- An automatic MLLM hallucination detection framework☆19Sep 26, 2023Updated 2 years ago
- ☆38Feb 8, 2024Updated 2 years ago
- ☆21Oct 10, 2023Updated 2 years ago
- [ACL'25] Mosaic-IT: Cost-Free Compositional Data Synthesis for Instruction Tuning☆20Sep 27, 2025Updated 10 months ago
- Codes and files for the paper Are Emergent Abilities in Large Language Models just In-Context Learning☆33Jan 9, 2025Updated last year
- Official repository of the paper "Unsupervised Audio-Visual Lecture Segmentation", WACV 2023☆13Mar 3, 2025Updated last year
- [CVPR 2023] Official code for "Learning Procedure-aware Video Representation from Instructional Videos and Their Narrations"☆56Aug 8, 2023Updated 3 years ago
- ☆46Dec 30, 2024Updated last year
- [EMNLP 2024] Official code for "Beyond Embeddings: The Promise of Visual Table in Multi-Modal Models"☆20Oct 17, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆19Jul 1, 2026Updated last month
- ☆37Oct 7, 2023Updated 2 years ago
- PyTorch code for Improving Commonsense in Vision-Language Models via Knowledge Graph Riddles (DANCE)☆22Nov 29, 2022Updated 3 years ago
- ☆27Jul 20, 2024Updated 2 years ago
- ☆28Oct 18, 2022Updated 3 years ago
- Video-R2: Reinforcing Consistent and Grounded Reasoning in Multimodal Language Models☆19Jan 21, 2026Updated 6 months ago
- ☆49Sep 5, 2024Updated last year
- Dataset and Code for Multimodal Fact Checking and Explanation Generation (Mocheg)☆68Nov 24, 2023Updated 2 years ago
- ☆11May 24, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Visualization tool for CUB-200-2011 part keypoints (Wah et al.).☆10Sep 17, 2021Updated 4 years ago
- Scaling Multi-modal Instruction Fine-tuning with Tens of Thousands Vision Task Types☆32Jul 16, 2025Updated last year
- [ICLR'24] Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning☆296Mar 13, 2024Updated 2 years ago
- Open-source code for ''Graph Neural Networks with Adaptive Frequency Response Filter''.☆25Jul 8, 2022Updated 4 years ago
- VisualOverload (CVPR 2026) is a VQA benchmark for image understanding in dense, high-resolution scenes.☆18May 31, 2026Updated 2 months ago
- Open-source datasets for paper "Fairness in Graph Mining: A Survey".☆19Nov 3, 2022Updated 3 years ago
- First explanation metric (diagnostic report) for text generation evaluation☆62Mar 3, 2025Updated last year