The official repo of our work "Pensieve: Retrospect-then-Compare mitigates Visual Hallucination"
☆15May 4, 2024Updated 2 years ago
Alternatives and similar repositories for Pensieve
Users that are interested in Pensieve are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Data pre-processing and training code on Open-X-Embodiment with pytorch☆11Jan 20, 2025Updated last year
- [NeurIPS 2024] Efficient Large Multi-modal Models via Visual Context Compression☆66Feb 19, 2025Updated last year
- Attention-Enhanced Cross-modal Localization Between Spherical Images and Point Clouds (IEEE Sensors Journal)☆12May 1, 2023Updated 3 years ago
- [CVPR 2024] This is official implementation of our CVPR 2024 paper "Building a Strong Pre-Training Baseline for Universal 3D Large-Scale …☆17Jun 11, 2024Updated 2 years ago
- [CVPR 2024] Retrieval-Augmented Image Captioning with External Visual-Name Memory for Open-World Comprehension☆66Apr 8, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Accelerating Vision-Language Pretraining with Free Language Modeling (CVPR 2023)☆31May 15, 2023Updated 3 years ago
- [EMNLP 2022] Fine-grained Category Discovery under Coarse-grained supervision with Hierarchical Weighted Self-contrastive Learning☆14Jun 22, 2024Updated 2 years ago
- GET-Zero: Graph Embodiment Transformer for Zero-shot Embodiment Generalization☆60Mar 5, 2026Updated 5 months ago
- Official Repository for NeurIPS'25 Paper "Tool-Augmented Spatiotemporal Reasoning for Streamlining Video Question Answering Task"☆23May 18, 2026Updated 3 months ago
- Source code for WWW 2019 paper "Efficient Path Prediction for Semi-Supervised and Weakly Supervised Hierarchical Text Classification"☆13May 3, 2019Updated 7 years ago
- We introduce new approach, Token Reduction using CLIP Metric (TRIM), aimed at improving the efficiency of MLLMs without sacrificing their…☆22Jan 11, 2026Updated 7 months ago
- [CVPR 2026] See Less, See Right: Bi-directional Perceptual Shaping For Multimodal Reasoning☆23Jun 28, 2026Updated 2 months ago
- EgoBlind: Towards Egocentric Visual Assistance for the Blind (NeurIPS'25, D&B Track)☆26Apr 20, 2026Updated 4 months ago
- Official Code of our AAAI-24 Paper: "Generative Multi-modal Knowledge Retrieval with Large Language Models".☆28Sep 15, 2025Updated 11 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Reinforcement Learning framework for Temporal Goals☆11Mar 6, 2023Updated 3 years ago
- ☆24Apr 26, 2026Updated 4 months ago
- [NeurIPS 2022] Official code for REVIVE: Regional Visual Representation Matters in Knowledge-Based Visual Question Answering☆105Apr 6, 2025Updated last year
- Code for "Learning Generalizable Robotic Reward Functions from "In-The-Wild" Human Videos"☆28Oct 25, 2021Updated 4 years ago
- [AAAI 2025] Does VLM Classification Benefit from LLM Description Semantics?☆26Aug 5, 2025Updated last year
- Code for the papers "Induction of Subgoal Automata for Reinforcement Learning" (AAAI-20) and "Induction and Exploitation of Subgoal Autom…☆14Aug 15, 2023Updated 3 years ago
- [Coursera] Natural Language Processing Specialization by "deeplearning.ai".☆15Jun 26, 2020Updated 6 years ago
- [CVPR 2024 Highlight] OPERA: Alleviating Hallucination in Multi-Modal Large Language Models via Over-Trust Penalty and Retrospection-Allo…☆415Aug 24, 2024Updated 2 years ago
- Official repository for "IntentQA: Context-aware Video Intent Reasoning" from ICCV 2023.☆26Nov 29, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official implementation of "Vision LLMs Are Bad at Hierarchical Visual Understanding, and LLMs Are the Bottleneck" [CVPR'26]☆16Nov 10, 2025Updated 9 months ago
- ☆11May 7, 2022Updated 4 years ago
- ☆18Aug 1, 2024Updated 2 years ago
- implementation for "learning weighted deterministic automata from queries and counterexamples", neurips 2019☆18Jan 8, 2020Updated 6 years ago
- This is the official repo for Densely-Anchored Sampling for Deep Metric Learning (ECCV 22).☆16May 24, 2024Updated 2 years ago
- DreamDance: Personalized Text-to-video Generation by Combining Text-to-Image Synthesis and Motion Transfer☆14Dec 16, 2022Updated 3 years ago
- Code for a model-based version of Constrained Policy Optimization☆11May 6, 2021Updated 5 years ago
- [IEEE IROS'25] GSPR: Multimodal Place Recognition using 3D Gaussian Splatting for Autonomous Driving☆54Apr 2, 2026Updated 5 months ago
- Official Codebase of "A Unified Audio-Visual Learning Framework for Localization, Separation, and Recognition" (ICML 2023)☆12Jun 1, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆10Nov 23, 2023Updated 2 years ago
- [CVPR 2024] "Towards Robust Audiovisual Segmentation in Complex Environments with Quantization-based Semantic Decomposition"☆11Feb 27, 2024Updated 2 years ago
- [ECCV 26'] Official codebase for the paper LaViT☆35Jul 30, 2026Updated last month
- ☆23Aug 9, 2025Updated last year
- [CVPR 2024] Code and datasets for 'Learning Spatial Features from Audio-Visual Correspondence in Egocentric Videos'☆14Jun 16, 2024Updated 2 years ago
- ☆45Dec 9, 2024Updated last year
- Cross-Layer Independent Deformable Description for Efficient and Discriminative Local Feature Representation☆30Aug 11, 2026Updated 3 weeks ago