[EMNLP 2023] Hi-ToM benchmark
β21Oct 11, 2025Updated 10 months ago
Alternatives and similar repositories for Hi-ToM_dataset
Users that are interested in Hi-ToM_dataset are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- π² Code and benchmark for our COLM 2025 paper - "Thought Tracing: Hypothesis-Driven Theory-of-Mind Reasoning for Large Language Models"β15Aug 8, 2025Updated last year
- [AAAI 2025 ππ«ππ₯] MuMA-ToM: Multi-modal Multi-Agent Theory of Mindβ41Jun 28, 2026Updated last month
- β40Jul 16, 2023Updated 3 years ago
- ToMBench: Benchmarking Theory of Mind in Large Language Models, ACL 2024.β69Jun 24, 2024Updated 2 years ago
- Machine Theory of Mind Reading List. Built upon EMNLP Findings 2023 Paper: Towards A Holistic Landscape of Situated Theory of Mind in Larβ¦β154Jun 11, 2026Updated 2 months ago
- Simple, predictable pricing with DigitalOcean hosting β’ AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Implementation of the Decrypto benchmark for multi-agent reasoning and theory of mind.β23Jan 19, 2026Updated 6 months ago
- [NeurIPS 2025 ππ©π¨ππ₯π’π π‘π] AutoToM: Scaling Model-based Mental Inference via Automated Agent Modelingβ46Jun 28, 2026Updated last month
- Hypothetical Minds is an autonomous LLM-based agent for diverse multi-agent settings, integrating a Theory of Mind module Theory of Mind β¦β65Jul 13, 2024Updated 2 years ago
- β12Jan 25, 2024Updated 2 years ago
- Examples of using Galileo for better ML data quality!!β13Feb 5, 2026Updated 6 months ago
- [EMNLP 2023] Official repository for Dialogue Chain-of-Thought Distillation (DONUT & DOCTOR)β11Nov 15, 2023Updated 2 years ago
- β36Mar 10, 2025Updated last year
- [ICLR 2026] Skill-Targeted Adaptive Trainingβ26Mar 12, 2026Updated 5 months ago
- [ICML 2024] Language Models Represent Beliefs of Self and Othersβ37Sep 26, 2024Updated last year
- Virtual machines for every use case on DigitalOcean β’ AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- β30Mar 11, 2025Updated last year
- β40Mar 20, 2017Updated 9 years ago
- Code for ACL 2023 paper "BOLT: Fast Energy-based Controlled Text Generation with Tunable Biases".β22Sep 7, 2023Updated 2 years ago
- Offline Policy Evaluation via Adaptive Weighting with Data from Contextual Banditsβ11Oct 21, 2024Updated last year
- Public repository for "Think Twice: Perspective-Taking Improves Large Language Modelsβ Theory-of-Mind Capabilities".β25Aug 16, 2023Updated 3 years ago
- Simple phoenix setup for padded window managementβ13Apr 25, 2018Updated 8 years ago
- Mental state inference from observable behaviorβ15Dec 3, 2021Updated 4 years ago
- β14May 9, 2024Updated 2 years ago
- Modeling Attention and Binding in the Brain through Bidirectional Recurrent Gatingβ16Jan 29, 2026Updated 6 months ago
- End-to-end encrypted email - Proton Mail β’ AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Code for the paper "Symmetric Machine Theory of Mind", presented at ICML 2022.β12Jul 18, 2022Updated 4 years ago
- This repo contains the ToMnet+ model for preference inference. Developed by Yun-Shiuan, Edwinn, Hsin-Yi, and Elaine.β10Feb 24, 2023Updated 3 years ago
- AAAI 2024-Controllable Mind Visual Diffusion Modelβ15Dec 18, 2023Updated 2 years ago
- β30Oct 29, 2024Updated last year
- Offline Multi-Agent Reinforcement Learning Implementations: Solving Overcooked Game with Data-Driven Methodβ48Sep 11, 2024Updated last year
- β12Aug 30, 2021Updated 4 years ago
- [CVPR'25] Attention IoU: Examining Biases in CelebA using Attention Mapsβ13Mar 26, 2025Updated last year
- possibly useful materials for learning RWKV language model.β27Jun 8, 2023Updated 3 years ago
- SeRL: Self-Play Reinforcement Learning for Large Language Models with Limited Dataβ24Jan 24, 2026Updated 6 months ago
- AI Agents on DigitalOcean Gradient AI Platform β’ AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Find context neurons in Pythia models.β13Jun 13, 2023Updated 3 years ago
- Code for our SIGGRAPH 2023 paper, "Acting as Inverse Inverse Planning"β20Apr 21, 2023Updated 3 years ago
- β11Dec 22, 2021Updated 4 years ago
- official repo for the paper "Learning From Mistakes Makes LLM Better Reasoner"β61Dec 20, 2023Updated 2 years ago
- Code for Engel, Grossmann & Ockenfelsβ20Jan 2, 2026Updated 7 months ago
- Luna is inspired by Lucid, a framework for Feature Visualization. However, Luna is built on Tensorflow 2, and thus supports modern modelsβ¦β11Aug 17, 2022Updated 4 years ago
- Code accompanying ICML 2021 paper "Few-shot Language Coordination by Modeling Theory of Mind"β18May 18, 2022Updated 4 years ago