ToMBench: Benchmarking Theory of Mind in Large Language Models, ACL 2024.
☆69Jun 24, 2024Updated 2 years ago
Alternatives and similar repositories for ToMBench
Users that are interested in ToMBench are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆21Oct 25, 2024Updated last year
- ToMATO: Verbalizing the Mental States of Role-Playing LLMs for Benchmarking Theory of Mind (AAAI2025)☆20Apr 16, 2025Updated last year
- [EMNLP 2023] Hi-ToM benchmark☆21Oct 11, 2025Updated last year
- ☆40Jul 16, 2023Updated 3 years ago
- ☆11May 6, 2024Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- The official repository of the OpenToM dataset☆34Feb 2, 2025Updated last year
- 👻 Code and benchmark for our EMNLP 2023 paper - "FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions"☆63May 31, 2024Updated 2 years ago
- ☆23Nov 8, 2023Updated 2 years ago
- Machine Theory of Mind Reading List. Built upon EMNLP Findings 2023 Paper: Towards A Holistic Landscape of Situated Theory of Mind in Lar…☆154Jun 11, 2026Updated 4 months ago
- The Implementation of "Machine Theory of Mind", ICML 2018☆27Mar 14, 2022Updated 4 years ago
- Code and data for the paper: On the Reliability of Psychological Scales on Large Language Models☆31Dec 15, 2025Updated 9 months ago
- [ACL24] EmoBench: Evaluating the Emotional Intelligence of Large Language Models☆118May 16, 2025Updated last year
- [ICML 2024] Language Models Represent Beliefs of Self and Others☆37Sep 26, 2024Updated 2 years ago
- Public repository for "Think Twice: Perspective-Taking Improves Large Language Models’ Theory-of-Mind Capabilities".☆26Aug 16, 2023Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [AAAI 2025 𝐎𝐫𝐚𝐥] MuMA-ToM: Multi-modal Multi-Agent Theory of Mind☆42Jun 28, 2026Updated 3 months ago
- This is a repository for sharing papers in the field of empathetic conversational AI. The related source code for each paper is linked if…☆274Apr 17, 2024Updated 2 years ago
- ☆13Aug 23, 2017Updated 9 years ago
- Official code for ICML 2024 paper on Persona In-Context Learning (PICLe)☆29Jun 27, 2024Updated 2 years ago
- Testing Theory of Mind (ToM) in language models with epistemic logic☆21Jul 3, 2026Updated 3 months ago
- ☆14Aug 13, 2020Updated 6 years ago
- ☆12Feb 22, 2021Updated 5 years ago
- [AAAI22] CEM: Commonsense-aware Empathetic Response Generation☆98May 16, 2025Updated last year
- Sotopia: an Open-ended Social Learning Environment (ICLR 2024 spotlight)☆335Jun 5, 2026Updated 4 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆12Mar 12, 2024Updated 2 years ago
- implementation of dualformer☆25Mar 1, 2025Updated last year
- Code of EMNLP 2025 paper 'UltraIF: Advancing Instruction Following from the Wild'.☆21Apr 3, 2025Updated last year
- Evaluating and improving the faithfulness of the interpretations offered by Neural Module Networks☆13Jun 12, 2023Updated 3 years ago
- Code and Data for the paper "Evaluating Character Understanding of Large Language Models via Character Profiling from Fictional Works".☆21Jul 24, 2024Updated 2 years ago
- A collection of works that investigate social agents, simulations and their real-world impact in text, embodied, and robotics contexts.☆114Jun 14, 2026Updated 3 months ago
- Simple phoenix setup for padded window management☆13Apr 25, 2018Updated 8 years ago
- Code used to run experiments for the ICLR 2023 paper "Computational Language Acquisition with Theory of Mind".☆15Apr 27, 2023Updated 3 years ago
- [ICLR 2026] VerifyBench: Benchmarking Reference-based Reward Systems for Large Language Models☆22Feb 18, 2026Updated 7 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆19Nov 15, 2023Updated 2 years ago
- Released code for「Self-Training with Pseudo-Label Scorer for Aspect Sentiment Quad Prediction」in ACL2024.☆23Feb 21, 2025Updated last year
- The Code Repo for Agent-Pro: Learning to Evolve via Policy-Level Reflection and Optimization☆134Sep 2, 2024Updated 2 years ago
- Code for ExploreTom☆95Jun 25, 2025Updated last year
- [ICCV 2025] MissRAG: Addressing the Missing Modality Challenge in Multimodal Large Language Models☆26May 12, 2026Updated 4 months ago
- A iterative feedback driven benchmark on LLM's instruction following ability☆58May 25, 2026Updated 4 months ago
- ☆15May 9, 2024Updated 2 years ago