[EMNLP 2023] Hi-ToM benchmark
☆21Oct 11, 2025Updated 10 months ago
Alternatives and similar repositories for Hi-ToM_dataset
Users that are interested in Hi-ToM_dataset are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The official repository of the OpenToM dataset☆34Feb 2, 2025Updated last year
- ☆40Jul 16, 2023Updated 3 years ago
- Machine Theory of Mind Reading List. Built upon EMNLP Findings 2023 Paper: Towards A Holistic Landscape of Situated Theory of Mind in Lar…☆154Jun 11, 2026Updated 2 months ago
- Implementation of the Decrypto benchmark for multi-agent reasoning and theory of mind.☆23Jan 19, 2026Updated 7 months ago
- [NeurIPS 2025 𝐒𝐩𝐨𝐭𝐥𝐢𝐠𝐡𝐭] AutoToM: Scaling Model-based Mental Inference via Automated Agent Modeling☆47Jun 28, 2026Updated 2 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Hypothetical Minds is an autonomous LLM-based agent for diverse multi-agent settings, integrating a Theory of Mind module Theory of Mind …☆64Jul 13, 2024Updated 2 years ago
- ☆12Jan 25, 2024Updated 2 years ago
- Examples of using Galileo for better ML data quality!!☆13Feb 5, 2026Updated 7 months ago
- ☆23Nov 8, 2023Updated 2 years ago
- 👻 Code and benchmark for our EMNLP 2023 paper - "FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions"☆63May 31, 2024Updated 2 years ago
- [EMNLP 2023] Official repository for Dialogue Chain-of-Thought Distillation (DONUT & DOCTOR)☆11Nov 15, 2023Updated 2 years ago
- ☆36Mar 10, 2025Updated last year
- [ICML 2024] Language Models Represent Beliefs of Self and Others☆37Sep 26, 2024Updated last year
- ☆30Mar 11, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆41Mar 20, 2017Updated 9 years ago
- Code for ACL 2023 paper "BOLT: Fast Energy-based Controlled Text Generation with Tunable Biases".☆22Sep 7, 2023Updated 3 years ago
- ☆19Jun 18, 2026Updated 2 months ago
- Offline Policy Evaluation via Adaptive Weighting with Data from Contextual Bandits☆11Oct 21, 2024Updated last year
- Public repository for "Think Twice: Perspective-Taking Improves Large Language Models’ Theory-of-Mind Capabilities".☆26Aug 16, 2023Updated 3 years ago
- Simple phoenix setup for padded window management☆13Apr 25, 2018Updated 8 years ago
- Mental state inference from observable behavior☆15Dec 3, 2021Updated 4 years ago
- Code used to run experiments for the ICLR 2023 paper "Computational Language Acquisition with Theory of Mind".☆15Apr 27, 2023Updated 3 years ago
- ☆10Mar 19, 2024Updated 2 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- A tool library for riichi mahjong written in Rust, made mostly to be used as a WASM component.☆12Aug 29, 2025Updated last year
- ☆14May 9, 2024Updated 2 years ago
- Code for the paper "Symmetric Machine Theory of Mind", presented at ICML 2022.☆12Jul 18, 2022Updated 4 years ago
- AAAI 2024-Controllable Mind Visual Diffusion Model☆16Dec 18, 2023Updated 2 years ago
- ☆31Oct 29, 2024Updated last year
- Official code repository for NeurIPS 2024 paper "Recurrent Complex-Weighted Autoencoders for Unsupervised Object Discovery"☆13Jan 8, 2025Updated last year
- Offline Multi-Agent Reinforcement Learning Implementations: Solving Overcooked Game with Data-Driven Method☆48Sep 11, 2024Updated last year
- ☆12Aug 30, 2021Updated 5 years ago
- [CVPR'25] Attention IoU: Examining Biases in CelebA using Attention Maps☆13Mar 26, 2025Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- SeRL: Self-Play Reinforcement Learning for Large Language Models with Limited Data☆26Jan 24, 2026Updated 7 months ago
- This repo is to demo the concept of lossless compression with Transformers as encoder and decoder.☆14May 2, 2024Updated 2 years ago
- Find context neurons in Pythia models.☆13Jun 13, 2023Updated 3 years ago
- Code for our SIGGRAPH 2023 paper, "Acting as Inverse Inverse Planning"☆20Apr 21, 2023Updated 3 years ago
- MrSteve: Instruction-Following Agents in Minecraft with What-Where-When Memory☆16May 1, 2025Updated last year
- EgoToM is an egocentric theory-of-mind benchmark built on Ego4D videos, containing multi-choice questions that evaluate multimodal large …☆17Apr 1, 2025Updated last year
- ☆11Dec 22, 2021Updated 4 years ago