β23Nov 8, 2023Updated 2 years ago
Alternatives and similar repositories for symbolictom
Users that are interested in symbolictom are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- π» Code and benchmark for our EMNLP 2023 paper - "FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions"β63May 31, 2024Updated 2 years ago
- β12May 6, 2024Updated 2 years ago
- Machine Theory of Mind Reading List. Built upon EMNLP Findings 2023 Paper: Towards A Holistic Landscape of Situated Theory of Mind in Larβ¦β154Jun 11, 2026Updated 2 months ago
- TyDiP Multilingual Politeness dataset and codeβ12Oct 15, 2023Updated 2 years ago
- [AAAI 2025 ππ«ππ₯] MuMA-ToM: Multi-modal Multi-Agent Theory of Mindβ41Jun 28, 2026Updated last month
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Testing Theory of Mind (ToM) in language models with epistemic logicβ22Jul 3, 2026Updated last month
- [EMNLP 2023] Hi-ToM benchmarkβ21Oct 11, 2025Updated 10 months ago
- Tasks for describing differences between text distributions.β17Aug 9, 2024Updated 2 years ago
- Scratchpad/Chain-of-Thought Promptsβ12Jun 6, 2022Updated 4 years ago
- ToMBench: Benchmarking Theory of Mind in Large Language Models, ACL 2024.β69Jun 24, 2024Updated 2 years ago
- Package for defining computation graphs and performing intervention experimentsβ16Oct 1, 2021Updated 4 years ago
- Train toy models using multi-token prediction objectiveβ14Apr 18, 2026Updated 4 months ago
- β40Mar 20, 2017Updated 9 years ago
- Code for ExploreTomβ93Jun 25, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [πOutstanding Paper Award at ACL 2024] MMToM-QA: Multimodal Theory of Mind Question Answeringβ159Jun 28, 2026Updated last month
- Code for the article "Shortcutted Commonsense: Data Spuriousness in Deep Learning of Commonsense Reasoning", Outstanding Paper at EMNLP20β¦β10Nov 7, 2021Updated 4 years ago
- β10Mar 19, 2024Updated 2 years ago
- Evaluation results for Machine Translation within the BigScience projectβ11May 15, 2023Updated 3 years ago
- State of What Art? A Call for Multi-Prompt LLM Evaluationβ15Apr 10, 2026Updated 4 months ago
- β10May 27, 2024Updated 2 years ago
- β15Nov 18, 2025Updated 9 months ago
- [ACL 2025 Findings] Implicit Reasoning in Transformers is Reasoning through Shortcutsβ18Mar 11, 2025Updated last year
- Repository for "I am a Strange Dataset: Metalinguistic Tests for Language Models"β46Jan 11, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Evaluation code for HAERAE-Vision benchmarkβ15Apr 29, 2026Updated 3 months ago
- Evaluating Reward Models in Multilingual Settings (ACL Main '25)β44May 16, 2025Updated last year
- This repo is to demo the concept of lossless compression with Transformers as encoder and decoder.β14May 2, 2024Updated 2 years ago
- implementation of paper "Large Language Models are In-Context Semantic Reasoners rather than Symbolic Reasoners"β20Aug 17, 2023Updated 3 years ago
- This is the code for neural-Jacana aligner, and the data for MultiMWA dataset.β20Feb 12, 2023Updated 3 years ago
- β20Jun 6, 2021Updated 5 years ago
- Code for our ACL '23 paper titled "Grokking of Hierarchical Structure in Vanilla Transformers"β26Oct 8, 2023Updated 2 years ago
- code for "GLEN: General-Purpose Event Detection for Thousands of Types"β13Nov 6, 2023Updated 2 years ago
- γιθδΈηδΊΊε·₯ζΊθ½γι ε₯代η β11Sep 20, 2022Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Code for Engel, Grossmann & Ockenfelsβ20Jan 2, 2026Updated 7 months ago
- Experiments on GPT-3's ability to fit numerical models in-context.β14Aug 11, 2022Updated 4 years ago
- π² Code and benchmark for our COLM 2025 paper - "Thought Tracing: Hypothesis-Driven Theory-of-Mind Reasoning for Large Language Models"β15Aug 8, 2025Updated last year
- Portable TCP/UDP/ICMP traceroute tool, written in Pythonβ17Apr 18, 2020Updated 6 years ago
- Code accompanying our EMNLP 2019 paper: "Revisiting the Evaluation of Theory of Mind through Question Answering"β29Aug 9, 2020Updated 6 years ago
- Korean Benchmark for Korean Legal Language Understandingβ19Nov 16, 2024Updated last year
- εε€§εθβ24Jan 2, 2023Updated 3 years ago