Code for ExploreTom
β95Jun 25, 2025Updated last year
Alternatives and similar repositories for ExploreToM
Users that are interested in ExploreToM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [AAAI 2025 ππ«ππ₯] MuMA-ToM: Multi-modal Multi-Agent Theory of Mindβ42Jun 28, 2026Updated 3 months ago
- [EMNLP 2023] Hi-ToM benchmarkβ21Oct 11, 2025Updated last year
- The official repository of the OpenToM datasetβ34Feb 2, 2025Updated last year
- β16Feb 5, 2026Updated 8 months ago
- β44May 29, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Evaluating Reward Models in Multilingual Settings (ACL Main '25)β45May 16, 2025Updated last year
- π» Code and benchmark for our EMNLP 2023 paper - "FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions"β63May 31, 2024Updated 2 years ago
- This library supports evaluating disparities in generated image quality, diversity, and consistency between geographic regions.β20Jun 3, 2024Updated 2 years ago
- Large Concept Models: Language modeling in a sentence representation spaceβ2,381Jan 29, 2025Updated last year
- π¦Ύ EvalGIM (pronounced as "EvalGym") is an evaluation library for generative image models. It enables easy-to-use, reproducible automaticβ¦β93Feb 5, 2026Updated 8 months ago
- Implementation of Monte Carlo Tree Searchβ17Aug 4, 2022Updated 4 years ago
- A framework to study AI models in Reasoning, Alignment, and use of Memory (RAM).β393Updated this week
- Code accompanying our EMNLP 2019 paper: "Revisiting the Evaluation of Theory of Mind through Question Answering"β28Aug 9, 2020Updated 6 years ago
- [NeurIPS 2025] Elevating Visual Perception in Multimodal LLMs with Visual Embedding Distillationβ75Oct 17, 2025Updated 11 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- β17Apr 7, 2025Updated last year
- Machine Theory of Mind Reading List. Built upon EMNLP Findings 2023 Paper: Towards A Holistic Landscape of Situated Theory of Mind in Larβ¦β154Jun 11, 2026Updated 4 months ago
- This is the repo for the paper "PANGEA: A FULLY OPEN MULTILINGUAL MULTIMODAL LLM FOR 39 LANGUAGES"β120Jun 27, 2025Updated last year
- Measuring and Controlling Persona Drift in Language Model Dialogsβ26Feb 26, 2024Updated 2 years ago
- Source code for GreaTer ICLR 2025 - Gradient Over Reasoning makes Smaller Language Models Strong Prompt Optimizersβ36Apr 18, 2025Updated last year
- Clue inspired puzzles for testing LLM deduction abilitiesβ48Mar 19, 2026Updated 6 months ago
- β25Mar 21, 2024Updated 2 years ago
- Official implementation for "Law of the Weakest Link: Cross capabilities of Large Language Models"β43Oct 1, 2024Updated 2 years ago
- β15Apr 26, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- WebGym: Web-browser-based tasks for RL Agentsβ24Feb 4, 2021Updated 5 years ago
- β24Aug 10, 2022Updated 4 years ago
- Source code for the collaborative reasoner research project at Meta FAIR.β116Mar 26, 2026Updated 6 months ago
- β11Oct 3, 2021Updated 5 years ago
- Dialog2Flow: convert your dialogs to flows. This repository accompanies the paper "Dialog2Flow: Pre-training Soft-Contrastive Sentence Emβ¦β20Jul 1, 2025Updated last year
- CUDA, CuDNN, NVIDIA Driver, and PyTorch Installation for Ubuntuβ12Feb 27, 2025Updated last year
- β22Jun 2, 2026Updated 4 months ago
- MOTIF: Modular Thinking via Reinforcement Fine-tuning in LLMs