Sotopia: an Open-ended Social Learning Environment (ICLR 2024 spotlight)
☆324Jun 5, 2026Updated 2 months ago
Alternatives and similar repositories for sotopia
Users that are interested in sotopia are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Sotopia-π: Interactive Learning of Socially Intelligent Language Agents (ACL 2024)☆85May 7, 2024Updated 2 years ago
- Sotopia-RL: Reward Design for Social Intelligence☆52Apr 1, 2026Updated 4 months ago
- A collection of works that investigate social agents, simulations and their real-world impact in text, embodied, and robotics contexts.☆113Jun 14, 2026Updated last month
- [EMNLP 2023] Official repository for Dialogue Chain-of-Thought Distillation (DONUT & DOCTOR)☆11Nov 15, 2023Updated 2 years ago
- [ACL2023, Findings] Source codes for the paper "Werewolf Among Us: Multimodal Resources for Modeling Persuasion Behaviors in Social Deduc…☆16Feb 22, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Social-AI papers across computing communities, courses, and dissertations.☆21Apr 8, 2026Updated 4 months ago
- [ICML 2026]: Building Social World Models with Large Language Models☆25Jun 9, 2026Updated 2 months ago
- ☆37Apr 22, 2025Updated last year
- website repo for agent-based social movement simulation☆27Jun 17, 2024Updated 2 years ago
- A platform to develop CTM-motivated AI architecture.☆19Jul 28, 2026Updated 2 weeks ago
- [ICML 2025] ResearchTown: Simulator of Human Research Community☆209Updated this week
- Machine Theory of Mind Reading List. Built upon EMNLP Findings 2023 Paper: Towards A Holistic Landscape of Situated Theory of Mind in Lar…☆155Jun 11, 2026Updated 2 months ago
- A collection of resources that investigate social agents.☆243Apr 22, 2025Updated last year
- A Collection of Competitive Text-Based Games for Language Model Evaluation and Reinforcement Learning☆416Aug 6, 2026Updated last week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [EMNLP 2023] Hi-ToM benchmark☆21Oct 11, 2025Updated 10 months ago
- Codebase describing experiments in Truncation Sampling as Language Model Desmoothing☆13Dec 6, 2022Updated 3 years ago
- This repository contains reference implementation for multi-LLM ToM paper (accepted to EMNLP 2023), Theory of Mind for Multi-Agent Collab…☆20Jun 11, 2024Updated 2 years ago
- ☆204Feb 8, 2026Updated 6 months ago
- Implementation of the Decrypto benchmark for multi-agent reasoning and theory of mind.☆23Jan 19, 2026Updated 6 months ago
- A Gym for Agentic LLMs☆505Jan 21, 2026Updated 6 months ago
- SPIRAL: Self-Play on Zero-Sum Games Incentivizes Reasoning via Multi-Agent Multi-Turn Reinforcement Learning☆202Mar 27, 2026Updated 4 months ago
- Official code and dataset for our paper: RefineBench: Evaluating Refinement Capability of Language Models via Checklists☆17Dec 1, 2025Updated 8 months ago
- Critique-out-Loud Reward Models☆76Oct 18, 2024Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Research Code for "ArCHer: Training Language Model Agents via Hierarchical Multi-Turn RL"☆208Apr 17, 2025Updated last year
- Official Repo of LangSuitE☆85Aug 15, 2024Updated last year
- An open-source, AI-driven Social Media Digital Twin powered by LLMs (Ollama, vLLM) for Computational Social Science simulations, network …☆27Updated this week
- Awesome papers involving LLMs in Social Science.☆644Aug 1, 2026Updated last week
- ToMBench: Benchmarking Theory of Mind in Large Language Models, ACL 2024.☆69Jun 24, 2024Updated 2 years ago
- This repository contains data, code and models for contextual noncompliance.☆26Jul 18, 2024Updated 2 years ago
- Code for our NeurIPS'24 Dataset and Benchmark paper: Cooperation, Competition, and Maliciousness: LLM-Stakeholders Interactive Negotiatio…☆54Nov 11, 2024Updated last year
- Text Adventure Learning Environment Suite - Benchmark to evaluate language models on interactive text environments.☆30Jul 17, 2026Updated 3 weeks ago
- The Social-IQ 2.0 Challenge Release for the Artificial Social Intelligence Workshop at ICCV '23☆38Oct 13, 2023Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- How to create rational LLM-based agents? Using game-theoretic workflows!☆110Jun 8, 2025Updated last year
- [ICML 2024] Language Models Represent Beliefs of Self and Others☆37Sep 26, 2024Updated last year
- 👻 Code and benchmark for our EMNLP 2023 paper - "FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions"☆62May 31, 2024Updated 2 years ago
- This repo contains code for our NeurIPS 2023 spotlight paper: Evaluating and Inducing Personality in Pre-trained Language Models☆59Dec 7, 2023Updated 2 years ago
- ☆28May 30, 2026Updated 2 months ago
- Code and data for Marked Personas (ACL 2023)☆30May 26, 2023Updated 3 years ago
- Code for Paper: Autonomous Evaluation and Refinement of Digital Agents [COLM 2024]☆149Nov 26, 2024Updated last year