HumanLM: Simulating Users with State Alignment Beats Response Imitation
☆87Jun 4, 2026Updated 2 months ago
Alternatives and similar repositories for humanlm
Users that are interested in humanlm are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆23May 12, 2026Updated 3 months ago
- Building Foundation Models for Human Behavior Simulation☆111Jul 8, 2026Updated last month
- ☆21Apr 3, 2026Updated 4 months ago
- ☆19Nov 5, 2025Updated 9 months ago
- https://scale.com/research/mrt☆20Mar 16, 2026Updated 4 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- (ICML'25 Outstanding) CollabLLM: From Passive Responders to Active Collaborators☆302Sep 25, 2025Updated 10 months ago
- Implementation of Not All Contexts Are Equal: Teaching LLMs Credibility-aware Generation. Paper: https://arxiv.org/abs/2404.06809☆22Oct 22, 2024Updated last year
- The official code repo of paper "Chasing the Tail: Effective Rubric-based Reward Modeling for Large Language Model Post-Training"☆30Feb 20, 2026Updated 5 months ago
- Training Proactive and Personalized LLM Agents☆113Jan 20, 2026Updated 6 months ago
- Uncertainty quantification for in-context learning of large language models☆15Apr 1, 2024Updated 2 years ago
- ☆10Jun 15, 2024Updated 2 years ago
- Example for a Monty-enabled RLM in DSPy☆20Feb 16, 2026Updated 5 months ago
- Code that accompanies the public release of the paper Lost in Conversation (https://arxiv.org/abs/2505.06120)☆295Jun 9, 2026Updated 2 months ago
- ☆40Oct 21, 2025Updated 9 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Code for "Towards Real-world Human Behavior Simulation: Benchmarking Large Language Models on Long-horizon, Cross-scenario, Heterogeneous…☆47May 18, 2026Updated 2 months ago
- ☆19Jun 21, 2025Updated last year
- Aligning Language Models from User Interactions via Self-Distillation☆29Mar 31, 2026Updated 4 months ago
- Molecular Explanation Generator☆17Jan 26, 2022Updated 4 years ago
- ☆22May 14, 2026Updated 3 months ago
- Extended Inductive Reasoning for Personalized Preference Inference from Behavioral Signals☆11Jan 8, 2026Updated 7 months ago
- [ICLR 2025] Code for the paper "Implicit Search via Discrete Diffusion: A Study on Chess"☆39Mar 3, 2025Updated last year
- [NAACL'25 Oral] Steering Knowledge Selection Behaviours in LLMs via SAE-Based Representation Engineering☆83Jun 20, 2026Updated last month
- [ICLR 2026] Official implementation of "ChatInject: Abusing Chat Templates for Prompt Injection in LLM Agents"☆21Mar 23, 2026Updated 4 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Official code for "How2Everything: Mining the Web for How-To Procedures to Evaluate and Improve LLMs"☆25Feb 10, 2026Updated 6 months ago
- ☆33Aug 9, 2024Updated 2 years ago
- [ICML 2026] Code for V1: Unifying Generation and Self-Verification for Parallel Reasoners.☆39Mar 5, 2026Updated 5 months ago
- ☆28Jun 1, 2026Updated 2 months ago
- List of learning-based PCC papers, welcome Pull Requests!☆25Nov 4, 2025Updated 9 months ago
- Code for the paper "Searching Privacy Risks in Multi-Agent Systems via Simulation"☆25Oct 13, 2025Updated 10 months ago
- A Mechanistic Understanding of Alignment Algorithms: A Case Study on DPO and Toxicity.☆90Mar 7, 2025Updated last year
- The code implementation of MAGDi: Structured Distillation of Multi-Agent Interaction Graphs Improves Reasoning in Smaller Language Models…☆40Feb 5, 2024Updated 2 years ago
- ☆17Feb 24, 2026Updated 5 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Detecting Corpus-Level Knowledge Inconsistencies in Wikipedia with Large Language Models☆18Sep 30, 2025Updated 10 months ago
- ☆19Mar 25, 2025Updated last year
- llms related stuff , including code, docs☆13Feb 25, 2025Updated last year
- A Python implementation of word2vec that allows custom sampling strategies☆10Jan 30, 2014Updated 12 years ago
- Repository for the paper: "TiC-LM: A Web-Scale Benchmark for Time-Continual LLM Pretraining" ACL Oral 2025☆24Apr 19, 2026Updated 3 months ago
- The best ChatGPT that $100 can buy.☆58Updated this week
- Code Release for the 2023 NeurIPS Paper How does GPT-2 compute greater-than?: Interpreting mathematical abilities in a pre-trained langua…☆17Dec 6, 2024Updated last year