PAHF Personalized Agent from Human Feedback
☆57Apr 26, 2026Updated 4 months ago
Alternatives and similar repositories for PAHF
Users that are interested in PAHF are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [NeurIPS 2024] Introspective Planning: Aligning Robots’ Uncertainty with Inherent Task Ambiguity☆23Nov 6, 2025Updated 9 months ago
- ☆22Jun 23, 2026Updated 2 months ago
- ☆22Apr 24, 2025Updated last year
- EgoToM is an egocentric theory-of-mind benchmark built on Ego4D videos, containing multi-choice questions that evaluate multimodal large …☆17Apr 1, 2025Updated last year
- PersonaMem-v2: Towards Personalized Intelligence via Learning Implicit User Personas and Agentic Memory☆43May 18, 2026Updated 3 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Does Socialization Emerge in AI Agent Society? A Case Study of Moltbook☆18Feb 17, 2026Updated 6 months ago
- The code for paper "MemGym: a Long-Horizon Memory Environment for LLM Agents".☆21Jun 2, 2026Updated 2 months ago
- Storybook for flowcharts. Auto-discovers Mermaid diagram files from your codebase, organizes them by category, and renders them in a brow…☆17Mar 9, 2026Updated 5 months ago
- FabScore: Fine-Grained Evaluation of Fabrications in Automated AI Research☆20Jul 24, 2026Updated last month
- [ICML 2026] Code for V1: Unifying Generation and Self-Verification for Parallel Reasoners.☆39Mar 5, 2026Updated 5 months ago
- ☆25Jun 10, 2025Updated last year
- ☆11Mar 11, 2025Updated last year
- [ICLR 2024] This is the official implementation for the paper: "Beyond imitation: Leveraging fine-grained quality signals for alignment"☆10May 5, 2024Updated 2 years ago
- The raw UserRL repo under construction☆118Jun 2, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official implementation for the paper "Video-Based Reward Modeling for Computer-Use Agents"☆17Mar 14, 2026Updated 5 months ago
- Quadruped Trajectory Optimization Stack (QTOS) is an optimization framework for legged locomotion that autonomously generates full-body t…☆28Sep 28, 2023Updated 2 years ago
- Code for experiments on self-prediction as a way to measure introspection in LLMs☆17Dec 10, 2024Updated last year
- A nonlinear MPC implementation for a quadcopter☆11Apr 11, 2018Updated 8 years ago
- Multi-agent coordination using game theory and nonlinear opinion dynamics - CDC 2023☆15Nov 29, 2023Updated 2 years ago
- A benchmark to measure AI progress on unsolved research problems in mathematics.☆33Aug 21, 2026Updated last week
- A research framework for evaluating proactive AI assistants through active user simulation☆39May 23, 2026Updated 3 months ago
- Repository for "Who Plays First? Optimizing the Order of Play in Stackelberg Games with Many Robots" - RSS 2024☆18Jun 25, 2024Updated 2 years ago
- This is the code for our paper: PLACES: Prompting Language Models for Social Conversation Synthesis☆11Feb 17, 2023Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [KDD'23] This is the code repo for our KDD'23 paper "DyGen: Learning from Noisy Labels via Dynamics-Enhanced Generative Modeling".☆11Jun 14, 2023Updated 3 years ago
- ☆23Dec 17, 2024Updated last year
- ALMA (Automated meta-Learning of Memory designs for Agentic systems) is a framework that meta-learns memory designs to replace human-engi…☆291Apr 8, 2026Updated 4 months ago
- Awesome Agent Environments☆17Apr 10, 2026Updated 4 months ago
- An interior point solver for mixed complementarity problems, supporting CPU/GPU batched-parallel execution.☆20Aug 4, 2026Updated 3 weeks ago
- Code for COLM 2026 Paper "Reinforcement Learning with Metacognitive Feedback Elicits Faithful Uncertainty Expression in LLMs"☆30Jul 1, 2026Updated last month
- [ICLR 2025] FLAT: LLM Unlearning via Loss Adjustment with Only Forget Data☆14Feb 26, 2025Updated last year
- OccuBench: Evaluating AI Agents on Real-World Professional Tasks via Language World Models☆22Apr 14, 2026Updated 4 months ago
- The official code for "Steering Dialogue Dynamics for Robustness against Multi-turn Jailbreaking Attacks".☆19Jun 24, 2026Updated 2 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆11Oct 12, 2021Updated 4 years ago
- MMA: Multimodal Memory Agent☆23Mar 30, 2026Updated 5 months ago
- multicast learning in network programming course☆10Oct 30, 2020Updated 5 years ago
- WAFR 2024: Multi-modal variational inference in multi-agent interaction enabled by VAE + differentiable Nash game solver.☆24Nov 10, 2025Updated 9 months ago
- Agentic System, Tool Use, Electronic Health Record, Large Language Models, Clinical Nature Language Processing☆24Apr 13, 2026Updated 4 months ago
- Personal Claude Code plugin marketplace☆16Updated this week
- Benchmarking Language Agents Under Controllable and Extreme Context Growth☆56Apr 29, 2026Updated 4 months ago