LLM-Empowered State Representation for Reinforcement Learning (ICML2024 Accepted paper)
☆42Jun 14, 2024Updated 2 years ago
Alternatives and similar repositories for LESR
Users that are interested in LESR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [AAAI-25] Latent Reward: LLM-Empowered Credit Assignment in Episodic Reinforcement Learning.☆34May 29, 2025Updated last year
- Official Repository for 'Promptable Behaviors: Personalizing Multi-Objective Rewards from Human Preferences' (CVPR 2024)☆17Mar 29, 2024Updated 2 years ago
- ☆20Nov 3, 2024Updated last year
- Code for NeurIPS paper "Self-Organized Group for Cooperative Multi-agentReinforcement Learning".☆22Feb 20, 2023Updated 3 years ago
- Official code for ACT: Empowering Decision Transformer with Dynamic Programming via Advantage Conditioning (AAAI'24)☆17Feb 10, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Code for GO4Align: Group Optimization for Multi-Task Alignment☆20Sep 25, 2024Updated last year
- Code for NeurIPS2023 accepted paper: Counterfactual Conservative Q Learning for Offline Multi-agent Reinforcement Learning.☆42Feb 18, 2025Updated last year
- A benchmark for evaluating reinforcement learning algorithms that train the policies using imaginary rollouts from LLMs.☆15Nov 4, 2025Updated 9 months ago
- Contains implementation of the DoubIL and ResiduIL algorithms from the ICML '22 paper Causal Imitation Learning under Temporally Correlat…☆11Dec 9, 2022Updated 3 years ago
- ☆13Apr 25, 2024Updated 2 years ago
- [ICLR 2024 Spotlight] Text2Reward: Reward Shaping with Language Models for Reinforcement Learning☆210Dec 17, 2024Updated last year
- ☆13Feb 13, 2024Updated 2 years ago
- ☆67Jan 22, 2025Updated last year
- Safe Model-Based RL HVAC Control Using Epistemic Uncertainty Estimation.☆13Feb 25, 2025Updated last year
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Robust Multi-Agent Reinforcement Learning with State Uncertainty☆12May 30, 2023Updated 3 years ago
- Code to reproduce results from the paper: Prediction and Control in Continual Reinforcement Learning, NeurIPS 2023.☆13May 10, 2024Updated 2 years ago
- Codes accompanying the paper "Offline Reinforcement Learning with Value-Based Episodic Memory" (ICLR 2022 https://arxiv.org/abs/2110.0979…☆15Mar 9, 2022Updated 4 years ago
- ☆14Apr 3, 2023Updated 3 years ago
- ☆50Jul 23, 2021Updated 5 years ago
- Implementation of CoDAIL in the ICLR 2020 paper <Multi-Agent Interactions Modeling with Correlated Policies>☆19Jun 17, 2021Updated 5 years ago
- ☆13May 2, 2019Updated 7 years ago
- Online Preference Alignment for Language Models via Count-based Exploration☆21Jan 14, 2025Updated last year
- [CVPR 2021] Official Implementation of VAI: Unsupervised Visual Attention and Invariance for Reinforcement Learning☆27May 3, 2022Updated 4 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- [NeurIPS 2025] Official codebase for T2DA: Offline Meta-RL from Natural Language Supervision☆17Jun 1, 2025Updated last year
- This is the source code of FUSION, a safety-aware causal representation for generalizable driving agents.☆29Oct 23, 2024Updated last year
- ☆19Jan 1, 2023Updated 3 years ago
- [ICML 2024] Fast Text-to-3D-Aware Face Generation and Manipulation via Direct Cross-modal Mapping and Geometric Regularization☆23Dec 20, 2024Updated last year
- Image-based gridworld experiment for learning Markov state abstractions☆20Sep 16, 2024Updated last year
- Official codebase for Exact Energy-Guided Diffusion Sampling via Contrastive Energy Prediction☆35Nov 3, 2023Updated 2 years ago
- DAC: Diffusion Actor-Critic: Formulating Constrained Policy Iteration as Diffusion Noise Regression for Offline Reinforcement Learning.☆30Jun 3, 2024Updated 2 years ago
- ☆17Aug 12, 2025Updated last year
- Codebase of paper "Self-Improving Safety Performance of Reinforcement Learning Based Driving with Black-Box Verification Algorithms" publ…☆12Jul 13, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [WWW 2023] The official code for the paper "Two-Stage Constrained Actor-Critic for Short Video Recommendation"☆15Jul 21, 2023Updated 3 years ago
- Monitoring recent cross-research on LLM & RL on arXiv for control. If there are good papers, PRs are welcome.☆558Nov 17, 2025Updated 9 months ago
- ☆22Jun 13, 2023Updated 3 years ago
- ☆18Jul 11, 2025Updated last year
- [AAAI 2024 (Oral)] Safety-MuJoCo Environments.☆12Jun 4, 2024Updated 2 years ago
- ☆12Mar 15, 2022Updated 4 years ago
- ICML'2024: Q-value Regularized Transformer for Offline Reinforcement Learning☆38Dec 30, 2024Updated last year