LLM-Empowered State Representation for Reinforcement Learning (ICML2024 Accepted paper)
☆42Jun 14, 2024Updated 2 years ago
Alternatives and similar repositories for LESR
Users that are interested in LESR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [AAAI-25] Latent Reward: LLM-Empowered Credit Assignment in Episodic Reinforcement Learning.☆34May 29, 2025Updated last year
- Official Repository for 'Promptable Behaviors: Personalizing Multi-Objective Rewards from Human Preferences' (CVPR 2024)☆17Mar 29, 2024Updated 2 years ago
- Code for NeurIPS paper "Self-Organized Group for Cooperative Multi-agentReinforcement Learning".☆22Feb 20, 2023Updated 3 years ago
- Official code for ACT: Empowering Decision Transformer with Dynamic Programming via Advantage Conditioning (AAAI'24)☆17Feb 10, 2024Updated 2 years ago
- Code for NeurIPS2023 accepted paper: Counterfactual Conservative Q Learning for Offline Multi-agent Reinforcement Learning.☆42Feb 18, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A benchmark for evaluating reinforcement learning algorithms that train the policies using imaginary rollouts from LLMs.☆15Nov 4, 2025Updated 8 months ago
- Contains implementation of the DoubIL and ResiduIL algorithms from the ICML '22 paper Causal Imitation Learning under Temporally Correlat…☆11Dec 9, 2022Updated 3 years ago
- ☆13Apr 25, 2024Updated 2 years ago
- ☆13Feb 13, 2024Updated 2 years ago
- ☆66Jan 22, 2025Updated last year
- ☆11Oct 3, 2022Updated 3 years ago
- Code to reproduce results from the paper: Prediction and Control in Continual Reinforcement Learning, NeurIPS 2023.☆13May 10, 2024Updated 2 years ago
- Codes accompanying the paper "Offline Reinforcement Learning with Value-Based Episodic Memory" (ICLR 2022 https://arxiv.org/abs/2110.0979…☆15Mar 9, 2022Updated 4 years ago
- ☆14Apr 3, 2023Updated 3 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆50Jul 23, 2021Updated 5 years ago
- Implementation of CoDAIL in the ICLR 2020 paper <Multi-Agent Interactions Modeling with Correlated Policies>☆19Jun 17, 2021Updated 5 years ago
- ☆13May 2, 2019Updated 7 years ago
- Online Preference Alignment for Language Models via Count-based Exploration☆21Jan 14, 2025Updated last year
- [CVPR 2021] Official Implementation of VAI: Unsupervised Visual Attention and Invariance for Reinforcement Learning☆27May 3, 2022Updated 4 years ago
- [NeurIPS 2025] Official codebase for T2DA: Offline Meta-RL from Natural Language Supervision☆17Jun 1, 2025Updated last year
- The repository is for Reinforcement-Learning Uncertainty research, in which we investigate various uncertain factors in RL.☆23Jun 16, 2023Updated 3 years ago
- This is the source code of FUSION, a safety-aware causal representation for generalizable driving agents.☆28Oct 23, 2024Updated last year
- ☆16Jul 16, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆19Jan 1, 2023Updated 3 years ago
- [ICML 2024] Fast Text-to-3D-Aware Face Generation and Manipulation via Direct Cross-modal Mapping and Geometric Regularization☆23Dec 20, 2024Updated last year
- Image-based gridworld experiment for learning Markov state abstractions☆20Sep 16, 2024Updated last year
- Official codebase for Exact Energy-Guided Diffusion Sampling via Contrastive Energy Prediction☆35Nov 3, 2023Updated 2 years ago
- DAC: Diffusion Actor-Critic: Formulating Constrained Policy Iteration as Diffusion Noise Regression for Offline Reinforcement Learning.☆30Jun 3, 2024Updated 2 years ago
- ☆17Aug 12, 2025Updated 11 months ago
- Codebase of paper "Self-Improving Safety Performance of Reinforcement Learning Based Driving with Black-Box Verification Algorithms" publ…☆12Jul 13, 2023Updated 3 years ago
- [WWW 2023] The official code for the paper "Two-Stage Constrained Actor-Critic for Short Video Recommendation"☆15Jul 21, 2023Updated 3 years ago
- Monitoring recent cross-research on LLM & RL on arXiv for control. If there are good papers, PRs are welcome.☆558Nov 17, 2025Updated 8 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆22Jun 13, 2023Updated 3 years ago
- ☆18Jul 11, 2025Updated last year
- [AAAI 2024 (Oral)] Safety-MuJoCo Environments.☆12Jun 4, 2024Updated 2 years ago
- ☆12Mar 15, 2022Updated 4 years ago
- ☆16Apr 6, 2022Updated 4 years ago
- Natural Language Reinforcement Learning☆101Jul 30, 2025Updated 11 months ago
- Code for paper Feasible Actor-Critic: Constrained Reinforcement Learning for Ensuring Statewise Safety.☆20May 22, 2022Updated 4 years ago