Code for Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning, AAAI 2025
☆15Dec 19, 2024Updated last year
Alternatives and similar repositories for CAPS
Users that are interested in CAPS are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official code for the paper: Continual Task Allocation in Meta-Policy Network via Sparse Prompting☆23Feb 10, 2025Updated last year
- ☆12Aug 13, 2022Updated 4 years ago
- Official Implementation of "Doubly Mixed-Effects Gaussian Process Regression" (Jun Ho Yoon, Daniel P. Jeong, Seyoung Kim) (AISTATS 2022, …☆12Jul 13, 2022Updated 4 years ago
- Official implementation of ICML'24 paper "Offline Multi-Objective Optimization".☆26May 24, 2026Updated 3 months ago
- Python implementation of the supervised graph prediction method proposed in http://arxiv.org/abs/2202.03813 using PyTorch library and POT…☆15Feb 25, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official Code for the L4DC 2023 conference paper and ICLR 2023 NeSy-GeMs workshop paper.☆15Oct 17, 2023Updated 2 years ago
- Code space for L4DC paper "State-wise Safe Reinforcement Learning With Pixel Observations"☆11Apr 5, 2024Updated 2 years ago
- [IROS2024] STAIR: Semantic-Targeted Active Implicit Reconstruction☆16Aug 3, 2024Updated 2 years ago
- This is the code for our paper: Increasing the Scope as You Learn: Adaptive Bayesian Optimization in Nested Subspaces (Leonard Papenmeier…☆22Jan 5, 2024Updated 2 years ago
- ☆15Jan 24, 2025Updated last year
- Implemention of the Decision-Pretrained Transformer (DPT) from the paper Supervised Pretraining Can Learn In-Context Reinforcement Learni…☆80May 28, 2024Updated 2 years ago
- A benchmark for evaluating reinforcement learning algorithms that train the policies using imaginary rollouts from LLMs.☆16Nov 4, 2025Updated 9 months ago
- Burstable Cloud Scheduler☆17Jun 6, 2024Updated 2 years ago
- (CEC2022) Fast Moving Natural Evolution Strategy for High-Dimensional Problems☆19Apr 13, 2026Updated 4 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Robust and safe deep reinforcement learning algorithms☆17Mar 27, 2024Updated 2 years ago
- [ICLR 2024] The official implementation of "Safe Offline Reinforcement Learning with Feasibility-Guided Diffusion Model"☆130Feb 11, 2025Updated last year
- Includes the implementation to train PPO, DDPG or TD3 agents (from Stable-Baselines3) in Isaac Lab. The considered task includes a UR5e o…☆16Jan 8, 2025Updated last year
- [NeurIPS 2024] Doubly Mild Generalization for Offline Reinforcement Learning☆17Oct 29, 2025Updated 10 months ago
- A library of discrete objectives☆27Oct 5, 2025Updated 10 months ago
- [ICLR 2026] AdaBlock-dLLM: Semantic-Aware Diffusion LLM Inference via Adaptive Block Size☆16Jan 28, 2026Updated 7 months ago
- The official implementation of "Rectified SpaAttn: Revisiting Attention Sparsity for Efficient Video Generation"☆25Feb 8, 2026Updated 6 months ago
- This is the source code of FUSION, a safety-aware causal representation for generalizable driving agents.☆29Oct 23, 2024Updated last year
- Code for "Re-evaluating Word Mover’s Distance" (ICML 2022)☆40Jun 15, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code for TRANSDREAMER: REINFORCEMENT LEARNING WITH TRANSFORMER WORLD MODELS☆32Oct 12, 2023Updated 2 years ago
- Online Preference Alignment for Language Models via Count-based Exploration☆21Jan 14, 2025Updated last year
- NeurIPS 2022: Tree Mover’s Distance: Bridging Graph Metrics and Stability of Graph Neural Networks☆37Aug 4, 2023Updated 3 years ago
- Official PyTorch implementation of "Query-Efficient and Scalable Black-Box Adversarial Attacks on Discrete Sequential Data via Bayesian O…☆26Sep 26, 2023Updated 2 years ago
- Safe Multi-Agent Robosuite benchmark for safe multi-agent reinforcement learning research.☆25Jun 13, 2024Updated 2 years ago
- A vision-based RL environment for the Franka Panda arm using NVIDIA Isaac Sim☆21Jan 3, 2025Updated last year
- The PackNet Continual Learning Method in Pytorch☆15Aug 19, 2021Updated 5 years ago
- Feasibility Consistent Representation Learning for Safe Reinforcement Learning (ICML 2024). Current SOTA model-free safe RL algorithm on …☆17Jul 12, 2024Updated 2 years ago
- Code accompanying the paper "Off-Policy Primal-Dual Safe Reinforcement Learning"☆22Mar 29, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- 🤖 Elegant implementations of offline safe RL algorithms in PyTorch☆251Sep 13, 2024Updated last year
- Lightweight Adapting for Black-Box Large Language Models☆26Feb 15, 2024Updated 2 years ago
- Fast Bayesian optimization, quadrature, inference over arbitrary domain with GPU parallel acceleration☆35Dec 10, 2025Updated 8 months ago
- Constrained Policy Optimization implementation on Safety Gym☆30Jan 8, 2022Updated 4 years ago
- A JAX-based Differentiable Optical and Radio Frequency Simulator for Multilayer Structures☆27Jun 1, 2026Updated 3 months ago
- 本项目基于 **NVIDIA Omniverse / Isaac Lab** 大 规模并行张量仿真框架,针对目前工业界最前沿的足式平台(**Unitree G1/H1 人形机器人, Unitree Go1 / ANYmal-C 机器狗**)进行了强化学习 (PPO) 步态策略…☆19Mar 4, 2026Updated 5 months ago
- [ICLR 2026] SERE: Similarity-Based Expert Re-routing for Efficient Batch Decoding in MoE Models☆21Feb 4, 2026Updated 6 months ago