Simple gym environments for safety in Reinforcement Learning Research
☆18Jul 17, 2024Updated 2 years ago
Alternatives and similar repositories for gym-safety
Users that are interested in gym-safety are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for the paper "AlwaysSafe: Reinforcement Learning Without Safety Constraint Violations During Training"☆17May 9, 2022Updated 4 years ago
- ☆12Nov 28, 2015Updated 10 years ago
- Pytorch Implementation for First Order Constrained Optimization in Policy Space (FOCOPS).☆29Dec 9, 2021Updated 4 years ago
- POMDP wrappers for OpenAI Gym☆15Nov 4, 2019Updated 6 years ago
- Comparing obstacle avoidance formulations☆11Oct 22, 2022Updated 3 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- 卒業論文・修士論文および英語論文のための自己チェックリスト☆12Nov 28, 2023Updated 2 years ago
- Density Constrained Reinforcement Learning☆12Mar 24, 2023Updated 3 years ago
- Guarantee_Learning_Control☆11Sep 5, 2019Updated 7 years ago
- Code for training and testing a Hidden Parameter Markov Decision Process, used to facilitate the transfer of learning☆28Dec 28, 2017Updated 8 years ago
- Safe SLAC, an algorithm for safe cost-constrained reinforcement learning in high-dimensional POMDPs.☆13Mar 1, 2023Updated 3 years ago
- Pytorch implementation of "Safe Exploration in Continuous Action Spaces" [Dalal et al.]☆75Jun 2, 2019Updated 7 years ago
- ☆14May 30, 2019Updated 7 years ago
- Official GitHub Repository for Efficient Off-Policy Safe Reinforcement Learning Using Trust Region Conditional Value At Risk.☆13Nov 24, 2025Updated 9 months ago
- Code repo for "Collapsing Bandits and Their Applications to Public Health Interventions", (NeurIPS'20)☆11Dec 3, 2025Updated 9 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆11Dec 3, 2020Updated 5 years ago
- Avoiding catastrophic failures in reinforcement learning by learning to shape rewards.☆10Nov 13, 2017Updated 8 years ago
- ☆49Dec 8, 2022Updated 3 years ago
- A (mixed integer) linear optimisation model for local energy systems☆13May 27, 2021Updated 5 years ago
- Probabilistic planning in continuous state-action MDPs in TensorFlow.☆13Jun 21, 2022Updated 4 years ago
- Constrained episodic reinforcement learning in concave-convex and knapsack settings☆11Oct 3, 2023Updated 2 years ago
- A tour of Pomdpland☆10Aug 10, 2022Updated 4 years ago
- ☆11Dec 27, 2021Updated 4 years ago
- Implementation of "POPCORN: Partially Observed Prediction Constrained Reinforcement Learning" (Futoma, Hughes, Doshi-Velez, AISTATS 2020)☆11May 19, 2021Updated 5 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Code associated with paper "High-Dimensional Contextual Policy Search with Unknown Context Rewards using Bayesian Optimization"☆16Dec 7, 2020Updated 5 years ago
- An open-source framework to benchmark and assess safety specifications of Reinforcement Learning problems.☆70Jul 7, 2023Updated 3 years ago
- ☆13Jul 9, 2022Updated 4 years ago
- Convergent Policy Optimization for Safe Reinforcement Learning☆11Oct 26, 2019Updated 6 years ago
- A web page to collect reproduced papers in one place with their codes☆14Mar 8, 2023Updated 3 years ago
- HIRE - HIgh Resolution Energy Demand simulation model☆16Dec 16, 2020Updated 5 years ago
- Implementation of MINIVAN (Mixed INteger InteractiVe plAnNing)☆20Jan 30, 2022Updated 4 years ago
- Code for SPIBB-DQN and Soft-SPIBB-DQN☆11May 5, 2020Updated 6 years ago
- ☆15May 11, 2022Updated 4 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Copy top-level annotations from one URL (and/or group) to another☆14Apr 29, 2021Updated 5 years ago
- The SSE version of the MCP service is modified from the Filesystem MCP server☆17May 21, 2025Updated last year
- Improved AppImage of opencode which is able to work on any linux system [Maintainer=@Samueru-sama]☆16Updated this week
- Simulator for online credit card transactions with multi-modal authentication☆21Nov 7, 2017Updated 8 years ago
- Quick definitions and intuitive explanations around machine learning.☆38May 19, 2026Updated 3 months ago
- Platform for training generalizable deep reinforcement learning agents☆15Mar 4, 2026Updated 6 months ago
- Implementation of ``Actor-Critic Alignment for Offline-to-Online Reinforcement Learning''☆13Oct 12, 2023Updated 2 years ago