A gym interface for AI safety gridworlds created in pycolab.
☆18May 12, 2022Updated 4 years ago
Alternatives and similar repositories for safe-grid-gym
Users that are interested in safe-grid-gym are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Training (hopefully) safe agents in gridworlds☆26May 12, 2019Updated 7 years ago
- Anomalous versions of OpenAI Gym and PyBullet3 environments☆15Oct 24, 2021Updated 4 years ago
- A collection of reading material for the Workshop on "Structure & Priors in Reinforcement Learning" (SPiRL) at ICLR 2019.☆13May 5, 2021Updated 5 years ago
- ☆13May 16, 2019Updated 7 years ago
- ☆26Jan 26, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A small bookmarks app for Solid☆11Jul 13, 2017Updated 9 years ago
- Web effectivethesis.com (and old version of efektivni-altruismus.cz)☆10Feb 22, 2022Updated 4 years ago
- Residual Quantization Autoencoder, used for interpreting LLMs☆14Jan 1, 2025Updated last year
- Checking D-separations and I-equivalence in Bayesian Networks.☆12Feb 11, 2017Updated 9 years ago
- A lightweight reimplementation of Adversarially Trained Actor Critic☆18Mar 19, 2026Updated 5 months ago
- Opinionated library for managing hyperparameters and mutable state of machine learning training systems.☆19Aug 4, 2023Updated 3 years ago
- Monte Carlo tree search for the travelling salesman problem (MCTS for the TSP)☆13Jun 18, 2022Updated 4 years ago
- Code for reproducing the results from the paper Avoiding Side Effects in Complex Environments☆12Jun 3, 2021Updated 5 years ago
- ☆11Jun 2, 2021Updated 5 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Code for the NeurIPS 2021 paper "Safe Reinforcement Learning by Imagining the Near Future"☆52Apr 8, 2022Updated 4 years ago
- Code to reproduce the experiments from the paper "Self-Compatibility: Evaluating Causal Discovery without Ground Truth"☆12Mar 9, 2024Updated 2 years ago
- A scalable Dreamer implementation in JAX☆10May 22, 2022Updated 4 years ago
- Interpretability dashboard for reinforcement learners☆16Jun 4, 2019Updated 7 years ago
- Experiments with representation engineering☆14Feb 28, 2024Updated 2 years ago
- SafeLife: safety benchmarks for reinforcement learning agents☆61May 13, 2021Updated 5 years ago
- Variational inference in Dirichlet process Gaussian mixture model (tensorflow implementation)☆13Oct 8, 2018Updated 7 years ago
- Brutaltester compatible referee for coders strike back☆13Jun 1, 2026Updated 3 months ago
- Scala Native 3 bindings for SFML library☆15Jul 9, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Code accompanying the paper "Information Directed Reward Learning for Reinforcement Learning" (NeurIPS 2021).☆13Nov 16, 2021Updated 4 years ago
- ☆14Aug 9, 2023Updated 3 years ago
- ☆14Oct 20, 2020Updated 5 years ago
- A tool for building Lean4 .olean files from Lean3 export data☆10Jul 28, 2021Updated 5 years ago
- An awesome list of Causality and Machine Learning related papers, books and other resources.☆12Nov 13, 2023Updated 2 years ago
- ☆16Nov 14, 2025Updated 9 months ago
- Code for the paper: Causal Action Influence Aware Counterfactual Data Augmentation @ICML2024☆13Jul 19, 2024Updated 2 years ago
- Unofficial and Partial Implementation of Fast AutoAugment in Pytorch☆10Oct 3, 2023Updated 2 years ago
- ☆13Jun 30, 2020Updated 6 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- EA Forum☆14Nov 19, 2018Updated 7 years ago
- Code that can be used to reproduce the experiments in our paper "Estimating Risk and Uncertainty in Deep Reinforcement Learning"☆31Nov 22, 2022Updated 3 years ago
- ☆12Apr 17, 2024Updated 2 years ago
- All code to download datasets for the three classification case studies, compute SPIs, fit SVMs, and visualise results for pyspi publicat…☆12Aug 14, 2023Updated 3 years ago
- Scaling safe exploration to vision control☆15Feb 19, 2025Updated last year
- Code for the ICLR 2021 Paper "In-N-Out: Pre-Training and Self-Training using Auxiliary Information for Out-of-Distribution Robustness"☆13Oct 23, 2021Updated 4 years ago
- Implementation of Johansson, Fredrik D., Shalit, Uri, and Sontag, David. Learning representations for counterfactual inference - ICML, 20…☆12Sep 30, 2020Updated 5 years ago