A reinforcement leaning environment for discrete MDPs.
☆25Nov 10, 2024Updated last year
Alternatives and similar repositories for matrix-mdp-gym
Users that are interested in matrix-mdp-gym are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Library that provides environments for planning problems☆17Apr 24, 2026Updated 3 months ago
- ☆20Apr 29, 2019Updated 7 years ago
- ☆15Sep 22, 2023Updated 2 years ago
- Simple Grid Environment for Gymnasium☆65Mar 1, 2026Updated 5 months ago
- NWB extension to store pose estimation data☆17Updated this week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Cost-aware Bayesian optimization via the Pandora's box Gittins index☆13Aug 8, 2025Updated last year
- This is a Human-like Upper-limb Motion Planner (HUMP) for the generation of arm-hand movements in humanoid robots.☆12Mar 4, 2022Updated 4 years ago
- Code for the paper "Learning to Do or Learning While Doing: Reinforcement Learning and Bayesian Optimisation for Online Continuous Tuning…☆14Nov 15, 2023Updated 2 years ago
- CookingZoo: a gym-cooking derivative to simulate a complex cooking environment☆22Dec 6, 2024Updated last year
- Implementation of "Active Exploration for Inverse Reinforcement Learning (AceIRL), NeurIPS 2022.☆14Oct 12, 2022Updated 3 years ago
- Residual Quantization Autoencoder, used for interpreting LLMs☆14Jan 1, 2025Updated last year
- Service Robot Simulator☆11May 3, 2020Updated 6 years ago
- ppx_system is a syntax extension to known operating system at compile time☆12May 9, 2023Updated 3 years ago
- ☆13Feb 24, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆13Mar 12, 2024Updated 2 years ago
- NWB Explorer is a web application to visualise and analyse the content of NWB:N 2 files☆27Aug 28, 2025Updated 11 months ago
- Deep universal probabilistic programming with Python and PyTorch☆14Apr 1, 2020Updated 6 years ago
- code for BINOCULARS and Multi-Step BO☆12Dec 7, 2020Updated 5 years ago
- Code for replicating experiments from the paper, Preference Exploration for Efficient Bayesian Optimization with Multiple Outcomes, publi…☆14Jun 22, 2023Updated 3 years ago
- ☆12Mar 17, 2024Updated 2 years ago
- NeurIPS 2019 Paper☆12Dec 9, 2019Updated 6 years ago
- ☆16May 11, 2023Updated 3 years ago
- Code for the paper "Harnessing Discrete Representations for Continual Reinforcement Learning"☆16Jun 16, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Tackling Non-Stationarity in Reinforcement Learning via Causal-Origin Representation (ICML 2024)☆12Aug 9, 2024Updated 2 years ago
- PreferenceNet: Encoding Human Preferences in Auction Design With Deep Learning☆17Aug 10, 2021Updated 5 years ago
- Change Mattermost font to Vazirmatn☆15Sep 29, 2023Updated 2 years ago
- Structure refinement software for total scattering data☆15Updated this week
- Metrics for spike sorting validation/quality control☆15Sep 8, 2021Updated 4 years ago
- ☆19Oct 2, 2023Updated 2 years ago
- An implementation of the RRTx Algorithm in Python☆11Apr 16, 2024Updated 2 years ago
- Create Custom GYM Environment for SUMO and reinforcement learning agant☆15May 5, 2023Updated 3 years ago
- Skew Gaussian Processes by Alessio Benavoli, Dario Azzimonti and Dario Piga☆16Aug 5, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Run adaptive decision making experiments☆16Nov 9, 2021Updated 4 years ago
- Tools, algorithms, and frameworks for managing and analyzing neural data.☆21Mar 21, 2022Updated 4 years ago
- Pytorch implementation on OpenAI's Procgen ppo-baseline, built from scratch.☆14May 17, 2024Updated 2 years ago
- This is the code repository for the paper "Zero-Sum Stochastic Stackelberg Games".☆18Oct 12, 2022Updated 3 years ago
- ☆13Jun 30, 2020Updated 6 years ago
- model based reinforcement learning algorithms for unstable baselines☆15May 9, 2023Updated 3 years ago
- The AI Arena: A framework for distributed multi-agent reinforcement learning☆14Aug 5, 2022Updated 4 years ago