Bandits Environments for the OpenAI Gym
☆89Jan 15, 2020Updated 6 years ago
Alternatives and similar repositories for gym-bandits
Users that are interested in gym-bandits are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Multi-armed bandits environments for OpenAI Gym☆10May 27, 2018Updated 8 years ago
- A simple Gridworld environment for Open AI gym☆25Jun 10, 2018Updated 8 years ago
- Library for Multi-Armed Bandit Algorithms☆57Apr 2, 2017Updated 9 years ago
- Motion imitation with deep reinforcement learning.☆13Jul 24, 2019Updated 6 years ago
- OpenAI Gym environment for DART robotics simulator.☆22Apr 17, 2018Updated 8 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Code for Tajima et al. (2019). Optimal policy for multi-alternative decisions. Nature Neuroscience.☆10Aug 23, 2019Updated 6 years ago
- ☆16Jan 24, 2020Updated 6 years ago
- 레이튼 교수화 최후의 시간여행☆10Apr 19, 2017Updated 9 years ago
- Simple grid-world environment compatible with OpenAI-gym☆50Mar 19, 2020Updated 6 years ago
- Gridworld environments for OpenAI gym.☆79Feb 15, 2024Updated 2 years ago
- A Tensorflow implementation of VGG-16 trained on CIFAR-100☆11May 25, 2018Updated 8 years ago
- A basic 2D maze environment where an agent start from the top left corner and try to find its way to the bottom left corner.☆375Oct 9, 2023Updated 2 years ago
- Hindsight policy gradients☆46Jan 31, 2020Updated 6 years ago
- ☆15Mar 21, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- R package for Partially Observable Markov Decision Processes☆21Aug 6, 2025Updated 11 months ago
- Multi-armed bandit simulation library☆140Nov 9, 2023Updated 2 years ago
- Actor-critic with experience replay☆257Oct 9, 2022Updated 3 years ago
- This repository contains the code used in the paper Evaluating the Performance of Reinformcent Learning Algorithms☆27Aug 14, 2021Updated 4 years ago
- Materials for the Practical Sessions of the Reinforcement Learning Summer School 2019: Bandits, RL & Deep RL (PyTorch).☆90Aug 21, 2019Updated 6 years ago
- This repo is intended as an extension for OpenAI Gym for auxiliary tasks (multitask learning, transfer learning, inverse reinforcement le…☆220Jul 22, 2019Updated 6 years ago
- This package allows to use PLE as a gym environment.☆71Jun 26, 2020Updated 6 years ago
- Approximate convex decomposition(ACD)☆10Sep 9, 2023Updated 2 years ago
- A framework for easy prototyping of distributed reinforcement learning algorithms☆97Dec 8, 2020Updated 5 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Python library for Multi-Armed Bandits☆770Feb 11, 2020Updated 6 years ago
- 🎮 A configurable Breakout environment for reinforcement learning☆11Mar 20, 2018Updated 8 years ago
- Official Project Webpage for paper "DiffSRL: Learning Dynamic-aware State Representation for Control via Differentiable Simulation"☆12Apr 4, 2022Updated 4 years ago
- Visualizations of Reinforcement Learning concepts including Value Iteration and Q-Learning☆18Jan 1, 2018Updated 8 years ago
- ☆28Apr 15, 2017Updated 9 years ago
- pdfplot is a Python library for easily managing your matplotlib figures as PDF files.☆13Jul 8, 2020Updated 6 years ago
- 1140 Dandisets, 965.1 TB total. DataLad super-dataset of all Dandisets from https://github.com/dandisets☆14May 28, 2026Updated last month
- A deep reinforcement learning multi-agent algorithm, where a team learns to complete a task and communicate between agents.☆16Jun 1, 2021Updated 5 years ago
- A package to study complex networks based on the temporal evolution of their Dynamic Communicability and Flow.☆11Jan 30, 2026Updated 5 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆10Jan 5, 2018Updated 8 years ago
- DNN-based small variant caller☆12May 2, 2022Updated 4 years ago
- Boilerplate Electron Application with Handlebars.js/Material Design CSS☆13Dec 11, 2015Updated 10 years ago
- Tinker is a parallel-by-default File/Directory Management System with additional interface to NLP and ML libraries☆10Jul 21, 2017Updated 9 years ago
- ☆11Nov 11, 2016Updated 9 years ago
- Reinforcement Learning -- Imitation Learning, Behavior Cloning, DAgger (Data Aggregation)☆22Apr 15, 2018Updated 8 years ago
- ☆38Mar 6, 2017Updated 9 years ago