Run OpenAI Gym on a Server
☆17Aug 25, 2017Updated 9 years ago
Alternatives and similar repositories for CartPole
Users that are interested in CartPole are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Monte Carlo Conterfactual Regret Minimization for imperfect information games☆13Mar 29, 2019Updated 7 years ago
- AlphaGo Zero Reinforcement Learning Sokoban Solver☆11Jun 20, 2018Updated 8 years ago
- ☆15Sep 22, 2023Updated 3 years ago
- Code for "Demonstration-free Autonomous Reinforcement Learning via Implicit and Bidirectional Curriculum" (ICML 2023)☆10Jul 6, 2023Updated 3 years ago
- Code for "Dynamic Discounted Counterfactual Regret Minimization", ICLR 2024 (Spotlight)☆19Apr 22, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Implementation of the Playground environment from the paper Language as a Cognitive Tool to Imagine Goals inCuriosity-Driven Exploration.☆11Mar 5, 2021Updated 5 years ago
- ☆47Feb 12, 2021Updated 5 years ago
- Which fellows cited my article?☆25Mar 6, 2022Updated 4 years ago
- A site comparing services of different Cloud Vendors☆10Jan 4, 2017Updated 9 years ago
- ☆10Mar 13, 2017Updated 9 years ago
- 采样FCRN: Fully-Convolutional Regression Network (全卷积回归网络),出自VGG 实验室这篇 CVPR2016的Paper:Synthetic Data for Text Localisation in Natural Image…☆10Jun 13, 2017Updated 9 years ago
- Model-Free-Episodic-Control implementation.☆18Jun 3, 2019Updated 7 years ago
- ☆13Jan 14, 2020Updated 6 years ago
- opencvprojects for android☆13Jan 27, 2013Updated 13 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [IJCAI 2021] Solving Continuous Control with Episodic Memory☆15Apr 10, 2022Updated 4 years ago
- Code for "Minimizing Weighted Counterfactual Regret with Optimistic Online Mirror Descent", IJCAI 2024 (Oral)☆16Aug 27, 2024Updated 2 years ago
- Code to reproduce the results in the "Unsupervised Learning of Goal Spaces for Intrinsically Motivated Exploration"☆21Feb 14, 2018Updated 8 years ago
- Convolutional Neural Network for Click-Through Rate prediction.☆15Sep 28, 2016Updated 9 years ago
- Thompson Sampling for Bandits using UCB policy☆10Jul 29, 2017Updated 9 years ago
- ☆17Oct 22, 2022Updated 3 years ago
- a python implementation of plsa☆25Oct 25, 2014Updated 11 years ago
- ☆16Jul 1, 2021Updated 5 years ago
- PyTorch implementation of "The Option Keyboard: Combining Skills in Reinforcement Learning" (NeurIPS 2019)☆12Jul 2, 2020Updated 6 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- A Reinforcement learning model which applies deterministic policy gradient algorithms to maximize return on portfolio management task☆14Jun 2, 2018Updated 8 years ago
- 利用图神经网络进行CTR预估☆15Nov 22, 2019Updated 6 years ago
- MoDem-V2 combines the sample efficiency of the original MoDem with conservative exploration in order to quickly and safely learn manipula…☆25Apr 1, 2024Updated 2 years ago
- A docker container that lets you run AirSim without building it.☆14Sep 20, 2017Updated 9 years ago
- A simple tool for labeling object bounding boxes in images☆12Oct 7, 2017Updated 8 years ago
- Episodic Control☆22Sep 20, 2022Updated 4 years ago
- Code for "AutoCFR: Learning to Design Counterfatual Regret Minimization Algorithms", AAAI 2022 (Oral)☆22Apr 22, 2024Updated 2 years ago
- Domain-Robust Visual Imitation Learning with Mutual Information Constraints code☆19Mar 1, 2021Updated 5 years ago
- Matlab toolbox containing algorithms from Gunnar Farnebäck's PhD and postdoc research.☆17Jul 22, 2014Updated 12 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Collection of game-theoretic algorithms for Poker☆30Apr 6, 2019Updated 7 years ago
- 微信朋友圈,QQ空间,微博等列表展示的功能实现☆15May 24, 2017Updated 9 years ago
- ☆24Feb 15, 2022Updated 4 years ago
- kafka学习实例demo☆12Aug 23, 2016Updated 10 years ago
- Automate Tensorflow Object Detection training and deployment with Azure Machine Learning Service☆15Apr 16, 2019Updated 7 years ago
- ☆34Oct 1, 2018Updated 7 years ago
- ☆24Jun 5, 2021Updated 5 years ago