Solving the Rubik's cube with deep reinforcement learning and Monte Carlo tree search
☆109Apr 15, 2019Updated 7 years ago
Alternatives and similar repositories for puzzle_cube
Users that are interested in puzzle_cube are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Tabula Rasa Tic-Tac-Toe☆10Jan 3, 2019Updated 7 years ago
- JAX implementation of Graph Attention Networks☆13Jan 29, 2022Updated 4 years ago
- Project 1 of Udacity's Deep Reinforcement Learning nanodegree program☆13Dec 2, 2018Updated 7 years ago
- Upper Confidence Tree Planner for ATARI games☆19Mar 9, 2016Updated 10 years ago
- A3C style Option-Critic with deliberation cost☆40Jan 9, 2018Updated 8 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Customisable 3D environment for assessing generalisation in Reinforcement Learning.☆72Jun 12, 2023Updated 3 years ago
- High level Lean 4 FFI for Rust☆14Mar 16, 2024Updated 2 years ago
- Stochastic Markov Games☆12Oct 5, 2017Updated 8 years ago
- Code for "Dream and Search to Control: Latent Space Planning for Continuous Control"☆12Jul 12, 2021Updated 5 years ago
- Actor critic reinforcement learning + motion and task planning under LTL tasks + wireless sensor network routing☆15Mar 6, 2021Updated 5 years ago
- Implement Google Deep Minds DQN for multiple agents for a grid world environment where vehicles must pick up customers.☆29Mar 7, 2018Updated 8 years ago
- CUDA extension for the SPORCO project☆18Jul 5, 2021Updated 5 years ago
- My Simple Implementation of AlphaGo Zero on Connect4☆18Apr 25, 2018Updated 8 years ago
- A standalone release of DeepMind Lab's maze generator with Python bindings.☆69Oct 3, 2023Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- MCM/ICM 2017 B☆10Jan 29, 2017Updated 9 years ago
- Reinforcement Learning (RL) is believe to be a more general approach towards Artificial Intelligence (AI). RL is the foundation for many …☆14Dec 22, 2022Updated 3 years ago
- Code for Continual Reinforcement Learning with Multi-Timescale Replay☆24Apr 16, 2020Updated 6 years ago
- ☆19Sep 20, 2018Updated 7 years ago
- Some hard problems for reinforcement learning.☆32Oct 5, 2018Updated 7 years ago
- A Software Defined Network (SDN) code with a load balancer that I implemented along with layer-3 routing application.☆11Jan 4, 2017Updated 9 years ago
- RC-NFQ: Regularized Convolutional Neural Fitted Q Iteration. A batch algorithm for deep reinforcement learning. Incorporates dropout regu…☆12Mar 17, 2021Updated 5 years ago
- Published by Packt☆11Jan 18, 2021Updated 5 years ago
- the solustion to https://openai.com/requests-for-research☆12Mar 23, 2017Updated 9 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Neural network definition models☆21Nov 15, 2016Updated 9 years ago
- Various DQN method with cartpole☆11May 30, 2018Updated 8 years ago
- (CoRL 2019 Spotlight) Asynchronous Methods for Model-Based Reinforcement Learning☆14Dec 27, 2022Updated 3 years ago
- for learning reinforcement learning using PyTorch.☆64Oct 2, 2019Updated 6 years ago
- Repository for code experimenting with RL and Solar Tracking☆13Apr 24, 2018Updated 8 years ago
- Code used for the master thesis at MIIS (UPF)☆16Dec 1, 2016Updated 9 years ago
- ☆15Apr 1, 2026Updated 5 months ago
- Label images with a mouse click, within a Jupyter Notebook☆15Feb 11, 2023Updated 3 years ago
- MongoDB-based storage engine for Openchain☆11Dec 26, 2016Updated 9 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Code for the paper "Deep FTRL-ORW: An Efficient Deep Reinforcement Learning Algorithm for Solving Imperfect Information Extensive-Form Ga…☆11Dec 1, 2022Updated 3 years ago
- Implementation of a maximum area coverage algorithm in MATLAB☆14Dec 27, 2020Updated 5 years ago
- matlab code for SEARCH: A Stochastic Election Approach for Heterogeneous Wireless Sensor Networks☆10Jun 11, 2016Updated 10 years ago
- Must-read papers on AI for Biology☆27Oct 4, 2023Updated 2 years ago
- Scalable Log Determinants for Gaussian Process Kernel Learning (https://arxiv.org/abs/1711.03481) (NIPS 2017)☆18Nov 10, 2017Updated 8 years ago
- ☆28Sep 3, 2019Updated 7 years ago
- Creating DRL infrastructure for Dynamic Beta with Zipline and Keras☆14Dec 8, 2022Updated 3 years ago