Solving the Rubik's cube with deep reinforcement learning and Monte Carlo tree search
☆109Apr 15, 2019Updated 7 years ago
Alternatives and similar repositories for puzzle_cube
Users that are interested in puzzle_cube are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- An implementation of the paper "Solving the Rubik's Cube without Human Knowledge"☆14Dec 9, 2018Updated 7 years ago
- Rubik's Cube solver using reinforcement learning☆56May 3, 2024Updated 2 years ago
- Python 2x2 Rubik's Cube representation & solver☆21Jul 29, 2022Updated 4 years ago
- Tabula Rasa Tic-Tac-Toe☆10Jan 3, 2019Updated 7 years ago
- Solving Rubik's Cube with Deep Reinforcement Learning, A* and visualize with PyQt5.☆24Aug 24, 2022Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- different AI algorithms to solve board games☆18Nov 4, 2018Updated 7 years ago
- Online demo of DRLViz, an interactive tool to understand decisions and memory in Deep Reinforcement Learning☆16Dec 8, 2022Updated 3 years ago
- Source Code for 'Beginning Game Programming with Pygame Zero: Coding Interactive Games on Raspberry Pi Using Python' by Stewart Watkiss☆15Aug 1, 2020Updated 5 years ago
- Project 1 of Udacity's Deep Reinforcement Learning nanodegree program☆13Dec 2, 2018Updated 7 years ago
- Code for the paper "Skynet: A Top Deep RL Agent in the Inaugural Pommerman Team Competition"☆38May 9, 2019Updated 7 years ago
- ☆14Jun 21, 2016Updated 10 years ago
- Code for "Dream and Search to Control: Latent Space Planning for Continuous Control"☆12Jul 12, 2021Updated 5 years ago
- Some baselines for Pommerman competition☆46Jul 18, 2018Updated 8 years ago
- This project was created for Unity ML-Agents Challenge - https://connect.unity.com/challenges/ml-agents-1☆12Aug 15, 2020Updated 5 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- CUDA extension for the SPORCO project☆18Jul 5, 2021Updated 5 years ago
- Implementation of SPW and DPW for Monte Carlo Tree Search in Continuous action/state space☆21Oct 3, 2023Updated 2 years ago
- A standalone release of DeepMind Lab's maze generator with Python bindings.☆69Oct 3, 2023Updated 2 years ago
- MCM/ICM 2017 B☆10Jan 29, 2017Updated 9 years ago
- Reinforcement Learning (RL) is believe to be a more general approach towards Artificial Intelligence (AI). RL is the foundation for many …☆14Dec 22, 2022Updated 3 years ago
- This is my personal wiki for hacking the router firmware used by (Sagemcom)F@ast Version 3.43.2 delivered from Sagemcom☆17Jun 1, 2019Updated 7 years ago
- Code for Continual Reinforcement Learning with Multi-Timescale Replay☆24Apr 16, 2020Updated 6 years ago
- ☆14May 10, 2021Updated 5 years ago
- Some hard problems for reinforcement learning.☆32Oct 5, 2018Updated 7 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- A Software Defined Network (SDN) code with a load balancer that I implemented along with layer-3 routing application.☆11Jan 4, 2017Updated 9 years ago
- [CHIL 2024] Interpretation of Intracardiac Electrograms Through Textual Representations☆12Sep 4, 2024Updated last year
- ☆10May 15, 2020Updated 6 years ago
- Various DQN method with cartpole☆11May 30, 2018Updated 8 years ago
- gui for board game hex (and Y) by broderick arneson☆15Dec 13, 2023Updated 2 years ago
- Open source Stable Diffusion examples and workflows for Ryzen AI software☆22Updated this week
- Astar and RRT implementation using matplotlib☆10May 24, 2020Updated 6 years ago
- (CoRL 2019 Spotlight) Asynchronous Methods for Model-Based Reinforcement Learning☆14Dec 27, 2022Updated 3 years ago
- for learning reinforcement learning using PyTorch.☆64Oct 2, 2019Updated 6 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Code used for the master thesis at MIIS (UPF)☆16Dec 1, 2016Updated 9 years ago
- Monte Carlo tree search (MCTS) on traveling salesman problem (TSP)☆22Apr 27, 2019Updated 7 years ago
- Display PDFs in your Dash apps.☆24Oct 2, 2024Updated last year
- matlab code for SEARCH: A Stochastic Election Approach for Heterogeneous Wireless Sensor Networks☆10Jun 11, 2016Updated 10 years ago
- Scalable Log Determinants for Gaussian Process Kernel Learning (https://arxiv.org/abs/1711.03481) (NIPS 2017)☆18Nov 10, 2017Updated 8 years ago
- Software for Open Source Ecology's MicroTrac☆12Oct 17, 2017Updated 8 years ago
- Codes of our team for the OpenAI Retro Contest of reinforcement learning☆98Jun 19, 2018Updated 8 years ago