Solves the Tower of Hanoi puzzle by Q-learning
☆27Nov 8, 2017Updated 8 years ago
Alternatives and similar repositories for Q-learning-Hanoi
Users that are interested in Q-learning-Hanoi are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Source code for Interpretable Reward Redistribution in Reinforcement Learning: A Causal Approach (NeurIPS 2023)☆10Dec 12, 2023Updated 2 years ago
- Randomized Linear Algebra in Python☆13Mar 21, 2017Updated 9 years ago
- flexible meta-learning in jax☆16Oct 19, 2023Updated 2 years ago
- JAX implementation of the Mistral 7b v0.1 model☆13Mar 27, 2024Updated 2 years ago
- Reimplementation of ToMNet with some extensions for RL as well☆14Apr 28, 2018Updated 8 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This library provides expression trees for representation of geometric expressions and automatic differentiation of these expressions. Th…☆14Jun 17, 2026Updated 3 months ago
- Illustration of counterfactual inference following Ferenc Huszar example☆13Aug 15, 2025Updated last year
- ☆19Apr 17, 2026Updated 5 months ago
- [TIP 2025] UniUIR: Considering Underwater Image Restoration as an All-in-One Learner☆17Jul 25, 2026Updated last month
- A2C training of Relational Deep Reinforcement Learning Architecture☆13Jun 22, 2022Updated 4 years ago
- SimPER: A Minimalist Approach to Preference Alignment without Hyperparameters (ICLR 2025)☆17Aug 22, 2025Updated last year
- A public repo for ICML 2021 "Shortest-Path Constrained Reinforcement Learning for Sparse Reward Tasks"☆13Jul 19, 2021Updated 5 years ago
- PyData San Luis 2017 Tutorial: An Introduction to Gaussian Processes in PyMC3☆14Nov 16, 2017Updated 8 years ago
- Convolutional Sparse Coding☆10Jul 18, 2014Updated 12 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Implementation of Variational Intrinsic Control in tensorflow☆11Apr 5, 2017Updated 9 years ago
- Learn to build neural networks from scratch, simply. No autograd, no deep learning libraries - just numpy.☆10Aug 10, 2022Updated 4 years ago
- Official implementation of Neural Episodic Control with State Abstraction☆13Aug 3, 2023Updated 3 years ago
- ☆10Dec 12, 2017Updated 8 years ago
- ☆19Sep 23, 2025Updated 11 months ago
- PyTorch implementation of linear and convolutional layers with fixed, random feedback weights.☆15Mar 14, 2021Updated 5 years ago
- A metaheuristic algorithm framework for solving discrete optimization problems☆20Jul 13, 2013Updated 13 years ago
- Few-shot Bayesian Imitation Learning with Policies as Logic over Programs☆22Oct 19, 2025Updated 11 months ago
- Codebase for Paper Reusing Embeddings: Reproducible Reward Model Research in Large Language Model Alignment without GPUs☆24Apr 24, 2025Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Drop-in environment replacements that make your RL algorithm train faster.☆22Jun 19, 2024Updated 2 years ago
- a library for deep reinforcement learning, with applications for navigation☆16Feb 6, 2018Updated 8 years ago
- CUDA extension for the SPORCO project☆18Jul 5, 2021Updated 5 years ago
- just a neater version of PointNet and PointNet++ in tensorflow☆13May 3, 2018Updated 8 years ago
- Discovering Quality-Diversity Algorithms via Meta-Black-Box Optimization☆28Dec 1, 2025Updated 9 months ago
- rail_surface_dataset☆31May 24, 2023Updated 3 years ago
- A Towers of Hanoi environment in OpenAI Gym Style☆14Jun 6, 2019Updated 7 years ago
- Code for recreating the figures in the brainrender paper (Claudi et al. 2020)☆12Dec 11, 2020Updated 5 years ago
- Minimal and Clean Reinforcement Learning Examples in PyTorch☆43Dec 25, 2018Updated 7 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- An awesome reading list for intuitive physics raning from cognitive studies to computational studies.☆23Aug 2, 2024Updated 2 years ago
- PyTorch implementation of Vanilla PG, TNPG, TRPO, PPO on Mujoco environment☆15Jul 1, 2018Updated 8 years ago
- hdnet - Hopfield denoising network☆15Oct 6, 2022Updated 3 years ago
- Game of Go implemented in Python☆10Jan 1, 2020Updated 6 years ago
- Automated Continuous Data Quality Measurement☆12Nov 15, 2023Updated 2 years ago
- VolFormer: Explore More Comprehensive Cube Interaction for Hyperspectral Image Restoration and Beyond☆28Mar 16, 2025Updated last year
- TensorFlow implementation for our paper "Learning Long-Term Reward Redistribution via Randomized Return Decomposition"☆19Mar 17, 2022Updated 4 years ago