A PyTorch implementation of DeepMind's MCTSnet
☆18Dec 8, 2022Updated 3 years ago
Alternatives and similar repositories for MCTSnet
Users that are interested in MCTSnet are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- (CoRL 2019 Spotlight) Asynchronous Methods for Model-Based Reinforcement Learning☆14Dec 27, 2022Updated 3 years ago
- PyTorch implementation of Munchausen Reinforcement Learning based on DQN and SAC. Handles discrete and continuous action spaces☆15Oct 3, 2021Updated 4 years ago
- Example implementation of Alpha Zero' s algotirhm on Jupyter notebook☆15Nov 21, 2019Updated 6 years ago
- Implementation for paper "A Consciousness-Inspired Planning Agent for Model-Based Reinforcement Learning".☆60Sep 25, 2024Updated last year
- Accelerating Exact Constrained Shortest Paths on GPUs☆15Dec 11, 2020Updated 5 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- The simple C/C++ library for hexapod (Robot spider with 6 legs) on Arduino.☆13Dec 27, 2018Updated 7 years ago
- Autonomous visual navigation using the depth images☆11Aug 15, 2019Updated 7 years ago
- Monte Carlo Tree Search (MCTS) ,realize using python☆12Mar 10, 2016Updated 10 years ago
- Code for NeurIPS 2021 paper "Curriculum Offline Imitation Learning"☆18Oct 21, 2022Updated 3 years ago
- Code for the paper 'Monte Carlo Tree Search for Asymmetric Trees'☆13May 24, 2018Updated 8 years ago
- Graph convolutional memory for reinforcement learning☆24Jul 10, 2021Updated 5 years ago
- Sokoban solver☆17Jun 11, 2026Updated 2 months ago
- Documentation and ressources of Kraby, an open-source hexapod robot☆15Aug 24, 2020Updated 5 years ago
- TD-VAE in PyTorch☆10May 28, 2019Updated 7 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Hexapod Robot Control☆10May 8, 2023Updated 3 years ago
- Simple C++ project that includes header only implementations of Monte Carlo Tree Search(MCTS), Temporal Difference Learning, Minimax, an…☆11Jan 29, 2026Updated 6 months ago
- A Unity WebGL project for a TicTacToe game, using Monte Carlo Tree Search (MCTS) for its AI decision making.☆13Mar 18, 2023Updated 3 years ago
- ☆10May 15, 2020Updated 6 years ago
- Planet: A unified sampling-based approach to integrated task and motion planning☆16Jul 9, 2020Updated 6 years ago
- Astar and RRT implementation using matplotlib☆10May 24, 2020Updated 6 years ago
- ☆15Sep 22, 2023Updated 2 years ago
- Self-normalizing neural network implemented in PyTorch.☆12Apr 3, 2019Updated 7 years ago
- PlaNet: Learning Latent Dynamics for Planning from Pixels☆10Feb 13, 2020Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code for "Demonstration-free Autonomous Reinforcement Learning via Implicit and Bidirectional Curriculum" (ICML 2023)☆10Jul 6, 2023Updated 3 years ago
- Example of android app written in Qt/Qml which uses MXNet for plant image recognition.☆10Nov 4, 2017Updated 8 years ago
- Free file storage options for Heroku hosted applications☆12Jan 27, 2025Updated last year
- Source code for experiments in "Identifying and Addressing Delusions for Target- Directed Decision Making"☆10Jun 1, 2025Updated last year
- Repository for the paper "Generative Adversarial Network to Learn Valid Distributions of Robot Configurations for Inverse Kinematics and …☆17Jul 24, 2022Updated 4 years ago
- ProdSim is a process-based discrete event simulation for production environments based on the SimPy framework☆40Dec 29, 2021Updated 4 years ago
- Python demo for the paper "Pareto Monte Carlo Tree Search for Multi-Objective Informative Planning".☆35Nov 9, 2022Updated 3 years ago
- A quadruped running machine in webots with three different gaits: trotting, pacing, and bounding.☆14May 22, 2022Updated 4 years ago
- Source code of Neural Logic Reinforcement Learning (https://arxiv.org/abs/1904.10729)☆78Jan 6, 2020Updated 6 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- An implementation of the paper "Solving the Rubik's Cube without Human Knowledge"☆14Dec 9, 2018Updated 7 years ago
- Repo for the paper: Learning with Muscles: Benefits for Data-Efficiency and Robustness in Anthropomorphic Tasks. https://al.is.mpg.de/pub…☆16Dec 1, 2022Updated 3 years ago
- Mobile manipulator Task and Motion Planning(TAMP) implementaion by using legacy Method (BasePlacement). This repository is Tested in C…☆19Oct 23, 2024Updated last year
- homework for shenlan's "Motion Planning For Mobile Robots "☆15May 14, 2020Updated 6 years ago
- This github gathers everything needed to reproduce AntBot, the hexapod robot I developped during my PhD.☆28Apr 6, 2020Updated 6 years ago
- Upper Confidence Tree Planner for ATARI games☆19Mar 9, 2016Updated 10 years ago
- Model-Free-Episodic-Control implementation.☆18Jun 3, 2019Updated 7 years ago