Adapting the AlphaZero algorithm to remove the need of execution traces to train NPI.
☆79Oct 3, 2023Updated 2 years ago
Alternatives and similar repositories for AlphaNPI
Users that are interested in AlphaNPI are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆26Jan 2, 2019Updated 7 years ago
- Extending the Neural Graph Algorithm Executor☆13Dec 8, 2022Updated 3 years ago
- Template for building 2D grid worlds with OpenAI Gym and Pycolab☆14Jun 12, 2019Updated 7 years ago
- 🐈⬛ Contextual bandits library for continuous action trees with smoothing in JAX☆72Jul 28, 2026Updated last month
- ☆25Nov 23, 2021Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Neural Programmer-Interpreter Implementation (Reed, de Freitas: https://arxiv.org/abs/1511.06279), in Tensorflow☆43Nov 17, 2018Updated 7 years ago
- Basic experiment framework for tensorflow.☆90Jun 24, 2021Updated 5 years ago
- Code for NeurIPS 2019 paper: "Symmetry-Based Disentangled Representation Learning requires Interaction with Environments" by H. Caselles-…☆34Dec 9, 2019Updated 6 years ago
- Tensorflow code for WACV 2019 paper "Attention Based Natural Language Grounding by Navigating Virtual Environment" - https://arxiv.org/ab…☆17Nov 7, 2018Updated 7 years ago
- Variational Walkback, NIPS'17☆28Oct 18, 2017Updated 8 years ago
- 🪐 The Sebulba architecture to scale reinforcement learning on Cloud TPUs in JAX☆61Aug 22, 2026Updated last month
- ☆19Nov 7, 2020Updated 5 years ago
- ☆85May 29, 2019Updated 7 years ago
- Karel dataset for program synthesis and program induction☆79Dec 24, 2017Updated 8 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- A TensorFlow Label Propagation library☆13Apr 7, 2018Updated 8 years ago
- Code for the paper Physics-as-Inverse-Graphics: Joint Unsupervised Learning of Objects and Physics from Video☆41May 22, 2023Updated 3 years ago
- A python implemenation of tabular MuZero for educational purposes☆21Dec 11, 2019Updated 6 years ago
- Code of the paper: Debiasing Meta-Gradient Reinforcement Learning by Learning the Outer Value Function☆13Apr 13, 2026Updated 5 months ago
- STRIPS Planning in Infinite Domains☆19Oct 19, 2020Updated 5 years ago
- [NeurIPS 2019] Code for the paper "Learning to Control Self-Assembling Morphologies: A Study of Generalization via Modularity"☆120Dec 13, 2019Updated 6 years ago
- Learning and Reasoning with Graph-Structured Data (ICML 2019 Workshop)☆26Jul 18, 2019Updated 7 years ago
- using information theory to encourage agents to cooperate and compete☆19Oct 4, 2018Updated 7 years ago
- This project was moved to: https://github.com/coax-dev/coax☆162Nov 28, 2022Updated 3 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- CompILE: Compositional Imitation Learning and Execution (ICML 2019)☆113May 12, 2019Updated 7 years ago
- **Sferes2 module** A unifying modular framework for Quality-Diversity algorithms☆22Nov 6, 2020Updated 5 years ago
- Upside-Down Reinforcement Learning (⅂ꓤ) implementation in PyTorch. Based on the paper published by Jürgen Schmidhuber.☆79Aug 13, 2020Updated 6 years ago
- StarCraft: BroodWars OpenAI Gym environment☆84Jan 8, 2019Updated 7 years ago
- Meta-Inverse Reinforcement Learning with Probabilistic Context Variables☆77Mar 16, 2023Updated 3 years ago
- Repository for the paper "Long-Horizon Visual Planning with Goal-Conditioned Hierarchical Predictors"☆46Nov 22, 2022Updated 3 years ago
- BabyAI platform. A testbed for training agents to understand and execute language commands.☆766Oct 1, 2023Updated 2 years ago
- Code for Harmonic Exponential Families on Manifolds☆10Jun 2, 2016Updated 10 years ago
- ☆24Dec 9, 2020Updated 5 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Code for Semantically Robust Unpaired Image Translation for Data with Unmatched Semantics Statistics (SRUNIT), ICCV 2021☆11Feb 10, 2022Updated 4 years ago
- Neural Turing Machine (NTM) & Differentiable Neural Computer (DNC) with pytorch & visdom☆279Feb 20, 2018Updated 8 years ago
- Qt-like event loops, signals and slots for communication across threads and processes in Python☆14Mar 26, 2024Updated 2 years ago
- Basic versions of agents from Spinning Up in Deep RL written in PyTorch☆208May 20, 2021Updated 5 years ago
- Code publication to the paper "Normalized Attention Without Probability Cage"☆17Aug 24, 2026Updated last month
- Code for "Learning Inductive Biases with Simple Neural Networks" (Feinman & Lake, 2018).☆22Jan 8, 2019Updated 7 years ago
- A highly-customisable gridworld game engine with some batteries included. Make your own gridworld games to test reinforcement learning ag…☆665Sep 6, 2019Updated 7 years ago