Code for our paper: Hierarchical RL Using an Ensemble of Proprioceptive Periodic Policies
☆15Feb 21, 2019Updated 7 years ago
Alternatives and similar repositories for hrl-ep3
Users that are interested in hrl-ep3 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Implementation of Random Expert Distillation☆29May 11, 2019Updated 7 years ago
- minimum viable experiment using evolution strategy to play catch☆21Jan 28, 2018Updated 8 years ago
- Pytorch implementation of BEAR in "Stabilizing Off-Policy Q-Learning via Bootstrapping Error Reduction"☆11Oct 29, 2019Updated 6 years ago
- A3C style Option-Critic with deliberation cost☆40Jan 9, 2018Updated 8 years ago
- A Toolkit for creating Peripheral Architectures (NIPS 2016)☆26Sep 7, 2017Updated 8 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Gamepad API Content Kit☆14Jun 1, 2016Updated 10 years ago
- Inferring beliefs about dynamics from behavior☆30May 24, 2018Updated 8 years ago
- Code for the paper "Skynet: A Top Deep RL Agent in the Inaugural Pommerman Team Competition"☆38May 9, 2019Updated 7 years ago
- ☆11Oct 19, 2018Updated 7 years ago
- Ant Gather and Ant Maze envs, separated from RLLab☆11Aug 2, 2018Updated 7 years ago
- A TensorFlow implementation of perceptual generative autoencoder (PGA).☆22Nov 2, 2020Updated 5 years ago
- Bayesian pragmatic models implemented in Python☆21May 11, 2025Updated last year
- Simple change of a3c to a2c☆15Jun 18, 2017Updated 9 years ago
- Model-based reinforcement learning (generative simulator models and planning agents)☆16Mar 13, 2026Updated 4 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Shared autonomy via deep reinforcement learning☆80Mar 24, 2023Updated 3 years ago
- Winning solution of the Microsoft Research "First TextWorld Problems: A Reinforcement and Language Learning Challenge"☆12Jun 21, 2022Updated 4 years ago
- Intermediate wrapper of MOI for some linear quadratic solvers☆16Apr 17, 2020Updated 6 years ago
- Auxiliary variable Markov chain Monte Carlo methods☆10Oct 24, 2017Updated 8 years ago
- Algorithmic Framework for Model-based Deep Reinforcement Learning with Theoretical Guarantees☆55Jul 26, 2019Updated 7 years ago
- Ranking Policy Gradient☆23Nov 27, 2019Updated 6 years ago
- ReaSCAN is a synthetic navigation task that requires models to reason about surroundings over syntactically difficult languages. (NeurIPS…☆19Nov 28, 2021Updated 4 years ago
- Code database for Fast Texform generation as proposed in the work of Deza, Chen, Long and Konkle (CCN 2019).☆12Jul 26, 2019Updated 7 years ago
- ☆15Jan 20, 2020Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆18Jan 3, 2022Updated 4 years ago
- ☆13May 24, 2020Updated 6 years ago
- Profile repository of Pietro Monticone.☆15Updated this week
- Implementation of the Playground environment from the paper Language as a Cognitive Tool to Imagine Goals inCuriosity-Driven Exploration.☆11Mar 5, 2021Updated 5 years ago
- ☆11Oct 19, 2020Updated 5 years ago
- Creating fixed-length vectors to describe RL/GA policies☆20Oct 23, 2021Updated 4 years ago
- Table Query with ML☆14Jan 3, 2023Updated 3 years ago
- Code for Goal-Aware Prediction: Learning to Model what Matters☆20Jul 15, 2020Updated 6 years ago
- DSTC8-AVSD: Sentence generation task for Audio Visual Scene-aware Dialog☆14Jun 10, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Ray Tracing☆21Apr 6, 2017Updated 9 years ago
- Template Tracker based on Correlation Filters☆13Mar 23, 2015Updated 11 years ago
- A composable container for Adaptive ROS 2 Node computations. Select between FPGA, CPU or GPU at run-time.☆12Apr 14, 2022Updated 4 years ago
- Official repository for our ICLR 2021 paper Evaluating the Disentanglement of Deep Generative Models with Manifold Topology☆37Mar 22, 2021Updated 5 years ago
- Emotion classification of speech using GMMHMMs☆10Jul 1, 2016Updated 10 years ago
- 📝 Papers I read and notes/reviews I made. Also useful links to courses (RL/NLP/Bio/QC/DevOps)☆10May 4, 2021Updated 5 years ago
- Python MUD/MUX/MUSH/MU* development system☆26Oct 30, 2015Updated 10 years ago