Train agents on MiniGrid from human demonstrations using Inverse Reinforcement Learning
☆13Apr 15, 2020Updated 6 years ago
Alternatives and similar repositories for Minigrid_HCI-project
Users that are interested in Minigrid_HCI-project are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for Diagnosing Bottlenecks in Deep Q-learning. Contains implementations of tabular environments plus solvers.☆17May 14, 2019Updated 7 years ago
- Chemistry AR project in Unity to simulate chemical reactions☆10Sep 5, 2022Updated 4 years ago
- A library for building reinforcement learning and imitation learning agents in Pytorch☆61Jun 13, 2020Updated 6 years ago
- High-quality reference implementations of various algorithms for Inverse Reinforcement Learning☆13Jun 20, 2018Updated 8 years ago
- Augmented Reality App☆10Nov 24, 2016Updated 9 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Automatic code generator for training Reinforcement Learning policies☆11Jan 3, 2021Updated 5 years ago
- Emotiv SDK Community Edition☆13Oct 9, 2015Updated 10 years ago
- Implementing the two pioneering IRL papers "Algorithms for Inverse Reinforcement Learning" - (Ng &Russell 2000) and "Maximum Entropy Inve…☆31Jul 6, 2023Updated 3 years ago
- 2018 RoboCup@Rescue China / 2018 China Robot Competition@Rescue☆13Oct 8, 2019Updated 6 years ago
- a modular reinforcement learning library with JAX agents☆27Mar 3, 2025Updated last year
- Official implementation of the paper "Approximating two value functions instead of one: towards characterizing a new family of Deep Reinf…☆11Jul 14, 2021Updated 5 years ago
- ☆10Dec 29, 2020Updated 5 years ago
- ☆12Mar 4, 2024Updated 2 years ago
- ☆11Apr 12, 2020Updated 6 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Official implementation of "Regularizing neural networks for future trajectory prediction via IRL framework" published in IET CV☆37Jul 20, 2022Updated 4 years ago
- ☆12May 27, 2023Updated 3 years ago
- Research on Inverse Reinforcement Learning for self driving vehicles at UCLA☆13Nov 7, 2018Updated 7 years ago
- Gridworld for MARL experiments☆146Jan 29, 2021Updated 5 years ago
- Fine-tune GPT2 to generate fake job experiences☆11Jan 17, 2023Updated 3 years ago
- Landing a rocket in unity3d simulation using python☆14Jul 18, 2021Updated 5 years ago
- Ant algorithm to solve vehicle routing problems with time windows☆10May 26, 2018Updated 8 years ago
- Privacy-preserving Voice Analysis via Disentangled Representations☆12Aug 30, 2021Updated 5 years ago
- This is a project based on OpenAI's multi-agent-emergence-environments (Emergent Tool Use from Multi-Agent Autocurricula, Baker et al.), …☆13Jan 5, 2021Updated 5 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆12Jun 8, 2020Updated 6 years ago
- This is the codebase for our ICRA 2020 submission, GraphRQI: Classifying Driver Behaviors Using Graph Spectrums.☆13Dec 8, 2019Updated 6 years ago
- City Metro Network Expansion with Reinforcement Learning☆13Feb 13, 2020Updated 6 years ago
- Learning 2-opt Heuristics for the TSP via Deep Reinforcement Learning☆58Oct 20, 2020Updated 5 years ago
- Reinforcement Learning Enhanced Quantum-inspired Algorithm for Combinatorial Optimization☆16Feb 19, 2020Updated 6 years ago
- Web-based page layout editor created for EMOP (Early Modern OCR Project).☆11May 21, 2021Updated 5 years ago
- Chorus: Heterogeneous GPU+CPU Multiple Protein Sequences Alignment Search for Large Database☆17Apr 7, 2025Updated last year
- Explore and Control with Adversarial Surprise☆10Jul 20, 2021Updated 5 years ago
- 在PyTorch上重构multi-agent deep deterministic policy gradient(MADDPG),将https://github.com/xuemei-ye/maddpg-mpe 修改到自己电脑上可运行。因为本人笔记本没有CUDA,实验速度…☆14May 10, 2019Updated 7 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Codepack accompanying "Internal models for interpreting neural population activity during sensorimotor control," by Matthew D. Golub, Byr…☆16Jul 3, 2018Updated 8 years ago
- character recognition, textline recognition☆10Aug 31, 2019Updated 7 years ago
- Testing different RL algorithms for multi-agent environments. From SARSA, QLearning to Independent Q-Learning, Joint Action Learning and …☆12Mar 29, 2019Updated 7 years ago
- Digital twins are created using data derived from IoT sensors that are attached to or embedded in an object. This data provides both stru…☆32Feb 14, 2026Updated 6 months ago
- Github for the TOPCONS2☆17Mar 6, 2025Updated last year
- ☆10Mar 16, 2023Updated 3 years ago
- Predictive Maintenance avoids the drawbacks of Preventive Maintenance (under utilization of a part's life) and Reactive Maintenance (unsc…☆17Apr 25, 2022Updated 4 years ago