This is the official repository for the paper "Guided Exploration with Proximal Policy Optimization using a Single Demonstration", https://arxiv.org/abs/2007.03328
☆19Oct 5, 2021Updated 4 years ago
Alternatives and similar repositories for ppo_D
Users that are interested in ppo_D are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Reimplementation of Policy Optimization with Demonstrations (POfD) from ICML 2018.☆16Jun 5, 2019Updated 7 years ago
- Diabetic classification based on retinal images☆11Aug 26, 2019Updated 6 years ago
- Independent Generative Adversarial Self-Imitation Learning In Cooperative Multiagent Systems☆32Oct 9, 2018Updated 7 years ago
- ☆10Jun 28, 2022Updated 4 years ago
- [ICLR 2023] The official code for paper "Guarded Policy Optimization with Imperfect Online Demonstrations"☆14Apr 30, 2023Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Source code for journal paper "Multiagent Reinforcement Learning With Sparse Interactions by Negotiation and Knowledge Transfer"☆13Dec 26, 2017Updated 8 years ago
- ☆10Jun 22, 2020Updated 6 years ago
- A memory efficient implementation of custom SWISH and MISH activation functions in Pytorch☆12Jun 29, 2020Updated 6 years ago
- Implementation of clipped action policy gradient (CAPG) with PPO and TRPO☆31Jun 24, 2018Updated 8 years ago
- PyTorch implementation of the paper Overcoming Exploration in Reinforcement Learning with Demonstrations in surgical robot manipulation t…☆12Aug 21, 2022Updated 3 years ago
- TensorFlow implementation of "Sample-efficient Imitation Learning via Generative Adversarial Nets"☆10Dec 8, 2022Updated 3 years ago
- ☆42Mar 19, 2021Updated 5 years ago
- ⚡️ Shockingly fast imitation learning algorithms via combining online and offline data engines. ⚡️☆14Sep 1, 2025Updated 10 months ago
- Unity Machine Learning Agents Toolkit☆19Feb 8, 2020Updated 6 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆14May 31, 2022Updated 4 years ago
- Notes on putting micropython on STM32F407VG bare board☆11Oct 7, 2019Updated 6 years ago
- DDPGfD: This is our implementation project for the Reinforcement Learning course in NCTU.☆36Feb 13, 2022Updated 4 years ago
- This repo contains the implementation of deep reinforcement learning (DRL) algorithms for virtual machine rescheduling in data centers.☆12Dec 2, 2022Updated 3 years ago
- Robot Learning from Expert Demonstration Using IRL☆13Mar 21, 2021Updated 5 years ago
- Resilient Multi-Agent Reinforcement Learning☆10Nov 4, 2022Updated 3 years ago
- Supporting code for the paper «Leveraging molecular structure and bioactivity with chemical language models for drug design»☆12Feb 22, 2022Updated 4 years ago
- BILIBILI.☆15Jan 6, 2019Updated 7 years ago
- [IEEE Transactions on Intelligent Transportation Systems] Curricular Subgoal for Inverse Reinforcement Learning☆18Jul 31, 2023Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- C implementation of RL and IRL algorithms☆19Jul 6, 2020Updated 6 years ago
- Intel Atom D2550 Embedded Motherboard☆13Dec 26, 2018Updated 7 years ago
- Testing Spark Structured Streaming anf Kafka with real data from traffic sensors☆17Nov 11, 2022Updated 3 years ago
- Codes for the paper "Multi-task Hierarchical Adversarial Inverse Reinforcement Learning"☆19May 20, 2023Updated 3 years ago
- Common molecule fragments for visualization in Avogadro☆17Apr 1, 2026Updated 3 months ago
- Gradient Boosting Models on Real-Time Sensor Data for AI-Enhanced Vehicle Predictive Maintenance. By using a web-based interface to forec…☆19Nov 17, 2024Updated last year
- MindSpore implementations of deep reinforcement learning algorithms and environments☆17Sep 3, 2023Updated 2 years ago
- Algorithms for Uni-Modal Inverse Reinforcement Learning☆22Sep 23, 2022Updated 3 years ago
- ☆55Jul 16, 2021Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Implementation of ResFit, Residual Off-Policy RL for Finetuning Behavior Cloning Policies☆17Sep 29, 2025Updated 9 months ago
- Bias-controlled 3D generative framework for structure-based ligand design☆17Nov 2, 2022Updated 3 years ago
- IJCAI 2019 - Regularized Opponent Model with Maximum Entropy Objective (ROMMEO)☆23Dec 8, 2022Updated 3 years ago
- M^3PC: Test-Time Model Predictive Control for Pretrained Masked Trajectory Model, ICLR 2025☆19Mar 17, 2025Updated last year
- An OpenAI gym environment for the Kuka arm.☆46Dec 8, 2022Updated 3 years ago
- ☆14May 18, 2020Updated 6 years ago
- ☆12Dec 13, 2023Updated 2 years ago