An implementation of Deep Q-Learning from Demonstrations (DQfD) for playing Atari 2600 video games
☆31Dec 10, 2022Updated 3 years ago
Alternatives and similar repositories for DQfD
Users that are interested in DQfD are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- An implement of DQfD(Deep Q-learning from Demonstrations) raised by DeepMind:Learning from Demonstrations for Real World Reinforcement Le…☆130Dec 5, 2017Updated 8 years ago
- PyTorch Implementation of Visual GAIL in Atari Games☆14Dec 7, 2022Updated 3 years ago
- ☆15Jul 4, 2022Updated 4 years ago
- This is pytorch implmentation project of Bootsrapped DQN☆13Dec 6, 2020Updated 5 years ago
- (ICLR 2021) Learning to Represent Action Values as a Hypergraph on the Action Vertices☆23Jun 22, 2021Updated 5 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Codes for the paper "HAVEN: Hierarchical Cooperative Multi-Agent Reinforcement Learning with Dual Coordination Mechanism"☆27Oct 22, 2022Updated 3 years ago
- This is a project using Pytorch to fulfill reinforcement learning on a simple game - Gridworld☆14Jul 13, 2020Updated 6 years ago
- ☆10Sep 21, 2020Updated 6 years ago
- Code for the NeurIPS 2021 paper "Deep Bandits Show-Off: Simple and Efficient Exploration with Deep Networkst"☆14Sep 12, 2022Updated 4 years ago
- Transfer Learning in Reinforcement Learning using Stable-Baseline3 | Transfer Reinforcement Learning for Differing Action Spaces via Q-Ne…☆21Feb 27, 2022Updated 4 years ago
- Editable scientific figures in PowerPoint and draw.io via Codex/MCP with Designer-Drawer-Reviewer-Corrector quality gates.☆854Sep 8, 2026Updated 2 weeks ago
- Predicting path with preference based on user demonstration using Maximum Entropy Deep Inverse Reinforcement Learning in a continuous env…☆25Jun 10, 2022Updated 4 years ago
- The ALL Arduino Nano 33 BLE Sense Classifier is an experiment to explore how low powered microcontrollers, specifically the Arduino Nano …☆10Jul 21, 2021Updated 5 years ago
- Integrate AutoRL into DQN to implement a single traffic signal control system.☆16Nov 16, 2023Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ROS wrapper for AirDet☆12Jul 24, 2022Updated 4 years ago
- Never fill a sockaddr_in struct by hand again!☆13Apr 10, 2020Updated 6 years ago
- DDPGfD: This is our implementation project for the Reinforcement Learning course in NCTU.☆35Feb 13, 2022Updated 4 years ago
- Segway Simulation Environment☆11Dec 31, 2020Updated 5 years ago
- Tensorflow Implementation for "Pre-trained Deep Convolution Neural Network Model With Attention for Speech Emotion Recognition"☆10Dec 19, 2021Updated 4 years ago
- This is an attempt to summarize feature engineering methods that I have learned over the course of my graduate school.☆11Mar 3, 2022Updated 4 years ago
- This repository is an implementation of "MASER: Multi-Agent Reinforcement Learning with Subgoals Generated from Experience Replay Buffer"…☆23Jul 6, 2023Updated 3 years ago
- Adaptive PI controller based on a reinforcement learning algorithm for speed control of a DC motor☆13Oct 5, 2023Updated 2 years ago
- This is the open source code of Cumulative Curriculum Reinforcement Learning (CCRL)☆20Mar 11, 2023Updated 3 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- arduino due code to control a 4 wheeled differential vehicle by a cmd_vel callback using rosserial pid_arduino_library and quadrature_enc…☆13Aug 3, 2019Updated 7 years ago
- Compare Laguerre-based MPC and Traditional MPC for platoon of vehicles.☆13Feb 14, 2023Updated 3 years ago
- Keras implementation of guide actor-critic for continuous control☆11Mar 12, 2018Updated 8 years ago
- The original version of Firefox for GNS3 is completely outdated, so I made a new version with the latest update☆12Feb 14, 2025Updated last year
- Latency Equalization Policy of End-to-End Network Slicing Based on Reinforcement Learning☆19Jul 15, 2023Updated 3 years ago
- ☆16May 5, 2022Updated 4 years ago
- Forked from: https://github.com/carcamdou/cr_grasper☆10Dec 7, 2020Updated 5 years ago
- ☆13May 25, 2026Updated 3 months ago
- Synthesis of Control Barrier Functions Using a Supervised Machine Learning Approach☆13Apr 1, 2020Updated 6 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Arduino TinyML trash classification example☆10Dec 21, 2020Updated 5 years ago
- PyTorch implementation of the paper Overcoming Exploration in Reinforcement Learning with Demonstrations in surgical robot manipulation t…☆12Aug 21, 2022Updated 4 years ago
- Official codebase for Improving Computational Efficiency in Visual Reinforcement Learning via Stored Embeddings.☆21Mar 5, 2021Updated 5 years ago
- A trivial raycaster using minifb for rendering/input☆13Jan 2, 2023Updated 3 years ago
- Nonparallel Emotional Speech Conversion with MUNIT. Introduction: This is a tensorflow implementation of paper(https://arxiv.org/pdf/1811…☆14Oct 13, 2021Updated 4 years ago
- ☆13Apr 26, 2025Updated last year
- gabor filter bank, sift and bag of visual words implementation☆11Jul 20, 2019Updated 7 years ago