Implement DQN and DDQN algorithm on Atari games,such as BreakoutNoFrameskip-v4, PongNoFrameskip-v4,BoxingNoFrameskip-v4.
☆15Jun 30, 2020Updated 6 years ago
Alternatives and similar repositories for DQN-pytorch-Atari
Users that are interested in DQN-pytorch-Atari are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Implement PPO algorithm on mujoco environment,such as Ant-v2, Humanoid-v2, Hopper-v2, Halfcheeth-v2.☆59Jun 30, 2020Updated 6 years ago
- Play Atari(Breakout) Game by DRL - DQN, Noisy DQN and A3C☆15May 30, 2020Updated 6 years ago
- use DQN(pytorch) to play pong☆12May 30, 2021Updated 5 years ago
- A simple human interface for human-in-the-loop machine learning research, which allows: 1. annote image on webpage, 2. collect human feed…☆14May 24, 2024Updated 2 years ago
- ☆17Oct 12, 2023Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A PyTorch implementation of PTSA-MCTS from [Accelerating Monte Carlo Tree Search with Probability Tree State Abstraction].☆16Oct 21, 2023Updated 2 years ago
- PyTorch implementation of Vanilla PG, TNPG, TRPO, PPO on Mujoco environment☆15Jul 1, 2018Updated 8 years ago
- A simple baseline for mountain-car @ gym☆12Jan 15, 2020Updated 6 years ago
- ☆12Apr 27, 2018Updated 8 years ago
- Deep Q-Learning (DQN) implementation for Atari pong.☆84Nov 22, 2022Updated 3 years ago
- `.torrent`文件解析器☆11Mar 27, 2021Updated 5 years ago
- Stock Price prediction for Yahoo Inc. using GRU (Gated Recurrant Units) in Keras. Predicting closing price for Yahoo stocks☆23Apr 13, 2018Updated 8 years ago
- 掼蛋AI☆13Oct 18, 2020Updated 5 years ago
- DQN with pytorch with on Breakout and SpaceInvaders☆27Aug 13, 2019Updated 6 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Source and solution codes for Professional CUDA C Programming book.☆15Aug 20, 2020Updated 5 years ago
- Solving the OpenAI Gym (MountainCarContinuous-v0) with DDPG☆21Jan 23, 2023Updated 3 years ago
- 一个强大的MCP(Model Context Protocol)开发框架,一个用于SSE对接的模块化工具框架。该框架允许开发者轻松创建和扩展自定义工具,支持JWT鉴权,并通过MCP协议与模型交互。☆17May 15, 2025Updated last year
- Playing Mountain-Car without reward engineering, by combining DQN and Random Network Distillation (RND)☆41Jan 28, 2019Updated 7 years ago
- Udacity Deep Reinforcement Learning Nanodegree Program☆11Jul 12, 2019Updated 7 years ago
- Code repo for the NeurIPS 2021 paper "Online Adaption to Label Distribution Shift".☆16Feb 15, 2023Updated 3 years ago
- Codebase of NeurIPS 2022 paper ''Planning for Sample Efficient Imitation Learning''☆41Oct 25, 2022Updated 3 years ago
- Implementation of MuZero with PyTorch, based on the pseudocode from DeepMind (https://arxiv.org/src/1911.08265v2/anc/pseudocode.py).☆33Aug 14, 2022Updated 3 years ago
- Official Code for "SoliReward: Mitigating Susceptibility to Reward Hacking and Annotation Noise in Video Generation Reward Models" [CVPR2…☆21Jul 13, 2026Updated 2 weeks ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Minimal example to apply Decision Transformer in Atari Pong☆16Feb 1, 2025Updated last year
- A webapp for monitoring GPU machines, written in Vue.js and Flask.☆23Nov 21, 2021Updated 4 years ago
- ☆12Feb 23, 2023Updated 3 years ago
- General-Purpose Reinforcement Learning☆18Oct 31, 2021Updated 4 years ago
- ARMv6 binary build for Trojan-GFW(树莓派用上trojan)☆17Mar 10, 2020Updated 6 years ago
- This is the web scraper and world file processor for the text2mc project☆18Nov 15, 2024Updated last year
- Code Tricks For Python☆21Jan 9, 2019Updated 7 years ago
- MydockFinder是一款极致模拟Mac OS的软件,这款软件不是黑苹果系统,也不需要重装系统,仅需要打开运行程序并完成个性化设置后即可伪装为Mac OS系统,设置为开机自启后即可随时体验到Mac OS的动画效果。☆21Nov 5, 2023Updated 2 years ago
- 极简版上海交大课程作业LaTex模板(支持多作者)。A XeLaTeX template of TERM PAPER or PROJECT for Shanghai Jiao Tong University (SJTU) students.☆20Dec 30, 2019Updated 6 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆17Mar 4, 2024Updated 2 years ago
- Official implementation of NeurIPS'24 paper "Reinforcement Learning Policy as Macro Regulator Rather than Macro Placer".☆20Aug 13, 2025Updated 11 months ago
- ☆10Dec 10, 2021Updated 4 years ago
- Interface definitions for the Compute@Edge platform in witx.☆15Feb 11, 2022Updated 4 years ago
- Q-learning and SARSA algorithms from Sutton's Reinforcement Learning book.☆21May 19, 2019Updated 7 years ago
- ☆40Jun 19, 2024Updated 2 years ago
- ☆18May 24, 2023Updated 3 years ago