Implement DQN and DDQN algorithm on Atari games,such as BreakoutNoFrameskip-v4, PongNoFrameskip-v4,BoxingNoFrameskip-v4.
☆15Jun 30, 2020Updated 6 years ago
Alternatives and similar repositories for DQN-pytorch-Atari
Users that are interested in DQN-pytorch-Atari are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Implement PPO algorithm on mujoco environment,such as Ant-v2, Humanoid-v2, Hopper-v2, Halfcheeth-v2.☆58Jun 30, 2020Updated 6 years ago
- Play Atari(Breakout) Game by DRL - DQN, Noisy DQN and A3C☆15May 30, 2020Updated 6 years ago
- use DQN(pytorch) to play pong☆12May 30, 2021Updated 5 years ago
- ☆17Oct 12, 2023Updated 2 years ago
- Materials for "Prompting is not a substitute for probability measurements in large language models" (EMNLP 2023)☆24Oct 24, 2023Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- PyTorch implementation of DQN, DDQN and Dueling DQN to solve Atari games including PongNoFrameskip-v4, BreakoutNoFrameskip-v4 and BoxingN…☆18Jun 1, 2023Updated 3 years ago
- [AAAI 2026] Few-step Flow for 3D Generation via Marginal-Data Transport Distillation☆57Apr 29, 2026Updated 4 months ago
- Implementation of Russo and Van Roy work on Information Directed Sampling (2017)☆21Jan 18, 2019Updated 7 years ago
- Deep Q-Learning (DQN) implementation for Atari pong.☆84Nov 22, 2022Updated 3 years ago
- Stock Price prediction for Yahoo Inc. using GRU (Gated Recurrant Units) in Keras. Predicting closing price for Yahoo stocks☆23Apr 13, 2018Updated 8 years ago
- Implementations of algorithms and protocols from Justin Thaler's "Proofs, Arguments, and Zero-knowledge"☆21Feb 19, 2023Updated 3 years ago
- 掼蛋AI☆14Oct 18, 2020Updated 5 years ago
- DQN with pytorch with on Breakout and SpaceInvaders☆27Aug 13, 2019Updated 7 years ago
- 一个强大的MCP(Model Context Protocol)开发框架,一个用于SSE对接的模块化工具框架。该框架允许开发者轻松创建和扩展自定义工具,支持JWT鉴权,并通过MCP协议与模型交互。☆17May 15, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Solving the OpenAI Gym (MountainCarContinuous-v0) with DDPG☆21Jan 23, 2023Updated 3 years ago
- This repo is for our submission for ICSE 2025.☆20Jun 12, 2024Updated 2 years ago
- Udacity Deep Reinforcement Learning Nanodegree Program☆11Jul 12, 2019Updated 7 years ago
- Playing Mountain-Car without reward engineering, by combining DQN and Random Network Distillation (RND)☆41Jan 28, 2019Updated 7 years ago
- A psycholinguistic modeling toolkit☆35Jul 30, 2026Updated last month
- For practice to using halo2☆22Jun 7, 2023Updated 3 years ago
- ☆12Feb 23, 2023Updated 3 years ago
- This is a concise Pytorch implementation of Rainbow DQN, including Double Q-learning, Dueling network, Noisy network, PER and n-steps Q-l…☆38Jul 23, 2022Updated 4 years ago
- General-Purpose Reinforcement Learning☆18Oct 31, 2021Updated 4 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- ARMv6 binary build for Trojan-GFW(树莓派用上trojan)☆17Mar 10, 2020Updated 6 years ago
- MydockFinder是一款极致模拟Mac OS的软件,这款软件不是黑苹果系统,也不需要重装系统,仅需要打开运行程序并完成个性化设置后即可伪装为Mac OS系统,设置为开机自启后即可随时体验到Mac OS的动画效果。☆21Nov 5, 2023Updated 2 years ago
- ☆17Mar 4, 2024Updated 2 years ago
- ☆10Dec 10, 2021Updated 4 years ago
- Interface definitions for the Compute@Edge platform in witx.☆15Feb 11, 2022Updated 4 years ago
- Incremental Convex Hull Algorithm and SAT Collision Detection for 3D Objects.☆19Jun 19, 2021Updated 5 years ago
- HAProxy combined with confd for HTTP load balancing with SSL offloading☆10Feb 5, 2017Updated 9 years ago
- 🎳 Environments for Reinforcement Learning☆66Feb 5, 2026Updated 7 months ago
- PoC of Swift for Compute@Edge☆12Feb 3, 2022Updated 4 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- 🎾 Multi-Agent Proximal Policy Optimization approach to a competitive reinforcement learning problem☆22Sep 25, 2022Updated 3 years ago
- PyTorch code to train and evaluate Procgen tasks☆26Nov 1, 2020Updated 5 years ago
- A repository for a Deep Q-Learning approach to intrusion detection for networks cyber-attacks.☆10Sep 3, 2021Updated 5 years ago
- An OpenAI gym multi-agent environment implementing the Commons Game proposed in "A multi-agent reinforcement learning model of common-poo…☆22Jun 20, 2020Updated 6 years ago
- This project shall be based on setting up of Google Football Research Environment as an OpenAI gym for code purposes.☆20Jul 4, 2019Updated 7 years ago
- ☆56Nov 10, 2022Updated 3 years ago
- Pytorch realization of multiple Deep Reinforcement Learning alogrithms(DQN,DDPG,TD3,PPO,A3C...) with openai gym☆58Aug 28, 2021Updated 5 years ago