Implement DQN and DDQN algorithm on Atari games,such as BreakoutNoFrameskip-v4, PongNoFrameskip-v4,BoxingNoFrameskip-v4.
☆15Jun 30, 2020Updated 6 years ago
Alternatives and similar repositories for DQN-pytorch-Atari
Users that are interested in DQN-pytorch-Atari are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- In this project, I explore some typical value-based and policy-based RL algorithms. I do experiments on DQN and its six variants and thei…☆12Nov 18, 2020Updated 5 years ago
- Play Atari(Breakout) Game by DRL - DQN, Noisy DQN and A3C☆15May 30, 2020Updated 6 years ago
- use DQN(pytorch) to play pong☆12May 30, 2021Updated 5 years ago
- ☆10May 15, 2020Updated 6 years ago
- ☆17Oct 12, 2023Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆26Mar 12, 2015Updated 11 years ago
- A PyTorch implementation of PTSA-MCTS from [Accelerating Monte Carlo Tree Search with Probability Tree State Abstraction].☆17Oct 21, 2023Updated 2 years ago
- PyTorch implementation of Vanilla PG, TNPG, TRPO, PPO on Mujoco environment☆15Jul 1, 2018Updated 8 years ago
- Tagbeat - Sensing Vibration through Backscatter Signals!☆17Jun 14, 2017Updated 9 years ago
- A simple baseline for mountain-car @ gym☆12Jan 15, 2020Updated 6 years ago
- ☆12Apr 27, 2018Updated 8 years ago
- Implementation of Russo and Van Roy work on Information Directed Sampling (2017)☆21Jan 18, 2019Updated 7 years ago
- ☆18Mar 28, 2023Updated 3 years ago
- Deep Q-Learning (DQN) implementation for Atari pong.☆84Nov 22, 2022Updated 3 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- A model problem for big data self adaptive systems using SUMO and TraCI☆17Feb 11, 2024Updated 2 years ago
- Repository for IROS 2019☆28Nov 16, 2019Updated 6 years ago
- An efficient interactive zero-knowledge proof scheme based on GKR in terms of unlayered circuit.☆18May 23, 2023Updated 3 years ago
- Implementations of algorithms and protocols from Justin Thaler's "Proofs, Arguments, and Zero-knowledge"☆21Feb 19, 2023Updated 3 years ago
- 掼蛋AI☆14Oct 18, 2020Updated 5 years ago
- DQN with pytorch with on Breakout and SpaceInvaders☆27Aug 13, 2019Updated 7 years ago
- The code for my TraCI tutorial found at https://www.youtube.com/watch?v=YntoPdPFFkU☆23Nov 20, 2019Updated 6 years ago
- 一个强大的MCP(Model Context Protocol)开发框架,一个用于SSE对接的模块化工具框架。该框架允许开发者轻松创建和扩展自定义工具,支持JWT鉴权,并通过MCP协议与模型交互。☆17May 15, 2025Updated last year
- Some templates for experimental plots.☆21May 22, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- A zk-SNARK for randomized algorithms + linear-size universal circuits (https://eprint.iacr.org/2020/278)☆19Jan 24, 2021Updated 5 years ago
- Playing Mountain-Car without reward engineering, by combining DQN and Random Network Distillation (RND)☆41Jan 28, 2019Updated 7 years ago
- Udacity Deep Reinforcement Learning Nanodegree Program☆11Jul 12, 2019Updated 7 years ago
- 👨💻 ❤️ 💻 本科阶段比较值得参加的与代码有关的比赛,或者活动,或者组织☆37Jun 17, 2025Updated last year
- Codebase of NeurIPS 2022 paper ''Planning for Sample Efficient Imitation Learning''☆41Oct 25, 2022Updated 3 years ago
- Official Code for "SoliReward: Mitigating Susceptibility to Reward Hacking and Annotation Noise in Video Generation Reward Models" [CVPR2…☆23Jul 13, 2026Updated 2 months ago
- Implementation of MuZero with PyTorch, based on the pseudocode from DeepMind (https://arxiv.org/src/1911.08265v2/anc/pseudocode.py).☆33Aug 14, 2022Updated 4 years ago
- Minimal example to apply Decision Transformer in Atari Pong☆16Feb 1, 2025Updated last year
- ☆12Feb 23, 2023Updated 3 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Implementations of protocols from the book 'Proofs, Arguments and Zero Knowledge'☆21Jun 18, 2026Updated 3 months ago
- This is a concise Pytorch implementation of Rainbow DQN, including Double Q-learning, Dueling network, Noisy network, PER and n-steps Q-l…☆38Jul 23, 2022Updated 4 years ago
- Timing prediction dataset download and instructions.☆19Jun 7, 2023Updated 3 years ago
- General-Purpose Reinforcement Learning☆18Oct 31, 2021Updated 4 years ago
- ARMv6 binary build for Trojan-GFW(树莓派用上trojan)☆17Mar 10, 2020Updated 6 years ago
- gkr-mimc is a POC-grad gnark gadget to accelerate the proving time of Mimc computation☆25Jun 24, 2024Updated 2 years ago
- MydockFinder是一款极致模拟Mac OS的软件,这款软件不是黑苹果系统,也不需要重装系统,仅需要打开运行程序并完成个性化设置后即可伪装为Mac OS系统,设置为开机自启后即可随时体验到Mac OS的动画效果。☆21Nov 5, 2023Updated 2 years ago