Tutorial series on how to implement DQN with PyTorch from scratch.
☆72Mar 20, 2025Updated last year
Alternatives and similar repositories for dqn_pytorch
Users that are interested in dqn_pytorch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- An Gymnasium environment for the Flappy Bird game☆115Jan 28, 2026Updated 7 months ago
- ☆12Oct 24, 2024Updated last year
- This is a custom project for WGU, the original project repo is https://github.com/udacity/nd0821-c2-build-model-workflow-starter☆14Feb 1, 2026Updated 7 months ago
- Pointax: PointMaze Environment for JAX☆28Oct 22, 2025Updated 10 months ago
- ☆15Oct 10, 2025Updated 11 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Flappy Bird as a Farama Gymnasium environment.☆39Aug 1, 2023Updated 3 years ago
- OpenAI Gym 课程练习笔记☆15Apr 16, 2024Updated 2 years ago
- The interface between probabilistic model checking and data-driven policy learning.☆20Sep 6, 2026Updated 2 weeks ago
- Informed Rapidly-exploring Random Tree-Star with C# Programming☆10Nov 6, 2021Updated 4 years ago
- Download files from web☆13Nov 9, 2016Updated 9 years ago
- A Reinforcement Learning / Neural Network library, written in Rust.☆21Mar 7, 2021Updated 5 years ago
- PyBullet CartPole and Quadrotor environments—with CasADi symbolic a priori dynamics—for learning-based control and RL☆26Jun 30, 2026Updated 2 months ago
- trading by Deep Q-Network☆15Oct 20, 2016Updated 9 years ago
- Fast computation of soft tissue thermal response under deformation based on fast explicit dynamics finite element algorithm for surgical …☆10May 30, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Quantum-Machine-Learning, published by Packt☆12Jan 18, 2021Updated 5 years ago
- A community build website for Software Developer Academy Pro☆11Apr 1, 2024Updated 2 years ago
- ☆10Aug 2, 2022Updated 4 years ago
- ☆11Nov 28, 2024Updated last year
- 这是一个菜鸟关于视觉SLAM的学习笔记☆10Sep 29, 2019Updated 6 years ago
- High-performance localization software for autonomous vehicles. A particle filter is combined with a map to localize a vehicle.☆14Jan 13, 2021Updated 5 years ago
- A resource hub for developers, PMs, and designers building LLM-forward products☆16Apr 12, 2026Updated 5 months ago
- Abstract 2D multi-robot multi-interface mobile robot simulator (fork of MobileSim simulator with various fixes, optimizations, and added …☆16May 30, 2023Updated 3 years ago
- ML from scratch in Jax☆12Aug 20, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Multi agent PPO implementation in Pytorch for Unity ML Agents environments.☆29Jul 25, 2024Updated 2 years ago
- Deep learning model zoo with PyTorch 1.X☆16Apr 13, 2019Updated 7 years ago
- Clustering Through Decision Tree Construction☆33Apr 14, 2019Updated 7 years ago
- ☆30Apr 20, 2024Updated 2 years ago
- A lightweight particle filter in C++☆13Jan 16, 2022Updated 4 years ago
- A* is a computer algorithm that is widely used in pathfinding and graph traversal, which is the process of finding a path between multipl…☆10Apr 29, 2019Updated 7 years ago
- A clean Pytorch implementation of DDPG on continuous action space.☆31Jun 8, 2024Updated 2 years ago
- 5GMdata project: Scripts to repeatedly invoke Remcom Wireless Insite and SUMO☆13Oct 3, 2025Updated 11 months ago
- ☆15Apr 21, 2022Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- We use reachability to ensure the safety of a decision agent acting on a dynamic system in real-time. We compute the Forward Reachable Se…☆34Jun 19, 2021Updated 5 years ago
- SAC, PPO, A2C implementation on Mujoco environments : Humanoid-v4, Ant-v4, Cheetah-v4 . Includes reward manipulation.☆37Sep 1, 2025Updated last year
- Benchmark codebase for 2D range finder based people detectors using the FROG dataset☆14Oct 20, 2025Updated 11 months ago
- ☆17Jan 3, 2025Updated last year
- An implementation of the paper ‘Channel Distribution Learning: Model-Driven GAN-Based Channel Modeling for IRS-Aided Wireless Communicati…☆16Oct 27, 2022Updated 3 years ago
- Contains MATLAB and Python codes and plots for deriving inferences for various concepts of Wireless Communications.☆16Jan 5, 2021Updated 5 years ago
- Simulation of RRT, RRT*, RRT*-FN and RRT*-FND algorithms.☆14May 15, 2023Updated 3 years ago