Proximal Policy Optimization (Continuous Version) in PyTorch.
☆28May 12, 2025Updated last year
Alternatives and similar repositories for Continuous-PPO
Users that are interested in Continuous-PPO are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official codebase for paper "Revisiting Some Common Practices in Cooperative Multi-Agent Reinforcement Learning" (ICML22)☆23Jul 16, 2022Updated 4 years ago
- My Submission for the OpenAI/NeurIPS ProcGen Competition☆11Nov 12, 2020Updated 5 years ago
- Modular Single-file Reinfocement Learning Algorithms Library☆38May 16, 2023Updated 3 years ago
- Émulateur Dofus 1.29.1 en Java☆14Dec 5, 2016Updated 9 years ago
- The core repository of the elsciRL framework.☆18Dec 8, 2025Updated 7 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A videogame made with PyGame turned into an Open AI Gym Learning Environment for Reinforcement Learning agents.☆14Jan 3, 2023Updated 3 years ago
- Codebase for the paper "How Crucial is Transformer in Decision Transformer?". Containing experiments on different pendulum tasks and code…☆28Mar 24, 2023Updated 3 years ago
- Connect 4 AI using Monte Carlo Tree Search algorithm.☆11Feb 10, 2024Updated 2 years ago
- Dreamer on JAX☆16Jan 19, 2022Updated 4 years ago
- Code for Discovered Policy Optimisation (NeurIPS 2022)☆12Jun 15, 2023Updated 3 years ago
- Deep Q-Learning (DQN) implementation for Atari pong.☆86Nov 22, 2022Updated 3 years ago
- The Laser Learning Environment (LLE) is a cooperative MARL grid-world☆13Updated this week
- 使用 rwkv 架構的大型語言模型運作的 Discord 聊天機器人,擁有辨識維基百科條目名並將其當作關鍵字抽取出的能力,並且還會將此關鍵字用於爬取維基百科,並將條目內容回傳給自身參考。近期更新串接API版本。☆12Aug 27, 2024Updated last year
- Proximal Policy Optimization(PPO) with Intrinsic Curiosity Module(ICM)☆18Apr 15, 2022Updated 4 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- A Gym env for propulsive rocket landing.☆23Jun 7, 2022Updated 4 years ago
- A concurrent library based on cooperative scheduling of user-level threads(fibers) implemented in C++☆27Jun 15, 2022Updated 4 years ago
- Learning Fair Policies in Decentralized Cooperative Multi-Agent Reinforcement Learning☆10Nov 14, 2021Updated 4 years ago
- Vivisecting Mobility Management in 5G Cellular Networks (SIGCOMM'22)☆13Jun 26, 2022Updated 4 years ago
- Fork of https://github.com/xbpeng/DeepMimic☆15Sep 10, 2020Updated 5 years ago
- Scalable Probabilistic Estimates of Electric Vehicle Charging (SPEECh)☆13Nov 12, 2024Updated last year
- ☆12Apr 26, 2022Updated 4 years ago
- A collection of matrix games in JAX☆14Apr 13, 2026Updated 3 months ago
- A simple userspace program to interact with Linux KVM☆23Aug 4, 2023Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- 一个注重创作的轻博客系统,选用python语言flask框架开发,前端采用bootstrap4轻量模板,注重内容创作与工具开发☆11May 1, 2023Updated 3 years ago
- ☆15Jul 10, 2019Updated 7 years ago
- This repository is an implementation of "MASER: Multi-Agent Reinforcement Learning with Subgoals Generated from Experience Replay Buffer"…☆23Jul 6, 2023Updated 3 years ago
- Implementation of Proximal Policy Optimization in Jax+Flax☆21May 18, 2023Updated 3 years ago
- Representing robots as graphs for reinforcement-learning in PyBullet locomotion environments.☆35Apr 11, 2021Updated 5 years ago
- My reproduction of various reinforcement learning algorithms (DQN variants, A3C, DPPO, RND with PPO) in Tensorflow.☆37Mar 24, 2023Updated 3 years ago
- https://x.com/BlinkDL_AI/status/1884768989743882276☆28May 4, 2025Updated last year
- This is a collection of interesting papers that I have read so far or want to read. Note that the list is not up-to-date. Topics: reinfor…☆11Apr 3, 2025Updated last year
- Web application where humans can play Overcooked with AI agents.☆60Dec 6, 2022Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A reinforcement learning agent that learns to solve mazes using Group Relative Policy Optimization (GRPO).☆12Feb 9, 2025Updated last year
- Pathfinding Using Reinforcement Learning☆12May 21, 2019Updated 7 years ago
- Implementation of Gated State Spaces, from the paper "Long Range Language Modeling via Gated State Spaces", in Pytorch☆101Feb 25, 2023Updated 3 years ago
- Acer Nitro 5 AN515-54 Hackintosh EFI☆21Feb 12, 2025Updated last year
- Build a Responsive Calendar App with HTML, CSS and Javascript | Tutorial 2024☆19Oct 11, 2024Updated last year
- The project to learn the QMIX.☆13Dec 19, 2019Updated 6 years ago
- gym environment for drone 2D active perception in dynamic environment☆15Feb 9, 2026Updated 5 months ago