Proximal Policy Optimization (Continuous Version) in PyTorch.
☆28May 12, 2025Updated last year
Alternatives and similar repositories for Continuous-PPO
Users that are interested in Continuous-PPO are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Modular Single-file Reinfocement Learning Algorithms Library☆38May 16, 2023Updated 3 years ago
- Émulateur Dofus 1.29.1 en Java☆14Dec 5, 2016Updated 9 years ago
- The core repository of the elsciRL framework.☆18Dec 8, 2025Updated 8 months ago
- A videogame made with PyGame turned into an Open AI Gym Learning Environment for Reinforcement Learning agents.☆14Jan 3, 2023Updated 3 years ago
- Codebase for the paper "How Crucial is Transformer in Decision Transformer?". Containing experiments on different pendulum tasks and code…☆28Mar 24, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Connect 4 AI using Monte Carlo Tree Search algorithm.☆11Feb 10, 2024Updated 2 years ago
- Dreamer on JAX☆16Jan 19, 2022Updated 4 years ago
- Code for Discovered Policy Optimisation (NeurIPS 2022)☆12Jun 15, 2023Updated 3 years ago
- The Laser Learning Environment (LLE) is a cooperative MARL grid-world☆13Aug 22, 2026Updated last week
- Proximal Policy Optimization(PPO) with Intrinsic Curiosity Module(ICM)☆18Apr 15, 2022Updated 4 years ago
- A Gym env for propulsive rocket landing.☆23Jun 7, 2022Updated 4 years ago
- 使用 rwkv 架構的大型語言模型運作的 Discord 聊天機器人,擁有辨識維基百科條目名並將其當作關鍵字抽取出的能力,並且還會將此關鍵字用於爬取維基百科,並將條目內容回傳給自身參考。近期更新串接API版本。☆12Aug 27, 2024Updated 2 years ago
- Learning Fair Policies in Decentralized Cooperative Multi-Agent Reinforcement Learning☆10Nov 14, 2021Updated 4 years ago
- Vivisecting Mobility Management in 5G Cellular Networks (SIGCOMM'22)☆13Jun 26, 2022Updated 4 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- A concurrent library based on cooperative scheduling of user-level threads(fibers) implemented in C++☆27Jun 15, 2022Updated 4 years ago
- Fork of https://github.com/xbpeng/DeepMimic☆15Sep 10, 2020Updated 5 years ago
- Scalable Probabilistic Estimates of Electric Vehicle Charging (SPEECh)☆14Nov 12, 2024Updated last year
- Demo Telegram Mini-App integrated with Turnkey.☆16May 12, 2026Updated 3 months ago
- ☆12Apr 26, 2022Updated 4 years ago
- [AAAI-2024] MATS-LP addresses the challenging problem of decentralized lifelong multi-agent pathfinding. The proposed approach utilizes a…☆31Jul 28, 2025Updated last year
- Distributed Deep Reinforcement Learning☆30Jan 21, 2021Updated 5 years ago
- A collection of matrix games in JAX☆14Apr 13, 2026Updated 4 months ago
- ☆15Jul 10, 2019Updated 7 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- This repository is an implementation of "MASER: Multi-Agent Reinforcement Learning with Subgoals Generated from Experience Replay Buffer"…☆23Jul 6, 2023Updated 3 years ago
- Implementation of Proximal Policy Optimization in Jax+Flax☆21May 18, 2023Updated 3 years ago
- Representing robots as graphs for reinforcement-learning in PyBullet locomotion environments.☆35Apr 11, 2021Updated 5 years ago
- My reproduction of various reinforcement learning algorithms (DQN variants, A3C, DPPO, RND with PPO) in Tensorflow.☆37Mar 24, 2023Updated 3 years ago
- https://x.com/BlinkDL_AI/status/1884768989743882276☆28May 4, 2025Updated last year
- Web application where humans can play Overcooked with AI agents.☆60Dec 6, 2022Updated 3 years ago
- A reinforcement learning agent that learns to solve mazes using Group Relative Policy Optimization (GRPO).☆12Feb 9, 2025Updated last year
- A simple userspace program to interact with Linux KVM☆23Aug 4, 2023Updated 3 years ago
- Implementation of Gated State Spaces, from the paper "Long Range Language Modeling via Gated State Spaces", in Pytorch☆101Feb 25, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Build a Responsive Calendar App with HTML, CSS and Javascript | Tutorial 2024☆20Oct 11, 2024Updated last year
- Acer Nitro 5 AN515-54 Hackintosh EFI☆22Feb 12, 2025Updated last year
- The project to learn the QMIX.☆13Dec 19, 2019Updated 6 years ago
- gym environment for drone 2D active perception in dynamic environment☆15Feb 9, 2026Updated 6 months ago
- 心理所筆記☆18Jun 25, 2025Updated last year
- Store and display microservice dependency graph☆18May 24, 2016Updated 10 years ago
- Code for magnetic mirror descent.☆20Oct 5, 2023Updated 2 years ago