A MARL PPO implementation with tf-agents, configured for the MultiCarRacing-v0 Gym environment.
☆19Jun 24, 2021Updated 5 years ago
Alternatives and similar repositories for marl_ppo
Users that are interested in marl_ppo are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Multi-Agent Deep Recurrent Q-Learning with Bayesian epsilon-greedy on AirSim simulator☆13Apr 1, 2022Updated 4 years ago
- This repository contains the Python implementation of our submitted paper titled "Deep Reinforcement Learning for Joint Trajectory and Co…☆16Jun 29, 2024Updated 2 years ago
- PyTorch implementation of Soft Actor-Critic (SAC), Twin Delayed DDPG (TD3), Actor-Critic (AC/A2C), Proximal Policy Optimization (PPO), QT…☆16Nov 18, 2020Updated 5 years ago
- An implementation for CVRP problem with A3C+Attention mechanism and GCN☆18May 17, 2020Updated 6 years ago
- Resilient Multi-Agent Reinforcement Learning☆10Nov 4, 2022Updated 3 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- An OpenAI Gym environment for multi-agent car racing based on Gym's original car racing environment.☆91Feb 20, 2026Updated 5 months ago
- BILIBILI.☆15Jan 6, 2019Updated 7 years ago
- SRL: Scaling Distributed Reinforcement Learning to Over Ten Thousand Cores☆15Apr 24, 2024Updated 2 years ago
- My implementation of common algorithms☆13Oct 6, 2019Updated 6 years ago
- 浙江大学Beamer模板☆16May 19, 2022Updated 4 years ago
- Classify the jamming pattern and predict the action of channel selection in the future time slots☆22Aug 28, 2021Updated 4 years ago
- Apple watch application to collect accelerometer and gyroscope data after detecting a golf swing.☆12Jan 16, 2019Updated 7 years ago
- PyTorch implementation of our paper Reinforcement Learning with Random Delays (ICLR 2020)☆44May 25, 2022Updated 4 years ago
- Simulation code for "Achievable Rate Maximization for Underlay Spectrum Sharing MIMO System with Intelligent Reflecting Surface," by V. K…☆26Nov 1, 2023Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆17Nov 16, 2020Updated 5 years ago
- ☆14May 18, 2020Updated 6 years ago
- Python demo for the paper "Pareto Monte Carlo Tree Search for Multi-Objective Informative Planning".☆35Nov 9, 2022Updated 3 years ago
- A robotics library for Python☆21Jun 12, 2026Updated 2 months ago
- an implementation of ATOC☆14Dec 6, 2021Updated 4 years ago
- ☆15Oct 26, 2022Updated 3 years ago
- In this repository there are the projects developed during the course of Advance Optimization-based Robot Control. The main topics are Ta…☆17Feb 20, 2023Updated 3 years ago
- ☆16Jun 30, 2019Updated 7 years ago
- ☆11Apr 29, 2018Updated 8 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A comparison of some conformal quantile regression methods.☆12Sep 14, 2019Updated 6 years ago
- [IJCAI 2021] Robust Adversarial Imitation Learning via Adaptively-Selected Demonstrations☆16Feb 17, 2023Updated 3 years ago
- Implementation of PILCO for the Model-Based Baselines Project☆18Jul 18, 2019Updated 7 years ago
- Multi agent PPO implementation in Pytorch for Unity ML Agents environments.☆29Jul 25, 2024Updated 2 years ago
- [ICLR 2026] SpikePingpong: Spike Vision-based Fast-Slow Pingpong Robot System☆21Mar 13, 2026Updated 4 months ago
- Reproducible research code for the experiments presented in our article "Kara1k: a karaoke dataset for cover song identification and sing…☆10Jan 9, 2018Updated 8 years ago
- Agile Quadruped Locomotion RL based on bullet☆23Jun 23, 2022Updated 4 years ago
- This is the official repository for the paper "Guided Exploration with Proximal Policy Optimization using a Single Demonstration", https:…☆19Oct 5, 2021Updated 4 years ago
- NeurIPS paper 'Censored Quantile Regression Neural Networks for Distribution-Free Survival Analysis'☆12Oct 28, 2022Updated 3 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆12Oct 20, 2023Updated 2 years ago
- PyTorch Implementation of COPA for coordinating teams that can dynamically change.☆24Apr 16, 2022Updated 4 years ago
- ☆12Jun 9, 2025Updated last year
- Multi-agent Reinforcement Learning Algorithms(COMA, VDN, QMIX)☆16May 24, 2020Updated 6 years ago
- A Fast, Portable Deep Reinforcement Learning Library for Continuous Control☆13Jul 26, 2023Updated 3 years ago
- ☆26Mar 11, 2026Updated 5 months ago
- [ICML 2021] Learning to Weight Imperfect Demonstrations☆20Nov 4, 2022Updated 3 years ago