Official implementation for "Towards Safe Reinforcement Learning via Constraining Conditional Value at Risk" (IJCAI 2022)
☆27Aug 29, 2024Updated 2 years ago
Alternatives and similar repositories for CPPO
Users that are interested in CPPO are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- (AAAI24 oral) Implementation of RPPO(Risk-sensitive PPO) and RPBT(Population-based self-play with RPPO)☆12May 22, 2023Updated 3 years ago
- Active learning☆28Dec 17, 2020Updated 5 years ago
- ☆18Nov 22, 2023Updated 2 years ago
- ☆26Aug 21, 2024Updated 2 years ago
- Various explorations into the game of Poker using MCTS, NFSP, and image-recognition/web-scraping☆13Oct 23, 2020Updated 5 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆10Oct 3, 2023Updated 2 years ago
- Codebase for BRDiv: Diverse teammate generation for ad hoc teamwork☆13May 2, 2024Updated 2 years ago
- ☆11Mar 5, 2024Updated 2 years ago
- Codebase of paper "Self-Improving Safety Performance of Reinforcement Learning Based Driving with Black-Box Verification Algorithms" publ…☆12Jul 13, 2023Updated 3 years ago
- 基于强化学习的游戏空战推演☆13May 8, 2021Updated 5 years ago
- [ICLR 2023] The official code for paper "Guarded Policy Optimization with Imperfect Online Demonstrations"☆14Apr 30, 2023Updated 3 years ago
- Code for Policy Bifurcation in Safe Reinforcement Learning☆10Jul 4, 2025Updated last year
- ☆15Jul 23, 2025Updated last year
- MiniMax Multi-Agent Deep Deterministic Policy Gradient (M3DDPG) pytorch implementation☆15Feb 19, 2021Updated 5 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Soccer toy example simulator used in Reinforcement Learning☆12Mar 11, 2018Updated 8 years ago
- Implementation of Proximal Policy Optimization using Transformer☆12Jul 4, 2023Updated 3 years ago
- Deep Reinforcement Learning for Multi Agent Soccer☆16Dec 15, 2016Updated 9 years ago
- Experimenting with meta-learning approaches to opponent modelling in MARL. Building upon previous public implementations of MADDPG and M3…☆14Apr 26, 2022Updated 4 years ago
- Gradient-based planning on RPZ subspace.☆10Jun 15, 2023Updated 3 years ago
- Robust and safe deep reinforcement learning algorithms☆17Mar 27, 2024Updated 2 years ago
- Implementation of vanilla stochaistic (categorical) policy gradient algorithm to play cartpole.☆16Apr 1, 2021Updated 5 years ago
- Very Short Term Load Forecasting☆15Nov 1, 2011Updated 14 years ago
- Official codebase for Exact Energy-Guided Diffusion Sampling via Contrastive Energy Prediction (ICML 2023)☆55Aug 26, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆14May 29, 2024Updated 2 years ago
- Author's PyTorch implementation of paper Reinforced Imitative Trajectory Planning for Urban Automated Driving☆17Aug 7, 2025Updated last year
- ☆19Aug 27, 2020Updated 6 years ago
- Implementing different learning algorithms and analyzing their performance in a Markov game model called the Soccer Game☆23Jan 29, 2023Updated 3 years ago
- EEG情感识别☆16Dec 9, 2020Updated 5 years ago
- ☆20Sep 14, 2019Updated 7 years ago
- ☆18Jul 24, 2023Updated 3 years ago
- ☆16Mar 25, 2024Updated 2 years ago
- Code for "FF-LOGO: Cross-Modality Registration with Feature Filtering and Local to Global Optimization"☆11Sep 14, 2023Updated 3 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Tool to read, write, and visualize CommonRoad scenarios and base for other tools from the CommonRoad framework.☆22Aug 6, 2026Updated last month
- This is a SLAM algorithm combining plane features and point features.☆12Apr 13, 2021Updated 5 years ago
- Adaptive Hypernetworks for Multi-Agent RL. NeurIPS 2025.☆25Apr 14, 2026Updated 5 months ago
- An environment based on JSBSIM aimed at one-to-one close air combat.☆20Sep 14, 2025Updated last year
- Code and additional information for our paper entitled 'Scene Augmentation Methods for Interactive Embodied AI Tasks'☆11Apr 25, 2023Updated 3 years ago
- About iSatCR aims at Joint Optimization of Computing and Routing in LEO Satellite Constellations with Distributed Deep Reinforcement Lear…☆24Mar 25, 2026Updated 6 months ago
- ☆27Aug 24, 2023Updated 3 years ago