GAIL learning to imitate PPO playing CartPole.
☆13May 27, 2021Updated 5 years ago
Alternatives and similar repositories for PPO-GAIL-cartpole
Users that are interested in PPO-GAIL-cartpole are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- End-to-End Driving via Generative Adversarial Imitation Learning☆29Apr 4, 2023Updated 3 years ago
- Results reproductions & comparisons between OpenSpiel implementations, associated paper & originating works☆18Mar 2, 2021Updated 5 years ago
- ☆10Apr 23, 2021Updated 5 years ago
- multiagent-gail working with multiagent-particle-env-v2 (which was modified by magail authors)☆13Aug 17, 2019Updated 6 years ago
- A retro multiplayer shooter☆12Jul 2, 2020Updated 6 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Trains an agent with Twin Delayed Deep Deterministic Policy Gradient (TD3) to solve the Bipedal Walker challenge from OpenAI☆12Sep 22, 2023Updated 2 years ago
- An attempt to add bots to teeworlds-0.5.1 =)☆16Nov 27, 2013Updated 12 years ago
- ☆13Jan 22, 2025Updated last year
- Kuhn poker implemented in accordance to OpenAI gym interface☆14Dec 5, 2019Updated 6 years ago
- ☆12Jun 17, 2022Updated 4 years ago
- ☆16Nov 7, 2020Updated 5 years ago
- CartPole-v0 via PPO with GAE, PyTorch☆22Feb 10, 2019Updated 7 years ago
- Soccer Trajectory Prediction Competition☆16Aug 28, 2025Updated 11 months ago
- ☆24Sep 27, 2024Updated last year
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Lecture notes for a course on Decision and Game Theory for undergraduates studying AI☆13Dec 14, 2018Updated 7 years ago
- ☆22May 20, 2021Updated 5 years ago
- ☆17Dec 13, 2019Updated 6 years ago
- Collection of OpenAI parametrized action-space environments.☆70Mar 19, 2025Updated last year
- Actor-Sharer-Learner training framework for off-policy DRL algorithms☆22Dec 29, 2024Updated last year
- A simple implementation of Generative Adversarial Imitation Learning with PyTorch☆175Mar 22, 2022Updated 4 years ago
- ☆16Mar 9, 2019Updated 7 years ago
- Code for magnetic mirror descent.☆20Oct 5, 2023Updated 2 years ago
- Parses a document (scanned or phone captured) and returns the underlying question - answer layout structured capture by LayoutXLM model☆10Jun 14, 2021Updated 5 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- 算法工程师技术栈学习笔记☆15Aug 22, 2022Updated 3 years ago
- ☆21Dec 22, 2020Updated 5 years ago
- ☆10Jun 4, 2024Updated 2 years ago
- Generalised UDRL☆37May 12, 2022Updated 4 years ago
- fork of gitlab.com/bpaassen/five_clique. Solution to Matt Parker's 5-clique problem☆10Aug 4, 2022Updated 3 years ago
- Pytorch Implementation of AAMAS 2021 paper <Energy-Based Imitation Learning>☆12Oct 8, 2021Updated 4 years ago
- Scaling Population-Based Reinforcement Learning with GPU Accelerated Simulation☆13Nov 5, 2025Updated 8 months ago
- [ICLR 2022] "Bayesian Modeling and Uncertainty Quantification for Learning to Optimize: What, Why, and How" by Yuning You, Yue Cao, Tianl…☆14Aug 19, 2022Updated 3 years ago
- A dataset for hockey player tracking, following the same format as the MOT challenge dataset.☆24Oct 31, 2023Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- JAX/Haiku implementation of "Auction Learning as a Two-Player Game"☆11Jul 6, 2024Updated 2 years ago
- Actor-Critic and openAI clipped PPO in gym cartpole-v0 and pendulum-v0 environment☆27Aug 2, 2020Updated 5 years ago
- Code for the paper "Deep FTRL-ORW: An Efficient Deep Reinforcement Learning Algorithm for Solving Imperfect Information Extensive-Form Ga…☆11Dec 1, 2022Updated 3 years ago
- Official Code Release for Pipeline PSRO: A Scalable Approach for Finding Approximate Nash Equilibria in Large Games☆57Aug 30, 2024Updated last year
- ☆84Dec 4, 2018Updated 7 years ago
- Inclined Drone landing using deep reinforcement learning☆24Feb 10, 2022Updated 4 years ago
- Evidential Calibration☆11Mar 8, 2022Updated 4 years ago