Average-Reward Reinforcement Learning with Trust Region Methods
☆11Oct 17, 2022Updated 3 years ago
Alternatives and similar repositories for apo
Users that are interested in apo are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A collection of heat engines, based on the OpenAI Gym environment framework for use with reinforcement learning applications.☆15Dec 20, 2021Updated 4 years ago
- many powerful tools for studying irreducible representations of SU(n), including making animations of hadron flavor-state multiplets☆13Jun 17, 2026Updated last month
- Optim4RL is a Jax framework of learning to optimize for reinforcement learning.☆28Nov 27, 2024Updated last year
- This is a quadruped simulated on pybullet physics engine, walking using trot and bound mechanisms☆16Feb 24, 2024Updated 2 years ago
- Controlled Online Optimization Learning (COOL): Finding the Ground State of Spin Hamiltonians with Reinforcement Learning (arXiv:2003.000…☆13Jun 18, 2020Updated 6 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A model predictive control based voltage source inverter☆11Jan 11, 2020Updated 6 years ago
- Code for analyzing the relationship between curvature and migration rate in meandering rivers☆14Dec 8, 2022Updated 3 years ago
- Repository for "Revisiting Non-Acyclic GFlowNets in Discrete Environments" (ICML 2025)☆14Oct 8, 2025Updated 9 months ago
- ☆15Oct 20, 2025Updated 9 months ago
- Python implementation of algorithms for multi-objective multi-agent path finding.☆13May 17, 2022Updated 4 years ago
- Diffusion Probabilistic Model in Jax☆13Apr 20, 2024Updated 2 years ago
- An open-source Reinforcement Learning (RL) harness written in Python to work with SimFire for training agents to fight wildfires on real …☆18Oct 8, 2024Updated last year
- Versions of hybrid pso algorithms for engineering optimization☆10Dec 21, 2017Updated 8 years ago
- Code for AAAI 2023 paper "Hypernetworks for Zero-shot Transfer in Reinforcement Learning"☆24Apr 26, 2023Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- This submission demonstrates modeling and simulation of a Two-Zone MVDC electric ship in Simscape Electrical, and considers modeling con…☆10Mar 23, 2026Updated 3 months ago
- An interactive visualization of convex duality (Fenchel conjugate)☆22Dec 2, 2020Updated 5 years ago
- Distributional Soft Actor Critic☆63Jun 6, 2020Updated 6 years ago
- ☆10Jan 22, 2023Updated 3 years ago
- ☆11Jan 20, 2023Updated 3 years ago
- paper <<Hierarchical Deep Reinforcement Learning: Integrating Temporal Abstraction and Intrinsic Motivation>> python implementation☆10Mar 27, 2018Updated 8 years ago
- Vision-Based Navigation for Auto-Docking☆13Apr 21, 2021Updated 5 years ago
- Solving the CVRPTW with geatpy2☆11Mar 24, 2020Updated 6 years ago
- ☆15Sep 4, 2025Updated 10 months ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- [NeurIPS 2020 Spotlight] State-adversarial PPO for robust deep reinforcement learning☆32Nov 18, 2021Updated 4 years ago
- Pose-disentangled Contrastive Learning☆14Jan 27, 2024Updated 2 years ago
- ☆18Jun 10, 2022Updated 4 years ago
- The goal of this design is to use the PYNQ-Z2 development board to design a general convolution neural network accelerator. And through r…☆11Sep 30, 2020Updated 5 years ago
- Code for "LifeLong Incremental Reinforcement Learning (LLIRL)"☆21Jan 28, 2021Updated 5 years ago
- source code for VESAEA, paper in CEC2019☆10Mar 27, 2019Updated 7 years ago
- Deep Reinforcement Learning - Implementations and Theory: A path to mastery☆13Nov 21, 2021Updated 4 years ago
- ☆15Nov 20, 2025Updated 8 months ago
- Thompson Sampling for Bandits using UCB policy☆10Jul 29, 2017Updated 8 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆17Oct 2, 2021Updated 4 years ago
- A collection of Fault Diagnosis python codes☆10Mar 13, 2022Updated 4 years ago
- Optimal solution of the Generalized Dubins Interval Problem (GDIP)☆23Oct 24, 2022Updated 3 years ago
- Webots simulation environment and a vision-based autonomous docking algorithm for robotic vessels with a novel latching system.☆16Oct 8, 2024Updated last year
- Everything about Transfer Learning and Domain Adaptation--迁移学习☆10Jun 5, 2019Updated 7 years ago
- A solution for multicopter vibration measurement, vibration isolator design and digital filter design☆20Jan 23, 2021Updated 5 years ago
- 【个人源码】Matlab环境下的改进PlatEMO☆12Apr 1, 2019Updated 7 years ago