soft q learning and soft actor critic
☆16Dec 23, 2018Updated 7 years ago
Alternatives and similar repositories for Entropy-Regularized-RL
Users that are interested in Entropy-Regularized-RL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- DQN with freezing target network in tensorflow on pygame FlappyBird☆11Dec 19, 2018Updated 7 years ago
- A small library intended for controlling KUKA robots using KRC4 over KUKA RSI (Robot Sensor Interface) from Simulink.☆11Mar 30, 2017Updated 9 years ago
- Modified versions of the Soft Actor-Critic algorithm for Atari games from https://github.com/ac-93/soft-actor-critic.☆20May 18, 2020Updated 6 years ago
- Javakuka is an open-source project for creating Kuka Robot Language (KRL) codes in Java.☆10May 15, 2024Updated 2 years ago
- Official pytorch implementation for our ICLR 2023 paper "Latent State Marginalization as a Low-cost Approach for Improving Exploration".☆24Feb 9, 2023Updated 3 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Dataset collection and training code for "Ask Your Humans: Using Human Instructions to Improve Generalization in Reinforcement Learning"☆11Apr 8, 2025Updated last year
- TF2 Implementation of the Soft Actor-Critic Algorithm☆43Dec 8, 2022Updated 3 years ago
- ☆16Sep 25, 2020Updated 5 years ago
- The continuous mountain car problem solved with DDPG☆13Apr 19, 2020Updated 6 years ago
- Jaxplorer is a Jax reinforcement learning (RL) framework for exploring new ideas.☆12Jul 19, 2024Updated 2 years ago
- Quadratic Programming for Continuous Control of Safety-Critical Multi-Agent Systems Under Uncertainty☆14Sep 7, 2024Updated last year
- Represented Value Function Approach for Large Scale Multi Agent Reinforcement Learning☆17Mar 11, 2020Updated 6 years ago
- Waste Sorting with Robot Arm Tossing☆10Sep 19, 2023Updated 2 years ago
- Actor Prioritized Experience Replay☆19Nov 20, 2023Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Reinforcement Learning with Deep Energy-Based Policies☆438Nov 28, 2023Updated 2 years ago
- Designing an optimized path for multiple robots in a warehouse for picking and delivery operations using A* algorithm (shortest path) and…☆11Jul 28, 2023Updated 3 years ago
- Pytorch implementation of large network design in continous control RL.☆19Jan 5, 2022Updated 4 years ago
- PyTorch implementation of D4PG with the SOTA IQN Critic instead of C51. Implementation includes also the extensions Munchausen RL and D2R…☆25Apr 7, 2021Updated 5 years ago
- discrete soft Q learning(SQL) and soft Q imitation learning(SQIL) implementation in pytorch, simple!☆57Oct 18, 2022Updated 3 years ago
- ☆20Feb 18, 2022Updated 4 years ago
- Aerial Combat environment build around PyFlyt☆12Aug 12, 2023Updated 3 years ago
- Code for paper "Learning to Guide: Guidance Law Based on Deep Meta-learning and Model Predictive Path Integral Control"☆11May 26, 2019Updated 7 years ago
- Implementation of CoDAIL in the ICLR 2020 paper <Multi-Agent Interactions Modeling with Correlated Policies>☆19Jun 17, 2021Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A PyTorch implementation of SVGD (Stein Variational Gradient Descent), contains all examples including bayesian inference in the paper☆12Jul 30, 2020Updated 6 years ago
- Codes for the paper "SIDE: State Inference for Partially Observable Cooperative Multi-Agent Reinforcement Learning"☆11Jun 24, 2022Updated 4 years ago
- reinforcement learning, navigation, unitree, velodyne, slam, collision avoidance☆30Jul 8, 2025Updated last year
- PyTorch implementation of Soft Actor-Critic (SAC), Twin Delayed DDPG (TD3), Actor-Critic (AC/A2C), Proximal Policy Optimization (PPO), QT…☆16Nov 18, 2020Updated 5 years ago
- The Ranking Cost algorithm for multi-path routing of gridworld.(多智能体路径规划,电路规划)☆19Dec 20, 2021Updated 4 years ago
- Deep RL agents with PyTorch☆35Sep 25, 2021Updated 4 years ago
- A python implementation of PROCLUS: PROjected CLUStering algorithm.☆10Jan 12, 2015Updated 11 years ago
- Recommendation engine and it's algorithms in python , R .☆12Oct 26, 2018Updated 7 years ago
- A ROS2-based framework for TurtleBot3-like DRL autonomous navigation☆27Jun 16, 2023Updated 3 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Reinforcement Learning Seminar at the Chinese University of Hong Kong, Shenzhen, China.☆21Nov 17, 2023Updated 2 years ago
- ArXiv'18 implementation of amortized maximum likelihood (AML) for high-quality, weakly-supervised shape completion.☆11Nov 30, 2018Updated 7 years ago
- Codes reproducing paper Jongeun Choi, Songhwai Oh, Roberto Horowitz, Distributed learning and cooperative control for multi-agent systems…☆27Sep 14, 2021Updated 4 years ago
- Modified versions of the SAC algorithm from spinningup for discrete action spaces and image observations.☆100Jun 22, 2020Updated 6 years ago
- Offline trajectory planning for a swarm of UAVs using Boids model.☆21May 28, 2017Updated 9 years ago
- Yet Another Academic Homepage Template☆24Feb 25, 2026Updated 6 months ago
- Hands-On TensorBoard for PyTorch Developers, Published by Packt☆11Dec 15, 2025Updated 8 months ago