PyTorch implementation of the discrete Soft-Actor-Critic algorithm.
☆57Oct 1, 2021Updated 4 years ago
Alternatives and similar repositories for SAC_discrete
Users that are interested in SAC_discrete are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆40Nov 17, 2021Updated 4 years ago
- PyTorch implementation of SAC-Discrete.☆316Jul 25, 2024Updated 2 years ago
- PyTorch implementation of discrete version of Soft Actor-Critic.☆37Sep 19, 2021Updated 4 years ago
- Single-file pytorch implementation of hybrid-SAC☆68Jun 25, 2021Updated 5 years ago
- Re-produce DQN, REINFORCE, REINFORCE with baseline, one-step AC, QAC, QAC with shared network, PPO2, DDPG, TD3, SAC, SAC discrete,A2C,A3C☆21Jul 27, 2020Updated 6 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A clean and robust Pytorch implementation of SAC on discrete action space☆43Oct 23, 2024Updated last year
- PyTorch implementation of the Offline Reinforcement Learning algorithm CQL. Includes the versions DQN-CQL and SAC-CQL for discrete and co…☆147May 6, 2024Updated 2 years ago
- Jax and Torch Multi-Agent SAC on PettingZoo API☆101Nov 23, 2024Updated last year
- Implementation of ICLR 2025 paper "Q-Adapter: Customizing Pre-trained LLMs to New Preferences with Forgetting Mitigation"☆18Oct 5, 2024Updated last year
- PyTorch implementations for Offline Preference-Based RL (PbRL) algorithms☆21Mar 24, 2025Updated last year
- Color: Train a Real-world Local Path Planner in One Hour via Partially Decoupled Reinforcement Learning and Vectorized Diversity☆23Dec 23, 2024Updated last year
- Minimal RLHF implementation built on top of minGPT.☆32Jul 4, 2024Updated 2 years ago
- ☆10Sep 19, 2023Updated 2 years ago
- Reference code for the paper ""Centroid-Guided Target-Driven Topology Control Method for UAV Ad-Hoc Networks Based on Tiny Deep Reinforce…☆13Oct 21, 2024Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Pytorch Implementation of AAMAS 2021 paper <Energy-Based Imitation Learning>☆12Oct 8, 2021Updated 4 years ago
- PyTorch implementation of Constrained Reinforcement Learning for Soft Actor Critic Algorithm☆63Jul 11, 2022Updated 4 years ago
- PyTorch implementation of Soft Actor-Critic (SAC), Twin Delayed DDPG (TD3), Actor-Critic (AC/A2C), Proximal Policy Optimization (PPO), QT…☆18Oct 18, 2022Updated 3 years ago
- Implement some algorithms of RL☆46Mar 28, 2023Updated 3 years ago
- Actor-Sharer-Learner training framework for off-policy DRL algorithms☆22Dec 29, 2024Updated last year
- Model predictive control–based value estimation for efficient reinforcement learning. This repository contains the code for the implement…☆12Nov 6, 2024Updated last year
- Random parameter environments using gym 0.7.4 and mujoco-py 0.5.7☆20Feb 14, 2019Updated 7 years ago
- Curiosity-driven Exploration by Self-supervised Prediction☆148Mar 12, 2023Updated 3 years ago
- "Adaptive Cruise Control for a Hybrid Vehicle with Deep Policy Gradients". Final project for ECE 517/414 Reinforcement Learning.☆13Dec 8, 2021Updated 4 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- D3QN Pytorch☆72Dec 13, 2021Updated 4 years ago
- A Reinforcement Learning Friendly Simulator for Mobile Robot☆17Jan 5, 2025Updated last year
- Various explorations into the game of Poker using MCTS, NFSP, and image-recognition/web-scraping☆13Oct 23, 2020Updated 5 years ago
- Code for Adapting Environment Sudden Changes by Learning Context Sensitive Policy☆21Jun 1, 2022Updated 4 years ago
- Multi Agent SAC and DDPG applied to path finding in a 3-dimensional grid☆15Aug 8, 2021Updated 5 years ago
- Reproducing several bandwidth-based traffic signal coordination models (including MaxBand, MultiBand, etc.)☆12Sep 18, 2020Updated 5 years ago
- Deep recurrent Q learning on CartPole-v1 environment☆96Jan 15, 2024Updated 2 years ago
- Flatland Multi Agent Reinforcement Learning☆16Aug 1, 2020Updated 6 years ago
- ☆19Apr 22, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- code for☆11Apr 10, 2021Updated 5 years ago
- PyTorch implementation of Soft Actor-Critic (SAC), Twin Delayed DDPG (TD3), Actor-Critic (AC/A2C), Proximal Policy Optimization (PPO), QT…☆1,360Mar 13, 2025Updated last year
- A novel preference-driven multi-objective reinforcement learning algorithm using a single policy network that covers the entire preferenc…☆44Nov 15, 2023Updated 2 years ago
- Study to test if Volume leak index (VLI) is a marker of severity of illness in sepsis.☆14Sep 29, 2022Updated 3 years ago
- A python module designed for agile RL algorithm developing.☆26Jul 11, 2024Updated 2 years ago
- Official code release for ICLR23 "Diminishing Return of Value Expansion Methods in Model-Based Reinforcement Learning"☆16Mar 8, 2023Updated 3 years ago
- This repository contains the R code used analyse the eICU and MIMIC-III databases for the Sarkar et al paper "Performance of intensive ca…☆10Nov 27, 2020Updated 5 years ago