PyTorch implementation of soft actor critic
☆943Jul 17, 2025Updated last year
Alternatives and similar repositories for pytorch-soft-actor-critic
Users that are interested in pytorch-soft-actor-critic are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Soft Actor-Critic☆1,292Nov 29, 2023Updated 2 years ago
- PyTorch implementation of Soft Actor-Critic (SAC)☆599Dec 5, 2021Updated 4 years ago
- Collection of reinforcement learning algorithms☆2,922Jun 17, 2024Updated 2 years ago
- Author's PyTorch implementation of TD3 for OpenAI gym tasks☆2,098Jul 14, 2023Updated 3 years ago
- Softlearning is a reinforcement learning framework for training maximum entropy policies in continuous domains. Includes the official imp…☆1,434Nov 29, 2023Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- PyTorch implementation of Soft-Actor-Critic and Prioritized Experience Replay (PER) + Emphasizing Recent Experience (ERE) + Munchausen RL…☆297Feb 24, 2021Updated 5 years ago
- PyTorch implementation of Soft Actor-Critic (SAC), Twin Delayed DDPG (TD3), Actor-Critic (AC/A2C), Proximal Policy Optimization (PPO), QT…☆1,356Mar 13, 2025Updated last year
- PyTorch implementation of Soft Actor-Critic(SAC).☆107Jun 9, 2020Updated 6 years ago
- PyTorch implementation of DQN, AC, ACER, A2C, A3C, PG, DDPG, TRPO, PPO, SAC, TD3 and ....☆4,646Mar 24, 2023Updated 3 years ago
- PyTorch implementation of Deep Reinforcement Learning: Policy Gradient methods (TRPO, PPO, A2C) and Generative Adversarial Imitation Lear…☆1,285Feb 9, 2021Updated 5 years ago
- PyTorch implementation of Advantage Actor Critic (A2C), Proximal Policy Optimization (PPO), Scalable trust-region method for deep reinfor…☆3,905May 29, 2022Updated 4 years ago
- PyTorch implementations of deep reinforcement learning algorithms and environments☆5,935Jul 25, 2024Updated 2 years ago
- PyTorch implementation of Soft Actor-Critic + Autoencoder(SAC+AE)☆257May 3, 2020Updated 6 years ago
- Code for conservative Q-learning☆486Dec 7, 2021Updated 4 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- PyTorch implementation of SAC-Discrete.☆316Jul 25, 2024Updated 2 years ago
- A collection of reference environments for offline reinforcement learning☆1,694Nov 18, 2024Updated last year
- Minimal implementation of clipped objective Proximal Policy Optimization (PPO) in PyTorch☆2,368Jul 9, 2024Updated 2 years ago
- Author's PyTorch implementation of TD3+BC, a simple variant of TD3 for offline RL☆410Dec 18, 2021Updated 4 years ago
- Deep Planning Network: Control from pixels by latent planning with learned dynamics☆379Oct 15, 2021Updated 4 years ago
- Code for the paper "When to Trust Your Model: Model-Based Policy Optimization"☆558Nov 22, 2022Updated 3 years ago
- Implementation of Efficient Off-policy Meta-learning via Probabilistic Context Variables (PEARL)☆512Dec 1, 2022Updated 3 years ago
- Reinforcement Learning in PyTorch☆2,278Jan 4, 2021Updated 5 years ago
- Author's PyTorch implementation of BCQ for continuous and discrete actions☆667Apr 6, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- CURL: Contrastive Unsupervised Representation Learning for Sample-Efficient Reinforcement Learning☆605Oct 28, 2020Updated 5 years ago
- A pytorch reprelication of the model-based reinforcement learning algorithm MBPO☆189Apr 12, 2022Updated 4 years ago
- PyTorch implementation of Trust Region Policy Optimization☆448Sep 13, 2018Updated 7 years ago
- Library for Model Based RL☆1,065Jul 12, 2024Updated 2 years ago
- Code for "Actor-Attention-Critic for Multi-Agent Reinforcement Learning" ICML 2019☆807May 29, 2022Updated 4 years ago
- PyTorch version of Stable Baselines, reliable implementations of reinforcement learning algorithms.☆13,624Updated this week
- Collections of robotics environments geared towards benchmarking multi-task and meta reinforcement learning☆1,863Jul 19, 2026Updated last week
- Reinforcement learning algorithms for MuJoCo tasks☆467Jul 20, 2026Updated last week
- Implicit Normalizing Flows + Reinforcement Learning☆62May 31, 2019Updated 7 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- An elegant PyTorch deep reinforcement learning library.☆10,892Apr 3, 2026Updated 3 months ago
- ☆399Jul 18, 2019Updated 7 years ago
- Stochastic Latent Actor-Critic: Deep Reinforcement Learning with a Latent Variable Model☆154Oct 26, 2020Updated 5 years ago
- Experiment code for "Deep Reinforcement Learning in a Handful of Trials using Probabilistic Dynamics Models"☆475Jul 6, 2023Updated 3 years ago
- Prioritized Experience Replay (PER) implementation in PyTorch☆360Feb 3, 2020Updated 6 years ago
- PyTorch implementation of Asynchronous Advantage Actor Critic (A3C) from "Asynchronous Methods for Deep Reinforcement Learning".☆1,332Sep 25, 2019Updated 6 years ago
- Fault-tolerant, highly scalable GPU orchestration, and a machine learning framework designed for training models with billions to trillio…☆3,997May 25, 2024Updated 2 years ago