PyTorch implementation of discrete version of Soft Actor-Critic.
☆37Sep 19, 2021Updated 5 years ago
Alternatives and similar repositories for Discrete-SAC-PyTorch
Users that are interested in Discrete-SAC-PyTorch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- PyTorch implementation of SAC-Discrete.☆317Jul 25, 2024Updated 2 years ago
- PyTorch implementation of the discrete Soft-Actor-Critic algorithm.☆57Oct 1, 2021Updated 4 years ago
- Code and results of the academic publication "Blockchain-enabled Network Sharing for O-RAN"☆11Jan 10, 2022Updated 4 years ago
- ☆10Sep 21, 2020Updated 6 years ago
- Actor-Sharer-Learner training framework for off-policy DRL algorithms☆22Dec 29, 2024Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Re-produce DQN, REINFORCE, REINFORCE with baseline, one-step AC, QAC, QAC with shared network, PPO2, DDPG, TD3, SAC, SAC discrete,A2C,A3C☆21Jul 27, 2020Updated 6 years ago
- ☆40Nov 17, 2021Updated 4 years ago
- Color: Train a Real-world Local Path Planner in One Hour via Partially Decoupled Reinforcement Learning and Vectorized Diversity☆23Dec 23, 2024Updated last year
- ☆45Dec 8, 2021Updated 4 years ago
- A Framework for Safe and Accelerated Reinforcement Learning-based Radio Resource Management☆20Oct 1, 2022Updated 3 years ago
- Single-Life Reinforcement Learning☆14Dec 17, 2022Updated 3 years ago
- ☆10May 10, 2023Updated 3 years ago
- GitHub for the article Deep Reinforcement Learning for URLLC data management on top of scheduled eMBB traffic (Fabio Saggese, Luca Pasqua…☆17Feb 18, 2021Updated 5 years ago
- PyTorch implementation of Distribution Correction(DisCor) based on Soft Actor-Critic.☆39Jun 22, 2022Updated 4 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- STL源码剖析学习笔记☆11Jun 28, 2022Updated 4 years ago
- ☆10Mar 22, 2021Updated 5 years ago
- An easy to understand implementation of the paper "Model-Based Reinforcement Learning for Atari"☆18Sep 27, 2019Updated 7 years ago
- Code Release for Task Agnostic Dynamics Priors for Deep Reinforcement Learning☆12Jun 13, 2019Updated 7 years ago
- ☆23Oct 1, 2020Updated 5 years ago
- A pytorch implementation of Constrained Reinforcement Learning Algorithm, including Constrained Soft Actor Critic (Soft Actor Critic Lagr…☆49May 30, 2023Updated 3 years ago
- ☆13Mar 14, 2023Updated 3 years ago
- ☆37Oct 10, 2025Updated 11 months ago
- Simulate one server for one user, use PPO.☆15Nov 21, 2021Updated 4 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- PyTorch implementation of the Offline Reinforcement Learning algorithm CQL. Includes the versions DQN-CQL and SAC-CQL for discrete and co…☆148May 6, 2024Updated 2 years ago
- [ISSN 0168-1699, COMPUT ELECTRON AGR 2024] FastSegFormer: A knowledge distillation-based method for real-time semantic segmentation of su…☆22Feb 6, 2024Updated 2 years ago
- Blockchain federated learning simulate by python☆10May 29, 2023Updated 3 years ago
- Implementation of the Discrete Soft Actor-Critic algorithm with RNN policy in PyTorch☆26Jan 7, 2023Updated 3 years ago
- In general, the way we think about handling multi-modal uncertainty is by maintaining some beliefs about how probable each potential mode…☆20Dec 10, 2019Updated 6 years ago
- PyTorch implementation of R2D2 (Recurrent Reply Distributed DQN)☆13Nov 14, 2019Updated 6 years ago
- PyTorch implementation of PtrNet to solve sorting problem.☆12Dec 19, 2017Updated 8 years ago
- Simulation code for “Joint Power Control and LSFD for Wireless-Powered Cell-Free Massive MIMO,” by Özlem Tuğfe Demir and Emil Björnson, I…☆43Feb 3, 2024Updated 2 years ago
- Simulation code for "Energy Efficiency Maximization in Large-Scale Cell-Free Massive MIMO: A Projected Gradient Approach," IEEE Trans. Wi…☆19Oct 2, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- An implementation of the paper ‘Channel Distribution Learning: Model-Driven GAN-Based Channel Modeling for IRS-Aided Wireless Communicati…☆16Oct 27, 2022Updated 3 years ago
- ☆11Dec 3, 2020Updated 5 years ago
- A blockchain-enable decentralized federated learning implement☆12Jul 10, 2023Updated 3 years ago
- This is an attempt to summarize feature engineering methods that I have learned over the course of my graduate school.☆11Mar 3, 2022Updated 4 years ago
- DSAC; Distributional Soft Actor-Critic☆141Feb 12, 2025Updated last year
- ☆13Sep 1, 2023Updated 3 years ago
- This repo implements Deep Q-Network (DQN) for solving the Frozenlake-v1 environment of the Gymnasium library using Python 3.8 and PyTorch…☆20Mar 19, 2024Updated 2 years ago