☆18Sep 7, 2023Updated 2 years ago
Alternatives and similar repositories for fast-rl-with-slow-updates
Users that are interested in fast-rl-with-slow-updates are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 用parl框架的DQN强化学习算法玩“合成大西瓜”☆14Mar 5, 2021Updated 5 years ago
- ☆11Oct 19, 2020Updated 5 years ago
- Implementing DQNClipped and DQNReg Algorithms☆10Mar 2, 2021Updated 5 years ago
- Source code for ICML 2023 paper "Competing for Shareable Arms in Multi-Player Multi-Armed Bandits"☆10May 14, 2024Updated 2 years ago
- ☆10Sep 21, 2020Updated 5 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Application of REINFORCE algorithm to downlink NOMA system☆13Jan 28, 2026Updated 6 months ago
- ☆12Jan 6, 2022Updated 4 years ago
- RL and DQN for easy JSP, FSP or others☆24Apr 23, 2022Updated 4 years ago
- MUX-PLMs: Pretraining LMs with Data Multiplexing☆15Jan 29, 2023Updated 3 years ago
- Code of Paper "Cooperative Sensing and Uploading for Quality-Cost Tradeoff of Digital Twins in VEC", IEEE TCE, 2024.☆13Jul 10, 2023Updated 3 years ago
- Multi-Agent Deep Recurrent Q-Learning with Bayesian epsilon-greedy on AirSim simulator☆13Apr 1, 2022Updated 4 years ago
- ☆15Sep 21, 2020Updated 5 years ago
- In this repository, we try to solve musculoskeletal tasks with `Double DQN reinforcement learning` by using a `transformer` model has bee…☆17Nov 7, 2023Updated 2 years ago
- ☆14Dec 9, 2021Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This is code of paper entitled "AI-based Radio Resource and Transmission Opportunity Allocation for 5G-V2X HetNets: NR and NR-U networks…☆16Sep 8, 2023Updated 2 years ago
- ☆12Feb 21, 2024Updated 2 years ago
- Verify MAPPO in task ‘simple_spread_v3‘☆15Aug 10, 2024Updated 2 years ago
- This repository contains all the projects, and necessary scripts and files developed for the anti-jamming project based on ns3-gym. You c…☆16Aug 14, 2023Updated 3 years ago
- ☆58Jan 20, 2023Updated 3 years ago
- An implementation of TRPO with GAE in PyTorch☆16Jul 22, 2023Updated 3 years ago
- This framework is for resource allocation in C-V2X Mode 4☆17Nov 14, 2025Updated 9 months ago
- ☆17May 13, 2024Updated 2 years ago
- ☆13Apr 11, 2022Updated 4 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆16May 20, 2025Updated last year
- Simulation code for “RIS-Assisted High-Speed Communications with Time-Varying Distance-Dependent Rician Channels,” by K. Wang, CT. Lam, a…☆14Sep 29, 2024Updated last year
- UAV offloading based on QMIX☆16Oct 12, 2023Updated 2 years ago
- A transformer-based deep RL trading bot built with PyTorch.☆13Jan 16, 2025Updated last year
- ☆10May 1, 2023Updated 3 years ago
- Efficient Exploration through Bayesian Deep-Q Networks.☆18Mar 22, 2022Updated 4 years ago
- ☆12Jun 1, 2026Updated 2 months ago
- Official PyTorch code for "Recurrent Off-policy Baselines for Memory-based Continuous Control" (DeepRL Workshop, NeurIPS 21)☆95Nov 21, 2023Updated 2 years ago
- Fuzzy Q-Learning Algorithm☆18Jun 3, 2023Updated 3 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- pytorch implementation of SAC, TD3 and TD7 with Mujoco Benchmark results from 4 seeds.☆15Jul 4, 2024Updated 2 years ago
- [ECCV 2022] "TALISMAN: Targeted Active Learning for Object Detection with Rare Classes and Slices using Submodular Mutual Information" by…☆11Sep 21, 2022Updated 3 years ago
- ☆10Sep 14, 2022Updated 3 years ago
- Tools for speech recognition☆11Jun 24, 2017Updated 9 years ago
- Universal Robot Description for Anki Robots☆18Feb 4, 2026Updated 6 months ago
- D3QN framework for distributed resource allocation☆19Jul 17, 2024Updated 2 years ago
- Code-base for the paper Spectral Normalisation for Deep Reinforcement Learning: An Optimisation Perspective.☆11Jun 26, 2021Updated 5 years ago