Heuristic Reinforcement Learning
☆11Aug 9, 2018Updated 7 years ago
Alternatives and similar repositories for HRL
Users that are interested in HRL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Dueling Double Deep Q-Network TensorFlow + TFLearn implementation☆11Mar 9, 2017Updated 9 years ago
- Research on Inverse Reinforcement Learning for self driving vehicles at UCLA☆13Nov 7, 2018Updated 7 years ago
- communication simulation (Communication Signal Processing LAB)☆15Mar 22, 2018Updated 8 years ago
- Coverage path planning with Reinforcement Learning☆11Mar 29, 2022Updated 4 years ago
- DQN for 5G RAN Slicing☆15May 28, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆34May 4, 2020Updated 6 years ago
- Dynamic mode decomposition in Python☆13Jun 9, 2015Updated 11 years ago
- Energy-Efficient Power and Subcarrier Allocation for OFDMA Systems with Value Function Approximation Approach. EI paper from march to sep…☆13Mar 13, 2017Updated 9 years ago
- Cross-platform libraries for the EM7180 Ultimate Sensor Fusion Solution☆31Aug 27, 2024Updated last year
- 利用python强大的可视化,加深对一些机器人运动规划算法的理解☆13Sep 28, 2019Updated 6 years ago
- We are developing a time series forecasting model using reinforcement learning, based on OneNet, for stock market data prediction.☆10Apr 19, 2024Updated 2 years ago
- ☆18Oct 6, 2021Updated 4 years ago
- 面试总结☆26Oct 20, 2021Updated 4 years ago
- Link to paper: https://www.ssrn.com/abstract=3804655☆14Jul 27, 2021Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Open source community's implementation of the model from "LANGUAGE MODEL BEATS DIFFUSION — TOKENIZER IS KEY TO VISUAL GENERATION"☆15Nov 11, 2024Updated last year
- add a Arg: label_smoothing for torch.nn.CrossEntropyLoss()☆14Jan 13, 2021Updated 5 years ago
- 九度题目代码,包含Java C++☆15Sep 23, 2017Updated 8 years ago
- A package containing neural network architectures based on Variational Autoencoders (VAE) and Restricted Boltzmann Machines (RBM) for lea…☆12Dec 23, 2021Updated 4 years ago
- pytorch☆14Dec 11, 2020Updated 5 years ago
- ☆12May 14, 2021Updated 5 years ago
- Convolutional Variational Autoencoder☆10Sep 7, 2018Updated 7 years ago
- Subjective consensus algorithm and value governance protocol 🌎☆16Dec 21, 2022Updated 3 years ago
- Label smoothed Aggregation cross entropy loss for generalisation in sequence to sequence tasks.☆14Dec 17, 2019Updated 6 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Using WoLF (win or learn fast) PHC (policy hill climbing) algorithm to implement stochastic games☆15Jun 14, 2019Updated 7 years ago
- Transformer for text summarization implemented in pytorch☆11Aug 17, 2019Updated 6 years ago
- This project implements a functional motion planning stack for autonomous vehicles to avoid both static and dynamic obstacles while track…☆53Aug 22, 2019Updated 6 years ago
- Reinforcement Learning Algorithms Based on PyTorch☆21Apr 28, 2022Updated 4 years ago
- [NeurIPS 2024] Official Implementation of "SDformer: Similarity-driven Discrete Transformer For Time Series Generation"☆17May 23, 2025Updated last year
- Bayesian Reward Shaping Framework for Deep Reinforcement Learning☆26Mar 29, 2019Updated 7 years ago
- ☆11Apr 2, 2021Updated 5 years ago
- ☆12Jul 6, 2023Updated 3 years ago
- ☆13Aug 17, 2020Updated 5 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Code for the paper On-Policy vs. Off-Policy Deep Reinforcement Learning for Resource Allocation in Open Radio Access Network☆34Jan 22, 2022Updated 4 years ago
- ☆12Oct 10, 2021Updated 4 years ago
- This project aims to foster the usage of Transformers as generative models to sample artificial multivariate time series. Thereby, the la…☆16Feb 20, 2025Updated last year
- Exploring the Dyna-Q reinforcement learning algorithm☆17Feb 27, 2018Updated 8 years ago
- This paper has been accepted by IEEE MASS 2022.☆15Oct 9, 2023Updated 2 years ago
- Course project of SJTU EE357: Computer Network, advised by Prof. Na Ruan. We implemented and improved "A Hierarchical Framework of Cloud …☆38Jun 26, 2019Updated 7 years ago
- Community Implementation of *Temporal Latent Auto-Encoder* as described in [Temporal Latent Auto-Encoder: A Method for Probabilistic Mult…☆15Jun 9, 2022Updated 4 years ago