Experiments with transformer based RL algorithms
☆22Nov 23, 2019Updated 6 years ago
Alternatives and similar repositories for Transformer-RL
Users that are interested in Transformer-RL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Experiments to train transformer network to master reinforcement learning environments.☆32Mar 14, 2021Updated 5 years ago
- Adaptive Attention Span for Reinforcement Learning☆136May 11, 2020Updated 6 years ago
- PyTorch implementation of R2D2 (Recurrent Replay Distributed DPG (not DQN))☆14Mar 22, 2019Updated 7 years ago
- A bipedal humanoid control system using a Physics-Informed Neural Network (PINN) and Reinforcement Learning (RL) for stability and manipu…☆13Mar 25, 2026Updated 5 months ago
- An implementation of the traffic simulation optimisation with reinforcement learning, with FLOW and SUMO.☆17Jan 15, 2021Updated 5 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Uncertainty on Asynchronous Time Event Prediction (Spotlight, Neurips 2019)☆20Oct 8, 2020Updated 5 years ago
- Implementing DQNClipped and DQNReg Algorithms☆10Mar 2, 2021Updated 5 years ago
- Benchmarking suite for dual arm manipulation☆11Aug 12, 2026Updated 3 weeks ago
- ☆32Dec 1, 2019Updated 6 years ago
- ☆23Dec 25, 2024Updated last year
- [Neurocomputing, 2023] Personalized Robotic Control via Constrained Multi-Objective Reinforcement Learning☆28Dec 25, 2023Updated 2 years ago
- Lightweight multi-agent PPO for IEEE field.☆15Mar 23, 2022Updated 4 years ago
- Convergent Policy Optimization for Safe Reinforcement Learning☆11Oct 26, 2019Updated 6 years ago
- This repository contains Python functions for predicting time series.☆15May 24, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆10Aug 16, 2022Updated 4 years ago
- NGSIM Driving RL/Imitation learning environment compatible with rllab☆13Feb 23, 2018Updated 8 years ago
- Transformer-based Multi-Agent Actor-Critic Framework☆46Jun 8, 2022Updated 4 years ago
- The CLI & python API for the well-known project gpt-academic.☆19Sep 22, 2024Updated last year
- Slay the Spire simulator using C++ with some reinforcement learning☆37Jul 25, 2020Updated 6 years ago
- SemiDefinite Programming Algorithm (SDPA) for Python☆12Jul 1, 2026Updated 2 months ago
- reproducible dev+test+production environments for java+javascript+clojure(script)☆13Feb 2, 2021Updated 5 years ago
- This is the unofficial implementation of LEMON (ICLR'2024).☆13Apr 14, 2024Updated 2 years ago
- Implementation of ``Actor-Critic Alignment for Offline-to-Online Reinforcement Learning''☆13Oct 12, 2023Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Within's Discord Bot☆18Apr 19, 2023Updated 3 years ago
- Run a Franka Panda simulation in MuJoCo with ROS☆15Jun 25, 2026Updated 2 months ago
- Scala Eventstore Client☆12Updated this week
- On-Policy Model-free Reinforcement Learning for simplified Blackjack (David Silver Assignement)☆11Nov 20, 2017Updated 8 years ago
- Work in progress save editor for Monster Hunter: World☆11Aug 15, 2018Updated 8 years ago
- Functional, tagless and lens-based, global state management. With scalajs-react fs2.Stream integrations.☆13Updated this week
- Keyed Semaphore Implementation☆11Jul 8, 2024Updated 2 years ago
- Demo/Hand-On: Sealed Secrets☆11Nov 21, 2019Updated 6 years ago
- Implementation of Deep Reinforcement Learning from Self-Play in Imperfect-Information Games (Heinrich and Silver, 2016)☆48Nov 30, 2018Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A fork of ns3 LTE module for reinforcement learning experiments☆13Feb 20, 2017Updated 9 years ago
- Simulating V2V and V2I connectivity in Matlab using car following, lane changing models and entry and exit ramps on a 4-Highways, 3-Ramps…☆22Jun 15, 2015Updated 11 years ago
- ☆15Oct 22, 2023Updated 2 years ago
- ☆16Apr 28, 2023Updated 3 years ago
- This repository contains the replication of the iGSM dataset generation process from the Physics of LLM paper by Zeyuan Zhu.☆17Sep 13, 2024Updated last year
- Lyrics for Spotify Android☆10Jun 17, 2019Updated 7 years ago
- 论文一体化写作神器(Python)☆17Apr 11, 2020Updated 6 years ago