Tabular methods for reinforcement learning
☆39Jul 3, 2020Updated 6 years ago
Alternatives and similar repositories for tabular-methods
Users that are interested in tabular-methods are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A videogame made with PyGame turned into an Open AI Gym Learning Environment for Reinforcement Learning agents.☆14Jan 3, 2023Updated 3 years ago
- Modular Single-file Reinfocement Learning Algorithms Library☆38May 16, 2023Updated 3 years ago
- D3PE (Deep Data-Driven Policy Evaluation) aims to evaluation a large set of candidate policies from a fixed dataset to select best ones.☆10Jun 2, 2022Updated 4 years ago
- Accelerated replay buffers in JAX☆47Sep 17, 2022Updated 4 years ago
- Code for building and experimenting on saliency maps for RL agents.☆12Feb 13, 2020Updated 6 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Autoregressive transformer in JAX from scratch☆23Jan 28, 2022Updated 4 years ago
- Code for Powderworld: A Platform for Understanding Generalization via Rich Task Distributions☆74Aug 31, 2024Updated 2 years ago
- Implementation of NeurIPS2021 paper <On Effective Scheduling of Model-based Reinforcement Learning>☆13Nov 16, 2021Updated 4 years ago
- JAX implementations of various deep reinforcement learning algorithms.☆25Feb 2, 2025Updated last year
- Core interface to design, solve, and simulate trajectory games.☆21Jul 3, 2026Updated 2 months ago
- JAX implementation of RL algorithms and vectorized environments☆50Dec 26, 2023Updated 2 years ago
- A Julia interface to the PATH solver☆14Jan 26, 2021Updated 5 years ago
- Opinionated library for managing hyperparameters and mutable state of machine learning training systems.☆19Aug 4, 2023Updated 3 years ago
- Minimal C++ implementation of GPT2☆41Jul 1, 2023Updated 3 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- A number of agents (PPO, MuZero) with a Perceiver-based NN architecture that can be trained to achieve goals in nethack/minihack environm…☆44Sep 19, 2022Updated 4 years ago
- [deprecated] Engine Agnostic Gym Environment for Robotics☆17Feb 10, 2022Updated 4 years ago
- A tiny reinforcement learning codebase for continuous control, built on top of JAX.☆15Mar 28, 2023Updated 3 years ago
- [AAAI 2024 (Oral)] Safety-MuJoCo Environments.☆12Jun 4, 2024Updated 2 years ago
- Official implementation of CoNSAL for analytical Lyapunov function discovery☆12Jun 26, 2024Updated 2 years ago
- The code for experiments conducted to verify the correctness of mirror learning.☆11Jun 3, 2022Updated 4 years ago
- A DP beam-search extension of Mitchell Stern's span-based neural constituency parser☆11Aug 24, 2022Updated 4 years ago
- Multi-agent coordination using game theory and nonlinear opinion dynamics - CDC 2023☆15Nov 29, 2023Updated 2 years ago
- Representation Learning (RepL) Methods in Reinforcement Learning and Causal Inference☆32Nov 24, 2025Updated 9 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆15Apr 5, 2023Updated 3 years ago
- Nature and nurture combined adapt faster than either alone. Darwinian evolution (agents selected across generations) plus Multi-Agent Dee…☆26Updated this week
- Implementation of Attentive Multi Task Deep Reinforcement Learning Architecture in Tensorflow☆15Apr 5, 2019Updated 7 years ago
- ☆15Nov 22, 2019Updated 6 years ago
- Reviving Any-Order Autoregressive Models via Principled Parallel Sampling and Speculative Decoding☆16Nov 16, 2025Updated 10 months ago
- ☆10Apr 13, 2023Updated 3 years ago
- Auto Differentiate from scratch based on Autograd☆11Jun 21, 2022Updated 4 years ago
- Logging library for JAX that is compatible with transformations and primitives such as vmap and scan.☆16Sep 1, 2026Updated 2 weeks ago
- Semi-Supervised Offline Reinforcement Learning with Action-Free Trajectories☆42Jul 16, 2023Updated 3 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- A 4-hour long tutorial session for learning to use LLMs and align them with custom data. We will also train a custom LLM.☆17Sep 12, 2024Updated 2 years ago
- Python Laboratory for Dislocation Dynamics☆13Jun 8, 2026Updated 3 months ago
- Standalone library of frequently-used wrappers for dm_env environments.☆19Jul 9, 2024Updated 2 years ago
- Project exploring Multi Task Deep Reinforcement Learning neural network architectures and algorithms with Open AI Gym and TensorFlow☆17Sep 5, 2018Updated 8 years ago
- ☆12Aug 26, 2025Updated last year
- A modular implementation of PPO, and soon hopefully other algorithms.☆27Jan 16, 2024Updated 2 years ago
- ☆15Oct 16, 2020Updated 5 years ago