Fast reinforcement learning 💨
☆29Jul 15, 2025Updated last year
Alternatives and similar repositories for flashrl
Users that are interested in flashrl are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆10Feb 20, 2024Updated 2 years ago
- Train an agent to play VizDoom with multi sensory inputs. Trained using sample factory☆14Jul 9, 2021Updated 5 years ago
- Modular Single-file Reinfocement Learning Algorithms Library☆38May 16, 2023Updated 3 years ago
- Reinforcement learning training framework for entity-gym environments.☆17Mar 18, 2024Updated 2 years ago
- Basic world models☆33Oct 30, 2025Updated 9 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Simple single file implementations of Reinforcement Learning algorithms in Julia☆24Feb 15, 2025Updated last year
- Creating fixed-length vectors to describe RL/GA policies☆20Oct 23, 2021Updated 4 years ago
- A TF2.0 implementation of RL baselines.☆10Sep 24, 2021Updated 4 years ago
- Collection of in-progress libraries for entity neural networks.☆29Jun 24, 2022Updated 4 years ago
- The NetHack Learning Environment☆137Jul 30, 2026Updated last week
- Curated list of Moroccans publishing in the most prestigious AI conferences☆11Jul 6, 2026Updated last month
- A collection of matrix games in JAX☆14Apr 13, 2026Updated 3 months ago
- An OSX print to pdf-file printer driver☆41Jul 8, 2020Updated 6 years ago
- ☆16Aug 7, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Simple single-file baselines for Q-Learning in pure-GPU setting☆244Nov 24, 2025Updated 8 months ago
- Official implementation of 'A Large-Scale Exploration of mu-Transfer' (CoRR 2024)☆31Jun 5, 2025Updated last year
- A programming project on automatic differentiation in OCaml☆11Dec 22, 2022Updated 3 years ago
- Reinforcement learning with RealAnt: an open-source low-cost quadruped☆43Jan 11, 2022Updated 4 years ago
- Flax (Jax) implementation of DeepSeek-R1-Distill-Qwen-1.5B with weights ported from Hugging Face.☆26Feb 20, 2025Updated last year
- ☆11Jul 12, 2021Updated 5 years ago
- 🪐 The Sebulba architecture to scale reinforcement learning on Cloud TPUs in JAX☆61Oct 23, 2023Updated 2 years ago
- Julia Implementation of the POMCP algorithm for solving POMDPs☆12Aug 6, 2021Updated 5 years ago
- High quality implementations of imitation and inverse reinforcement learning algorithms☆24Aug 19, 2025Updated 11 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- The source code for mastering the game of Chutes and Ladders☆19Apr 2, 2021Updated 5 years ago
- Hands-on tutorial about Meta RL and GP-MPC at the RL4AA'24 workshop.☆15Apr 20, 2026Updated 3 months ago
- streaming deep reinforcement learning but 4x faster with jax!☆19Jan 4, 2026Updated 7 months ago
- ☆21Jul 14, 2020Updated 6 years ago
- Repository for Iterated Relearning: The Impact of Non-stationarity on Generalisation in Deep Reinforcement Learning☆11Jun 8, 2020Updated 6 years ago
- GCRL in JAX. Official repository for LEO (ICML 2026).☆29Jun 20, 2026Updated last month
- Implementation of Interpretable and Editable Programmatic Tree Policies for Reinforcement Learning (Kohler, Delfosse, et. al. 2024).☆15Sep 10, 2024Updated last year
- A LLM-friendly framework for translating dynamical equations to gymnasium-compatible RL environments.☆33Mar 18, 2026Updated 4 months ago
- Re-implementations of SOTA RL algorithms.☆137Sep 7, 2023Updated 2 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- NeurIPS 2019: DQN(λ) = Deep Q-Network + λ-returns.☆25May 20, 2024Updated 2 years ago
- Implementing scalable LLMs in pure JAX (no third-party libraries)☆54Jun 11, 2026Updated last month
- A curated list of awesome projects applying reinforcement learning (RL) to building control.☆18Nov 11, 2022Updated 3 years ago
- ☆12Jul 18, 2024Updated 2 years ago
- paper link: https://jcheminf.biomedcentral.com/articles/10.1186/s13321-022-00643-2☆22Sep 27, 2022Updated 3 years ago
- Bayes-Adaptive Monte-Carlo Planning algorithm☆19Mar 5, 2013Updated 13 years ago
- This repository contains code and data of the paper **On the Limitations of Continual Learning for Malware Classification**, accepted to …☆20Dec 29, 2023Updated 2 years ago