πͺ The Sebulba architecture to scale reinforcement learning on Cloud TPUs in JAX
β61Aug 22, 2026Updated last week
Alternatives and similar repositories for sebulba
Users that are interested in sebulba are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code of the paper: Debiasing Meta-Gradient Reinforcement Learning by Learning the Outer Value Functionβ13Apr 13, 2026Updated 4 months ago
- A collection of matrix games in JAXβ15Apr 13, 2026Updated 4 months ago
- Accelerated replay buffers in JAXβ47Sep 17, 2022Updated 3 years ago
- β‘ Flashbax: Accelerated Replay Buffers in JAXβ281Updated this week
- Simple single file implementations of Reinforcement Learning algorithms in Juliaβ24Feb 15, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A categorised list of Multi-Agent Reinforcemnt Learning (MARL) papersβ58Jan 20, 2023Updated 3 years ago
- CleanRL's implementation of DeepMind's Podracer Sebulba Architecture for Distributed DRLβ124Aug 22, 2024Updated 2 years ago
- Reinforcement learning training framework for entity-gym environments.β17Mar 18, 2024Updated 2 years ago
- A tool for aggregating and plotting MARL experiment data.β86Apr 13, 2026Updated 4 months ago
- A neural network library written in jaxβ14Feb 3, 2025Updated last year
- πΉοΈ A diverse suite of scalable reinforcement learning environments in JAXβ859Aug 24, 2026Updated last week
- Population-Based Reinforcement Learning for Combinatorial Optimizationβ88Feb 12, 2024Updated 2 years ago
- Baselines for gymnax π€β78Apr 3, 2023Updated 3 years ago
- ποΈA research-friendly codebase for fast experimentation of single-agent reinforcement learning in JAX β’ End-to-End JAX RLβ420Mar 18, 2026Updated 5 months ago
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Parallel hyperparameter tuning with JAXβ40Jul 18, 2026Updated last month
- Modular Single-file Reinfocement Learning Algorithms Libraryβ38May 16, 2023Updated 3 years ago
- 𧬠ManyFold: An efficient and flexible library for training and validating protein folding modelsβ80Dec 14, 2022Updated 3 years ago
- Vectorization techniques for fast population-based training.β57Apr 26, 2026Updated 4 months ago
- Datasets with baselines for Offline MARL.β228Nov 2, 2025Updated 9 months ago
- πββ¬ Contextual bandits library for continuous action trees with smoothing in JAXβ71Jul 28, 2026Updated last month
- Opinionated library for managing hyperparameters and mutable state of machine learning training systems.β19Aug 4, 2023Updated 3 years ago
- COMPASS: Combinatorial Optimization with Policy Adaptation using Latent Space Searchβ47Jun 21, 2024Updated 2 years ago
- Corax: Core RL in JAXβ41Feb 22, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A LLM-friendly framework for translating dynamical equations to gymnasium-compatible RL environments.β33Mar 18, 2026Updated 5 months ago
- π¦ A research-friendly codebase for fast experimentation of multi-agent reinforcement learning in JAXβ933Updated this week
- A toolkit for practical Human-AI cooperation researchβ15Apr 19, 2024Updated 2 years ago
- (Crafter + NetHack) in JAX. ICML 2024 Spotlight.β444Jun 20, 2026Updated 2 months ago
- flexible meta-learning in jaxβ16Oct 19, 2023Updated 2 years ago
- βοΈ Vectorized RL game environments in JAXβ641Mar 6, 2025Updated last year
- Library for efficient training and application of Machine Learning Interatomic Potentials (MLIP)β131Aug 4, 2026Updated 3 weeks ago
- JAX-accelerated Meta-Reinforcement Learning Environments Inspired by XLand and MiniGrid ποΈβ344Dec 16, 2025Updated 8 months ago
- A library for deploying App on deepchain.bioβ31Sep 24, 2021Updated 4 years ago
- GPUs on demand by Runpod - Special Offer Available β’ AdRun AI, ML, and HPC workloads on powerful cloud GPUsβwithout limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Minimal Decision Transformer Implementation written in Jax (Flax).β18Aug 8, 2022Updated 4 years ago
- Accelerated minigrid environments with JAXβ179Updated this week
- β21Dec 22, 2020Updated 5 years ago
- β101Jan 21, 2026Updated 7 months ago
- Standard interface for entity based reinforcement learning environments.β40Feb 28, 2024Updated 2 years ago
- Code for Discovered Policy Optimisation (NeurIPS 2022)β12Jun 15, 2023Updated 3 years ago
- AbBFN2: A flexible antibody foundation model based on Bayesian Flow Networksβ39Jun 4, 2025Updated last year