Experimentation with Regularized Nash Dynamics on a GPU accelerated game
☆55Apr 21, 2023Updated 3 years ago
Alternatives and similar repositories for R-NaD
Users that are interested in R-NaD are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official code of Nash-DQN for paper: Nash-DQN algorithm for two-player zero-sum Markov games, details see our paper: A Deep Reinforcement…☆22Aug 26, 2022Updated 4 years ago
- Code for magnetic mirror descent.☆20Oct 5, 2023Updated 2 years ago
- Partially Observable Benchmarks in JAX☆27Apr 30, 2026Updated 4 months ago
- Classic MCTS example with mctx☆25May 25, 2023Updated 3 years ago
- Official Code Release for Pipeline PSRO: A Scalable Approach for Finding Approximate Nash Equilibria in Large Games☆58Aug 30, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Web version of “Neuroevolution of Self-Interpretable Agents” (https://arxiv.org/abs/2003.08165)☆22Jan 12, 2022Updated 4 years ago
- Control barrier functions (CBFs) in Julia.☆14Sep 19, 2024Updated last year
- Official code for "A General Learning Framework for Open Ad Hoc Teamwork Using Graph-based Policy Learning"☆15Mar 1, 2023Updated 3 years ago
- V-MPO torch version with DMLab30 and GTrXL☆13Mar 1, 2021Updated 5 years ago
- Code accompanying paper "Models as Agents: Optimizing Multi-Step Predictions of Interactive Local Models in Model-Based Multi-Agent Reinf…☆15Dec 2, 2023Updated 2 years ago
- Leniax is a search and rendering engine for Lenia.☆28Apr 3, 2024Updated 2 years ago
- A collection of deep reinforcement learning algorithm implementations☆11Jan 9, 2020Updated 6 years ago
- ☆10Jan 24, 2022Updated 4 years ago
- I added selfplay functionality to openai gyms☆10Jan 16, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Sample for training an agent which mimics a cab driver to gain maximum profits by picking the correct rides. The agent is trained using d…☆10Jul 5, 2022Updated 4 years ago
- A number of agents (PPO, MuZero) with a Perceiver-based NN architecture that can be trained to achieve goals in nethack/minihack environm…☆44Sep 19, 2022Updated 3 years ago
- A simple 2D solar system explorer written in C++, powered by SDL and ImGui☆11Sep 30, 2019Updated 6 years ago
- flexible meta-learning in jax☆16Oct 19, 2023Updated 2 years ago
- ☆43Apr 27, 2022Updated 4 years ago
- ☆16Nov 26, 2013Updated 12 years ago
- ☆28Feb 17, 2020Updated 6 years ago
- ☆11Feb 8, 2026Updated 6 months ago
- A minimal Pytorch Implementation of Stochastically Quantized Variational AutoEncoder (SQ-VAE) by Sony☆33Oct 16, 2023Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- A deep learning agent for The Legend of Zelda (nes)☆27Aug 4, 2026Updated last month
- A Julia package for constrained trajectory optimization using direct methods.☆29Jun 12, 2022Updated 4 years ago
- A rust implementation of Counterfacutual Regret Minimization☆10Nov 27, 2021Updated 4 years ago
- Translation of, and commentary on, Joyal's classic paper "Une théorie combinatoire des séries formelles" (A combinatorial theory of forma…☆30Jul 20, 2024Updated 2 years ago
- Themes for Makie☆33Nov 24, 2025Updated 9 months ago
- Deep Reinforcement Learning - Pong☆14Feb 14, 2022Updated 4 years ago
- Collect orderbook data from crypto exchanges and publish as GRPC☆13Jun 19, 2022Updated 4 years ago
- Deep Reinforcement Learning for Nash Equilibria☆44Oct 25, 2022Updated 3 years ago
- ☆11Oct 19, 2020Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Implementation of POMDP algorithms on the tiger example, as described in Littman, Cassandra and Kaelbling (1994).☆17Aug 8, 2017Updated 9 years ago
- c++ implementation of alphagozero☆15May 29, 2018Updated 8 years ago
- Code for NeurIPS 2022 paper Exploiting Reward Shifting in Value-Based Deep RL☆29Oct 29, 2023Updated 2 years ago
- (Crafter + NetHack) in JAX. ICML 2024 Spotlight.☆446Jun 20, 2026Updated 2 months ago
- Deep Learning the Sorting Algorithm☆12Dec 11, 2016Updated 9 years ago
- SBX: Stable Baselines Jax (SB3 + Jax) RL algorithms☆608Aug 24, 2026Updated last week
- A Julia package for constrained iterative LQR (iLQR)☆44Mar 17, 2023Updated 3 years ago