The Laser Learning Environment (LLE) is a cooperative MARL grid-world
☆13Sep 10, 2026Updated last week
Alternatives and similar repositories for lle
Users that are interested in lle are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Safe SLAC, an algorithm for safe cost-constrained reinforcement learning in high-dimensional POMDPs.☆13Mar 1, 2023Updated 3 years ago
- Bayes-Adaptive Monte-Carlo Planning algorithm☆19Mar 5, 2013Updated 13 years ago
- Code for the paper "AlwaysSafe: Reinforcement Learning Without Safety Constraint Violations During Training"☆17May 9, 2022Updated 4 years ago
- Émulateur Dofus 1.29.1 en Java☆14Dec 5, 2016Updated 9 years ago
- PDiT: Interleaving Perception and Decision-making Transformers for Deep Reinforcement Learning. AAMAS 2024 (full paper with oral presenta…☆10Dec 27, 2023Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- The solutions to the various CTF challenges I've taken part in, and that I've come up with myself.☆12Dec 15, 2025Updated 9 months ago
- This small project uses OpenAI's whisper AI to generate captions for videos.☆18Nov 10, 2022Updated 3 years ago
- Stochastic gradient descent Haskell library☆13Jun 24, 2024Updated 2 years ago
- Connect 4 AI using Monte Carlo Tree Search algorithm.☆11Feb 10, 2024Updated 2 years ago
- Density Constrained Reinforcement Learning☆12Mar 24, 2023Updated 3 years ago
- A student platform for ULB focused on real student collaboration☆53Updated this week
- Électronique numérique / Digital Electronics☆11Feb 16, 2022Updated 4 years ago
- ☆12Sep 8, 2022Updated 4 years ago
- Peter Selinger's LaTeX macros for Fitch style natural deduction☆20Updated this week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Code for paper, A Multimodal Graph Neural Network Framework for Cancer Molecular Subtype Classification☆16Apr 14, 2025Updated last year
- Connect your Obsidian to Wakatime or Wakapi to track the time spent while browsing or writing notes.☆21Sep 14, 2026Updated last week
- My 2023 list of useful websites, online apps and utilities (small or big) for devs, artists, writers and just everyone.☆11Jan 11, 2023Updated 3 years ago
- Proximal Policy Optimization(PPO) with Intrinsic Curiosity Module(ICM)☆18Apr 15, 2022Updated 4 years ago
- FireCommander2020: A Multiagent, Interactive Joint Perception-Action Reconnaissance Environment☆17Sep 18, 2022Updated 4 years ago
- Bridging State and History Representations: Understanding Self-Predictive RL, ICLR 2024☆28Apr 26, 2026Updated 4 months ago
- ☆12Jun 5, 2023Updated 3 years ago
- ☆14May 30, 2019Updated 7 years ago
- Code repo for "Collapsing Bandits and Their Applications to Public Health Interventions", (NeurIPS'20)☆11Dec 3, 2025Updated 9 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Avoiding catastrophic failures in reinforcement learning by learning to shape rewards.☆10Nov 13, 2017Updated 8 years ago
- This repository is an implementation of "MASER: Multi-Agent Reinforcement Learning with Subgoals Generated from Experience Replay Buffer"…☆23Jul 6, 2023Updated 3 years ago
- LAMBDA is a model-based reinforcement learning agent that uses Bayesian world models for safe policy optimization☆41Jan 16, 2023Updated 3 years ago
- Non-stationary Off-policy Evaluation☆13Nov 8, 2018Updated 7 years ago
- Probabilistic planning in continuous state-action MDPs in TensorFlow.☆13Jun 21, 2022Updated 4 years ago
- Constrained episodic reinforcement learning in concave-convex and knapsack settings☆11Oct 3, 2023Updated 2 years ago
- A tour of Pomdpland☆10Aug 10, 2022Updated 4 years ago
- ☆11Dec 27, 2021Updated 4 years ago
- Implementation of "POPCORN: Partially Observed Prediction Constrained Reinforcement Learning" (Futoma, Hughes, Doshi-Velez, AISTATS 2020)☆11May 19, 2021Updated 5 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- ☆16Jun 1, 2023Updated 3 years ago
- A collection of environments and reference agents for planning and reinforcement learning research in partially observable, multi-agent …☆34Jun 2, 2025Updated last year
- Probabilistic Programming in Python. Uses Theano as a backend and includes the NUTS sampler.☆12Apr 19, 2017Updated 9 years ago
- Build a Responsive Calendar App with HTML, CSS and Javascript | Tutorial 2024☆20Oct 11, 2024Updated last year
- Convergent Policy Optimization for Safe Reinforcement Learning☆11Oct 26, 2019Updated 6 years ago
- Pymojang is a full wrapper around de Mojang API and Mojang Authentication API☆12Jun 22, 2026Updated 3 months ago
- A web page to collect reproduced papers in one place with their codes☆14Mar 8, 2023Updated 3 years ago