The Laser Learning Environment (LLE) is a cooperative MARL grid-world
☆13Jul 23, 2026Updated this week
Alternatives and similar repositories for lle
Users that are interested in lle are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Safe SLAC, an algorithm for safe cost-constrained reinforcement learning in high-dimensional POMDPs.☆11Mar 1, 2023Updated 3 years ago
- Bayes-Adaptive Monte-Carlo Planning algorithm☆19Mar 5, 2013Updated 13 years ago
- Code for the paper "AlwaysSafe: Reinforcement Learning Without Safety Constraint Violations During Training"☆17May 9, 2022Updated 4 years ago
- Émulateur Dofus 1.29.1 en Java☆14Dec 5, 2016Updated 9 years ago
- PDiT: Interleaving Perception and Decision-making Transformers for Deep Reinforcement Learning. AAMAS 2024 (full paper with oral presenta…☆10Dec 27, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- The solutions to the various CTF challenges I've taken part in, and that I've come up with myself.☆12Dec 15, 2025Updated 7 months ago
- This small project uses OpenAI's whisper AI to generate captions for videos.☆17Nov 10, 2022Updated 3 years ago
- Stochastic gradient descent Haskell library☆13Jun 24, 2024Updated 2 years ago
- Connect 4 AI using Monte Carlo Tree Search algorithm.☆11Feb 10, 2024Updated 2 years ago
- Density Constrained Reinforcement Learning☆12Mar 24, 2023Updated 3 years ago
- A student platform for ULB focused on real student collaboration☆53Jun 30, 2026Updated 3 weeks ago
- Électronique numérique / Digital Electronics☆11Feb 16, 2022Updated 4 years ago
- ☆12Sep 8, 2022Updated 3 years ago
- Peter Selinger's LaTeX macros for Fitch style natural deduction☆19Jul 15, 2026Updated last week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- My 2023 list of useful websites, online apps and utilities (small or big) for devs, artists, writers and just everyone.☆11Jan 11, 2023Updated 3 years ago
- Code for paper, A Multimodal Graph Neural Network Framework for Cancer Molecular Subtype Classification☆16Apr 14, 2025Updated last year
- Connect your Obsidian to Wakatime or Wakapi to track the time spent while browsing or writing notes.☆21Updated this week
- Proximal Policy Optimization(PPO) with Intrinsic Curiosity Module(ICM)☆18Apr 15, 2022Updated 4 years ago
- FireCommander2020: A Multiagent, Interactive Joint Perception-Action Reconnaissance Environment☆17Sep 18, 2022Updated 3 years ago
- Bridging State and History Representations: Understanding Self-Predictive RL, ICLR 2024☆27Apr 26, 2026Updated 2 months ago
- ☆12Jun 5, 2023Updated 3 years ago
- ☆14May 30, 2019Updated 7 years ago
- Code repo for "Collapsing Bandits and Their Applications to Public Health Interventions", (NeurIPS'20)☆11Dec 3, 2025Updated 7 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Avoiding catastrophic failures in reinforcement learning by learning to shape rewards.☆10Nov 13, 2017Updated 8 years ago
- This repository is an implementation of "MASER: Multi-Agent Reinforcement Learning with Subgoals Generated from Experience Replay Buffer"…☆23Jul 6, 2023Updated 3 years ago
- LAMBDA is a model-based reinforcement learning agent that uses Bayesian world models for safe policy optimization☆39Jan 16, 2023Updated 3 years ago
- Non-stationary Off-policy Evaluation☆13Nov 8, 2018Updated 7 years ago
- Probabilistic planning in continuous state-action MDPs in TensorFlow.☆13Jun 21, 2022Updated 4 years ago
- Constrained episodic reinforcement learning in concave-convex and knapsack settings☆11Oct 3, 2023Updated 2 years ago
- A tour of Pomdpland☆10Aug 10, 2022Updated 3 years ago
- ☆11Dec 27, 2021Updated 4 years ago
- Implementation of "POPCORN: Partially Observed Prediction Constrained Reinforcement Learning" (Futoma, Hughes, Doshi-Velez, AISTATS 2020)☆11May 19, 2021Updated 5 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆16Jun 1, 2023Updated 3 years ago
- A collection of environments and reference agents for planning and reinforcement learning research in partially observable, multi-agent …☆33Jun 2, 2025Updated last year
- Probabilistic Programming in Python. Uses Theano as a backend and includes the NUTS sampler.☆12Apr 19, 2017Updated 9 years ago
- Build a Responsive Calendar App with HTML, CSS and Javascript | Tutorial 2024☆19Oct 11, 2024Updated last year
- Convergent Policy Optimization for Safe Reinforcement Learning☆11Oct 26, 2019Updated 6 years ago
- Pymojang is a full wrapper around de Mojang API and Mojang Authentication API☆13Jun 22, 2026Updated last month
- A web page to collect reproduced papers in one place with their codes☆14Mar 8, 2023Updated 3 years ago