A reinforcement learning agent that learns to solve mazes using Group Relative Policy Optimization (GRPO).
☆12Feb 9, 2025Updated last year
Alternatives and similar repositories for grpo-maze-solver
Users that are interested in grpo-maze-solver are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Implementation of PPO for CartPole-v1☆10Jan 1, 2019Updated 7 years ago
- 🐭 A tiny single-file implementation of Group Relative Policy Optimization (GRPO) as introduced by the DeepSeekMath paper☆43Jun 28, 2025Updated last year
- Snake's Food Hunt" is a competitive AI-driven game where two snakes learn to navigate, collect food, and avoid collisions using Deep Q-Le…☆10Nov 18, 2025Updated 8 months ago
- Hdl21 Schematics☆17Jan 24, 2024Updated 2 years ago
- 基于`Git`仓库存储的`Markdown`笔记应用☆22Nov 28, 2019Updated 6 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆18Jun 26, 2026Updated 3 weeks ago
- The official code and model of HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding.☆15Sep 19, 2025Updated 10 months ago
- Public code for implementation and experiments with differentiable decision trees.☆32Oct 17, 2024Updated last year
- Optimising electricity expenditure in an HVAC system under dynamic electricity pricing scheme and weather conditions using a DDPG model.☆26Feb 6, 2022Updated 4 years ago
- Build and deploy agentic finance applications on the Alva platform. Access 250+ financial data sources, run cloud-side analytics, backtes…☆33Updated this week
- Face Recognition Door Lock☆16Nov 22, 2022Updated 3 years ago
- Vision-driven Autonomous Flight of UAV Along River Using Deep Reinforcement Learning with Dynamic Expert Guidance☆15Mar 8, 2025Updated last year
- Open-source code for paper CDT: Cascading Decision Trees for Explainable Reinforcement Learning☆41Oct 31, 2025Updated 8 months ago
- ☆12Nov 12, 2022Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- An open source robot reinforcement learing plantform using stable-baselines and OpenAI Gym☆10Mar 24, 2023Updated 3 years ago
- LLM Prompting for Text2SQL via Gradual SQL Reffnement☆15Feb 19, 2025Updated last year
- 4WD Mecanum Mobile Robot ROS 1&2 Ready☆19Jun 6, 2024Updated 2 years ago
- Joint Pedestrian and Vehicle Traffic Optimization in Urban Environments using Reinforcement Learning☆17Sep 23, 2025Updated 9 months ago
- Solutions to neuralnetworksanddeeplearning.com☆14Dec 21, 2016Updated 9 years ago
- 汽车出租小项目,使用ssm框架以及layui☆12Dec 16, 2022Updated 3 years ago
- ☆15Dec 29, 2020Updated 5 years ago
- the Pytorch implementation of A Dynamic Multi-Modal Deep Reinforcement Learning Framework for 3D Bin Packing Problem☆11Sep 6, 2025Updated 10 months ago
- Code for Sibling Rivalry and experiments presented in associated paper☆18May 1, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- 此项目创建的初衷是为了帮助人工智能、自然语言处理和大语言模型相关背景的同学找工作使用,欢迎加入项目的建设和维护☆18Mar 30, 2025Updated last year
- 基于 Android Studio 与 Java 的 Android 端游戏应用,是一个结合 RPG 与 GalGame 模式的解密攻略类游戏, 包含背包系统、地图系统、交易系统、存档系统等。☆21Mar 11, 2024Updated 2 years ago
- 百度语音示例☆50Feb 28, 2018Updated 8 years ago
- Speech corpora for the speech recognition evaluation system☆21Mar 20, 2018Updated 8 years ago
- Systems Modeling. Learn a variety of systems, such as those involving mechanical, electrical, hydraulic, pneumatic systems, and mixtures …☆16Dec 20, 2017Updated 8 years ago
- ☆14Oct 28, 2023Updated 2 years ago
- A command line tool for comparing JSON files by degree of similarity.☆12Oct 28, 2019Updated 6 years ago
- Cross-State Transition Attention Transformer for improved robotic manipulation with better temporal modeling; https://arxiv.org/abs/2510.…☆19Mar 8, 2026Updated 4 months ago
- One-Shot Unsupervised Cross Domain Detection☆13Nov 22, 2022Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A multi-agent reinforcement learning solution to Flatland3 challenge.☆18Feb 16, 2024Updated 2 years ago
- SPICE Netlist Datasets: https://symbench.github.io/spice-datasets/☆40Oct 10, 2023Updated 2 years ago
- A Python package for building and cutting sparse layered s-t graphs.☆13Nov 6, 2023Updated 2 years ago
- Providing the answer to "How to do patching on all available SAEs on GPT-2?". It is an official repository of the implementation of the p…☆13Jan 26, 2025Updated last year
- ☆12Jan 9, 2025Updated last year
- Inverse Reinforcement Learning via State Marginal Matching, CoRL 2020☆45Jul 19, 2023Updated 3 years ago
- Value & Policy Iteration for the frozenlake environment of OpenAI☆15May 14, 2019Updated 7 years ago