基于stablebaseline3强化学习框架和gym-super-mario-bros马里奥游戏包,训练马里奥通关。
☆223Dec 8, 2025Updated 8 months ago
Alternatives and similar repositories for RL_SuperMario
Users that are interested in RL_SuperMario are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Training a humanoid robot for locomotion using Reinforcement Learning☆1,210May 3, 2026Updated 4 months ago
- Python code for Reinforcement Learning of Tic Tac Toe☆36Mar 16, 2025Updated last year
- Reproduce of MPC-D-CBF☆26Oct 9, 2024Updated last year
- ☆331Oct 9, 2024Updated last year
- ☆14Dec 22, 2025Updated 8 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- (NeurIPS 2025) COLA: Towards Efficient Multi-Objective Reinforcement Learning with Conflict Objective Regularization in Latent Space☆18Dec 12, 2025Updated 8 months ago
- ☆21Mar 2, 2026Updated 6 months ago
- 收录一些数学书籍,便于查阅(工科生必须夯实数理基础啊!)☆12Oct 29, 2022Updated 3 years ago
- [RA-L 2025] RT-GuIDE: Real-Time Gaussian Splatting for Information-Driven Exploration☆25Nov 30, 2025Updated 9 months ago
- "Personal site documenting my journey and idae in CV and SLAM.☆16May 30, 2026Updated 3 months ago
- A non-embedded AI for Clash Royale based on RL and CV.☆468Jun 6, 2024Updated 2 years ago
- 自动驾驶学习资料/书籍/感知/规划/控制/SLAM/入门/Automatic driving learning materials☆13Dec 14, 2021Updated 4 years ago
- A MARL method for motion planning of free-floating space robot.☆36Mar 15, 2024Updated 2 years ago
- Extending PRD to MAPPO with soft and semi-hard attention mechanisms☆13May 26, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆490Feb 25, 2026Updated 6 months ago
- ☆641Oct 31, 2024Updated last year
- ☆17Jan 14, 2025Updated last year
- An environment based on JSBSIM aimed at one-to-one close air combat.☆20Sep 14, 2025Updated 11 months ago
- Text-to-Speech for ROS 2☆23Dec 8, 2025Updated 8 months ago
- 使用DDQN算法控制单个路口的红绿灯☆15Mar 25, 2022Updated 4 years ago
- ☆222May 13, 2025Updated last year
- A simulation project on Dynamic Movement Primitive (DMP) , containing 1D numerical simulation, 2D learning from demonstration and 3D simu…☆12Feb 27, 2025Updated last year
- MSHub: Medical Image Segmentation Hub with Pre-trained nnUNets☆24Feb 28, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆20Dec 24, 2023Updated 2 years ago
- ☆19Sep 12, 2022Updated 3 years ago
- Repository containing variety of model files and program scripts for Nvidia Isaac -enviroments☆16Updated this week
- Hierarchical reinforcement learning framework which uses a directed graph to define the hierarchy.☆16Aug 5, 2022Updated 4 years ago
- ☆31Apr 25, 2026Updated 4 months ago
- Code for ACM MM 2024 paper "A Picture Is Worth a Graph: A Blueprint Debate Paradigm for Multimodal Reasoning"☆19Dec 5, 2024Updated last year
- VL53L1x激光测距传感器源码☆11Jul 26, 2021Updated 5 years ago
- Integrating opencv with mujoco.☆12Mar 25, 2025Updated last year
- This project enables the efficient extraction of structured data from unstructured text using large language models (LLMs). It provides a…☆22Updated this week
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- CURLA: CURL x CARLA -- Robust end-to-end Autonomous Driving by combining Contrastive Learning and Reinforcement Learning☆17Feb 6, 2024Updated 2 years ago
- Reinforcement Learning for Autonomous Satellite Constellation Scheduling and Earth Observation Optimization☆24May 31, 2026Updated 3 months ago
- ☆14Jul 30, 2024Updated 2 years ago
- 常见的自动驾驶控制算法 | 纵向控制算法:PID 横向控制算法:Pure pursuit、Stanley、MPC、LQR☆19Feb 28, 2024Updated 2 years ago
- User-specified ICP.☆12Sep 14, 2021Updated 4 years ago
- A Multi-Agent Approach Integrating Socratic Guidance for Automated Prompt Optimization☆18Dec 15, 2025Updated 8 months ago
- ☆13Oct 22, 2024Updated last year