强化学习贪吃蛇
☆18Oct 19, 2023Updated 2 years ago
Alternatives and similar repositories for RL-snack
Users that are interested in RL-snack are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆10Dec 12, 2023Updated 2 years ago
- Cooperative Multi Agent Reinforcement Learning with Human in the Loop☆14Apr 25, 2023Updated 3 years ago
- Implementation of ICCV 2025 paper "Growing a Twig to Accelerate Large Vision-Language Models".☆31May 23, 2026Updated 4 months ago
- how to build a sentence embedding application using BentoML☆15Jul 14, 2026Updated 2 months ago
- code for `A Hybrid Human-in-the-Loop Deep Reinforcement Learning Method for UAV Motion Planning'☆14Jan 15, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- A Bevy friendly wrapper of Box2D and LiquidFun.☆10Aug 14, 2024Updated 2 years ago
- 程序员的数学三本书☆17Apr 12, 2019Updated 7 years ago
- Procedurally generated starfield inside a WebGL shader☆10Sep 4, 2016Updated 10 years ago
- gym 框架下的多智能体追逃博弈强化学习平台☆17Jun 20, 2023Updated 3 years ago
- The official repo for ”[WACV2025] Towards Accurate Unified Anomaly Segmentation“☆16Apr 14, 2025Updated last year
- An reconstruction of RL Introduction and its course materials for a more efficient entry☆20Mar 4, 2026Updated 6 months ago
- Games and AI in the browser! Demo site here: http://rl-games-ai.netlify.app/☆10Oct 27, 2020Updated 5 years ago
- Ailanxier's note of Database Systems☆11Jan 18, 2022Updated 4 years ago
- A simple 1-d diffusion/flow model tutorial for LeCAR group meeting☆16Sep 27, 2025Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Simulation for DJI Pupper v2 robot☆17Oct 17, 2023Updated 2 years ago
- Multi-Layer Key-Value sharing experiments on Pythia models☆35Jun 14, 2024Updated 2 years ago
- The old version of Hugo academic theme | 旧版Hugo学术主题☆15Aug 23, 2021Updated 5 years ago
- ☆19May 16, 2025Updated last year
- Understanding Large Language Transformer Architecture like a child☆38Apr 3, 2024Updated 2 years ago
- 这是2023华为软件精英挑战赛 初赛阶段319万分的代码,广西省第一名,粤港澳区排名第8。该比赛要求选手在一个50m*50m的地图上,控制4台机器人进入任务调度,设计机器人的运动算法、路径规划算法、任务调度算法,去分布在地图上的各种类型的工作台购买或者出售商品,赚取差价,以…☆17Sep 2, 2023Updated 3 years ago
- An end-to-end (E2E) reinforcement learning model for autonomous vehicle collision avoidance in the CARLA simulator, using a recurrent PPO…☆44Aug 20, 2026Updated last month
- Energy-based models in PyTorch☆13Nov 1, 2020Updated 5 years ago
- CRUD Word documents with Python☆13Sep 10, 2026Updated 2 weeks ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Generating Emoji with Conditional GAN☆17May 14, 2020Updated 6 years ago
- Code for paper: Traffic expertise meets residual RL: Knowledge-informed model-based residual reinforcement learning for CAV trajectory co…☆26Mar 1, 2025Updated last year
- A simplified port of DIStage phased dependency injection for Typescript☆22Aug 17, 2026Updated last month
- STAR-PólyaMath: Multi-Agent Reasoning under Persistent Meta-Strategic Supervision☆24Jun 8, 2026Updated 3 months ago
- 本意是想做一个直接调用kimi的API帮我读论文的程序,然后发现API太贵了,但kimi网页版免费,就结合chrome和python写了这么个东西☆12Mar 7, 2024Updated 2 years ago
- Panda3D tutorials translated from Python to C++☆24Jan 12, 2013Updated 13 years ago
- OmiAD: One-Step Adaptive Masked Diffusion Model for Multi-class Anomaly Detection via Adversarial Distillation(ICML 2025)☆19May 29, 2025Updated last year
- 由于BAAI/bge-large-zh 在Hugging Face Clone不下来,手动下载下来,便于使用☆11Sep 16, 2023Updated 3 years ago
- ☆29Jan 9, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Robot Learning Algorithms☆25Aug 19, 2024Updated 2 years ago
- This is official repository of Physics-AD☆24Feb 24, 2026Updated 7 months ago
- Collection of LLM completions for reasoning-gym task datasets☆31Jul 4, 2025Updated last year
- Generate java source code based on existing classes using templates☆13Jun 18, 2024Updated 2 years ago
- Code Nimble is a light code editor dedicated for competitive programming.☆16Jan 9, 2025Updated last year
- A WYSIWYG markdown editor☆16May 6, 2026Updated 4 months ago
- https://wiki.openjdk.org/display/CodeTools/jcov☆18Aug 20, 2026Updated last month