强化学习贪吃蛇
☆17Oct 19, 2023Updated 2 years ago
Alternatives and similar repositories for RL-snack
Users that are interested in RL-snack are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Authors's code for "Variational Causal Inference Network for Explanatory Visual Question Answering" and "Integrating Neural-Symbolic Reas…☆13Apr 13, 2026Updated 3 months ago
- ☆10Dec 12, 2023Updated 2 years ago
- Cooperative Multi Agent Reinforcement Learning with Human in the Loop☆14Apr 25, 2023Updated 3 years ago
- Implementation of ICCV 2025 paper "Growing a Twig to Accelerate Large Vision-Language Models".☆30May 23, 2026Updated 2 months ago
- code for `A Hybrid Human-in-the-Loop Deep Reinforcement Learning Method for UAV Motion Planning'☆14Jan 15, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- The official repo for ”[WACV2025] Towards Accurate Unified Anomaly Segmentation“☆15Apr 14, 2025Updated last year
- An reconstruction of RL Introduction and its course materials for a more efficient entry☆19Mar 4, 2026Updated 4 months ago
- Ailanxier's note of Database Systems☆11Jan 18, 2022Updated 4 years ago
- A simple 1-d diffusion/flow model tutorial for LeCAR group meeting☆16Sep 27, 2025Updated 10 months ago
- [CVPR 2026] Where MLLMs Attend and What They Rely On: Explaining Autoregressive Token Generation☆44Jun 18, 2026Updated last month
- Multi-Layer Key-Value sharing experiments on Pythia models☆34Jun 14, 2024Updated 2 years ago
- 多智能体学习库☆22Dec 28, 2021Updated 4 years ago
- The official implementation of Natural Language Fine-Tuning☆54Jan 7, 2025Updated last year
- The old version of Hugo academic theme | 旧版Hugo学术主题☆15Aug 23, 2021Updated 4 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆16May 16, 2025Updated last year
- A collection of papers and libraries for performing multi-agent optimization☆20Jul 12, 2026Updated 2 weeks ago
- [AAAI 26'] This is the official pytorch implementation for paper: Filter, Correlate, Compress: Training-Free Token Reduction for MLLM Acc…☆47Nov 13, 2025Updated 8 months ago
- 这是2023华为软件精英挑战赛 初赛阶段319万分的代码,广西省第一名,粤港澳区排名第8。该比赛要求选手在一个50m*50m的地图上,控制4台机器人进入任务调度,设计机器人的运动算法、路径规划算法、任务调度算法,去分布在地图上的各种类型的工作台购买或者出售商品,赚取差价,以…☆17Sep 2, 2023Updated 2 years ago
- ☆19Nov 7, 2025Updated 8 months ago
- Generating Emoji with Conditional GAN☆17May 14, 2020Updated 6 years ago
- Code for paper: Traffic expertise meets residual RL: Knowledge-informed model-based residual reinforcement learning for CAV trajectory co…☆24Mar 1, 2025Updated last year
- STAR-PólyaMath: Multi-Agent Reasoning under Persistent Meta-Strategic Supervision☆21Jun 8, 2026Updated last month
- ☆189Mar 27, 2026Updated 4 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- OmiAD: One-Step Adaptive Masked Diffusion Model for Multi-class Anomaly Detection via Adversarial Distillation(ICML 2025)☆19May 29, 2025Updated last year
- ☆28Jan 9, 2025Updated last year
- Robot Learning Algorithms☆25Aug 19, 2024Updated last year
- This is official repository of Physics-AD☆23Feb 24, 2026Updated 5 months ago
- LightwheelOcc: A 3D Occupancy Synthetic Dataset in Autonomous Driving☆106Jul 2, 2025Updated last year
- ☆25Mar 7, 2026Updated 4 months ago
- EMIT: Enhancing MLLMs for Industrial Anomaly Detection via Difficulty-Aware GRPO☆27Jan 24, 2026Updated 6 months ago
- ☆39Jun 17, 2024Updated 2 years ago
- Official implementation of TailedCore(CVPR25)☆33Jun 12, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- My research papers reading list. Animation and Robotics.☆40Nov 2, 2024Updated last year
- a novel self-evolving paradigm, without task, reward, or complex workflow☆36May 12, 2026Updated 2 months ago
- [TNNLS] Anomaly Detection on Attributed Networks via Contrastive Self-Supervised Learning☆106May 12, 2021Updated 5 years ago
- [NeurIPS 2023] H-InDex: Visual Reinforcement Learning with Hand-Informed Representations for Dexterous Manipulation☆44Nov 6, 2023Updated 2 years ago
- Context-Guided Prompt Learning and Attention Refinement for Zero-Shot Anomaly Detection☆35Aug 9, 2025Updated 11 months ago
- ☆38Jun 26, 2025Updated last year
- [ECCV 2024] Official Implementation of An Incremental Unified Framework for Small Defect Inspection☆50Feb 17, 2025Updated last year