Tackling the UNO card game with reinforcement learning
☆36May 6, 2023Updated 3 years ago
Alternatives and similar repositories for uno-card-game-rl
Users that are interested in uno-card-game-rl are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆10Mar 19, 2024Updated 2 years ago
- A tool library for riichi mahjong written in Rust, made mostly to be used as a WASM component.☆12Aug 29, 2025Updated last year
- This repo is to demo the concept of lossless compression with Transformers as encoder and decoder.☆14May 2, 2024Updated 2 years ago
- Code for Engel, Grossmann & Ockenfels☆20Jan 2, 2026Updated 8 months ago
- ☆30Aug 25, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Bayesian Reward Shaping Framework for Deep Reinforcement Learning☆26Mar 29, 2019Updated 7 years ago
- ☆21Jun 27, 2024Updated 2 years ago
- [COLM'24] How Easily do Irrelevant Inputs Skew the Responses of Large Language Models?☆23Oct 13, 2024Updated last year
- Code for magnetic mirror descent.☆20Oct 5, 2023Updated 2 years ago
- ☆23Nov 8, 2023Updated 2 years ago
- This repo contains code for paper: "Uncertainty Estimation and Quantification for LLMs: A Simple Supervised Approach".☆26Oct 21, 2024Updated last year
- Code for the paper "CoS: Enhancing Personalization and Mitigating Bias with Context Steering"☆20Dec 13, 2024Updated last year
- A rewrite of scambier/markov-strings to utilize a relational SQL database rather than an in-memory object. The goal is to reduce memory u…☆10Aug 27, 2026Updated 3 weeks ago
- A gym game for Contra that for reinforcement learning☆10Oct 18, 2021Updated 4 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Hands-on with popular deep learning datasets and tasks☆13Apr 4, 2023Updated 3 years ago
- Code accompanying HAAR paper, NeurIPS 2019 - Hierarchical Reinforcement Learning with Advantage-Based Auxiliary Rewards☆31Jan 19, 2023Updated 3 years ago
- 华为软件精英挑战赛 2023 年年旅游嘎嘎开心代码☆15Mar 31, 2024Updated 2 years ago
- ☆25Mar 4, 2024Updated 2 years ago
- Efficiently discovering algorithms via LLMs with evolutionary search and reinforcement learning.☆17Apr 22, 2025Updated last year
- Code for the "Cultural evolution in populations of Large Language Models" paper☆35Sep 9, 2026Updated last week
- A recommendation model kernel optimizing system☆14Jun 5, 2025Updated last year
- The benchmark and datasets of the ICML 2024 paper "VisionGraph: Leveraging Large Multimodal Models for Graph Theory Problems in Visual C…☆17May 27, 2024Updated 2 years ago
- ☆40Jul 16, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆41Mar 20, 2017Updated 9 years ago
- A software UART implementation for the RP2040 using PIO.☆18Sep 9, 2024Updated 2 years ago
- Extension for colcon to support CMake packages☆18Jun 16, 2026Updated 3 months ago
- The standard template to create a lean game☆56Jun 12, 2026Updated 3 months ago
- Blog post☆17Feb 16, 2024Updated 2 years ago
- Alex Graves' Adaptive Computation Time in PyTorch☆14Jan 9, 2018Updated 8 years ago
- Generating diverse and realistic datasets for computer vision training using AI.☆18Dec 28, 2024Updated last year
- Pure-Python interface for WIZNET 5k Ethernet modules☆17Apr 23, 2026Updated 4 months ago
- Pytorch implementation of an energy transformer - an energy-based reccurrent variant of the transformer.☆17Jul 11, 2023Updated 3 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Documentation regarding Ygopro card creation☆16Dec 29, 2014Updated 11 years ago
- LlaMA3-SFT, Meta-Llama-3-8B/Meta-Llama-3-8B-Instruct微调(transformers)/LORA(peft)/推理, 支持中文(chinese, zh)☆34May 17, 2024Updated 2 years ago
- 문장단위로 분절된 나무위키 데이터셋. Releases에서 다운로드 받거나, tfds-korean을 통해 다운로드 받으세요.☆19Jun 16, 2021Updated 5 years ago
- A Transformer approach for polyphonic Audio-to-Score (A2S) transcription (ICASSP 2024)☆15Jun 21, 2025Updated last year
- Utility tools for tenhou.net log☆31Jan 23, 2024Updated 2 years ago
- The implementation of Discriminator Soft Actor Critic☆15Jan 25, 2020Updated 6 years ago
- The official code release for Don't Trust: Verify -- Grounding LLM Quantitative Reasoning with Autoformalization☆39Mar 9, 2025Updated last year