Tackling the UNO card game with reinforcement learning
☆36May 6, 2023Updated 3 years ago
Alternatives and similar repositories for uno-card-game-rl
Users that are interested in uno-card-game-rl are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆10Mar 19, 2024Updated 2 years ago
- ☆14Mar 15, 2019Updated 7 years ago
- A tool library for riichi mahjong written in Rust, made mostly to be used as a WASM component.☆12Aug 29, 2025Updated 10 months ago
- A complete Python framework to perform real-time fMRI decoded neurofeedback experiments☆17Sep 26, 2022Updated 3 years ago
- Code for SaGe subword tokenizer (EACL 2023)☆28Nov 30, 2024Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Repository for ACL 2020 Paper: "It Takes Two to Lie: One to Lie, and One to Listen"☆20Nov 4, 2022Updated 3 years ago
- 「速習 強化学習 -基礎理論とアルゴリズム-」サポートページ☆15Nov 25, 2017Updated 8 years ago
- [ ICCV CVAMD 2023] Official implementation of "CheXFusion: Effective Fusion of Multi-View Features using Transformers for Long-Tailed Che…☆50Aug 3, 2024Updated last year
- Code for Engel, Grossmann & Ockenfels☆20Jan 2, 2026Updated 6 months ago
- Code for the paper "CoS: Enhancing Personalization and Mitigating Bias with Context Steering"☆20Dec 13, 2024Updated last year
- Portable TCP/UDP/ICMP traceroute tool, written in Python☆17Apr 18, 2020Updated 6 years ago
- ☆30Aug 25, 2022Updated 3 years ago
- Bayesian Reward Shaping Framework for Deep Reinforcement Learning☆26Mar 29, 2019Updated 7 years ago
- [COLM'24] How Easily do Irrelevant Inputs Skew the Responses of Large Language Models?☆23Oct 13, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆23Nov 8, 2023Updated 2 years ago
- ☆45Oct 21, 2022Updated 3 years ago
- This repo contains code for paper: "Uncertainty Estimation and Quantification for LLMs: A Simple Supervised Approach".☆26Oct 21, 2024Updated last year
- ☆25Mar 4, 2024Updated 2 years ago
- Brainwave is a state-of-the-art neural decoder that transforms electroencephalogram (EEG) and brain signals into multimodal outputs inclu…☆14Oct 6, 2025Updated 9 months ago
- A gym game for Contra that for reinforcement learning☆10Oct 18, 2021Updated 4 years ago
- A rewrite of scambier/markov-strings to utilize a relational SQL database rather than an in-memory object. The goal is to reduce memory u…☆10May 25, 2026Updated last month
- Official repository for Paper "Offline Goal-Conditioned Reinforcement Learning via f-Advantage Regression" (NeurIPS 2022)☆39Oct 19, 2023Updated 2 years ago
- ☆21Oct 11, 2025Updated 9 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Hands-on with popular deep learning datasets and tasks☆13Apr 4, 2023Updated 3 years ago
- Bandit algorithms☆30Oct 12, 2017Updated 8 years ago
- Code accompanying HAAR paper, NeurIPS 2019 - Hierarchical Reinforcement Learning with Advantage-Based Auxiliary Rewards☆31Jan 19, 2023Updated 3 years ago
- 华为软件精英挑战赛 2023 年年旅游嘎嘎开心代码☆15Mar 31, 2024Updated 2 years ago
- An AI for Hearthstone using deep reinforcement learning☆10Oct 6, 2017Updated 8 years ago
- Simple Python Socket-based Split Learning technique using PyTorch☆14Mar 13, 2020Updated 6 years ago
- Efficiently discovering algorithms via LLMs with evolutionary search and reinforcement learning.☆17Apr 22, 2025Updated last year
- Code for the "Cultural evolution in populations of Large Language Models" paper☆35Jul 7, 2026Updated last week
- A Convolutional Variational Autoencoder (CVAE) for 3D CFD data reconstruction and generation.☆45Mar 22, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A recommendation model kernel optimizing system☆12Jun 5, 2025Updated last year
- Pytorch implementation of an energy transformer - an energy-based reccurrent variant of the transformer.☆16Jul 11, 2023Updated 3 years ago
- The benchmark and datasets of the ICML 2024 paper "VisionGraph: Leveraging Large Multimodal Models for Graph Theory Problems in Visual C…☆17May 27, 2024Updated 2 years ago
- ☆40Jul 16, 2023Updated 3 years ago
- ☆40Mar 20, 2017Updated 9 years ago
- Inverse Reinforcement learning proof-of-concept using the Guided Cost/Reward Learning approach☆10Mar 23, 2020Updated 6 years ago
- A software UART implementation for the RP2040 using PIO.☆17Sep 9, 2024Updated last year