2048 environment for Reinforcement Learning and DQN algorithm
☆40May 27, 2022Updated 4 years ago
Alternatives and similar repositories for 2048_env
Users that are interested in 2048_env are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICLR 2024] DMBP: Diffusion Model-Based Predictor for Robust Offline Reinforcement Learning against State Observations Perturbations.☆17May 24, 2024Updated 2 years ago
- Code for ICLR 2022 paper Rethinking Goal-Conditioned Supervised Learning and Its Connection to Offline RL.☆27Feb 21, 2022Updated 4 years ago
- Robust Reinforcement Learning Benchmark☆14Sep 22, 2024Updated last year
- Implement many Sparse Reward algorithms in Gym Fetch environment☆89Jul 9, 2020Updated 6 years ago
- Code for https://arxiv.org/abs/1811.00145☆12Feb 13, 2021Updated 5 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Model-based Hindsight Experience Replay☆10Jun 8, 2022Updated 4 years ago
- code of IJCAI submission "Soft Hindsight Experience Replay"☆13Mar 23, 2020Updated 6 years ago
- SVIP: Towards Verifiable Inference of Open-Source Large Language Models☆15Jun 3, 2025Updated last year
- ☆29Oct 10, 2018Updated 7 years ago
- Implementation of DyMA-CL, MARL algorithm☆30Apr 18, 2020Updated 6 years ago
- ☆25Nov 30, 2020Updated 5 years ago
- Policy Transfer across Visual and Dynamics Domain Gaps via Iterative Grounding (RSS 2021)☆12Oct 22, 2021Updated 4 years ago
- Language independent SSL-based Speaker Anonymization system☆20May 28, 2024Updated 2 years ago
- ☆17Apr 14, 2026Updated 4 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Repository for running LLMs efficiently on Mac silicon (M1, M2, M3). Features Jupyter notebook for Meta-Llama-3 setup using MLX framework…☆11May 4, 2024Updated 2 years ago
- Reinforcement Learning (PPO) applied to a multiplayer simple card game (Witches)☆10Jun 7, 2020Updated 6 years ago
- ☆10Nov 23, 2020Updated 5 years ago
- OpenAI Gym 课程练习笔记☆15Apr 16, 2024Updated 2 years ago
- KiCAD plugin written in Python for programatically placing clusters of components onto a PCB from a layout file.☆10Jun 30, 2021Updated 5 years ago
- An app for human activities recognition(HAR), which is based on PaddlePaddle framework of Baidu☆22Mar 12, 2021Updated 5 years ago
- Code for "Minimizing Weighted Counterfactual Regret with Optimistic Online Mirror Descent", IJCAI 2024 (Oral)☆16Aug 27, 2024Updated 2 years ago
- Reinforcement learning - Batched Impala - PyTorch - Mario Kart☆13Jul 21, 2020Updated 6 years ago
- [WWW '24] UnifiedSSR: A Unified Framework of Sequential Search and Recommendation☆12Feb 16, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- 从今天起,每天看一小时的ROS源代码,写写笔记和注释☆18Sep 29, 2017Updated 8 years ago
- [ICLR 2024 Spotlight] Code for ICLR 2024 paper "Towards Robust Offline Reinforcement Learning under Diverse Data Corruption"☆22Nov 25, 2024Updated last year
- Learning from Guided Play: A Scheduled Hierarchical Approach for Improving Exploration in Adversarial Imitation Learning Source Code☆17Aug 23, 2024Updated 2 years ago
- A simple highway traffic simulation for self-driving car agents in occupancy grid world☆16May 28, 2019Updated 7 years ago
- Deep Learning 2021 in School of Data Science, USTC☆12May 17, 2023Updated 3 years ago
- Fixed version of tg-cli with support of channels and groups.☆13Jul 7, 2017Updated 9 years ago
- 爬取各大OJ题目☆10Aug 28, 2017Updated 9 years ago
- Pathfinding Using Reinforcement Learning☆12May 21, 2019Updated 7 years ago
- A minimalistic header only C++11 Neural Network library based on Eigen::Tensor☆20Jan 15, 2018Updated 8 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆14Jul 7, 2019Updated 7 years ago
- Recurrent Network-based Deterministic Policy Gradient for Solving Bipedal Walking Challenge on Rugged Terrains☆12Oct 16, 2017Updated 8 years ago
- ☆17Apr 22, 2026Updated 4 months ago
- Global Router Built for ICCAD Contest 2019☆35Mar 20, 2020Updated 6 years ago
- 基于point_to_line的icp算法☆19May 18, 2019Updated 7 years ago
- Lightweight speaker anonymization [IEEE SLT2021]☆27Jun 6, 2022Updated 4 years ago
- Code for InfoDeepSeek: Benchmarking Agentic Information Seeking for Retrieval-Augmented Generation☆19May 29, 2025Updated last year