☆25Jun 30, 2022Updated 4 years ago
Alternatives and similar repositories for Competition_Olympics-Integrated
Users that are interested in Competition_Olympics-Integrated are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆14Jul 5, 2021Updated 5 years ago
- Codes accompanying the paper "Offline Reinforcement Learning with Value-Based Episodic Memory" (ICLR 2022 https://arxiv.org/abs/2110.0979…☆15Mar 9, 2022Updated 4 years ago
- Code repository for On the interaction between supervision and self-play in emergent communication (ICLR 2020)☆15Feb 4, 2020Updated 6 years ago
- ☆74Feb 4, 2024Updated 2 years ago
- Scalable Multi-Agent Reinforcement Learning☆15Dec 25, 2021Updated 4 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Additional environments compatible with OpenAI gym☆23Mar 11, 2021Updated 5 years ago
- Logarithmic Reinforcement Learning☆28Apr 7, 2023Updated 3 years ago
- original source code of the ASE 2019 paper: Wuji: Automatic Online Combat Game Testing Using Evolutionary Deep Reinforcement Learning☆28Jun 8, 2020Updated 6 years ago
- ☆12Apr 1, 2025Updated last year
- 该项目可以根据用户给出的上文自动生成下文 该项目是本人的本科毕业设计。项目主要基于GPT-2 Chinese实现。本人的工作主要是用新的语料库进行了几次训练,得出来了一个还凑合的模型。该项目已经初步完成,不再进行进一步的更新。☆12Jun 9, 2020Updated 6 years ago
- Extreme Q-Learning: Max Entropy RL without Entropy☆87Feb 14, 2023Updated 3 years ago
- Making maps from DOOM in Rust☆13Mar 12, 2018Updated 8 years ago
- ☆174Oct 9, 2023Updated 2 years ago
- PLM: Efficient Peripheral Language Models Hardware-Co-Designed for Ubiquitous Computing☆21Mar 18, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [TPAMI] "Symbolic Visual Reinforcement Learning: A Scalable Framework with Object-Level Abstraction and Differentiable Expression Search"…☆18Jan 4, 2023Updated 3 years ago
- ☆12May 12, 2026Updated 3 months ago
- Pytorch implementation of "Maximum a Posteriori Policy Optimization" with Retrace for Discrete gym environments☆29Sep 10, 2020Updated 5 years ago
- ☆13Jul 25, 2023Updated 3 years ago
- Code for "Goal-Conditioned Predictive Coding for Offline Reinforcement Learning" (NeurIPS 2023)☆15Dec 8, 2023Updated 2 years ago
- ☆18Aug 3, 2022Updated 4 years ago
- We investigate the effect of populations on finding good solutions to the robust MDP☆29Mar 27, 2021Updated 5 years ago
- [WIP] Rust implementation of Gymnasium API☆15Updated this week
- Level-based Foraging (LBF): A multi-agent environment for RL☆214Sep 15, 2024Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- A repository for code of reinforcement learning algorithms with PyTorch☆30Sep 20, 2021Updated 4 years ago
- A Qwen .5B reasoning model trained on OpenR1-Math-220k☆14Aug 26, 2026Updated last week
- Jax implementation of VIT-VQGAN☆10Jan 25, 2024Updated 2 years ago
- DiWA: Diverse Weight Averaging for Out-of-Distribution Generalization☆31Jan 31, 2023Updated 3 years ago
- A set of competitive environments for Reinforcement Learning research.☆31Dec 1, 2022Updated 3 years ago
- ☆13Jul 9, 2021Updated 5 years ago
- A Simple, Distributed and Asynchronous Multi-Agent Reinforcement Learning Framework for Google Research Football AI.☆120Jan 16, 2024Updated 2 years ago
- 统计微信朋友圈送出的赞票与得到的赞票人员比例☆11May 3, 2016Updated 10 years ago
- PyTorch implementation of Count-Based Exploration with Neural Density Models☆10Mar 22, 2018Updated 8 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Okapi BM25 with Go☆13Mar 28, 2023Updated 3 years ago
- Applying PBT optimization technique to different domains☆10Oct 16, 2019Updated 6 years ago
- 中国常用大地测量(投影)坐标系相互转换☆11Feb 14, 2020Updated 6 years ago
- A random map generator for Doom (community maintenance repo)☆17Jun 28, 2021Updated 5 years ago
- PyTorch implementations of Reinforcement Learning algorithms in less than 200 lines☆10Apr 3, 2020Updated 6 years ago
- [ICLR 2022 Spotlight] Code for Reinforcement Learning with Sparse Rewards using Guidance from Offline Demonstration☆28Feb 10, 2022Updated 4 years ago
- MVE: model-based value estimation☆11Jul 30, 2018Updated 8 years ago