Files from the published Alpha Star paper by DeepMind
☆18Nov 14, 2019Updated 6 years ago
Alternatives and similar repositories for learning_alpha_star
Users that are interested in learning_alpha_star are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆10Aug 17, 2022Updated 4 years ago
- This repository contains the data used for the paper "Entity Recognition at First Sight: Improving NER with Eye Movement Information" by …☆12Jan 22, 2020Updated 6 years ago
- Latin texts annotated for named entities and NER tagger used for the Herodotos Project (Ohio State University / Ghent University)☆12Sep 26, 2022Updated 3 years ago
- Simple Rust text editor in the spirit of kilo☆16Feb 2, 2025Updated last year
- Categorial Variation Database☆11Nov 27, 2021Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A Straightforward Pytorch Implementation of Gated Feedback RNNs☆12May 8, 2017Updated 9 years ago
- Reinforcement learning approach to playing competitive Pokémon.☆17Mar 17, 2017Updated 9 years ago
- ☆14Dec 26, 2022Updated 3 years ago
- Adding Dreamer-v3's implementation tricks to CleanRL's PPO☆16May 19, 2023Updated 3 years ago
- Learning bisimulation metrics for control, particularly suited to sparse reward settings☆11Feb 28, 2023Updated 3 years ago
- ☆14Apr 11, 2022Updated 4 years ago
- In this repository I'll be programming the cool exercises of the Book Reinforcement-Learning: An introduction by Sutton☆14Apr 15, 2018Updated 8 years ago
- Reproduction of the paper "Soft Q-Learning with Mutual Information Regularization" CoRL 2019.☆10Jan 10, 2019Updated 7 years ago
- A few versions of auction algorithms using Python and Google's Optimization Tools (OR-Tools) package for Python!☆14May 5, 2018Updated 8 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Get the aligned BERT embedding for sequence labeling tasks☆18Jun 6, 2019Updated 7 years ago
- An implementation of Factoid Question Answering presented in Large-scale Simple Question Answering with Memory Networks☆15Oct 15, 2019Updated 6 years ago
- The SMAPH system for query entity linking.☆20Jul 29, 2018Updated 8 years ago
- ☆14Mar 6, 2020Updated 6 years ago
- A curated list of awesome machine learning resources in the context of digital media and (interactive) computer graphics.☆29Aug 14, 2022Updated 4 years ago
- AlphaGo代码☆11Apr 25, 2016Updated 10 years ago
- PyCon Taiwan 2017 猜謎機器人☆12Jun 11, 2017Updated 9 years ago
- Creating a Pokemon battling and analyzing C++ library designed for speed.☆31Updated this week
- github-trending☆12Aug 18, 2018Updated 8 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Equal Loudness Filter☆11Mar 4, 2019Updated 7 years ago
- ☆13Nov 8, 2022Updated 3 years ago
- Bi-directional LSTM with attention for question answering☆15Nov 28, 2025Updated 9 months ago
- Rules used in Neural Rule Engine.☆27Aug 31, 2018Updated 8 years ago
- This repository is a collection of widely used self-supervised auxiliary losses used for learning representations in reinforcement learni…☆14Feb 27, 2023Updated 3 years ago
- HTML5 3D RTS Game (Warcraft like in the browser)☆14Dec 11, 2022Updated 3 years ago
- papers about reinforcement learning☆13Jan 4, 2021Updated 5 years ago
- Code to accompany the paper "The Information Geometry of Unsupervised Reinforcement Learning"☆20Oct 6, 2021Updated 4 years ago
- Code associated with the "Natural Language Rationales with Full-Stack Visual Reasoning" EMNLP Findings 2020 paper☆24Jan 15, 2021Updated 5 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Unofficial PyTorch Implementation of OpenAI's GPT-3☆13Apr 11, 2022Updated 4 years ago
- ☆10May 15, 2021Updated 5 years ago
- Codes for 'Deep Deterministic Information Bottleneck with Matrix-based entropy functional' in ICASSP 2021☆12Jul 27, 2022Updated 4 years ago
- ☆22Dec 20, 2018Updated 7 years ago
- Adversarial Soft Advantage Fitting: Imitation Learning without Policy Optimization☆16Dec 10, 2020Updated 5 years ago
- Data for the paper "A Dataset for Learning University STEM Courses at Scale" by Zhang et al., 2022.☆15Nov 22, 2022Updated 3 years ago
- Hybrid Linear UCB Multi-arm Bandit library☆14Oct 5, 2016Updated 9 years ago