Interpretability dashboard for reinforcement learners
☆16Jun 4, 2019Updated 7 years ago
Alternatives and similar repositories for agent
Users that are interested in agent are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Training (hopefully) safe agents in gridworlds☆26May 12, 2019Updated 7 years ago
- This project includes various scripts for Ensage.☆11Jan 5, 2015Updated 11 years ago
- ☆18May 12, 2025Updated last year
- Implementation of a new Quantum Oracle for solving the Max-Cut Problem with Grover Search Algorithm☆11Sep 16, 2024Updated last year
- This repository contains the dataset and code for our ACL'23 publication: "MatSci-NLP: Evaluating Scientific Language Models on Materials…☆17Nov 21, 2023Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- A dashboard for exploring timm learning rate schedulers☆20Nov 22, 2024Updated last year
- Displays cricket score as notification. OS X Only.☆11Mar 26, 2015Updated 11 years ago
- ☆14Aug 9, 2023Updated 3 years ago
- Repo for the paper on Escalation Risks of AI systems☆44Apr 12, 2024Updated 2 years ago
- ☆14Dec 10, 2017Updated 8 years ago
- ☆14Jun 4, 2026Updated 2 months ago
- A gym environment for Stuart Armstrong's model of a treacherous turn.☆18Jul 28, 2018Updated 8 years ago
- ☆19Dec 4, 2025Updated 8 months ago
- ☆14May 4, 2021Updated 5 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- EvalDNN: A Toolbox for Evaluating Deep Neural Network Models☆14Mar 9, 2020Updated 6 years ago
- A curated list of awesome resources for Artificial Intelligence Alignment research☆82Jul 14, 2023Updated 3 years ago
- A python implementation of PROCLUS: PROjected CLUStering algorithm.☆10Jan 12, 2015Updated 11 years ago
- An implementation of the Escape Room domain for Hierarchical Reinforcement Learning.☆25May 15, 2019Updated 7 years ago
- ArXiv'18 implementation of amortized maximum likelihood (AML) for high-quality, weakly-supervised shape completion.☆11Nov 30, 2018Updated 7 years ago
- Mac port of Torcs, The Open Racing Car Simulator☆11Jun 16, 2010Updated 16 years ago
- Library that provides environments for planning problems☆17Apr 24, 2026Updated 3 months ago
- Hands-On TensorBoard for PyTorch Developers, Published by Packt☆11Dec 15, 2025Updated 7 months ago
- Copy code from IDE or vim. Paste to make pretty. Paste the pretty code into your Keynote or PowerPoint slides.☆15Jul 10, 2015Updated 11 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Minimal (truly) muP implementation, consistent with TP4 and TP5 papers notation☆14Jan 2, 2026Updated 7 months ago
- [AAAI26] Trade-offs in Large Reasoning Models: An Empirical Analysis of Deliberative and Adaptive Reasoning over Foundational Capabilitie…☆11Feb 7, 2026Updated 6 months ago
- A modern, minimal argument parser for Zig.☆15Apr 5, 2026Updated 4 months ago
- ☆15Dec 12, 2022Updated 3 years ago
- The Hybrid Public Key Encryption (HPKE) standard in Python☆12Apr 29, 2024Updated 2 years ago
- Gomoku AI based AlphaZero Algorithm☆10Feb 27, 2019Updated 7 years ago
- Fast interpolative decompositions in Python☆10Jan 4, 2021Updated 5 years ago
- Lua scripts for Ensage (Outdated)☆29Oct 9, 2015Updated 10 years ago
- uct tree search + supervised lerning for atari games☆12Feb 14, 2017Updated 9 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Code used in our paper "Robust Deep Reinforment Learning through Adversarial Loss"☆33Oct 3, 2023Updated 2 years ago
- An attempt to apply reinforcement learning to graph signal recovery problem☆11Aug 25, 2021Updated 4 years ago
- Pytorch implementation of Stable Opponent Shaping (https://openreview.net/pdf?id=SyGjjsC5tQ).☆21Jan 15, 2020Updated 6 years ago
- Make quick and dirty data mining made easier in Sublime Text☆11Feb 24, 2021Updated 5 years ago
- CTC beam search☆12Oct 26, 2016Updated 9 years ago
- C#的GUI五子棋大作业 包括禁手 AI 简单直播功能☆10Dec 14, 2018Updated 7 years ago
- Connect6 (Korean: 육목) for Python.☆11May 15, 2017Updated 9 years ago