☆14Jun 17, 2022Updated 4 years ago
Alternatives and similar repositories for hoad
Users that are interested in hoad are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A Continual Multi-agent RL testbed based on Hanabi☆31Aug 1, 2021Updated 5 years ago
- Implementation of the Off Belief Learning algorithm.☆49Aug 18, 2022Updated 4 years ago
- Framework for writing bots that play Hanabi.☆37May 16, 2019Updated 7 years ago
- Implementation of Tactical Optimistic and Pessimistic value estimation☆25Jul 18, 2023Updated 3 years ago
- Dockerfile for RL research. Including MuJoCo / DMC / PyTorch / Tensoflow / Atari support.☆16Jan 5, 2022Updated 4 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Repository for ML Reproducibility Challenge 2020 for the Neurips paper, "The Value Equivalence Principle for Model-Based Reinforcement Le…☆18Apr 13, 2021Updated 5 years ago
- Docker containers of baseline agents for the Crafter environment☆30Dec 14, 2021Updated 4 years ago
- Code for ICRA2018 - Intent-aware Multi-agent Reinforcement Learning.☆22Feb 22, 2018Updated 8 years ago
- ♊ Minimal PyTorch Twin Delayed DDPG (TD3) implementation☆10Jun 20, 2021Updated 5 years ago
- ☆28Nov 22, 2019Updated 6 years ago
- PyTorch implementation for "On the Critical Role of Conventions in Adaptive Human-AI Collaboration", ICLR 2021☆15Mar 9, 2021Updated 5 years ago
- ☆14Feb 6, 2024Updated 2 years ago
- 🔍 Codebase for the ICML '20 paper "Ready Policy One: World Building Through Active Learning" (arxiv: 2002.02693)☆18Jul 6, 2023Updated 3 years ago
- Release code for ICML2020 Knowing The What But Not The Where in Bayesian Optimization☆15Mar 7, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Example of android app written in Qt/Qml which uses MXNet for plant image recognition.☆10Nov 4, 2017Updated 8 years ago
- ☆17Jun 30, 2022Updated 4 years ago
- Code of "Towards Skilled Population Curriculum for MARL" + Implementation of Curriculum MARL algorithms based on Ray☆13Feb 20, 2023Updated 3 years ago
- Anti exploration in offline reinforcement learning☆11May 17, 2021Updated 5 years ago
- ☆20Jun 14, 2022Updated 4 years ago
- An implementation of the paper "Solving the Rubik's Cube without Human Knowledge"☆14Dec 9, 2018Updated 7 years ago
- Code from the paper "Effective Diversity in Population Based Reinforcement Learning", presented as a spotlight at NeurIPS 2020. This is t…☆45Oct 29, 2020Updated 5 years ago
- Web Interface for gaze recording: CVPR 2018☆10Jul 10, 2018Updated 8 years ago
- Repo for the paper: Learning with Muscles: Benefits for Data-Efficiency and Robustness in Anthropomorphic Tasks. https://al.is.mpg.de/pub…☆16Dec 1, 2022Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [AutoML'22] Bayesian Generational Population-based Training (BG-PBT)☆31Sep 16, 2022Updated 3 years ago
- Installer for Rhino 3D on Wine.☆14Oct 19, 2024Updated last year
- ☆48Dec 8, 2022Updated 3 years ago
- Research code implementing the search AI agent for Hanabi, as well as a web server so people can play against it☆129Jul 18, 2023Updated 3 years ago
- Open source demo for the paper Learning to Score Behaviors for Guided Policy Optimization☆24Jun 24, 2020Updated 6 years ago
- Code for ICLR 2024 paper "When should we prefer Decision Transformers for Offline Reinforcement Learning?"☆17Jan 31, 2024Updated 2 years ago
- AISTATS 2019: Reference-based Adversarial Sampling & Its applications to Soft Q-learning☆15Jan 21, 2019Updated 7 years ago
- PAIRED in PyTorch 🔥☆65Mar 8, 2023Updated 3 years ago
- A simple RNN meta-learner☆10Dec 17, 2018Updated 7 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- PyTorch Implementation of "NDDR-CNN: Layerwise Feature Fusing in Multi-Task CNNs by Neural Discriminative Dimensionality Reduction"☆14Jun 29, 2019Updated 7 years ago
- Minimizing Control for Credit Assignment with Strong Feedback☆14Nov 3, 2024Updated last year
- DSTC8-AVSD: Sentence generation task for Audio Visual Scene-aware Dialog☆14Jun 10, 2021Updated 5 years ago
- EVM in python from scratch because why not☆11Aug 22, 2022Updated 4 years ago
- Code for [NeurIPS'2019 Spotlight] Policy Continuation with Hindsight Inverse Dynamics☆15Jan 7, 2020Updated 6 years ago
- [ICLR'20] Learning to Learn by Zeroth-Order Oracle☆14Feb 7, 2020Updated 6 years ago
- Code for the paper Watch-And-Help: A Challenge for Social Perception and Human-AI Collaboration☆106Jul 15, 2022Updated 4 years ago