PyTorch implementation of both discrete and continuous ACER
☆25Jan 27, 2019Updated 7 years ago
Alternatives and similar repositories for acer
Users that are interested in acer are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆15Oct 29, 2018Updated 7 years ago
- Actor-critic with experience replay☆257Oct 9, 2022Updated 3 years ago
- Implement Conditional VAE and train on MNIST by tensorflow 1.3.0.☆10Nov 7, 2017Updated 8 years ago
- ☆13Mar 26, 2019Updated 7 years ago
- paper <<Hierarchical Deep Reinforcement Learning: Integrating Temporal Abstraction and Intrinsic Motivation>> python implementation☆10Mar 27, 2018Updated 8 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Reinforcement learning training project for a SLG game☆13Dec 21, 2017Updated 8 years ago
- Implementation of point-based value iteration (for POMDPs)☆12Mar 31, 2020Updated 6 years ago
- ☆12May 21, 2017Updated 9 years ago
- PyTorch implementation of Sample Efficient Actor-Critic with Experience Replay(ACER)☆16Oct 7, 2020Updated 5 years ago
- ☆15Sep 19, 2017Updated 9 years ago
- General implementation of Advantage Actor Critic using Pytorch☆28Dec 7, 2021Updated 4 years ago
- This repo is "NTHU Parallel Programing" course project.☆10Dec 5, 2017Updated 8 years ago
- Framework for generating adversarial examples using formal methods and for analyzing robustness of DNNs.☆21Jul 21, 2017Updated 9 years ago
- [NeurIPS, 2020 - Reproducibility Challenge]: [RE] Towards Interpretable Reinforcement Learning Using Attention Augmented Agents☆13Apr 26, 2021Updated 5 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- This is the source code for our (Matthias Jasny, Lasse Thostrup, Tobias Ziegler and Carsten Binnig) published paper at SIGMOD’22: P4DB - …☆14Jan 24, 2023Updated 3 years ago
- Minimal working example of deployment of a jupyter notebook using voila, jupyter/docker-stacks and nginx☆12Jul 3, 2020Updated 6 years ago
- Implementation of PPO in Pytorch☆41Dec 6, 2017Updated 8 years ago
- A hash work for SBIR☆20Oct 9, 2018Updated 7 years ago
- [AAMAS 2023] Code for the paper "Automatic Noise Filtering with Dynamic Sparse Training in Deep Reinforcement Learning"☆12Feb 22, 2024Updated 2 years ago
- Source code for "A deep dive into reinforcement learning"☆13Dec 17, 2019Updated 6 years ago
- PyTorch implementation of Trust Region Policy Optimization☆450Sep 13, 2018Updated 8 years ago
- RDMA programming examples using Soft-RoCE☆14Aug 13, 2021Updated 5 years ago
- Implementation of Deepmind's LaserTag-v0 game in A Unified Game-Theoretic Approach to Multiagent Reinforcement Learning(2017)☆20Nov 30, 2018Updated 7 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆15Oct 6, 2019Updated 6 years ago
- my public website☆12Jul 8, 2026Updated 2 months ago
- ☆13Jan 23, 2021Updated 5 years ago
- An implementation of DreamerV2 written in JAX, with support for running multiple random seeds of an experiment on a single GPU.☆18Jan 16, 2023Updated 3 years ago
- Material de apoyo de la Materia Análisis Predictivo de la Licenciatura en Analítica (Data Science) ITBA☆21Aug 14, 2026Updated last month
- Ilúvatar is an open Serverless platform built with the goal of jumpstarting and streamlining FaaS research. It provides a system that is …☆26Sep 4, 2026Updated 2 weeks ago
- Pytorch2Jax is a small Python library that provides functions that wraps PyTorch models into Jax functions and Flax modules.☆21Feb 20, 2023Updated 3 years ago
- Use pytorch the right way http://pytorch.org/docs/☆22Nov 15, 2017Updated 8 years ago
- Docker for the minerl gym environment with Jupyter, PyTorch and CUDA drivers installed☆21Aug 15, 2019Updated 7 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- An implementation of TRPO with GAE in PyTorch☆16Jul 22, 2023Updated 3 years ago
- ☆12May 8, 2025Updated last year
- A small package used to visualize gradient descent of test functions.☆19Aug 1, 2022Updated 4 years ago
- Distributed Graph Mining on a Massive "Single" Graph☆15Mar 28, 2020Updated 6 years ago
- Supporting material for Princeton ORF522☆15Aug 27, 2025Updated last year
- Collection of reinforcement learning algorithms☆16Sep 29, 2025Updated 11 months ago
- Source code of "PathEnum: Towards Real-Time Hop-Constrained s-t Path Enumeration", published in SIGMOD'2021 - By Shixuan Sun, Yuhang Chen…☆17Mar 23, 2021Updated 5 years ago