unofficial code reproducing Agent57
☆40Mar 9, 2024Updated 2 years ago
Alternatives and similar repositories for agent57_pytorch
Users that are interested in agent57_pytorch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A collection of Deep Reinforcement Learning algorithms implemented with PyTorch to solve Atari games and classic control tasks like CartP…☆125Feb 21, 2024Updated 2 years ago
- Distributed & asynchronous DQN implementation using gRPC and PyTorch.☆10Feb 15, 2021Updated 5 years ago
- Author implementation of Monte Carlo Augmented Actor Critic in PyTorch☆18Oct 24, 2022Updated 3 years ago
- 根据《电子科技大学研究生学位论文(研究报告)撰写格式规范》修改的zotero引用style文件☆10Dec 2, 2021Updated 4 years ago
- A2C training of Relational Deep Reinforcement Learning Architecture☆13Jun 22, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Exploration by Random Network Distillation☆15Dec 30, 2018Updated 7 years ago
- ☆11Oct 3, 2022Updated 3 years ago
- PyTorch implementation of "Sample-efficient Imitation Learning via Generative Adversarial Nets"☆10Nov 22, 2019Updated 6 years ago
- TensorFlow implementation of "Sample-efficient Imitation Learning via Generative Adversarial Nets"☆10Dec 8, 2022Updated 3 years ago
- Codebase for "Causal Induction from Visual Observations for Goal-Directed Tasks"☆14Feb 25, 2020Updated 6 years ago
- Avenue is a simulator designed to test and prototype reinforcement learning algorithms. Avenue is a ServiceNow Research project that was …☆14Jul 15, 2022Updated 4 years ago
- slowly building a set of infinite riddle generators for data-hungry methods☆14Nov 15, 2022Updated 3 years ago
- G-HER algorithm☆18May 24, 2019Updated 7 years ago
- metaTextGrad: Automatically optimizing language model optimizers. Published in NeurIPS 2025.☆15Nov 5, 2025Updated 10 months ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Curiosity-driven Exploration by Self-supervised Prediction☆148Mar 12, 2023Updated 3 years ago
- A mirror of the Open Risk white paper collection☆10Updated this week
- Source code for the paper "Positional Attention: Expressivity and Learnability of Algorithmic Computation"☆14May 26, 2025Updated last year
- Official implementation of Neural Episodic Control with State Abstraction☆13Aug 3, 2023Updated 3 years ago
- Implementation of Neurips 2023 Paper "Multi Time Scale World Models"☆22Nov 8, 2024Updated last year
- Implementation of spiking DQN training using different conversion techniques and backpropagation with surrogate gradients employed on the…☆11Feb 11, 2023Updated 3 years ago
- MuJoCo benchmark for Deep Reinforcement Learning as provided by Tianshou framework.☆14Jan 12, 2025Updated last year
- ☆13Mar 18, 2020Updated 6 years ago
- ☆15Jun 30, 2019Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Modified Fetch Robotics environments from OpenAI gym☆11Nov 27, 2021Updated 4 years ago
- curriculum☆27Feb 7, 2023Updated 3 years ago
- This is the code for my published paper: Improved Reinforcement Learning through Imitation Learning Pretraining Towards Image-based Auton…☆17Apr 28, 2021Updated 5 years ago
- Replication package for ESEC/FSE-2019 submission titled Diversity Web Test Generation☆15Feb 13, 2025Updated last year
- CoOP: V2V-based Cooperative Overtaking for Platoons on Freeways☆10Oct 29, 2021Updated 4 years ago
- The implement of GAIL with pytorch☆14Mar 11, 2020Updated 6 years ago
- [NeuIPS2024 DTQL] Diffusion Trusted Q-Learning for Offline RL — Official PyTorch Implementation☆28May 31, 2024Updated 2 years ago
- Simple but Useful Layers based on Tensorflow☆14Mar 29, 2020Updated 6 years ago
- Simulation code of paper "H. Zhu, Z. Wang, F. Yang, Y. Zhou and X. Luo, "Intelligent Traffic Network Control in the Era of Internet of Ve…☆11Mar 30, 2022Updated 4 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Linear Relational Embeddings (LREs) and Linear Relational Concepts (LRCs) for LLMs in PyTorch☆11Aug 7, 2024Updated 2 years ago
- ☆13Sep 18, 2024Updated last year
- Code files for HoloLens Beginner's Guide by Packt☆12Jan 30, 2023Updated 3 years ago
- 技術書典2で出店する書籍の紹介ページです☆11Apr 9, 2017Updated 9 years ago
- JQuantLib is a free, open-source, comprehensive framework for quantitative finance, written in 100% Java.☆10Aug 10, 2012Updated 14 years ago
- Multi-Agent Context Learning (MACOL): A new machine learning algorithm for multi-agent cooperation in competing environment☆13Sep 25, 2024Updated last year
- Code of Paper "Vehicular Fog Computing Enabled Real-time Collision Warning via Trajectory Calibration", MONET, 2019.☆10Oct 21, 2022Updated 3 years ago