Proximal Policy Optimization(PPO) with Intrinsic Curiosity Module(ICM)
☆18Apr 15, 2022Updated 4 years ago
Alternatives and similar repositories for icmppo
Users that are interested in icmppo are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆17Dec 23, 2024Updated last year
- Émulateur Dofus 1.29.1 en Java☆14Dec 5, 2016Updated 9 years ago
- Connect 4 AI using Monte Carlo Tree Search algorithm.☆11Feb 10, 2024Updated 2 years ago
- Model-based Policy Gradients☆32Mar 12, 2020Updated 6 years ago
- The Laser Learning Environment (LLE) is a cooperative MARL grid-world☆13Updated this week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆17Oct 18, 2022Updated 3 years ago
- A clean and minimal implementation of PPO (Proximal Policy Optimization) algorithm in Pytorch, for continuous action spaces.☆20Jan 3, 2023Updated 3 years ago
- Modified versions of the Soft Actor-Critic algorithm for Atari games from https://github.com/ac-93/soft-actor-critic.☆20May 18, 2020Updated 6 years ago
- ☆13Dec 12, 2022Updated 3 years ago
- This repository is an implementation of "MASER: Multi-Agent Reinforcement Learning with Subgoals Generated from Experience Replay Buffer"…☆23Jul 6, 2023Updated 3 years ago
- ☆28Jun 16, 2026Updated 2 months ago
- ☆20Jan 9, 2025Updated last year
- ☆18Oct 6, 2021Updated 4 years ago
- Memory Augmented Neural Networks (Pytorch)☆14Sep 2, 2018Updated 8 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Epi: An Open Humanoid Platform☆17Jun 18, 2023Updated 3 years ago
- M.Sc Thesis: Robotic Navigation under Partial Observability with Actor-Critic Methods DDPG, SAC, PPO. Environment, Lidar, and Kinematic M…☆20Jun 9, 2025Updated last year
- Code for simulations in "Computational mechanisms of curiosity and goal-directed exploration"☆11May 22, 2020Updated 6 years ago
- ☆12Dec 8, 2022Updated 3 years ago
- Build a Responsive Calendar App with HTML, CSS and Javascript | Tutorial 2024☆20Oct 11, 2024Updated last year
- Transcribing long blocks of speech using Watson Speech To Text.☆11Sep 24, 2020Updated 5 years ago
- Promoss Topic Modelling Toolbox☆11Jan 21, 2019Updated 7 years ago
- 🤣咦?好像这所学校也不错>v<☆17Sep 25, 2020Updated 5 years ago
- ☆11Mar 9, 2018Updated 8 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code for the paper "Unlocking Slot Attention by Changing Optimal Transport Costs"☆13Sep 19, 2023Updated 2 years ago
- Reinforcement Learning papers on exploration methods.☆19Jun 27, 2021Updated 5 years ago
- Codes for the paper "HAVEN: Hierarchical Cooperative Multi-Agent Reinforcement Learning with Dual Coordination Mechanism"☆27Oct 22, 2022Updated 3 years ago
- Exploring techniques to generate diverse conventions in multi-agent settings☆16Nov 14, 2023Updated 2 years ago
- Python implement of paper "PD-FAC: Probability Density Factorized Multi-Agent Distributional Reinforcement Learning for Multi-Robot Relia…☆12Mar 5, 2022Updated 4 years ago
- 检测视频中振动结构的振动频率,像素振幅。可检测亚像素级别振动。☆23Mar 14, 2019Updated 7 years ago
- project that aims to run on raspberry pi and take voice commands and answering them using chatgpt api☆19May 24, 2023Updated 3 years ago
- COOM: Benchmarking Continual Reinforcement Learning on Doom☆27Mar 5, 2026Updated 5 months ago
- PaCMAP in pure MLX for Apple Silicon. Pure GPU, no scipy/numba.☆21Mar 5, 2026Updated 5 months ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- A Reinforcement Learning agent in Diablo environment☆28Aug 25, 2026Updated last week
- ☆13Jun 18, 2024Updated 2 years ago
- Code release for the paper "Goal Representations for Instruction Following: A Semi-Supervised Language Interface to Control"☆17Apr 9, 2024Updated 2 years ago
- Proximal Policy Optimization (Continuous Version) in PyTorch.☆28May 12, 2025Updated last year
- A Semi-Supervised VAE Based Active Anomaly Detection Framework in Multivariate Time Series for Online Systems☆26Feb 15, 2023Updated 3 years ago
- [ICML2023] Instant Soup Cheap Pruning Ensembles in A Single Pass Can Draw Lottery Tickets from Large Models. Ajay Jaiswal, Shiwei Liu, Ti…☆11Nov 28, 2023Updated 2 years ago
- neuralpy - neural network library written in python☆12Jun 25, 2023Updated 3 years ago