A technical exploration comparing standard deep RL (PPO) against biologically plausible learning rules on a custom Pong environment. Everything is implemented from scratch with no Stable-Baselines, no Gymnasium, no pre-built algorithms.
β29May 19, 2026Updated 3 months ago
Alternatives and similar repositories for Biologically-Plausible-RL-Plays-Pong
Users that are interested in Biologically-Plausible-RL-Plays-Pong are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- π¦ The World's First Sensor-Agnostic Tactile Data Annotation Toolkit β Load any tactile sensor, annotate visually, export a unified schemβ¦β44Updated this week
- A prototype implementation of the "dataset as a queue" pattern for processing web pages into interleaved image/text content.β30Nov 16, 2025Updated 9 months ago
- Implementation of the fast weight product key memory from Sakana AIβ20Aug 19, 2026Updated 3 weeks ago
- Synchronized Curriculum Learning for RL Agentsβ123Jul 31, 2026Updated last month
- Implementation and explorations into Blackbox Gradient Sensing (BGS), an evolutionary strategies approach proposed in a Google Deepmind pβ¦β20Apr 17, 2026Updated 4 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Implementation of Kalmanformer, modeling the Kalman gain with a transformerβ46Jun 2, 2026Updated 3 months ago
- Open source code combining implementations of Upside Down Reinforcement Learning and Reward Conditioned Policiesβ19Mar 10, 2021Updated 5 years ago
- Implementation of Fast Weight Attentionβ34Aug 20, 2026Updated 3 weeks ago
- Stealth LLM inference engineβ40Updated this week
- Implementation of GenMimic, "From Generated Human Videos to Physically Plausible Robot Trajectories"β18Sep 1, 2026Updated last week
- β13Oct 19, 2023Updated 2 years ago
- Score-Based Diffusion Policy Compatible with Reinforcement Learning via Optimal Transportβ15Feb 26, 2025Updated last year
- Score-based Diffusion models in JAX.β19Dec 29, 2025Updated 8 months ago
- Personal websiteβ16Jul 22, 2026Updated last month
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- An implementation of AlphaZero and MCTS with neural networks for Tetrisβ22Jul 7, 2026Updated 2 months ago
- Self contained pytorch implementation of a sinkhorn based router, for mixture of experts or otherwiseβ40Aug 29, 2024Updated 2 years ago
- code for paper "Entropy-regularized Diffusion Policy with Q-Ensembles for Offline Reinforcement Learning"β21Feb 24, 2024Updated 2 years ago
- Self-Expanding Neural Networksβ44Feb 9, 2024Updated 2 years ago
- redux es5β12May 12, 2016Updated 10 years ago
- This is a high performance stub server.β14Sep 3, 2024Updated 2 years ago
- β13Jul 6, 2020Updated 6 years ago
- Generic MCP Client to use any MCP tool in a chatβ44May 11, 2025Updated last year
- Reinforcement Learning example in Nim, playing tic tac toe. Based off original C version from the great Antirezβ15Apr 2, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Official repo for Offline RL for Online RLβ19Oct 14, 2023Updated 2 years ago
- β28Jul 14, 2024Updated 2 years ago
- β10Sep 13, 2021Updated 5 years ago
- Official implementation of the paper: "ZClip: Adaptive Spike Mitigation for LLM Pre-Training".β153Jul 20, 2026Updated last month
- Rust implementation of the Loopy Belief propagation algorithm for inference in Bayesian Networksβ36Mar 13, 2022Updated 4 years ago
- Official implementation of HEAD CoRL 2025β18Aug 9, 2025Updated last year
- un nnUnet on Windowsβ13Feb 27, 2024Updated 2 years ago
- Python wrapper for the Universal Robots ROS driver with Robotiq gripper support.β14May 21, 2024Updated 2 years ago
- Probabilistic inference for models of behaviourβ13Mar 5, 2026Updated 6 months ago
- Simple, predictable pricing with DigitalOcean hosting β’ AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- [IROS2023]Learning to Solve Tasks with Exploring Prior Behavioursβ13Mar 3, 2024Updated 2 years ago
- β17Mar 10, 2026Updated 6 months ago
- Repository for the code of the paper "Neural Networks Regularization Through Class-wise Invariant Representation Learning".β12Oct 1, 2017Updated 8 years ago
- Minimal code for extracting structured Insights from Sustainability Reports via Large Language Modelsβ12Jul 9, 2025Updated last year
- Sequential Monte Carlo sampler for PyMC2 models.β14Apr 4, 2018Updated 8 years ago
- open-arms-mini: cheap human like teleoperation device that supports human in the loop correctionsβ129Mar 26, 2026Updated 5 months ago
- Official code for On Path Integration of Grid Cells: Group Representation and Isotropic Scaling (NeurIPS 2021)β53Nov 10, 2021Updated 4 years ago