☆17Aug 20, 2025Updated 11 months ago
Alternatives and similar repositories for weight-clipping
Users that are interested in weight-clipping are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- streaming deep reinforcement learning but 4x faster with jax!☆19Jan 4, 2026Updated 7 months ago
- Vintix II: Decision Pre-Trained Transformer is a Scalable In-Context Reinforcement Learner - - = ICLR 2026☆16Apr 8, 2026Updated 4 months ago
- ☆16Jul 16, 2024Updated 2 years ago
- Object-Centric-Representation Library (OCRL): This repo is to explore OCR on various downstream tasks from supervised learning tasks to R…☆12Feb 23, 2024Updated 2 years ago
- ☆10Jun 27, 2024Updated 2 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Code for the paper "FIRE: Frobenius-Isometry Reinitialization for Balancing the Stability–Plasticity Tradeoff" (ICLR 2026 Oral)☆30Apr 27, 2026Updated 3 months ago
- A fast and robust algorithm for temporal difference learning☆25Mar 16, 2026Updated 4 months ago
- A lightweight PyTorch implementation of the Transformer-XL architecture proposed by Dai et al. (2019)☆37Feb 7, 2023Updated 3 years ago
- COOM: Benchmarking Continual Reinforcement Learning on Doom☆27Mar 5, 2026Updated 5 months ago
- [ICML 2026 GenBio Workshop] Official Implementation for "Harmonic Torsional Diffusion for Protein-Ligand Flexible Docking"☆15Jun 30, 2026Updated last month
- Random parameter environments using gym 0.7.4 and mujoco-py 0.5.7☆20Feb 14, 2019Updated 7 years ago
- ☆19Apr 22, 2024Updated 2 years ago
- Official Implementation of NeurIPS'23 Paper "Cross-Episodic Curriculum for Transformer Agents"☆32Oct 12, 2023Updated 2 years ago
- Code of the Paper "Time-Efficient Reinforcement Learning with Stochastic Stateful Policies"☆25May 5, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 🔀 yet another mixture of experts☆23Jun 5, 2026Updated 2 months ago
- A scalable benchmark for state representation learning in visual reinforcement learning.☆17Jun 23, 2025Updated last year
- High-performance JAX-powered simulator for robotic navigation in 2D mazes, optimized for Quality-Diversity algorithm research and benchma…☆21Jun 19, 2025Updated last year
- Gym wrapper for pysc2☆10Sep 16, 2022Updated 3 years ago
- [ICML 2025 GenBio Workshop] Official Implementation for "Electrostatics from Laplacian Eigenbasis for Neural Network Interatomic Potentia…☆18Jun 12, 2025Updated last year
- Mitigating Partial Observability in Sequential Decision Processes via the Lambda Discrepancy☆24Oct 28, 2024Updated last year
- Next-gen Foundation Model for Embodied AI☆33Apr 7, 2026Updated 4 months ago
- Standardized Minecraft Diamond Environment for Reinforcement Learning☆40May 19, 2023Updated 3 years ago
- Official codebase for "Sampling For Learnability", published at NeurIPS 2024☆25Oct 21, 2025Updated 9 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- MR.Q is a general-purpose model-free reinforcement learning algorithm.☆154Apr 7, 2026Updated 4 months ago
- reinforcement learning for bridge☆23Jul 25, 2024Updated 2 years ago
- Reinforcement Learning inside a 3D soccer simulation☆37Sep 15, 2024Updated last year
- High quality implementations of imitation and inverse reinforcement learning algorithms☆24Aug 19, 2025Updated 11 months ago
- Jax like function transformation engine but micro, microjax☆34Oct 25, 2024Updated last year
- Atari-style POMDPs☆34Aug 4, 2026Updated last week
- Multi-agent simulator in Jax for research and teaching in AI & ALife☆32Apr 11, 2026Updated 4 months ago
- Official Pytorch Implementation of "Zero-Shot Off-Policy Learning" (ICML 2026)☆25Feb 16, 2026Updated 5 months ago
- A dataloader, but for JAX☆20May 17, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Fast and memory efficient PyTorch implementation of the Perceiver with FlashAttention.☆32Nov 4, 2024Updated last year
- Convertible environment for reinforcement learning with Kerbal Space Program☆11Dec 8, 2022Updated 3 years ago
- ☆24Aug 9, 2024Updated 2 years ago
- [ICLR 2024] Closing the Gap between TD Learning and Supervised Learning - A Generalisation Point of View.☆25Apr 19, 2024Updated 2 years ago
- Measuring the Signal to Noise Ratio in Language Model Evaluation☆31Aug 19, 2025Updated 11 months ago
- VC-FB and MC-FB algorithms from "Zero-Shot Reinforcement Learning from Low Quality Data" (NeurIPS 2024)☆29Jan 14, 2025Updated last year
- ☆10Jul 14, 2018Updated 8 years ago