A plug-in debugger and visualizer for RL reward functions. Detects reward hacking, tracks training health, and renders a live terminal dashboard.
☆80Jul 7, 2026Updated 3 weeks ago
Alternatives and similar repositories for rewardspy
Users that are interested in rewardspy are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆11Dec 13, 2013Updated 12 years ago
- Estimating barycenters of distributions by Neural Optimal Transport (ICML 2024)☆21Mar 6, 2025Updated last year
- High-performance batched Top-K selection for CPU inference. Up to 80x faster than PyTorch, optimized for LLM sampling with AVX2 SIMD.☆18Mar 20, 2026Updated 4 months ago
- ☆14Jun 29, 2023Updated 3 years ago
- The official repository for the paper "Statler: State-Maintaining Language Models for Embodied Reasoning"☆13Jun 10, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [ICML 2025] Official PyTorch implementation of "NegMerge: Sign-Consensual Weight Merging for Machine Unlearning"☆16Nov 25, 2025Updated 8 months ago
- Kernel-level security & attack response for Linux servers.☆15Jun 12, 2026Updated last month
- Лекции по алгоритмам и структуре данных☆10Dec 20, 2023Updated 2 years ago
- Tutorials on how to use EAGERx☆16Aug 14, 2025Updated 11 months ago
- A scanner for DKIM selectors.☆14Nov 5, 2018Updated 7 years ago
- 🏗 forkable Ethereum dev stack focused on fast product iterations☆10Aug 28, 2021Updated 4 years ago
- Official inference implementation of the paper "DON'T SETTLE TOO EARLY: SELF-REFLECTIVE REMASKING FOR DIFFUSION LANGUAGE MODELS". [ICLR 2…☆15Jan 28, 2026Updated 6 months ago
- The original code for the data providers and the datasets of the paper "Defining Benchmarks for Continual Few-Shot Learning".☆16Apr 15, 2020Updated 6 years ago
- The CocoNuts project is an open-source humanoid robotics research platform. Its primary design philosophy is the minimization of leg iner…☆23Jun 30, 2026Updated last month
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Stateless LLM runtime that dynamically routes, loads, executes, and unloads models per request with bounded VRAM caching and intelligent …☆41Apr 12, 2026Updated 3 months ago
- Rust derive macros for automating the boring stuff.☆14Aug 3, 2025Updated last year
- World's first zk explorer☆10Nov 12, 2023Updated 2 years ago
- ☆10Jan 23, 2025Updated last year
- [EMNLP 2025] The official implementation of "Zero-shot Multimodal Document Retrieval via Cross-Modal Question Generation"☆15Aug 26, 2025Updated 11 months ago
- Turing machine ZKVM☆10Nov 12, 2023Updated 2 years ago
- Bleeding edge low level Rust binding for GGML☆18Jun 26, 2024Updated 2 years ago
- A zero-knowledge virtual machine with only one instruction☆13Mar 3, 2022Updated 4 years ago
- ☆13Updated this week
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆14Oct 31, 2023Updated 2 years ago
- Rust library for ZKBoogie, ZKBoo-based ZKP scheme optimized for client-side proving and its recursion.☆11Jun 5, 2024Updated 2 years ago
- ☆11Dec 26, 2018Updated 7 years ago
- Rex is a JAX-powered framework for sim-to-real robotics.☆53Jun 11, 2025Updated last year
- [ICML 2026] Code for Equilibrium Reasoners: learning attractor dynamics for scalable reasoning☆45Jun 1, 2026Updated 2 months ago
- Verify DomainKeys Identified Mail signatures☆10Jan 24, 2023Updated 3 years ago
- Project page for the RSS 2025 paper Resolving Conflicting Constraints in Multi-Agent Reinforcement Learning with Layered Safety☆16Aug 7, 2025Updated 11 months ago
- Developing, training, and assessing the performance of a Proximal Policy Optimization (PPO) Stock Trading Agent.☆14Aug 20, 2025Updated 11 months ago
- A tiny package supporting distributed computation of COCO metrics for PyTorch models.☆15Feb 28, 2023Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Alternate LLM Architecture Experiment with O(n) time complexity☆11Jul 18, 2026Updated 2 weeks ago
- Tools for quick-and-dirty comparisons of popular robotics simulators☆14Aug 8, 2025Updated 11 months ago
- ☆10Sep 15, 2023Updated 2 years ago
- Reinforcement Learning framework for Bipedal robots using the OpenAI Gym environment and PyBullet☆12Jan 5, 2020Updated 6 years ago
- The repository is for Reinforcement-Learning Uncertainty research, in which we investigate various uncertain factors in RL.☆23Jun 16, 2023Updated 3 years ago
- ☆13Nov 10, 2023Updated 2 years ago
- Machine Learning framework built from scratch in CPP. Uses tensors for base operations.☆16May 23, 2026Updated 2 months ago