a rubric driven prioritized replay rl algo to maximise continual learning
☆16Oct 12, 2025Updated 9 months ago
Alternatives and similar repositories for infinite
Users that are interested in infinite are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- rl from zero pretrain, can it be done? yes.☆295Sep 28, 2025Updated 10 months ago
- ☆26Jan 22, 2026Updated 6 months ago
- An extention to the GaLore paper, to perform Natural Gradient Descent in low rank subspace☆19Oct 21, 2024Updated last year
- Safe Model-Based RL HVAC Control Using Epistemic Uncertainty Estimation.☆13Feb 25, 2025Updated last year
- Beating OpenZeppelin's Ethernaut in Pure Assembly. Masochists Only.☆24May 9, 2023Updated 3 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆11Apr 5, 2024Updated 2 years ago
- Notes from Andrew Blumberg's class on equivariant homotopy theory☆17Aug 19, 2025Updated 11 months ago
- Dynamic Telegram Trading Bot☆19Feb 21, 2025Updated last year
- a platform for monitoring the chip situation☆16Jul 19, 2025Updated last year
- This repository contains code for the paper "Learning Decision Trees as Amortized Structure Inference"☆16Mar 25, 2025Updated last year
- Score-Based Diffusion Policy Compatible with Reinforcement Learning via Optimal Transport☆15Feb 26, 2025Updated last year
- ☆12Oct 7, 2020Updated 5 years ago
- ☆14Sep 17, 2020Updated 5 years ago
- Play YouTube videos directly in your terminal with synchronized audio using ASCII rendering or ANSI truecolor.☆17Jul 23, 2026Updated 2 weeks ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Convert CVXPY expressions to PyTorch expressions☆18Jul 8, 2025Updated last year
- Powered by ChatterboxTTS | Transformer | Llama | Gradio☆18Sep 7, 2025Updated 11 months ago
- The official code implementation of the Autodiff algorithm.☆16Nov 10, 2023Updated 2 years ago
- Custom triton kernels for training Karpathy's nanoGPT.☆19Oct 21, 2024Updated last year
- Lab Cookbook☆41Updated this week
- The open source implementation of "AnyMAL: An Efficient and Scalable Any-Modality Augmented Language Model"☆22Jan 27, 2025Updated last year
- My Solutions to Sutton and Barto exercises, 2nd edition☆14Apr 27, 2018Updated 8 years ago
- Ludic – an LLM-RL library for the era of experience☆68Updated this week
- [ICML 2024] The algorithm of Reinforcement Learning with an Assistant Reward Agent (ReLara)☆18Aug 2, 2024Updated 2 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Subscriptions using Porto.☆28Jul 5, 2025Updated last year
- A framework for optimizing DSPy programs with RL☆341Jan 12, 2026Updated 6 months ago
- ☆32Nov 11, 2025Updated 8 months ago
- Generalized Scientific Compute on BIttensor☆30Sep 29, 2025Updated 10 months ago
- Sample PySpark code for interacting with the Microsoft Academic Graph☆22Mar 12, 2021Updated 5 years ago
- [NeurIPS 2022] Official codebase for "Meta-Reward-Net: Implicitly Differentiable Reward Learning for Preference-based Reinforcement Learn…☆26Feb 15, 2025Updated last year
- ☆29Dec 20, 2025Updated 7 months ago
- Write a fast kernel and see how you compare against the best humans and AI on gpumode.com☆109Jul 29, 2026Updated last week
- ☆11Dec 23, 2023Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Use Hermes-2-Pro-Mistral-7B function calling with your OpenAI API compatible code.☆18May 7, 2024Updated 2 years ago
- Deformable DETR in Pytorch Lightning☆12Sep 15, 2021Updated 4 years ago
- Logic architecture inspired by Toulmin. Forces LLMs into structured, sequential reasoning through Toulmin's argumentation model.☆17Dec 2, 2025Updated 8 months ago
- A playground for Uniswap v4 hooks☆30Sep 28, 2023Updated 2 years ago
- Action Value Gradient Algorithm☆29May 18, 2025Updated last year
- CCLoco: Scaling Up Top-K Error Feedback with Local Optimizers☆27Aug 22, 2025Updated 11 months ago
- ☆27Mar 25, 2026Updated 4 months ago