Target Policy Optimization (JAX)
☆31Apr 18, 2026Updated 4 months ago
Alternatives and similar repositories for tpo
Users that are interested in tpo are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- RL models to play Sokoban. The fastest recipe wins.☆31Jul 25, 2026Updated last month
- ☆30Jan 31, 2026Updated 7 months ago
- A minimal hackable implementation of policy gradient methods (GRPO, PPO, REINFORCE)☆17Feb 20, 2026Updated 6 months ago
- Collection of LLM completions for reasoning-gym task datasets☆31Jul 4, 2025Updated last year
- [EMNLP2022] Transformer-based Entity Typing in Knowledge Graphs☆15Nov 26, 2024Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Follow the Mean: controlling flow-matching generative models by shifting endpoint means with reference examples☆25May 21, 2026Updated 3 months ago
- Official code for "How2Everything: Mining the Web for How-To Procedures to Evaluate and Improve LLMs"☆25Feb 10, 2026Updated 7 months ago
- Official implementation of Stackelberg PPO for morphology–control co-design.☆18Mar 17, 2026Updated 5 months ago
- Run autonomous AI agents across machines using a distributed filesystem.☆17Apr 3, 2026Updated 5 months ago
- An implementation of ESM2 in Equinox+JAX☆36Apr 20, 2026Updated 4 months ago
- ☆18Mar 10, 2026Updated 6 months ago
- [arxiv: 2604.14142] From P(y|x) to P(y): Investigating Reinforcement Learning in Pre-train Space☆17Apr 16, 2026Updated 4 months ago
- Propose CDR single-point mutations that rescue antibody expression without disrupting antigen binding, using bound vs. unbound ProteinMPN…☆15May 26, 2026Updated 3 months ago
- [ICML 24 NGSM workshop] Associative Recurrent Memory Transformer implementation and scripts for training and evaluation☆68Aug 14, 2026Updated 3 weeks ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Brain Interpreter and Visualizer Online.☆10Sep 1, 2016Updated 10 years ago
- Gym wrapper for pysc2☆10Sep 16, 2022Updated 3 years ago
- ☆15May 19, 2024Updated 2 years ago
- ☆25Jan 28, 2026Updated 7 months ago
- ☆51Aug 19, 2026Updated 3 weeks ago
- Rethinking the Trust Region in LLM Reinforcement Learning☆74Mar 2, 2026Updated 6 months ago
- A simple implementation of ReasonGenRM.☆19Apr 21, 2025Updated last year
- Decoupled Q-Chunking☆75May 3, 2026Updated 4 months ago
- This repositories contains the reference implementation for the Sparse Delta Memory paper.More precisely, it contains the model definitio…☆38Jul 9, 2026Updated 2 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Parallelized Cross Entropy Method☆14Jul 26, 2023Updated 3 years ago
- ☆12Mar 31, 2024Updated 2 years ago
- Symbol-Equivariant Recurrent Reasoning Model☆18Mar 4, 2026Updated 6 months ago
- [CVPR 2026 Findings] V-GRPO: Online Reinforcement Learning for Denoising Generative Models Is Easier than You Think☆56Apr 28, 2026Updated 4 months ago
- This repository is a package to provide SAIT Machine Learning Force Field(MLFF) Framework☆40Oct 25, 2023Updated 2 years ago
- Framework for modified sampling from biomolecular generative models☆27Updated this week
- World-Gymnast: Training Robots with Reinforcement Learning in a World Model☆52Feb 11, 2026Updated 7 months ago
- A JAX-native High Performance Eval Metrics Library☆64Aug 31, 2026Updated last week
- ☆10Jul 14, 2018Updated 8 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Material for tutorial "Hybrid techniques for knowledge-based NLP: Knowledge graphs meet machine learning and all their friends" at KCAP…☆16Dec 4, 2017Updated 8 years ago
- This repository enables inference and sampling for Dyno Psi-1, a de novo miniprotein binder design model.☆24Aug 25, 2026Updated 2 weeks ago
- Archives for Triton Inference Server Practices☆15Feb 28, 2022Updated 4 years ago
- Recommendation models that use binary rather than floating point operations at prediction time.☆21Sep 18, 2017Updated 8 years ago
- experiment☆12Jan 1, 2023Updated 3 years ago
- Chroma Protein Diffusion Model☆12Nov 16, 2023Updated 2 years ago
- ☆13Sep 24, 2024Updated last year