SGLang model provider of Strands Agents for on-policy agentic RL training.
☆75Jul 18, 2026Updated last week
Alternatives and similar repositories for strands-sglang
Users that are interested in strands-sglang are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A framework for building agent environments for RL training and evaluation with Strands Agents.☆57Updated this week
- [ICML2026] Reproduce Kimi K1.5/K2 RL algorithm and rollout system☆19Apr 9, 2026Updated 3 months ago
- A Multi-Policy, Multi-Agent RL Training Framework☆30Jun 16, 2026Updated last month
- [NeurIPS 2025] Think-RM: Enabling Long-Horizon Reasoning in Generative Reward Models☆17Nov 2, 2025Updated 8 months ago
- Dataflow-Oriented Reinforcement Learning for (Multi-)Agentic LLMs☆97Updated this week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Evolve-AI☆40Feb 7, 2025Updated last year
- [ICML'26] Beyond Test-Time Memory: State-Space Optimal Control for LLM Reasoning☆15Jun 1, 2026Updated last month
- APRIL: Active Partial Rollouts in Reinforcement Learning to Tame Long-tail Generation. A system-level optimization for scalable LLM tra…☆60Oct 11, 2025Updated 9 months ago
- 💻 SETA: Scaling Environments for Terminal Agents☆127Jul 17, 2026Updated last week
- A lightweight post-training framework for LLMs and VLMs. 51 algorithms, 38 verified models. Scales with DeepSpeed, vLLM, and Ray.☆19Updated this week
- CI-native agent CLI tool for deterministic pipeline gating.☆79Updated this week
- Toolathlon-Gym for testing AI agents real-world tool-use capabilities across diverse MCP servers.☆140Updated this week
- Yushio (夕潮) — An AI collaborator persona for Claude Code and beyond. Three layered skills: base + art director + code auditor. Distilled …☆217Jun 16, 2026Updated last month
- ☆22Oct 3, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Language Models as Hierarchy Encoders☆43Jan 6, 2026Updated 6 months ago
- [ICLR 25 Oral] RM-Bench: Benchmarking Reward Models of Language Models with Subtlety and Style☆84Jul 18, 2025Updated last year
- An LLM post-training framework with vLLM for RL Scaling☆385Updated this week
- ☆18Jul 1, 2023Updated 3 years ago
- Provide performance insight capabilities for RL frameworks.☆47Updated this week
- ☆168Mar 27, 2026Updated 3 months ago
- ☆17Aug 1, 2025Updated 11 months ago
- Path planning using Q-learning and DQN with experience replay☆34Apr 7, 2026Updated 3 months ago
- ☆1,055Mar 8, 2026Updated 4 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Bridge Megatron-Core to Hugging Face/Reinforcement Learning☆226Jun 15, 2026Updated last month
- SkyRL: A Modular Full-stack RL Library for LLMs☆2,093Updated this week
- Energetic GraphNeural Networks (EGNN) implementation based on Dirichlet Energy Constrained Learning.☆27Nov 1, 2021Updated 4 years ago
- Data-centric LLM training with dynamic sample selection, domain mixture optimization, and example reweighting inside the LLaMA-Factory tr…☆1,686Jun 17, 2026Updated last month
- Code for "Are “Hierarchical” Visual Representations Hierarchical?" in NeurIPS Workshop for Symmetry and Geometry in Neural Representation…☆23Nov 8, 2023Updated 2 years ago
- Code for our ACL '20 paper "Representation Engineering with Natural Language Explanations"☆30Jun 15, 2020Updated 6 years ago
- ☆84Feb 24, 2025Updated last year
- ☆126May 7, 2025Updated last year
- Code of LeCoRE☆14Feb 15, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Official repository for On Over-Squashing in Message Passing Neural Networks (ICML 2023)☆16Sep 4, 2023Updated 2 years ago
- TTDS coursework 1&2.☆11Nov 17, 2019Updated 6 years ago
- An adaptive training algorithm for residual network☆17Aug 22, 2020Updated 5 years ago
- Code for the ACL2023 paper: CAT: A Contextualized Conceptualization and Instantiation Framework for Commonsense Reasoning (https://aclant…☆11May 9, 2023Updated 3 years ago
- ☆10Jun 23, 2018Updated 8 years ago
- A comprehensive, production-ready framework for building intelligent AI agents with advanced capabilities including tool calling, persist…☆164Aug 23, 2025Updated 11 months ago
- VibeLaTeX 是一个轻量、好看的 LaTeX 公式编辑与导出工具:左侧输入公式,右侧实时渲染预览(默认 KaTeX,可切换 MathJax 兼容模式),并支持一键导出透明背景的 SVG/PNG(可调缩放、边距、紧裁剪)。项目内置 LLM 动作面板(Format/Fix…☆39Mar 10, 2026Updated 4 months ago