Scalable and extensible reinforcement learning for LM agents.
☆122Sep 17, 2026Updated this week
Alternatives and similar repositories for AgentFly
Users that are interested in AgentFly are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆23May 20, 2025Updated last year
- [NAACL 2025] The official implementation of paper "Learning From Failure: Integrating Negative Examples when Fine-tuning Large Language M…☆28Mar 14, 2024Updated 2 years ago
- A version of verl to support diverse tool use [TMLR 2026]☆1,044Jul 15, 2026Updated 2 months ago
- Python script to download conference paper automatically☆16Sep 10, 2024Updated 2 years ago
- Verlog: A Multi-turn RL framework for LLM agents☆74Aug 8, 2026Updated last month
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- The respository describing a novel datasets for word association explanations☆13Sep 21, 2023Updated 3 years ago
- Code for "Semantic Perturbations with Normalizing Flows for Improved Generalization"☆11Jul 13, 2021Updated 5 years ago
- [ACL 2026 Oral] From Word to World: Can Large Language Models be Implicit Text-based World Models?☆74Apr 13, 2026Updated 5 months ago
- A repo for open research on building large reasoning models☆153Jul 3, 2026Updated 2 months ago
- verl-agent is an extension of veRL, designed for training LLM/VLM agents via RL. verl-agent is also the official code for paper "Group-in…☆2,334Jun 9, 2026Updated 3 months ago
- Training and evaluating with OpenReward☆33Apr 28, 2026Updated 4 months ago
- A Gym for Agentic LLMs☆510Jan 21, 2026Updated 8 months ago
- ☆20Apr 30, 2024Updated 2 years ago
- Latent Large Language Models☆19Aug 24, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A MemAgent framework that can be extrapolated to 3.5M, along with a training framework for RL training of any agent workflow.☆1,111May 12, 2026Updated 4 months ago
- [ICLR 2025] The source code of "MolSpectra: Pre-training 3D Molecular Representation with Multi-modal Energy Spectra"☆20Apr 19, 2025Updated last year
- Scaling Agentic Reinforcement Learning with a Multi-Turn, Multi-Task Framework☆354Jan 17, 2026Updated 8 months ago
- Agent-Omit: Training Efficient LLM Agents for Adaptive Thought and Observation Omission via Reinforcement Learning☆34May 11, 2026Updated 4 months ago
- [ICLR 2026] Agentic Reinforced Policy Optimization (ARPO)☆1,125Sep 12, 2026Updated last week
- A Practitioner's Guide to M(eow)ti Turn Agentic ReinfOrcement learning☆85Jan 16, 2026Updated 8 months ago
- Code and dataset for the paper "IsarStep: a Benchmark for High-level Mathematical Reasoning"☆12Mar 15, 2021Updated 5 years ago
- ☆18Jul 1, 2023Updated 3 years ago
- The official implementation of "EnvScaler: Scaling Tool-Interactive Environments for LLM Agent via Programmatic Synthesis".☆198Sep 3, 2026Updated 3 weeks ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆21Mar 28, 2022Updated 4 years ago
- RLAnything (ICML 2026) & AutoTool (ICML 2026), DemyAgent: Open-Source RL for LLMs and Agentic Scenarios☆641Jun 12, 2026Updated 3 months ago
- ☆40Jan 1, 2026Updated 8 months ago
- [WWW 2025 Oral] Large Language Models Empowered Personalized Web Agents.☆23Nov 11, 2025Updated 10 months ago
- Reinforcement Learning from Text Feedback☆49Feb 17, 2026Updated 7 months ago
- Agent RL framework for LLM agents: multi-turn reinforcement learning with StarPO and reasoning-collapse diagnostics☆2,807Aug 23, 2026Updated last month
- TreeRL: LLM Reinforcement Learning with On-Policy Tree Search in ACL'25☆105Jun 16, 2025Updated last year
- [ICLR 2026] Official repository of "Beyond Fixed: Training-Free Variable-Length Denoising for Diffusion Large Language Models"☆175Feb 16, 2026Updated 7 months ago
- [NeurIPS 2025] Reasoning MLLM, Share-GRPO, advantage vanishing, sparse reward☆38Sep 19, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- REALM-Bench: A Real-World Planning Benchmark for LLMs and Multi-Agent Systems☆47Jul 20, 2026Updated 2 months ago
- [ACL'26 Oral] AgentOCR is a token-efficient framework that compresses multi-turn agent history by rendering it into images and adopting R…☆49Mar 1, 2026Updated 6 months ago
- ☆21May 30, 2025Updated last year
- Scaling Test-time Training for LLM Reasoning☆37Apr 14, 2026Updated 5 months ago
- SWE-Swiss: A Multi-Task Fine-Tuning and RL Recipe for High-Performance Issue Resolution☆105Sep 24, 2025Updated last year
- ☆1,453Feb 12, 2026Updated 7 months ago
- The official implementations of the Paper "G1: Teaching LLMs to Reason on Graphs with Reinforcement Learning"☆45Jun 14, 2025Updated last year