Scalable and extensible reinforcement learning for LM agents.
☆122May 6, 2026Updated 3 months ago
Alternatives and similar repositories for AgentFly
Users that are interested in AgentFly are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [NAACL 2025] The official implementation of paper "Learning From Failure: Integrating Negative Examples when Fine-tuning Large Language M…☆28Mar 14, 2024Updated 2 years ago
- A version of verl to support diverse tool use [TMLR 2026]☆1,031Jul 15, 2026Updated 3 weeks ago
- Python script to download conference paper automatically☆16Sep 10, 2024Updated last year
- Verlog: A Multi-turn RL framework for LLM agents☆73Updated this week
- The respository describing a novel datasets for word association explanations☆13Sep 21, 2023Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Code for "Semantic Perturbations with Normalizing Flows for Improved Generalization"☆11Jul 13, 2021Updated 5 years ago
- The official implementation of Bi-Mamba☆17Oct 22, 2025Updated 9 months ago
- A repo for open research on building large reasoning models☆152Jul 3, 2026Updated last month
- verl-agent is an extension of veRL, designed for training LLM/VLM agents via RL. verl-agent is also the official code for paper "Group-in…☆2,218Jun 9, 2026Updated 2 months ago
- Training and evaluating with OpenReward☆33Apr 28, 2026Updated 3 months ago
- A Gym for Agentic LLMs☆505Jan 21, 2026Updated 6 months ago
- Latent Large Language Models☆19Aug 24, 2024Updated last year
- A MemAgent framework that can be extrapolated to 3.5M, along with a training framework for RL training of any agent workflow.☆1,093May 12, 2026Updated 3 months ago
- Event-QA is a Dataset for answering complex Event-Centric questions over Knowledge Graphs (KGs). We target EventKG, a recently proposed E…☆19Apr 14, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [ICLR 2025] The source code of "MolSpectra: Pre-training 3D Molecular Representation with Multi-modal Energy Spectra"☆20Apr 19, 2025Updated last year
- Scaling Agentic Reinforcement Learning with a Multi-Turn, Multi-Task Framework☆338Jan 17, 2026Updated 6 months ago
- Agent-Omit: Training Efficient LLM Agents for Adaptive Thought and Observation Omission via Reinforcement Learning☆33May 11, 2026Updated 3 months ago
- [ICLR 2026] Agentic Reinforced Policy Optimization (ARPO)☆1,105Jul 13, 2026Updated last month
- A Practitioner's Guide to M(eow)ti Turn Agentic ReinfOrcement learning☆83Jan 16, 2026Updated 6 months ago
- Official code repo for the paper "ChemToolAgent: The Impact of Tools on Language Agents for Chemistry Problem Solving" (previously "Tooli…☆20Jun 7, 2025Updated last year
- ☆18Jul 1, 2023Updated 3 years ago
- Code and dataset for the paper "IsarStep: a Benchmark for High-level Mathematical Reasoning"☆12Mar 15, 2021Updated 5 years ago
- The official implementation of "EnvScaler: Scaling Tool-Interactive Environments for LLM Agent via Programmatic Synthesis".☆184Feb 12, 2026Updated 6 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [EMNLP 2022] This is the code repo for our EMNLP‘22 paper "Dimension Reduction for Efficient Dense Retrieval via Conditional Autoencoder"…☆13Oct 20, 2022Updated 3 years ago
- ☆22Mar 28, 2022Updated 4 years ago
- RLAnything (ICML 2026) & AutoTool (ICML 2026), DemyAgent: Open-Source RL for LLMs and Agentic Scenarios☆615Jun 12, 2026Updated 2 months ago
- TreeRL: LLM Reinforcement Learning with On-Policy Tree Search in ACL'25☆102Jun 16, 2025Updated last year
- Reinforcement Learning from Text Feedback☆47Feb 17, 2026Updated 5 months ago
- RAGEN leverages reinforcement learning to train LLM reasoning agents in interactive, stochastic environments.☆2,768Jul 24, 2026Updated 2 weeks ago
- [ICLR 2026] Official repository of "Beyond Fixed: Training-Free Variable-Length Denoising for Diffusion Large Language Models"☆174Feb 16, 2026Updated 5 months ago
- [NeurIPS 2025] Reasoning MLLM, Share-GRPO, advantage vanishing, sparse reward☆38Sep 19, 2025Updated 10 months ago
- REALM-Bench: A Real-World Planning Benchmark for LLMs and Multi-Agent Systems☆46Jul 20, 2026Updated 3 weeks ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Official Code of Memento: Fine-tuning LLM Agents without Fine-tuning LLMs☆2,564Oct 5, 2025Updated 10 months ago
- [ACL'26 Oral] AgentOCR is a token-efficient framework that compresses multi-turn agent history by rendering it into images and adopting R…☆43Mar 1, 2026Updated 5 months ago
- ☆21May 30, 2025Updated last year
- A symbolic benchmark for verifiable chain-of-thought financial reasoning. Includes executable templates, 58 topics across 12 domains, and…☆30Dec 26, 2025Updated 7 months ago
- Scaling Test-time Training for LLM Reasoning☆28Apr 14, 2026Updated 3 months ago
- ☆1,441Feb 12, 2026Updated 6 months ago
- [NeurIPS 24] Can LLMs Solve Molecule Puzzles? A Multimodal Benchmark for Molecular Structure Elucidation☆20Jan 2, 2026Updated 7 months ago