Scalable and extensible reinforcement learning for LM agents.
☆122Aug 26, 2026Updated last week
Alternatives and similar repositories for AgentFly
Users that are interested in AgentFly are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official Code of Paper: MolAct: An Agentic RL Framework for Molecular Editing and Property Optimization☆17Apr 13, 2026Updated 4 months ago
- ☆23May 20, 2025Updated last year
- [NAACL 2025] The official implementation of paper "Learning From Failure: Integrating Negative Examples when Fine-tuning Large Language M…☆28Mar 14, 2024Updated 2 years ago
- A version of verl to support diverse tool use [TMLR 2026]☆1,037Jul 15, 2026Updated last month
- Python script to download conference paper automatically☆16Sep 10, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Verlog: A Multi-turn RL framework for LLM agents☆73Aug 8, 2026Updated 3 weeks ago
- The respository describing a novel datasets for word association explanations☆13Sep 21, 2023Updated 2 years ago
- Code for "Semantic Perturbations with Normalizing Flows for Improved Generalization"☆11Jul 13, 2021Updated 5 years ago
- [ACL 2026 Oral] From Word to World: Can Large Language Models be Implicit Text-based World Models?☆72Apr 13, 2026Updated 4 months ago
- A repo for open research on building large reasoning models☆153Jul 3, 2026Updated last month
- verl-agent is an extension of veRL, designed for training LLM/VLM agents via RL. verl-agent is also the official code for paper "Group-in…☆2,274Jun 9, 2026Updated 2 months ago
- Training and evaluating with OpenReward☆33Apr 28, 2026Updated 4 months ago
- A Gym for Agentic LLMs☆505Jan 21, 2026Updated 7 months ago
- Latent Large Language Models☆19Aug 24, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- A MemAgent framework that can be extrapolated to 3.5M, along with a training framework for RL training of any agent workflow.☆1,102May 12, 2026Updated 3 months ago
- Event-QA is a Dataset for answering complex Event-Centric questions over Knowledge Graphs (KGs). We target EventKG, a recently proposed E…☆19Apr 14, 2023Updated 3 years ago
- [ICLR 2025] The source code of "MolSpectra: Pre-training 3D Molecular Representation with Multi-modal Energy Spectra"☆20Apr 19, 2025Updated last year
- Scaling Agentic Reinforcement Learning with a Multi-Turn, Multi-Task Framework☆348Jan 17, 2026Updated 7 months ago
- The official implementation of the paper **LVChat: Facilitating Long Video Comprehension**☆14Apr 15, 2024Updated 2 years ago
- Agent-Omit: Training Efficient LLM Agents for Adaptive Thought and Observation Omission via Reinforcement Learning☆34May 11, 2026Updated 3 months ago
- [ICLR 2026] Agentic Reinforced Policy Optimization (ARPO)☆1,113Aug 20, 2026Updated last week
- A Practitioner's Guide to M(eow)ti Turn Agentic ReinfOrcement learning☆85Jan 16, 2026Updated 7 months ago
- Official code repo for the paper "ChemToolAgent: The Impact of Tools on Language Agents for Chemistry Problem Solving" (previously "Tooli…☆20Jun 7, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆18Jul 1, 2023Updated 3 years ago
- The official implementation of "EnvScaler: Scaling Tool-Interactive Environments for LLM Agent via Programmatic Synthesis".☆191Feb 12, 2026Updated 6 months ago
- [EMNLP 2022] This is the code repo for our EMNLP‘22 paper "Dimension Reduction for Efficient Dense Retrieval via Conditional Autoencoder"…☆13Oct 20, 2022Updated 3 years ago
- ☆21Mar 28, 2022Updated 4 years ago
- RLAnything (ICML 2026) & AutoTool (ICML 2026), DemyAgent: Open-Source RL for LLMs and Agentic Scenarios☆631Jun 12, 2026Updated 2 months ago
- Reinforcement Learning from Text Feedback☆48Feb 17, 2026Updated 6 months ago
- Agent RL framework for LLM agents: multi-turn reinforcement learning with StarPO and reasoning-collapse diagnostics☆2,786Aug 23, 2026Updated last week
- TreeRL: LLM Reinforcement Learning with On-Policy Tree Search in ACL'25☆105Jun 16, 2025Updated last year
- [ICLR 2026] Official repository of "Beyond Fixed: Training-Free Variable-Length Denoising for Diffusion Large Language Models"☆174Feb 16, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [NeurIPS 2025] Reasoning MLLM, Share-GRPO, advantage vanishing, sparse reward☆38Sep 19, 2025Updated 11 months ago
- REALM-Bench: A Real-World Planning Benchmark for LLMs and Multi-Agent Systems☆46Jul 20, 2026Updated last month
- Official Code of Memento: Fine-tuning LLM Agents without Fine-tuning LLMs☆2,568Oct 5, 2025Updated 10 months ago
- [ACL'26 Oral] AgentOCR is a token-efficient framework that compresses multi-turn agent history by rendering it into images and adopting R…☆47Mar 1, 2026Updated 6 months ago
- ☆21May 30, 2025Updated last year
- A symbolic benchmark for verifiable chain-of-thought financial reasoning. Includes executable templates, 58 topics across 12 domains, and…☆31Dec 26, 2025Updated 8 months ago
- SWE-Swiss: A Multi-Task Fine-Tuning and RL Recipe for High-Performance Issue Resolution☆105Sep 24, 2025Updated 11 months ago