☆86May 14, 2026Updated 2 months ago
Alternatives and similar repositories for MARFT
Users that are interested in MARFT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆20May 14, 2026Updated 2 months ago
- ☆54Sep 6, 2025Updated 10 months ago
- A Framework for LLM-based Multi-Agent Reinforced Training and Inference☆538Apr 14, 2026Updated 3 months ago
- Official implementation of the NeurIPS 2024 paper CORY☆33Mar 4, 2026Updated 4 months ago
- [ICLR'26] Stronger-MAS: A RL Framework for multi LLM agent system; [arxiv] MetaAgent-X: End-to-End Reinforcement Learning Automatic Mult…☆205May 15, 2026Updated 2 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Dr. MAS is an end-to-end RL training framework for multi-agent LLM systems, supporting the co-training of multiple (heterogeneous) LLMs.☆143Apr 1, 2026Updated 3 months ago
- Benchmark and research code for the paper SWEET-RL Training Multi-Turn LLM Agents onCollaborative Reasoning Tasks☆271May 5, 2025Updated last year
- LITEN: Learning from Inference Time Execution for VLAs☆27Oct 23, 2025Updated 8 months ago
- ☆25Jan 29, 2026Updated 5 months ago
- Source code for our paper: "ARIA: Training Language Agents with Intention-Driven Reward Aggregation".☆30Aug 9, 2025Updated 11 months ago
- Official implementation of MATPO: Multi-Agent Tool-Integrated Policy Optimization.☆82Oct 31, 2025Updated 8 months ago
- [NeurIPS 2025 Spotlight] Co-Evolving LLM Coder and Unit Tester via Reinforcement Learning☆167Sep 19, 2025Updated 10 months ago
- The code for paper "EPO: Entropy-regularized Policy Optimization for LLM Agents Reinforcement Learning"☆40Jul 13, 2026Updated last week
- ☆21Apr 3, 2026Updated 3 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Learning to Skip the Middle Layers of Transformers☆17Aug 7, 2025Updated 11 months ago
- Models, data, and codes for the paper: MetaAligner: Towards Generalizable Multi-Objective Alignment of Language Models☆24Sep 26, 2024Updated last year
- Companion code to https://arxiv.org/abs/2409.03797v2☆19Sep 18, 2025Updated 10 months ago
- ☆13Jan 22, 2025Updated last year
- [NAACL 2025] The official implementation of paper "Learning From Failure: Integrating Negative Examples when Fine-tuning Large Language M…☆28Mar 14, 2024Updated 2 years ago
- Code for the work "Adaptive Sample Scheduling for Direct Preference Optimization", which was accepted to the 2025 Conference on Neural In…☆33Sep 30, 2025Updated 9 months ago
- The repo for our paper: Enhancing LLM-Based Agents via Global Planning and Hierarchical Execution (NCIIP 2025 Best Paper)☆17Aug 18, 2025Updated 11 months ago
- This is the code of MMOA-RAG.☆111May 11, 2025Updated last year
- Math24o: 高中奥林匹克数学竞赛测评集 High School Olympiad Mathematics Chinese Benchmark☆14Mar 27, 2025Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Scaling Deep Research via Reinforcement Learning in Real-world Environments.☆782May 10, 2026Updated 2 months ago
- ☆166Jan 21, 2025Updated last year
- Codebase for [Order Matters: Agent-by-agent Policy Optimization](https://openreview.net/forum?id=Q-neeWNVv1)☆32Nov 22, 2025Updated 7 months ago
- Research Code for "ArCHer: Training Language Model Agents via Hierarchical Multi-Turn RL"☆208Apr 17, 2025Updated last year
- GisPy: A Tool for Measuring Gist Inference Score in Text https://aclanthology.org/2022.wnu-1.5/☆13Jul 1, 2024Updated 2 years ago
- Cooperation and Fairness in Multi-Agent Reinforcement Learning☆16Aug 6, 2025Updated 11 months ago
- ☆49Jun 26, 2026Updated 3 weeks ago
- (ACL2025 Findings) Official code for the paper "STeCa: Step-level Trajectory Calibration for LLM Agent Learning"☆29Mar 2, 2026Updated 4 months ago
- ☆23Apr 5, 2026Updated 3 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This is the reading list for the survey "A Survey on the Optimization of LLM-based Agents ". We will keep adding papers and improving the…☆241Jun 17, 2026Updated last month
- ☆46Sep 27, 2025Updated 9 months ago
- Behavior Injection: Preparing Language Models for Reinforcement Learning (NeurIPS 2025)☆17Jul 1, 2025Updated last year
- siiRL: Shanghai Innovation Institute RL Framework for Advanced LLMs and Multi-Agent Systems☆366Jan 30, 2026Updated 5 months ago
- [NAACL 2025] Representing Rule-based Chatbots with Transformers☆23Feb 9, 2025Updated last year
- [ACL 2026] Feedback-Driven Tool-Use Improvements in Large Language Models via Automated Build Environments☆51Jul 10, 2026Updated last week
- verl-agent is an extension of veRL, designed for training LLM/VLM agents via RL. verl-agent is also the official code for paper "Group-in…☆2,140Jun 9, 2026Updated last month