☆87May 14, 2026Updated 4 months ago
Alternatives and similar repositories for MARFT
Users that are interested in MARFT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆21May 14, 2026Updated 4 months ago
- ☆56Sep 6, 2025Updated last year
- [ICLR 2026] A Framework for LLM-based Multi-Agent Reinforced Training and Inference☆565Aug 20, 2026Updated last month
- Official implementation of the NeurIPS 2024 paper CORY☆33Mar 4, 2026Updated 7 months ago
- [ICLR'26] Stronger-MAS: A RL Framework for multi LLM agent system; [arxiv] MetaAgent-X: End-to-End Reinforcement Learning Automatic Mult…☆223May 15, 2026Updated 4 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Dr. MAS is an end-to-end RL training framework for multi-agent LLM systems, supporting the co-training of multiple (heterogeneous) LLMs.☆168Sep 26, 2026Updated 2 weeks ago
- Benchmark and research code for the paper SWEET-RL Training Multi-Turn LLM Agents onCollaborative Reasoning Tasks☆272May 5, 2025Updated last year
- LITEN: Learning from Inference Time Execution for VLAs☆27Oct 23, 2025Updated 11 months ago
- ☆25Jan 29, 2026Updated 8 months ago
- Source code for our paper: "ARIA: Training Language Agents with Intention-Driven Reward Aggregation".☆31Aug 9, 2025Updated last year
- Official implementation of MATPO: Multi-Agent Tool-Integrated Policy Optimization.☆83Oct 31, 2025Updated 11 months ago
- The code for paper "EPO: Entropy-regularized Policy Optimization for LLM Agents Reinforcement Learning"☆40Jul 13, 2026Updated 2 months ago
- [NeurIPS 2025 Spotlight] Co-Evolving LLM Coder and Unit Tester via Reinforcement Learning☆169Sep 19, 2025Updated last year
- ☆21Apr 3, 2026Updated 6 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Learning to Skip the Middle Layers of Transformers☆17Aug 7, 2025Updated last year
- Models, data, and codes for the paper: MetaAligner: Towards Generalizable Multi-Objective Alignment of Language Models☆25Sep 26, 2024Updated 2 years ago
- ☆13Jan 22, 2025Updated last year
- [NAACL 2025] The official implementation of paper "Learning From Failure: Integrating Negative Examples when Fine-tuning Large Language M…☆28Mar 14, 2024Updated 2 years ago
- The repo for our paper: Enhancing LLM-Based Agents via Global Planning and Hierarchical Execution (NCIIP 2025 Best Paper)☆17Aug 18, 2025Updated last year
- This is the code of MMOA-RAG.☆113May 11, 2025Updated last year
- Math24o: 高中奥林匹克数学竞赛测评集 High School Olympiad Mathematics Chinese Benchmark☆14Mar 27, 2025Updated last year
- Scaling Deep Research via Reinforcement Learning in Real-world Environments.☆803May 10, 2026Updated 5 months ago
- ☆167Jan 21, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Codebase for [Order Matters: Agent-by-agent Policy Optimization](https://openreview.net/forum?id=Q-neeWNVv1)☆33Nov 22, 2025Updated 10 months ago
- Research Code for "ArCHer: Training Language Model Agents via Hierarchical Multi-Turn RL"☆210Apr 17, 2025Updated last year
- GisPy: A Tool for Measuring Gist Inference Score in Text https://aclanthology.org/2022.wnu-1.5/☆13Jul 1, 2024Updated 2 years ago
- Cooperation and Fairness in Multi-Agent Reinforcement Learning☆16Aug 6, 2025Updated last year
- (ACL2025 Findings) Official code for the paper "STeCa: Step-level Trajectory Calibration for LLM Agent Learning"☆29Mar 2, 2026Updated 7 months ago
- Behavior Injection: Preparing Language Models for Reinforcement Learning (NeurIPS 2025)☆17Jul 1, 2025Updated last year
- ☆23Aug 25, 2026Updated last month
- This is the reading list for the survey "A Survey on the Optimization of LLM-based Agents ". We will keep adding papers and improving the…☆247Sep 12, 2026Updated 3 weeks ago
- ☆46Sep 27, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- siiRL: Shanghai Innovation Institute RL Framework for Advanced LLMs and Multi-Agent Systems☆377Jan 30, 2026Updated 8 months ago
- [NAACL 2025] Representing Rule-based Chatbots with Transformers☆23Feb 9, 2025Updated last year
- [ACL 2026] Feedback-Driven Tool-Use Improvements in Large Language Models via Automated Build Environments☆53Jul 10, 2026Updated 3 months ago
- verl-agent is an extension of veRL, designed for training LLM/VLM agents via RL. verl-agent is also the official code for paper "Group-in…☆2,365Jun 9, 2026Updated 4 months ago
- ☆76Jun 10, 2025Updated last year
- MemGen: Weaving Generative Latent Memory for Self-Evolving Agents☆424Sep 24, 2026Updated 2 weeks ago
- [ACL 2025] Agentic Knowledgeable Self-awareness☆94Jun 15, 2025Updated last year