[ACL 2026] Feedback-Driven Tool-Use Improvements in Large Language Models via Automated Build Environments
☆52Jul 10, 2026Updated last month
Alternatives and similar repositories for FTRL
Users that are interested in FTRL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [EMNLP 2025] Official Implement of "CRITICTOOL: Evaluating Self-Critique Capabilities of Large Language Models in Tool-Calling Error Scen…☆18Sep 2, 2025Updated 11 months ago
- OneEdit: A Neural-Symbolic Collaboratively Knowledge Editing System.☆20Oct 14, 2024Updated last year
- ☆25Aug 20, 2025Updated 11 months ago
- Official implementation of the paper "Instructions are all you need: Self-supervised Reinforcement Learning for Instruction Following"☆40Jan 11, 2026Updated 7 months ago
- [EMNLP 2024] RoTBench: A Multi-Level Benchmark for Evaluating the Robustness of Large Language Models in Tool Learning☆15May 13, 2025Updated last year
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- [ACL 2026] A Multi-Dimensional Constraint Framework for Evaluating and Improving Instruction Following in Large Language Models☆23Jul 10, 2026Updated last month
- [ACL 2024] ToolSword: Unveiling Safety Issues of Large Language Models in Tool Learning Across Three Stages☆15Sep 12, 2024Updated last year
- MUA-RL: MULTI-TURN USER-INTERACTING AGENT REINFORCEMENT LEARNING FOR AGENTIC TOOL USE☆67Nov 5, 2025Updated 9 months ago
- Pushing Test-Time Scaling Limits of Deep Search with Asymmetric Verification☆22Oct 8, 2025Updated 10 months ago
- The implementation of RAGSynth: Synthetic Data for Robust and Faithful RAG Component Optimization☆21May 26, 2025Updated last year
- Implementation of AdaCQR(COLING 2025)☆15Dec 30, 2024Updated last year
- The implementation for ACL 2026: MatchTIR: Fine-Grained Supervision for Tool-Integrated Reasoning via Bipartite Matching.☆20Jul 27, 2026Updated 2 weeks ago
- CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning☆37Aug 28, 2025Updated 11 months ago
- Unofficial implementation of Chain of Hindsight (https://arxiv.org/abs/2302.02676) using pytorch and huggingface Trainers.☆11Apr 5, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [AAAI 2026] ReCode: Reinforced Code Knowledge Editing for API Updates☆25Jul 1, 2025Updated last year
- [NAACL 2025] Representing Rule-based Chatbots with Transformers☆23Feb 9, 2025Updated last year
- ☆13Aug 13, 2024Updated 2 years ago
- [COLING 2025] ToolEyes: Fine-Grained Evaluation for Tool Learning Capabilities of Large Language Models in Real-world Scenarios☆74May 13, 2025Updated last year
- WideSearch: Benchmarking Agentic Broad Info-Seeking☆151Oct 9, 2025Updated 10 months ago
- ToolBridge: An Open-Source Dataset to Equip LLMs with External Tool Capabilities☆14Feb 11, 2025Updated last year
- The code for paper: Decoupled Planning and Execution: A Hierarchical Reasoning Framework for Deep Search [SIGIR 2026]☆65Jul 4, 2025Updated last year
- ASTRA is an end-to-end system for synthesizing agentic trajectories and rule-verifiable environments for SFT and RL training, developed b…☆152Jan 30, 2026Updated 6 months ago
- ☆33May 8, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆27Aug 13, 2025Updated last year
- Code Repo for EfficientRAG: Efficient Retriever for Multi-Hop Question Answering☆69Mar 4, 2025Updated last year
- ☆513Oct 16, 2025Updated 9 months ago
- Code for the "Long Context Needs Some R&R" paper.☆12Mar 11, 2024Updated 2 years ago
- 🔧Tool-Star: Empowering LLM-brained Multi-Tool Reasoner via Reinforcement Learning☆408Apr 3, 2026Updated 4 months ago
- Transfer Learning in Dialogue Benchmarking Toolkit☆14Mar 31, 2023Updated 3 years ago
- TreeRL: LLM Reinforcement Learning with On-Policy Tree Search in ACL'25☆102Jun 16, 2025Updated last year
- ☆23May 17, 2026Updated 2 months ago
- code for paper Sparse Structure Search for Delta Tuning☆11Oct 16, 2022Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- This is the official repo for Towards Uncertainty-Aware Language Agent.☆31Aug 15, 2024Updated last year
- MiroRL is an MCP-first reinforcement learning framework for deep research agent.☆247Aug 27, 2025Updated 11 months ago
- ☆36May 24, 2025Updated last year
- Repo for "MaskSearch: A Universal Pre-Training Framework to Enhance Agentic Search Capability"☆155May 27, 2025Updated last year
- Implementation code for ACL2024:Advancing Parameter Efficiency in Fine-tuning via Representation Editing☆15Apr 20, 2024Updated 2 years ago
- Implementation for OAgents: An Empirical Study of Building Effective Agents☆328Oct 13, 2025Updated 10 months ago
- a-m-team's exploration in large language modeling☆196May 29, 2025Updated last year