Toward Effective Tool-Integrated Reasoning via Self-Evolved Preference Learning
☆35Sep 30, 2025Updated 11 months ago
Alternatives and similar repositories for Tool-Light
Users that are interested in Tool-Light are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- GISA: A Benchmark for General Information-Seeking Assistant☆37Mar 20, 2026Updated 5 months ago
- Official repository for ToolScope: An Agentic Framework for Vision-Guided and Long-Horizon Tool Use☆31Nov 4, 2025Updated 10 months ago
- ☆20Jan 18, 2026Updated 8 months ago
- Including 12+ cutting-edge agent systems across multiple research directions☆36Nov 10, 2025Updated 10 months ago
- Some example codes for drawing figures in research paper☆36Mar 3, 2022Updated 4 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- From Prompt Injection to Persistent Control: Defending Agentic Workspaces Against Trojan Backdoors☆19Jun 1, 2026Updated 3 months ago
- ☆89May 2, 2026Updated 4 months ago
- HierSearch: A Hierarchical Enterprise Deep Search Framework Integrating Local and Web Searches☆41Oct 9, 2025Updated 11 months ago
- The demo, code and data of FollowRAG☆75Jun 30, 2025Updated last year
- MemSifter: Offloading LLM Memory Retrieval via Outcome-Driven Proxy Reasoning☆69Jun 14, 2026Updated 3 months ago
- ☆39Apr 6, 2026Updated 5 months ago
- [ACL 2026 Main] Repo for paper "ReasonRank: Empowering Passage Ranking with Strong Reasoning Ability"☆183Apr 9, 2026Updated 5 months ago
- ☆65May 7, 2026Updated 4 months ago
- 🔍 Awesome Agentic Search is a curated list of papers, tools, and resources on agentic search—where AI agents plan, search, and reason to…☆63Aug 28, 2025Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆25Jul 8, 2026Updated 2 months ago
- 🔧Tool-Star: Empowering LLM-brained Multi-Tool Reasoner via Reinforcement Learning☆418Apr 3, 2026Updated 5 months ago
- ☆69Aug 14, 2025Updated last year
- 知予人工智能:从学习者到研究者☆14Jan 20, 2025Updated last year
- ☆32May 23, 2024Updated 2 years ago
- (ICLR 2025) AgentRefine: Enhancing Agent Generalization through Refinement Tuning☆21Nov 22, 2025Updated 9 months ago
- This repository contains the code for the paper “Neuro-Symbolic Query Compiler”, accepted to the Findings of ACL 2025.☆19Oct 20, 2025Updated 10 months ago
- Harness for deep search agent☆110Jun 16, 2026Updated 3 months ago
- Arbitrary Entropy Policy Optimization: Entropy Is Controllable in Reinforcement Fine-tuning☆18Jan 19, 2026Updated 8 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- [AAAI'25] SPRING: Learning Scalable and Pluggable Virtual Tokens for Retrieval-Augmented Large Language Models☆26Sep 24, 2025Updated 11 months ago
- The official implementation of "EnvScaler: Scaling Tool-Interactive Environments for LLM Agent via Programmatic Synthesis".☆197Sep 3, 2026Updated 2 weeks ago
- ☆59Feb 27, 2025Updated last year
- ☆26May 14, 2026Updated 4 months ago
- [ICLR 2026] Agentic Reinforced Policy Optimization (ARPO)☆1,123Sep 12, 2026Updated last week
- A curated list of papers, tools, and benchmarks on LLM-based computer-use agents, covering both terminal/CLI and GUI approaches.☆18Aug 17, 2026Updated last month
- CIKM 2021: Contrastive Learning of User Behavior Sequence for Context-Aware Document Ranking☆20Sep 28, 2022Updated 3 years ago
- ☆15Aug 13, 2026Updated last month
- Code and Data for Paper "AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning"☆54Sep 4, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- SIGIR 2021: Proactive Retrieval-based Chatbots based on Relevant Knowledge and Goals☆11Jul 30, 2021Updated 5 years ago
- ☆43Jun 30, 2026Updated 2 months ago
- Executive Memory for Coherent Long-Horizon Reasoning!☆86Jan 14, 2026Updated 8 months ago
- Official Implementation of PL-FMS☆11Sep 30, 2023Updated 2 years ago
- Implementation of self-certainty as an extention of ZeroEval Project☆38May 31, 2025Updated last year
- ☆52Mar 6, 2026Updated 6 months ago
- Vehicle to Vehicle Communication in Self-Driving Car☆12May 14, 2018Updated 8 years ago