Evaluating tool-augmented LLMs in conversation settings
☆89May 31, 2024Updated 2 years ago
Alternatives and similar repositories for ToolTalk
Users that are interested in ToolTalk are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Companion code to https://arxiv.org/abs/2402.15491☆22Sep 18, 2025Updated 10 months ago
- [SIGIR '22] Code for our SIGIR 2022 accepted paper : P3 Ranker: Mitigating the Gaps between Pre-training and Ranking Fine-tuning with Pr…☆18Sep 24, 2023Updated 2 years ago
- [ICLR'24] MetaTool Benchmark for Large Language Models: Deciding Whether to Use Tools and Which to Use☆115Mar 21, 2024Updated 2 years ago
- m&ms: A Benchmark to Evaluate Tool-Use for multi-step multi-modal tasks☆46Sep 26, 2024Updated last year
- ☆922Jul 24, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆21May 22, 2023Updated 3 years ago
- Conversations with Search Engines☆14Jun 12, 2023Updated 3 years ago
- [ACL2024] Planning, Creation, Usage: Benchmarking LLMs for Comprehensive Tool Utilization in Real-World Complex Scenarios☆71Aug 5, 2025Updated 11 months ago
- A new tool learning benchmark aiming at well-balanced stability and reality, based on ToolBench.