Advancing and Benchmarking Personalized Tool Invocation for LLMs
Huang, Xu, Huang, Yuefeng, Liu, Weiwen, Zeng, Xingshan, Wang, Yasheng, Tang, Ruiming, Xie, Hong, Lian, Defu
–arXiv.org Artificial Intelligence
Tool invocation is a crucial mechanism for extending the capabilities of Large Language Models (LLMs) and has recently garnered significant attention. It enables LLMs to solve complex problems through tool calls while accessing up-to-date world knowledge. However, existing work primarily focuses on the fundamental ability of LLMs to invoke tools for problem-solving, without considering personalized constraints in tool invocation. In this work, we introduce the concept of Personalized Tool Invocation and define two key tasks: Tool Preference and Profile-dependent Query. Tool Preference addresses user preferences when selecting among functionally similar tools, while Profile-dependent Query considers cases where a user query lacks certain tool parameters, requiring the model to infer them from the user profile. To tackle these challenges, we propose PTool, a data synthesis framework designed for personalized tool invocation. Additionally, we construct \textbf{PTBench}, the first benchmark for evaluating personalized tool invocation. We then fine-tune various open-source models, demonstrating the effectiveness of our framework and providing valuable insights. Our benchmark is public at https://github.com/hyfshadow/PTBench.
arXiv.org Artificial Intelligence
May-8-2025
- Country:
- North America
- United States
- Minnesota > Hennepin County
- Minneapolis (0.14)
- Florida > Miami-Dade County
- Miami (0.04)
- California > Los Angeles County
- Los Angeles (0.04)
- Minnesota > Hennepin County
- Mexico > Mexico City
- Mexico City (0.04)
- United States
- Europe > France
- Île-de-France > Paris > Paris (0.04)
- Asia
- North America
- Genre:
- Research Report (0.50)
- Industry:
- Health & Medicine (0.68)
- Information Technology (0.46)
- Technology: