A Survey on Effective Invocation Methods of Massive LLM Services

Wang, Can; Zhang, Bolin; Sui, Dianbo; Tum, Zhiying; Liu, Xiaoyu; Kang, Jiabao

Computer Science > Software Engineering

arXiv:2402.03408v1 (cs)

[Submitted on 5 Feb 2024 (this version), latest version 1 Mar 2024 (v2)]

Title:A Survey on Effective Invocation Methods of Massive LLM Services

Authors:Can Wang, Bolin Zhang, Dianbo Sui, Zhiying Tum, Xiaoyu Liu, Jiabao Kang

View PDF HTML (experimental)

Abstract:Language models as a service (LMaaS) enable users to accomplish tasks without requiring specialized knowledge, simply by paying a service provider. However, numerous providers offer massive large language model (LLM) services with variations in latency, performance, and pricing. Consequently, constructing the cost-saving LLM services invocation strategy with low-latency and high-performance responses that meet specific task demands becomes a pressing challenge. This paper provides a comprehensive overview of the LLM services invocation methods. Technically, we give a formal definition of the problem of constructing effective invocation strategy in LMaaS and present the LLM services invocation framework. The framework classifies existing methods into four different components, including input abstract, semantic cache, solution design, and output enhancement, which can be freely combined with each other. Finally, we emphasize the open challenges that have not yet been well addressed in this task and shed light on future research.

Subjects:	Software Engineering (cs.SE); Distributed, Parallel, and Cluster Computing (cs.DC)
Cite as:	arXiv:2402.03408 [cs.SE]
	(or arXiv:2402.03408v1 [cs.SE] for this version)
	https://doi.org/10.48550/arXiv.2402.03408

Submission history

From: Can Wang [view email]
[v1] Mon, 5 Feb 2024 15:10:42 UTC (1,567 KB)
[v2] Fri, 1 Mar 2024 03:33:21 UTC (1,568 KB)

Computer Science > Software Engineering

Title:A Survey on Effective Invocation Methods of Massive LLM Services

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Software Engineering

Title:A Survey on Effective Invocation Methods of Massive LLM Services

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators