Efficient Prompt Optimization Through the Lens of Best Arm Identification

Shi, Chengshuai; Yang, Kun; Chen, Zihan; Li, Jundong; Yang, Jing; Shen, Cong

Statistics > Machine Learning

arXiv:2402.09723 (stat)

[Submitted on 15 Feb 2024 (v1), last revised 30 May 2024 (this version, v3)]

Title:Efficient Prompt Optimization Through the Lens of Best Arm Identification

Authors:Chengshuai Shi, Kun Yang, Zihan Chen, Jundong Li, Jing Yang, Cong Shen

View PDF HTML (experimental)

Abstract:The remarkable instruction-following capability of large language models (LLMs) has sparked a growing interest in automatically finding good prompts, i.e., prompt optimization. Most existing works follow the scheme of selecting from a pre-generated pool of candidate prompts. However, these designs mainly focus on the generation strategy, while limited attention has been paid to the selection method. Especially, the cost incurred during the selection (e.g., accessing LLM and evaluating the responses) is rarely explicitly considered. To overcome this limitation, this work provides a principled framework, TRIPLE, to efficiently perform prompt selection under an explicit budget constraint. TRIPLE is built on a novel connection established between prompt optimization and fixed-budget best arm identification (BAI-FB) in multi-armed bandits (MAB); thus, it is capable of leveraging the rich toolbox from BAI-FB systematically and also incorporating unique characteristics of prompt optimization. Extensive experiments on multiple well-adopted tasks using various LLMs demonstrate the remarkable performance improvement of TRIPLE over baselines while satisfying the limited budget constraints. As an extension, variants of TRIPLE are proposed to efficiently select examples for few-shot prompts, also achieving superior empirical performance.

Subjects:	Machine Learning (stat.ML); Artificial Intelligence (cs.AI); Computation and Language (cs.CL); Machine Learning (cs.LG)
Cite as:	arXiv:2402.09723 [stat.ML]
	(or arXiv:2402.09723v3 [stat.ML] for this version)
	https://doi.org/10.48550/arXiv.2402.09723

Submission history

From: Chengshuai Shi [view email]
[v1] Thu, 15 Feb 2024 05:31:13 UTC (2,609 KB)
[v2] Tue, 20 Feb 2024 06:35:36 UTC (3,044 KB)
[v3] Thu, 30 May 2024 19:40:21 UTC (1,391 KB)

Statistics > Machine Learning

Title:Efficient Prompt Optimization Through the Lens of Best Arm Identification

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Statistics > Machine Learning

Title:Efficient Prompt Optimization Through the Lens of Best Arm Identification

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators