cs.LGMay 14, 2026

Efficient Multi-objective Prompt Optimization via Pure-exploration Bandits

Authors: Donghao Li, Chengshuai Shi, Weijuan Ou, Cong Shen, Jing Yang

Organizations: University of Virginia, Charlottesville, VA 22904, USA · Princeton University, Princeton, NJ 08544, USA · Southern University of Science and Technology, Shenzhen, Guangdong 518055, China

Abstract

Prompt engineering has become central to eliciting the capabilities of large language models (LLMs). At its core lies prompt selection -- efficiently identifying the most effective prompts. However, most prior investigations overlook a key challenge: the inherently multi-faceted nature of prompt performance, which cannot be captured by a single metric. To fill this gap, we study the multi-objective prompt selection problem under two practical settings: Pareto prompt set recovery and best feasible prompt identification. Casting the problem into the pure-exploration bandits framework, we adapt provably efficient algorithms from multi-objective bandits and further introduce a novel design for best feasible arm identification in structured bandits, with theoretical guarantees on the identification error in the linear case. Extensive experiments across multiple LLMs show that the bandit-based approaches yield significant improvements over baselines, establishing a principled and efficient framework for multi-objective prompt optimization.

Explore similar work

CardsList
  1. RLMOpt: Adaptive Prompt Optimization via Recursive Language Models

    Aug 11, 2026Subhash Bangalore Satheesha, Nirvik Pande, Deepthi Duddempudi +1Automatic Prompt OptimizationPrompt Engineering

  2. MO-CAPO: Multi-Objective Cost-Aware Prompt Optimization

    May 15, 2026Jan Büssing, Moritz Schlager, Timo Heiß +2Inference CostPrompt Engineering