REVIEW 4 cited by
Localized Zeroth-Order Prompt Optimization
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
The efficacy of large language models (LLMs) in understanding and generating natural language has aroused a wide interest in developing prompt-based methods to harness the power of black-box LLMs. Existing methodologies usually prioritize a global optimization for finding the global optimum, which however will perform poorly in certain tasks. This thus motivates us to re-think the necessity of finding a global optimum in prompt optimization. To answer this, we conduct a thorough empirical study on prompt optimization and draw two major insights. Contrasting with the rarity of global optimum, local optima are usually prevalent and well-performed, which can be more worthwhile for efficient prompt optimization (Insight I). The choice of the input domain, covering both the generation and the representation of prompts, affects the identification of well-performing local optima (Insight II). Inspired by these insights, we propose a novel algorithm, namely localized zeroth-order prompt optimization (ZOPO), which incorporates a Neural Tangent Kernel-based derived Gaussian process into standard zeroth-order optimization for an efficient search of well-performing local optima in prompt optimization. Remarkably, ZOPO outperforms existing baselines in terms of both the optimization performance and the query efficiency, which we demonstrate through extensive experiments.
Forward citations
Cited by 4 Pith papers
-
Aviary: training language agents on challenging scientific tasks
A small open-source LLM trained in the new Aviary environments with expert iteration and majority voting matches or exceeds a frontier LLM agent on SeqQA and LitQA2 at far lower inference cost.
-
GReaTer: Gradients over Reasoning Makes Smaller Language Models Strong Prompt Optimizers
A gradient-based discrete prompt optimizer that uses reasoning chains to let small LMs self-optimize prompts, outperforming text-feedback baselines on reasoning benchmarks.
-
ACING: Actor-Critic for Instruction Learning in Black-Box LLMs
ACING uses off-policy actor-critic RL over continuous latent vectors, decoded by a frozen white-box model, to optimize discrete instructions for black-box LLMs from reward feedback alone.
-
Boosting Private Domain Understanding of Efficient MLLMs: A Tuning-free, Adaptive, Universal Prompt Optimization Framework
IDEALPrompt combines strategy search with self-reflection to craft prompts that let a 2B multimodal model match or beat fine-tuning on private e-commerce data, without changing model weights.
Discussion (0). Continue with ORCID to comment.