Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T18:00:22.243698Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 2 inbound Pith citation observations for arXiv:2502.00728.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T18:00:22.243698Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T19:22:25.370012Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-20T19:08:53.951599Z
28 of 28 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation af21dd5b-02d6-43f6-8c9d-181a57701e23 · outbound
Meta-Prompt Optimization for LLM-Based Sequential Decision Making A Survey on Data Selection for Language Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 462f690f-de3f-43a1-971b-d5307930b9d7 · outbound
Meta-Prompt Optimization for LLM-Based Sequential Decision Making 3.2, there are two major differences compared to the way in which our EXPO algorithm optimizes the task description and meta-instruction (Algo
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9359591a-19dd-43a1-b28f-42ead42bd4af · outbound
Meta-Prompt Optimization for LLM-Based Sequential Decision Making Exploring Large Language Model based Intelligent Agents: Definitions, Methods, and Prospects
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd1a4c34-747c-4588-85a7-88626223d23c · outbound
Meta-Prompt Optimization for LLM-Based Sequential Decision Making In-context Exploration-Exploitation for Reinforcement Learning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 654bebc2-e3da-48c5-a748-d3d4e50ee02e · outbound
Meta-Prompt Optimization for LLM-Based Sequential Decision Making Ambiguity-Aware In-Context Learning with Large Language Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4eceadef-b432-4c94-af6d-1a2370874ca4 · outbound
Meta-Prompt Optimization for LLM-Based Sequential Decision Making Task Facet Learning: A Structured Approach to Prompt Optimization
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c285bda6-aaf2-44ec-b807-bf17045ea1ae · outbound
Meta-Prompt Optimization for LLM-Based Sequential Decision Making Unresolved cited work
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ff14cd29-9328-44eb-87b2-62ad0667ac1a · outbound
Meta-Prompt Optimization for LLM-Based Sequential Decision Making Can large language models explore in-context?
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb9eb3b4-3c5e-4249-ae6f-3d199b64e5bb · outbound
Meta-Prompt Optimization for LLM-Based Sequential Decision Making The texts we have modified are highlighted in red
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9b2461c9-9827-4a92-aa5a-38452fb26acb · outbound
Meta-Prompt Optimization for LLM-Based Sequential Decision Making AgentBench: Evaluating LLMs as Agents
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6cc5a2ef-5aef-4f63-92dd-903ee2cf5cfc · outbound
Meta-Prompt Optimization for LLM-Based Sequential Decision Making P., Xie, Q., and Nowak, R
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2fb9a376-1f31-46f4-8570-bf2bd42546a8 · outbound
Meta-Prompt Optimization for LLM-Based Sequential Decision Making In-context Example Selection with Influences
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f5dde57-31f6-4e41-bdf7-e4b1d6d36c36 · outbound
Meta-Prompt Optimization for LLM-Based Sequential Decision Making Optimizing Instructions and Demonstrations for Multi-Stage Language Model Programs
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 784bf2ce-e540-4eed-957f-ceb6b1fe9cba · outbound
Meta-Prompt Optimization for LLM-Based Sequential Decision Making Hyperband-based Bayesian Optimization for Black-box Prompt Selection
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00c6afe2-3a42-4875-a0c0-26ca0322022e · outbound
Meta-Prompt Optimization for LLM-Based Sequential Decision Making Efficient Prompt Optimization Through the Lens of Best Arm Identification
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0a90b81-627e-42c7-b363-c5761d2ae73b · outbound
Meta-Prompt Optimization for LLM-Based Sequential Decision Making Transformers Can Learn Temporal Difference Methods for In-Context Reinforcement Learning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee64b18d-af2c-4530-8083-d969a5375018 · outbound
Meta-Prompt Optimization for LLM-Based Sequential Decision Making The Rise and Potential of Large Language Model Based Agents: A Survey
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eed582d5-f821-472c-ad03-5c6ff4c80787 · outbound
Meta-Prompt Optimization for LLM-Based Sequential Decision Making AgentGym: Evolving Large Language Model-based Agents across Diverse Environments
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da3d5e68-d7d4-4ad2-89cb-67ae10560aa3 · outbound
Meta-Prompt Optimization for LLM-Based Sequential Decision Making Beyond Numeric Rewards: In-Context Dueling Bandits with LLM Agents
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7242a53b-f5cf-48fb-99d8-336053bdde34 · outbound
Meta-Prompt Optimization for LLM-Based Sequential Decision Making Unlock- ing black-box prompt tuning efficiency via zeroth-order optimization
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7b5e3e55-8000-4d4f-8427-1f81d90481f0 · outbound
Meta-Prompt Optimization for LLM-Based Sequential Decision Making The different components in the prompt are explained in detail in App
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 754201a6-4ff8-495f-818b-5b3206d925dd · outbound
Meta-Prompt Optimization for LLM-Based Sequential Decision Making Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 80a28cc4-8303-444d-b549-dfc0310ba3c0 · outbound
Meta-Prompt Optimization for LLM-Based Sequential Decision Making PRewrite: Prompt Rewriting with Reinforcement Learning
Reference 1995
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 352d58f4-4954-4d01-9453-d9f5a2d16eba · outbound
Meta-Prompt Optimization for LLM-Based Sequential Decision Making Teach Better or Show Smarter? On Instructions and Exemplars in Automatic Prompt Optimization
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0f5423e-9fe3-47f0-8768-6d997c86d5ad · outbound
Meta-Prompt Optimization for LLM-Based Sequential Decision Making Prompt Optimization with Human Feedback
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d270bcf-542e-4ff5-92df-397c10ffe41f · outbound
Meta-Prompt Optimization for LLM-Based Sequential Decision Making Promptbreeder: Self-Referential Self-Improvement Via Prompt Evolution
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b52bdfb2-11c0-46d3-b3f9-7803b343a3a7 · outbound
Meta-Prompt Optimization for LLM-Based Sequential Decision Making Efficient Sequential Decision Making with Large Language Models
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75df76e5-3748-4981-8f55-a211fb62f66d · outbound
Meta-Prompt Optimization for LLM-Based Sequential Decision Making InstructZero: Efficient Instruction Optimization for Black-Box Large Language Models
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 217ebc3a-46c6-4f14-ba92-3dc281c7eb16 · inbound
MASPOB: Bandit-Based Prompt Optimization for Multi-Agent Systems with Graph Neural Networks Meta-Prompt Optimization for LLM-Based Sequential Decision Making
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f2cfcd6-14b5-451b-b840-9da92574f633 · inbound
ALSO: Adversarial Online Strategy Optimization for Social Agents Meta-Prompt Optimization for LLM-Based Sequential Decision Making
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.