Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T14:35:46.612050Z
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 1 inbound Pith citation observation for arXiv:2507.18618.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T14:35:46.612050Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-18T01:22:24.750044Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-18T01:25:35.064596Z
22 of 22 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 6f33332a-0516-4f7d-9498-fa8e5bb0c015 · outbound
TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards Llama 3 model card
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6296edb4-af6d-4e32-ab5b-1ad97188e58c · outbound
TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards Training Verifiers to Solve Math Word Problems
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8344a737-9e84-4b6f-b48d-f63975534ed0 · outbound
TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards RLPrompt: Optimizing Discrete Text Prompts with Reinforcement Learning
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76266c98-a71b-4ed2-9484-b401b7d2576b · outbound
TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards PAL: Program-aided Language Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3c4b92e-6ada-4f3c-8db2-0c3ed141a70a · outbound
TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards Measuring mathematical problem solving with the math dataset
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83dd4e17-26f4-4e34-bd3d-eedec8a5a4a3 · outbound
TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards Large language models are zero-shot reasoners
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5cbbbc3-0e24-4f63-b75c-615f955499db · outbound
TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards QPO: Query-dependent Prompt Optimization via Multi-Loop Offline Reinforcement Learning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec94d46f-565c-4263-ad3a-7b9b41f0a9cb · outbound
TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards GSM-Plus: A Comprehensive Benchmark for Evaluating the Robustness of LLMs as Mathematical Problem Solvers
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37df1f07-04c0-4273-a16f-c936b5f43567 · outbound
TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards Test-Time Preference Optimization: On-the-Fly Alignment via Iterative Textual Feedback
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0403aedc-c0ea-4daa-b1ea-898a546205aa · outbound
TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards Guiding large language models via directional stimulus prompting
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation d352ddc5-1a6c-4c75-821f-74b79dfd677d · outbound
TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards Large language models as evolutionary optimizers
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation d7fce49b-7cf9-4e1d-9e00-d29ba594963f · outbound
TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards GPT-4 Technical Report
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4071650e-385f-4f46-81e5-67f3bcb6e6a5 · outbound
TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards Training language models to follow instructions with human feedback
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35e7ddcc-6812-40fe-ac18-ea223bc4bb2e · outbound
TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards Automatic Prompt Optimization with "Gradient Descent" and Beam Search
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b668c31b-1d49-4d8d-b04a-1f26549da0f2 · outbound
TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards Learning Performance-Improving Code Edits
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc10323f-8216-4edc-b237-cf8e162669af · outbound
TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards Query-dependent prompt evaluation and optimization with offline inverse rl
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 7f9cf577-3ccb-44e6-bf91-2d2d88ea5050 · outbound
TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards Chain-of-thought prompting elicits reasoning in large language models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 666e019f-72e8-4a95-9c53-e87a7e1cefd1 · outbound
TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards Improving Reward Models with Synthetic Critiques
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb151b7c-ce8a-4e0c-81c6-e9f2a5de8ffe · outbound
TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards TextGrad: Automatic "Differentiation" via Text
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f7d78b6-fc10-43ad-8ed1-9ccd5a3510d3 · outbound
TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards Gonzalez
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 54b1263a-dcb4-4797-ad2b-1c9f11c2f0fb · outbound
TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards TEMPERA: Test-Time Prompting via Reinforcement Learning
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afce6f4e-457e-43e1-9b3d-ed5e6664f6d0 · outbound
TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards Large Language Models Are Human-Level Prompt Engineers
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe82830d-55f4-4d8a-89cf-0cab59cc8683 · inbound
Prompt-R1: Collaborative Automatic Prompting Framework via End-to-end Reinforcement Learning TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.