Pith. sign in

Paper Citation Record · LEDGER

TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards

As of 18 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 1 inbound Pith citation observation for arXiv:2507.18618.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.18618 v1

Coverage vector

measured 22 of 22 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T14:35:46.612050Z

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-18T01:22:24.750044Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T01:25:35.064596Z

Reference resolution

22 of 22 outbound references displayed

  • verified exact0
  • verified fuzzy4
  • unresolved18
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6f33332a-0516-4f7d-9498-fa8e5bb0c015 · outbound

This paper cites Llama 3 model card.

TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards Llama 3 model card

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T14:35:46.525111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:35:46.525111Z digest=sha256:53213e3aea360a33427b9945554c6367f4df31a5ff47a20f481180fdf28c0846

Observation 6296edb4-af6d-4e32-ab5b-1ad97188e58c · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards Training Verifiers to Solve Math Word Problems

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T14:35:46.529756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:35:46.529756Z digest=sha256:1045175bdaaa2d6e885d518f06911e4508161d69355ac03a1808628bead5cbe7

Observation 8344a737-9e84-4b6f-b48d-f63975534ed0 · outbound

This paper cites RLPrompt: Optimizing Discrete Text Prompts with Reinforcement Learning.

TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards RLPrompt: Optimizing Discrete Text Prompts with Reinforcement Learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T14:35:46.534226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:35:46.534226Z digest=sha256:0475e905873c564f2d453354ca39e53da8c0bdd9438c55ba70cecbf4f719f6a5

Observation 76266c98-a71b-4ed2-9484-b401b7d2576b · outbound

This paper cites PAL: Program-aided Language Models.

TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards PAL: Program-aided Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T14:35:46.538854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:35:46.538854Z digest=sha256:6d4aa79c557be9490929df975bc993caaa9881e42b1b34e5c7a2ee4141f0274e

Observation a3c4b92e-6ada-4f3c-8db2-0c3ed141a70a · outbound

This paper cites Measuring mathematical problem solving with the math dataset.

TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards Measuring mathematical problem solving with the math dataset

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T14:35:46.543097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:35:46.543097Z digest=sha256:b90e6e9aa12725a8a9c932f28fb4fed7ee017eeaf4349bc5ba8f0e526206e1d1

Observation 83dd4e17-26f4-4e34-bd3d-eedec8a5a4a3 · outbound

This paper cites Large language models are zero-shot reasoners.

TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards Large language models are zero-shot reasoners

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T14:35:46.547654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:35:46.547654Z digest=sha256:6f90dba7da13608f088a1816816431c9a91ca6d60333d84254cc119e3d5578e5

Observation f5cbbbc3-0e24-4f63-b75c-615f955499db · outbound

This paper cites QPO: Query-dependent Prompt Optimization via Multi-Loop Offline Reinforcement Learning.

TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards QPO: Query-dependent Prompt Optimization via Multi-Loop Offline Reinforcement Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T14:35:46.552148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:35:46.552148Z digest=sha256:990213a03d6da58735bef2d1adb27412faf6b926e8f84281e5f1261429f7326b

Observation ec94d46f-565c-4263-ad3a-7b9b41f0a9cb · outbound

This paper cites GSM-Plus: A Comprehensive Benchmark for Evaluating the Robustness of LLMs as Mathematical Problem Solvers.

TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards GSM-Plus: A Comprehensive Benchmark for Evaluating the Robustness of LLMs as Mathematical Problem Solvers

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T14:35:46.556756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:35:46.556756Z digest=sha256:91ec5d06fc1fd85991ed86796ccb0b5293f15b791b8707d9e2eb4fb8d6e1c47e

Observation 37df1f07-04c0-4273-a16f-c936b5f43567 · outbound

This paper cites Test-Time Preference Optimization: On-the-Fly Alignment via Iterative Textual Feedback.

TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards Test-Time Preference Optimization: On-the-Fly Alignment via Iterative Textual Feedback

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T14:35:46.561276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:35:46.561276Z digest=sha256:256f90a011d43ba5c8d63cf34006c8f526b8c519a07eccbce80074009390edaa

Observation 0403aedc-c0ea-4daa-b1ea-898a546205aa · outbound

This paper cites Guiding large language models via directional stimulus prompting.

TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards Guiding large language models via directional stimulus prompting

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:35:46.857731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T14:35:46.565344Z digest=sha256:b8a1191258c995869602ebd0949911a2c85057a81ffc48ea7a5b03e7be3e435e

Observation d352ddc5-1a6c-4c75-821f-74b79dfd677d · outbound

This paper cites Large language models as evolutionary optimizers.

TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards Large language models as evolutionary optimizers

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:35:46.845124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T14:35:46.568951Z digest=sha256:054fed9a93b96ba5f13babb6e25fc506cfb51562f8d1e181fcad999e4ac85b0c

Observation d7fce49b-7cf9-4e1d-9e00-d29ba594963f · outbound

This paper cites GPT-4 Technical Report.

TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards GPT-4 Technical Report

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T14:35:46.572930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:35:46.572930Z digest=sha256:b24e4cbfdcbda46f2dd8b4bdad1fd6b3204049c8bdf4668f9685ce334f4b519f

Observation 4071650e-385f-4f46-81e5-67f3bcb6e6a5 · outbound

This paper cites Training language models to follow instructions with human feedback.

TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards Training language models to follow instructions with human feedback

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T14:35:46.576923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:35:46.576923Z digest=sha256:70a686b1c6b8e8bce8ea82a61347f1e92de7101f467b8443513b461c0c8082ef

Observation 35e7ddcc-6812-40fe-ac18-ea223bc4bb2e · outbound

This paper cites Automatic Prompt Optimization with "Gradient Descent" and Beam Search.

TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards Automatic Prompt Optimization with "Gradient Descent" and Beam Search

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T14:35:46.580669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:35:46.580669Z digest=sha256:b46db974492b81d466cabe32f45d949cb026e88b33c48c1270edbe265c66e2ff

Observation b668c31b-1d49-4d8d-b04a-1f26549da0f2 · outbound

This paper cites Learning Performance-Improving Code Edits.

TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards Learning Performance-Improving Code Edits

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T14:35:46.584738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:35:46.584738Z digest=sha256:d701512ef22883f3e99b75661f7ee33e4ccbe5879c22bd58d70feee2a659f3db

Observation dc10323f-8216-4edc-b237-cf8e162669af · outbound

This paper cites Query-dependent prompt evaluation and optimization with offline inverse rl.

TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards Query-dependent prompt evaluation and optimization with offline inverse rl

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:35:46.824655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T14:35:46.588697Z digest=sha256:1d6a0910b314a21ffb9cba54e63fae304d38f1c9d86608297563a712f3d29ad3

Observation 7f9cf577-3ccb-44e6-bf91-2d2d88ea5050 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards Chain-of-thought prompting elicits reasoning in large language models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T14:35:46.592339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:35:46.592339Z digest=sha256:21d341c710ebe502ea20e12b8eba68c1e598cb8b3b2e2d344a58333afc1859c6

Observation 666e019f-72e8-4a95-9c53-e87a7e1cefd1 · outbound

This paper cites Improving Reward Models with Synthetic Critiques.

TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards Improving Reward Models with Synthetic Critiques

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T14:35:46.596042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:35:46.596042Z digest=sha256:374f6fa3762c2e4ca622195f05cf6bf325fc915cd191ecd0eb9fc9b9e7ec01fd

Observation fb151b7c-ce8a-4e0c-81c6-e9f2a5de8ffe · outbound

This paper cites TextGrad: Automatic "Differentiation" via Text.

TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards TextGrad: Automatic "Differentiation" via Text

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T14:35:46.600036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:35:46.600036Z digest=sha256:821969f3c46b012ad4b7e2e42d2a5681320beb45b5a0e0cb3da7d8bd21bcfc4c

Observation 0f7d78b6-fc10-43ad-8ed1-9ccd5a3510d3 · outbound

This paper cites Gonzalez.

TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards Gonzalez

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:35:46.803447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T14:35:46.604373Z digest=sha256:200d1eba1d44715a8ae5f9b3ffe2caaf3c4dd68d4ff6ffe626e70ab999917a0d

Observation 54b1263a-dcb4-4797-ad2b-1c9f11c2f0fb · outbound

This paper cites TEMPERA: Test-Time Prompting via Reinforcement Learning.

TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards TEMPERA: Test-Time Prompting via Reinforcement Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T14:35:46.608164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:35:46.608164Z digest=sha256:c69d453c18cc2cb332f1f597828ca9cebd1dbfd1312bacba3c1a99f065af8cb8

Observation afce6f4e-457e-43e1-9b3d-ed5e6664f6d0 · outbound

This paper cites Large Language Models Are Human-Level Prompt Engineers.

TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards Large Language Models Are Human-Level Prompt Engineers

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T14:35:46.612050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:35:46.612050Z digest=sha256:aef7db0495826922c88b10259404e013699178d4d0814af00068690e083aa034

Pith citing papers

Observation fe82830d-55f4-4d8a-89cf-0cab59cc8683 · inbound

Prompt-R1: Collaborative Automatic Prompting Framework via End-to-end Reinforcement Learning cites this paper.

Prompt-R1: Collaborative Automatic Prompting Framework via End-to-end Reinforcement Learning TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-18T01:25:35.067022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T01:22:24.750044Z digest=sha256:6bab365587c29ff0c6ed377453e75d9c4c0325bb7953f6ce87217d94424b7a09