Pith. sign in

Paper Citation Record · LEDGER

Test-Time Preference Optimization: On-the-Fly Alignment via Iterative Textual Feedback

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2501.12895.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.12895 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T14:35:46.561276Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 37df1f07-04c0-4273-a16f-c936b5f43567 · inbound

TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards cites this paper.

TRPrompt: Bootstrapping Query-Aware Prompt Optimization from Textual Rewards Test-Time Preference Optimization: On-the-Fly Alignment via Iterative Textual Feedback

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T14:35:46.561276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:35:46.561276Z digest=sha256:8ed6d62cf1442db7a82040a7168dbf63b00373aecfc2302fa2abc17c355c9483

Observation 27ce3375-f590-421a-ae23-8a0be3fc2df3 · inbound

Temporal Self-Rewarding Language Models: Decoupling Chosen-Rejected via Past-Future cites this paper.

Temporal Self-Rewarding Language Models: Decoupling Chosen-Rejected via Past-Future Test-Time Preference Optimization: On-the-Fly Alignment via Iterative Textual Feedback

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T23:06:49.634866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:06:49.634866Z digest=sha256:ff0479aff047ab5ae0b8331c43224d370bfb03594f1223003d975b6f6e4bd800

Observation d672e86c-c8eb-4896-be7a-1cf2216cfc51 · inbound

Self-Reflective Generation at Test Time cites this paper.

Self-Reflective Generation at Test Time Test-Time Preference Optimization: On-the-Fly Alignment via Iterative Textual Feedback

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T12:41:42.645044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T12:41:42.645044Z digest=sha256:d71a9cba5462a6a2ba1d2d6e4fac94d8a07c2b958986470385c5a226ccf7c5c0

Observation 0254192a-1b87-4d76-9642-589dfc430098 · inbound

Test-Time Personalization: A Diagnostic Framework and Probabilistic Fix for Scaling Failures cites this paper.

Test-Time Personalization: A Diagnostic Framework and Probabilistic Fix for Scaling Failures Test-Time Preference Optimization: On-the-Fly Alignment via Iterative Textual Feedback

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:37:27.141583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T06:32:58.314552Z digest=sha256:49569ebd7d3492d92eaa6dd8abe64d5d7fa70126cf3160c8959565ea8b925ef5

Observation a1490d2d-09b9-478a-b525-c210e506a057 · inbound

Query-Conditioned Test-Time Self-Training for Large Language Models cites this paper.

Query-Conditioned Test-Time Self-Training for Large Language Models Test-Time Preference Optimization: On-the-Fly Alignment via Iterative Textual Feedback

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:07:54.157505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-14T20:04:51.248797Z digest=sha256:e8a4cefeecad75e6e33e9fe3ddb8abc6944101f270f56f46269e730a0dc046ab

Observation bf0c43dc-e9d1-40de-be3e-efd3ecbc565e · inbound

Query-Conditioned Test-Time Self-Training for Large Language Models cites this paper.

Query-Conditioned Test-Time Self-Training for Large Language Models Test-Time Preference Optimization: On-the-Fly Alignment via Iterative Textual Feedback

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-15T05:55:04.892930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T05:53:42.235498Z digest=sha256:3ad39775b854bcb14a055285b2b96bed6487ae65f376373205363b719363264f

Observation 498b22ae-5642-4616-9407-c67476c08194 · inbound

Prompt Governance? On Governing Technologies Governed by Natural Language cites this paper.

Prompt Governance? On Governing Technologies Governed by Natural Language Test-Time Preference Optimization: On-the-Fly Alignment via Iterative Textual Feedback

Reference 194

Resolution
verified exact
arxiv_id, observed 2026-07-01T08:25:32.896087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-01T08:17:10.481202Z digest=sha256:cecfbde14890c163b77121a3538f1bfcadfeecb4889d233d02fe80412355e42d

Observation d39d5845-4cc3-475f-9f63-1e07632e6e0b · inbound

Robust Critics: Defending LLMs Against Multi-Turn Attacks cites this paper.

Robust Critics: Defending LLMs Against Multi-Turn Attacks Test-Time Preference Optimization: On-the-Fly Alignment via Iterative Textual Feedback

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T13:17:51.602308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:17:51.602308Z digest=sha256:db3e0aee2aedf3b2f4d2232caf3c3765f66b78f61a5b07215c81b8f217a21dc4