Pith. sign in

Paper Citation Record · LEDGER

PARM: Multi-Objective Test-Time Alignment via Preference-Aware Autoregressive Reward Model

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2505.06274.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.06274 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:52:11.728494Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T20:57:23.591970Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5d3526ba-172c-4804-a7a6-c6175ca99b12 · inbound

Multi-objective Large Language Model Alignment with Hierarchical Experts cites this paper.

Multi-objective Large Language Model Alignment with Hierarchical Experts PARM: Multi-Objective Test-Time Alignment via Preference-Aware Autoregressive Reward Model

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T13:52:11.728494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:52:11.728494Z digest=sha256:22a0a79475147eacc9e916fc1c56d52a1d29c46a74642d6e48ecd2f3400c347a

Observation a7d21660-4ca9-4677-95f0-283d7ccf09ed · inbound

Simultaneous Multi-objective Alignment Across Verifiable and Non-verifiable Rewards cites this paper.

Simultaneous Multi-objective Alignment Across Verifiable and Non-verifiable Rewards PARM: Multi-Objective Test-Time Alignment via Preference-Aware Autoregressive Reward Model

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T13:15:44.901150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T13:15:44.901150Z digest=sha256:e85b6cba9606612500afff026af5f2f07d546195783d870dc595bc0bfe4425be

Observation 6ecc8cbb-578e-4582-a2fe-d1ecb16fa803 · inbound

RVPO: Risk-Sensitive Alignment via Variance Regularization cites this paper.

RVPO: Risk-Sensitive Alignment via Variance Regularization PARM: Multi-Objective Test-Time Alignment via Preference-Aware Autoregressive Reward Model

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:36:08.500675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T15:00:11.237293Z digest=sha256:400a5d4c7435ea986981170750e11945f0d34568fbae7951ed856d7e08ca347a

Observation ebb5b8e6-6e07-4919-a53e-6533adb932cb · inbound

Common-agency Games for Multi-Objective Test-Time Alignment cites this paper.

Common-agency Games for Multi-Objective Test-Time Alignment PARM: Multi-Objective Test-Time Alignment via Preference-Aware Autoregressive Reward Model

Reference 213

Resolution
verified exact
arxiv_id, observed 2026-05-15T06:15:06.656263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-15T06:14:53.685486Z digest=sha256:194f611133a723efbd328e3a51bb15e9f35704a87fb7411b455b77da6a5c620a

Observation 2dead4b7-254c-4dff-980a-6ab7be548dd1 · inbound

SURF: Steering the Scalarization Weight to Uniformly Traverse the Pareto Front cites this paper.

SURF: Steering the Scalarization Weight to Uniformly Traverse the Pareto Front PARM: Multi-Objective Test-Time Alignment via Preference-Aware Autoregressive Reward Model

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-21T06:44:00.905436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-21T06:42:15.135148Z digest=sha256:7187199a17e11c08772545f74e89dbbdcef55b845c48935133dd068957096b25

Observation 53b22dc3-7e38-48e9-a5a8-9618cb494251 · inbound

PAFO: Pareto Fairness Optimization for Personalized Reward Modeling cites this paper.

PAFO: Pareto Fairness Optimization for Personalized Reward Modeling PARM: Multi-Objective Test-Time Alignment via Preference-Aware Autoregressive Reward Model

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-07-02T20:57:23.593457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-27T20:00:05.900814Z digest=sha256:5dc9eeb7d7dff5ab2d69b146031b4842d4b13bf6800d28f0d9c37721efc0b121

Observation 0aa80fd7-ec0e-4e0f-8614-04a2223b75f1 · inbound

Inference-Time Policy Alignment for Fair Reinforcement Learning cites this paper.

Inference-Time Policy Alignment for Fair Reinforcement Learning PARM: Multi-Objective Test-Time Alignment via Preference-Aware Autoregressive Reward Model

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-04T01:10:43.876458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:10:43.876458Z digest=sha256:5faa0811637f8f673527a5c92871c830fefb5e28fce347e57cd0a125f7c5df2a