Pith. sign in

Paper Citation Record · LEDGER

Value Augmented Sampling for Language Model Alignment and Personalization

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2405.06639.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.06639 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:57:56.712904Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T21:18:58.014691Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 28d7c55c-27d9-4503-ad9d-13a8f3f271d3 · inbound

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs cites this paper.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs Value Augmented Sampling for Language Model Alignment and Personalization

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-22T01:20:52.085102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:e67222683cbfa38da9f5c2af7823b2a504179ca962b029bc0c43cf7ae8c70418

Observation 750e2506-feb8-4427-b82d-5e28a9a52517 · inbound

From Outcomes to Processes: Guiding PRM Learning from ORM for Inference-Time Alignment cites this paper.

From Outcomes to Processes: Guiding PRM Learning from ORM for Inference-Time Alignment Value Augmented Sampling for Language Model Alignment and Personalization

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T00:57:56.712904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:57:56.712904Z digest=sha256:5dbe792b69cb1305fd6086316b0f92ef196ff10646a95ad208b7fba2d11acaeb

Observation 20763f08-9e12-42df-a381-4d89d1010d84 · inbound

PICACO: Pluralistic In-Context Value Alignment of LLMs via Total Correlation Optimization cites this paper.

PICACO: Pluralistic In-Context Value Alignment of LLMs via Total Correlation Optimization Value Augmented Sampling for Language Model Alignment and Personalization

Reference 838

Resolution
unresolved
no resolver link, observed 2026-08-06T15:12:44.020713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:12:44.020713Z digest=sha256:9952b68cd5a53e9a195bdf00ff34a42c1238aadf8a90c6162a86c3251d43874e

Observation 76f764cb-197a-4e02-8bac-ad017610ec3a · inbound

Controlling Multimodal LLMs via Reward-guided Decoding cites this paper.

Controlling Multimodal LLMs via Reward-guided Decoding Value Augmented Sampling for Language Model Alignment and Personalization

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T19:52:50.217899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:52:50.217899Z digest=sha256:0222c4d80962750537f2535f676dc5140101abb5102157cbbf1bd4fedeb8074a

Observation 3271f3c0-6c7b-4189-9bf2-161b90471adc · inbound

Test-time reward-guided alignment of language models by importance sampling on pre-logit space cites this paper.

Test-time reward-guided alignment of language models by importance sampling on pre-logit space Value Augmented Sampling for Language Model Alignment and Personalization

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T07:23:56.418435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T07:23:56.418435Z digest=sha256:3bc30b124c622519e02c373fac44acdfeaefb22a43c56d71f0b938618ab0b9fe

Observation 2a5b5c44-7609-47d2-a03f-3aa6438459ea · inbound

Selective Safety Steering via Value-Filtered Decoding cites this paper.

Selective Safety Steering via Value-Filtered Decoding Value Augmented Sampling for Language Model Alignment and Personalization

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-06-30T21:05:03.894587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T21:03:28.381687Z digest=sha256:c883873e96445e7af4f4e7102aea134f38e29856a9da75cefc370b2cf993f8a9

Observation a8d36622-6492-4532-98fa-b1ca879b1058 · inbound

Selective Safety Steering via Value-Filtered Decoding cites this paper.

Selective Safety Steering via Value-Filtered Decoding Value Augmented Sampling for Language Model Alignment and Personalization

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-14T18:59:01.049397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T18:59:01.049397Z digest=sha256:16bef567f33177a84343bff8d78565ebf26060c6d2faddb3533f0608222f90ee

Observation a95c06d5-05af-4733-8f2e-31ca0590582c · inbound

Multi-Objective Exploration and Preference Optimization via Mutual Information cites this paper.

Multi-Objective Exploration and Preference Optimization via Mutual Information Value Augmented Sampling for Language Model Alignment and Personalization

Reference 36

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T21:18:58.016386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-07-03T21:17:46.551850Z digest=sha256:34ed5f5d95eb42efbf697d9d08099505f74369fdc892dd97dfba804b4ca9bf24

Observation d4c7758f-5e0d-4663-8a03-d149eaa72f8a · inbound

Safe Inference-Time Alignment via Lagrangian Reward Augmentation cites this paper.

Safe Inference-Time Alignment via Lagrangian Reward Augmentation Value Augmented Sampling for Language Model Alignment and Personalization

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-12T07:05:47.150308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-12T07:05:47.150308Z digest=sha256:8801fdfb8626f1a55ebc4b3dbd8892bb9a62d1b15b1363ae1b464788fab2c3ea