Pith. sign in

Paper Citation Record · LEDGER

Soft Best-of-n Sampling for Model Alignment

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2505.03156.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.03156 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T10:22:38.298599Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T20:21:09.732556Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b4018f26-4601-459e-a03d-bbf657a72bf4 · inbound

Connections between reinforcement learning with feedback,test-time scaling, and diffusion guidance: An anthology cites this paper.

Connections between reinforcement learning with feedback,test-time scaling, and diffusion guidance: An anthology Soft Best-of-n Sampling for Model Alignment

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T10:22:38.298599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:22:38.298599Z digest=sha256:60d5ba0ca0ef999b74ea30b7c17ccebdcbb6c4fe08a1715bcdd596ae1cb19c69

Observation 709c5b54-ebe8-4c6f-b65e-954fa42b7bde · inbound

VIGOR: VIdeo Geometry-Oriented Reward for Temporal Generative Alignment cites this paper.

VIGOR: VIdeo Geometry-Oriented Reward for Temporal Generative Alignment Soft Best-of-n Sampling for Model Alignment

Reference 40

Resolution
unresolved
no resolver link, observed 2026-07-13T23:54:26.281624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T23:54:26.281624Z digest=sha256:b7c8c5c4e2924ef9b43edc5c8c8e7eff21146bd05f22f3af3e06d9dc15df8e64

Observation a7e70538-ab06-4cfb-a7f6-7b697f5e8681 · inbound

Multimodal Diffusion to Mutually Enhance Polarized Light and Low Resolution EBSD Data cites this paper.

Multimodal Diffusion to Mutually Enhance Polarized Light and Low Resolution EBSD Data Soft Best-of-n Sampling for Model Alignment

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:21:09.735874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T09:27:46.993773Z digest=sha256:98a5e3dd59c1f186b42eef15399322d51ce3934aef200de9a4c0904897a1ae3c

Observation 8c0b00f4-c033-4652-ab2c-fe81225ccd6e · inbound

VLA-ATTC: Adaptive Test-Time Compute for VLA Models with Relative Action Critic Model cites this paper.

VLA-ATTC: Adaptive Test-Time Compute for VLA Models with Relative Action Critic Model Soft Best-of-n Sampling for Model Alignment

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:46:06.318110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-09T15:10:16.533927Z digest=sha256:124f112d052a32c0e3ae9b0266c9cbc40873b2f2a58592717e0a27681c5e1a91

Observation ef5d32b2-c04c-47a7-b003-7bfa0784736e · inbound

Theoretical Limits of Language Model Alignment cites this paper.

Theoretical Limits of Language Model Alignment Soft Best-of-n Sampling for Model Alignment

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:30:56.867358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T01:18:37.614335Z digest=sha256:a58c796033b940f538698f3b752565187bcf897797866e357dfce8fdbc4d54c5

Observation 9ed9df29-4211-437d-97f5-2d7e42b33539 · inbound

Best-of-Better-$N$: Generating Pre-Aligned Responses with In-Context Learning cites this paper.

Best-of-Better-$N$: Generating Pre-Aligned Responses with In-Context Learning Soft Best-of-n Sampling for Model Alignment

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-12T02:23:04.515705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T02:23:04.515705Z digest=sha256:bd01c4fceeb98dacacbf67a561d0c1048241df8e5a1ca932ec263b8fc6c70851