Pith. sign in

Paper Citation Record · LEDGER

Cascade Reward Sampling for Efficient Decoding-Time Alignment

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2406.16306.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.16306 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:41:59.798320Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T21:18:58.020417Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b0ca6ada-dfb6-47b6-8910-fc9d13ec4de4 · inbound

Well Begun is Half Done: Low-resource Preference Alignment by Weak-to-Strong Decoding cites this paper.

Well Begun is Half Done: Low-resource Preference Alignment by Weak-to-Strong Decoding Cascade Reward Sampling for Efficient Decoding-Time Alignment

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:59.798320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:41:59.798320Z digest=sha256:945701ad15cda4e5f57ce935bec253f7d5fc1f7edcb882155673a71edaa67656

Observation 62c9e407-df19-4acd-a17f-b68465c7fbc0 · inbound

A Survey on Training-free Alignment of Large Language Models cites this paper.

A Survey on Training-free Alignment of Large Language Models Cascade Reward Sampling for Efficient Decoding-Time Alignment

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-05T21:18:43.127270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:18:43.127270Z digest=sha256:e2a32d8946a6e1ed9688cc78b3442512f5372e42e05427e751eabdbeaee536f0

Observation 48a19fec-95b3-4543-87ab-9fb66f07a8c8 · inbound

Simultaneous Multi-objective Alignment Across Verifiable and Non-verifiable Rewards cites this paper.

Simultaneous Multi-objective Alignment Across Verifiable and Non-verifiable Rewards Cascade Reward Sampling for Efficient Decoding-Time Alignment

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-04T13:15:44.323192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T13:15:44.323192Z digest=sha256:4f501da2525c0598f85405e22cae1722787717ceecf255eaaa97cf785b690465

Observation 001a68b9-b1f8-48bb-978c-6db05d923765 · inbound

Test-time reward-guided alignment of language models by importance sampling on pre-logit space cites this paper.

Test-time reward-guided alignment of language models by importance sampling on pre-logit space Cascade Reward Sampling for Efficient Decoding-Time Alignment

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-04T07:23:56.545588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T07:23:56.545588Z digest=sha256:fcbcf30861f29baf32ae8959202068f72ccdbc4ab8fcedd8e6f260bc0b3eb662

Observation d176487c-b34d-4627-a9d1-497da2ae4148 · inbound

Common-agency Games for Multi-Objective Test-Time Alignment cites this paper.

Common-agency Games for Multi-Objective Test-Time Alignment Cascade Reward Sampling for Efficient Decoding-Time Alignment

Reference 208

Resolution
verified exact
arxiv_id, observed 2026-05-15T06:15:06.574422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-15T06:14:53.685486Z digest=sha256:2e1349b219a01a509380dbefa19583658f3ed0ee5a2585da8a232d81f2e03694

Observation 9e2c6f30-50c2-4c7c-8324-6952ca20e787 · inbound

Gradient-Guided Reward Optimization for Inference-time Alignment cites this paper.

Gradient-Guided Reward Optimization for Inference-time Alignment Cascade Reward Sampling for Efficient Decoding-Time Alignment

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T01:17:30.641460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-27T16:43:04.984866Z digest=sha256:4279a23605651c7abeda037fa2546891136a5f3351cccb0905016a3d4cc970ee

Observation e3707ca7-1eda-43b4-b4a6-2ed0113b15ff · inbound

Multi-Objective Exploration and Preference Optimization via Mutual Information cites this paper.

Multi-Objective Exploration and Preference Optimization via Mutual Information Cascade Reward Sampling for Efficient Decoding-Time Alignment

Reference 56

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T21:18:58.021917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-07-03T21:17:46.551850Z digest=sha256:df1ffed0179d5b3a85e7cce2ecf2af1fa03818d006b6cd70291d054a728d3c11