Pith. sign in

Paper Citation Record · LEDGER

Preference Optimization with Multi-Sample Comparisons

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 4 inbound Pith citation observations for arXiv:2410.12138.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.12138 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 4 of 4 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T10:23:15.759043Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-10T19:55:50.936521Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a327a773-3a93-4ab5-98b1-f6ace7e9e464 · inbound

GRAPE: Generalizing Robot Policy via Preference Alignment cites this paper.

GRAPE: Generalizing Robot Policy via Preference Alignment Preference Optimization with Multi-Sample Comparisons

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T10:23:15.759043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T10:23:15.759043Z digest=sha256:110d433c82f86a4f1c64424f367c802547c0b6441468975291b712a21eeaae08

Observation 40dc9c4e-02ea-4380-8cd1-2ee0716fdc1a · inbound

An Overview and Discussion on Using Large Language Models for Implementation Generation of Solutions to Open-Ended Problems cites this paper.

An Overview and Discussion on Using Large Language Models for Implementation Generation of Solutions to Open-Ended Problems Preference Optimization with Multi-Sample Comparisons

Reference 149

Resolution
unresolved
no resolver link, observed 2026-08-10T22:51:52.745518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:51:52.745518Z digest=sha256:7569b7785c982190e97f7cf07630984a544701e6d7ed1247d90266d7ec3fe014

Observation 17a0ddab-9425-4747-922f-b6e90b486444 · inbound

Beyond Reward Hacking: Causal Rewards for Large Language Model Alignment cites this paper.

Beyond Reward Hacking: Causal Rewards for Large Language Model Alignment Preference Optimization with Multi-Sample Comparisons

Reference 62

Resolution
verified exact
local_arxiv, observed 2026-08-10T19:55:50.944111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-10T19:55:50.776578Z digest=sha256:5869e91b8d8da4659ced172ce46b269f1f2c63e93c771f3db94d3c9b9293f9a4

Observation 1500d34c-0fbd-4106-9fd3-3215c5a3d052 · inbound

Test-Time Scaling via Error Localization cites this paper.

Test-Time Scaling via Error Localization Preference Optimization with Multi-Sample Comparisons

Reference 136

Resolution
unresolved
no resolver link, observed 2026-08-01T07:28:31.881359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:28:31.881359Z digest=sha256:5c3e0f49b143fc4015807c7d8e43323172dbf6af1662621e9b33be96496151f7