Pith. sign in

Paper Citation Record · LEDGER

Assessing Robustness to Spurious Correlations in Post-Training Language Models

As of 20 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 1 inbound Pith citation observation for arXiv:2505.05704.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.05704 v1

Coverage vector

measured 14 of 14 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T23:03:24.526349Z

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-28T23:30:21.543618Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-28T23:32:47.032240Z

Reference resolution

14 of 14 outbound references displayed

  • verified exact0
  • verified fuzzy4
  • unresolved10
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4f9c4a02-e199-4b10-8de7-e51a3d92bd68 · outbound

This paper cites Q u AC : Question answering in context.

Assessing Robustness to Spurious Correlations in Post-Training Language Models Q u AC : Question answering in context

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T23:03:24.421499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:03:24.421499Z digest=sha256:dd701ee66fb4df1196f3297e6221b5a87d31639287c8d14cf1bff350beb9be81

Observation 918390ff-934f-43ba-a792-59120cd80f46 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Assessing Robustness to Spurious Correlations in Post-Training Language Models Training Verifiers to Solve Math Word Problems

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T23:03:24.428327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:03:24.428327Z digest=sha256:693684566cdc82aa908531271490a1b10a2a697c528fdfa3c43ca06422b2ab6b

Observation 28c6e3ad-7d0c-4613-bf0e-b511fb617407 · outbound

This paper cites The Llama 3 Herd of Models.

Assessing Robustness to Spurious Correlations in Post-Training Language Models The Llama 3 Herd of Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T23:03:24.437603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:03:24.437603Z digest=sha256:038e272b4da003536fb628d8135c54077641a193c217c748f7b75339360fd082

Observation 5047f68b-77f5-4f58-b76e-91a4845b435d · outbound

This paper cites Length-controlled alpacaeval: A simple way to debias automatic evaluators.

Assessing Robustness to Spurious Correlations in Post-Training Language Models Length-controlled alpacaeval: A simple way to debias automatic evaluators

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:03:24.968227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T23:03:24.450034Z digest=sha256:ee38f6fabba3c086a9cee0a5f78089c362ab5b0bcd6eebe045a3c8993855b025

Observation cf23ad11-e014-40af-a4fb-16e7acaa3587 · outbound

This paper cites KTO: Model Alignment as Prospect Theoretic Optimization.

Assessing Robustness to Spurious Correlations in Post-Training Language Models KTO: Model Alignment as Prospect Theoretic Optimization

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T23:03:24.458096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:03:24.458096Z digest=sha256:ebafd182ef267c8d3eb2295b68f216946daf8388f1160e992c89a03c4a25e36c

Observation 9482c36c-b441-4587-b731-92c1c353c56b · outbound

This paper cites RewardBench: Evaluating Reward Models for Language Modeling.

Assessing Robustness to Spurious Correlations in Post-Training Language Models RewardBench: Evaluating Reward Models for Language Modeling

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T23:03:24.467171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:03:24.467171Z digest=sha256:c99b7382722dc63dbcaa41bb93f29553db03b4a1b6eeda9c3c16338908fe3353

Observation ddffd884-7243-461a-8941-2e1d74a005e6 · outbound

This paper cites Thomas McCoy, Ellie Pavlick, and Tal Linzen.

Assessing Robustness to Spurious Correlations in Post-Training Language Models Thomas McCoy, Ellie Pavlick, and Tal Linzen

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:03:24.942626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T23:03:24.473293Z digest=sha256:8e4c4ef98b7e2097c58bd584da7abc5f039d932ef4e338e7814cb2ef4320d6b6

Observation 435b756c-2610-4c45-a814-0b3c88600217 · outbound

This paper cites Disentangling length from quality in direct preference optimization.

Assessing Robustness to Spurious Correlations in Post-Training Language Models Disentangling length from quality in direct preference optimization

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:03:24.919607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T23:03:24.482226Z digest=sha256:6d52340cf2b0ec788b82072ee518296d9fa4b493a9a93ecc1fc3ab20619f2972

Observation 586f679e-d00d-4b9b-bd81-d2bb6b833ea4 · outbound

This paper cites Direct preference optimization: Your language model is secretly a reward model.

Assessing Robustness to Spurious Correlations in Post-Training Language Models Direct preference optimization: Your language model is secretly a reward model

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T23:03:24.491313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:03:24.491313Z digest=sha256:756904273aba599074e3a32b81e0c3a47669e920dd3c80e449b5ce5dc50bf47e

Observation b7168d6a-0c0a-4d89-9c6a-ef24c4e98ed4 · outbound

This paper cites A long way to go: Investigating length correlations in rlhf.

Assessing Robustness to Spurious Correlations in Post-Training Language Models A long way to go: Investigating length correlations in rlhf

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:03:24.878302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T23:03:24.497244Z digest=sha256:514fed350927b55365a563d63d6d18425a9e1d4307021d16c03e2f75ed3f9a9b

Observation a6c6e277-36dc-44e3-a165-36c5d13088a6 · outbound

This paper cites Learning to summarize from human feedback.

Assessing Robustness to Spurious Correlations in Post-Training Language Models Learning to summarize from human feedback

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T23:03:24.505753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:03:24.505753Z digest=sha256:0ba3a583a7fcc9cc66903c9857b2bd7b88296c2b5da3763cba3062444b03a112

Observation 7fa0350d-603e-4085-b5c6-aac76cd6ea04 · outbound

This paper cites Finetuned Language Models Are Zero-Shot Learners.

Assessing Robustness to Spurious Correlations in Post-Training Language Models Finetuned Language Models Are Zero-Shot Learners

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T23:03:24.512009Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:03:24.512009Z digest=sha256:af6f578d7c50bab4411316c466fcc1b7261e0a2f862c3e245632a29050ff1d9f

Observation c059594d-c1e2-4a7e-b0dc-09f8e694395e · outbound

This paper cites COLLIE: Systematic Construction of Constrained Text Generation Tasks.

Assessing Robustness to Spurious Correlations in Post-Training Language Models COLLIE: Systematic Construction of Constrained Text Generation Tasks

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T23:03:24.520176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:03:24.520176Z digest=sha256:7f233913662ff78f1154fcfd928767aec1846ba1d3a4792f6e4b92a1fbc720f2

Observation 773afcba-de58-4edc-a8f5-8ab5169ee436 · outbound

This paper cites Instruction tuning for large language models: A survey.

Assessing Robustness to Spurious Correlations in Post-Training Language Models Instruction tuning for large language models: A survey

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T23:03:24.526349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:03:24.526349Z digest=sha256:ab0ecfefb65838a6180d9d6531d61612f67d759531ab6051391aaca8ca119228

Pith citing papers

Observation e2412212-bf33-4c1c-9df9-af392fe57745 · inbound

Shortcuts in the Tail: Debiasing via Post-Hoc Spectral Compression of Fine-Tuning Updates cites this paper.

Shortcuts in the Tail: Debiasing via Post-Hoc Spectral Compression of Fine-Tuning Updates Assessing Robustness to Spurious Correlations in Post-Training Language Models

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-06-28T23:32:47.033791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T23:30:21.543618Z digest=sha256:7baffd8f6b028a73be762845ec037fc946b224c50e77df54d0b729f7e2eae525