Pith. sign in

Paper Citation Record · LEDGER

Learning When to Trust via Selective Context Preference Optimization

As of 7 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 0 inbound Pith citation observations for arXiv:2608.06377.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.06377 v1

Coverage vector

measured 12 of 12 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:10:44.120366Z

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

12 of 12 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation de5fff26-e8b1-47a2-bcc3-f095138614c9 · outbound

This paper cites SmolLM2: When Smol Goes Big -- Data-Centric Training of a Small Language Model.

Learning When to Trust via Selective Context Preference Optimization SmolLM2: When Smol Goes Big -- Data-Centric Training of a Small Language Model

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:43.127049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:43.127049Z digest=sha256:5303407177ef23ccc5f24df7b3485a563b8baf3f39dcf766ce7d42632d25b0e5

Observation 6e938e8c-db6b-4dc0-ac0c-5762692f62be · outbound

This paper cites Gemma 3 Technical Report.

Learning When to Trust via Selective Context Preference Optimization Gemma 3 Technical Report

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:43.503102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:43.503102Z digest=sha256:5d3fe957524a5dc27c76002c856501097d9ae2327892df8a44611b25251e2066

Observation ebb65bb2-77a6-48d0-b336-fc58a81573cc · outbound

This paper cites an unresolved cited work.

Learning When to Trust via Selective Context Preference Optimization Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:43.574558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:43.574558Z digest=sha256:7644004b117607a587bae58d37a0d5b818254ebfbdd0cb9e342b97991a065657

Observation 4d82cb14-4d33-4725-bd02-00f969e27112 · outbound

This paper cites User-Assistant Bias in LLMs.

Learning When to Trust via Selective Context Preference Optimization User-Assistant Bias in LLMs

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:43.689310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:43.689310Z digest=sha256:7e975a1bc73572edbfc96e676b6f023272145a8c6f5e8891939b16bbdcff2325

Observation 89561fec-e23d-4a63-a641-293eab41b16d · outbound

This paper cites Ignore Previous Prompt: Attack Techniques For Language Models.

Learning When to Trust via Selective Context Preference Optimization Ignore Previous Prompt: Attack Techniques For Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:43.752073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:43.752073Z digest=sha256:a26bf4757faa09d9cff3fe34effba193b8c29175ac967c7338cea0e6a3b7bff4

Observation f6d5ff6d-3413-4f9d-87c4-f67d0d645c0f · outbound

This paper cites an unresolved cited work.

Learning When to Trust via Selective Context Preference Optimization Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:10:44.536920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:10:43.838984Z digest=sha256:ec1ad7dfb120ad678b5a45c96c849b41128caee01378448ded308591afbd9427

Observation 6757a1d3-49ee-44ec-a07c-bcd9837e1cbe · outbound

This paper cites Sycophancy under Pressure: Evaluating and Mitigating Sycophantic Bias via Adversarial Dialogues in Scientific QA.

Learning When to Trust via Selective Context Preference Optimization Sycophancy under Pressure: Evaluating and Mitigating Sycophantic Bias via Adversarial Dialogues in Scientific QA

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:44.016598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:44.016598Z digest=sha256:85f59f0c2e00881bdbec46639b482f9786346cd4301aa0ea2b1e134391e14bac

Observation 11c4f194-8b8c-4dc9-8b93-a3cd0d378853 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Learning When to Trust via Selective Context Preference Optimization Training Verifiers to Solve Math Word Problems

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:43.266892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:43.266892Z digest=sha256:8555a9ae0cec2af6b4933191bd333485bb0dc93e234f64c643edbdb270e1ce4b

Observation c5bd237e-3b6e-4762-9592-bbeaf23a7bb4 · outbound

This paper cites an unresolved cited work.

Learning When to Trust via Selective Context Preference Optimization Unresolved cited work

Reference 2023

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:10:44.758253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:10:43.423326Z digest=sha256:7b14b258270d4567b5ca1cf631d004c94579f69e50f1d47a8747bea08609e853

Observation a2ffcca9-4f3c-4ee8-9a5e-e6bd2b924f5c · outbound

This paper cites Purified OPSD: On-Policy Self-Distillation Without Losing How to Think.

Learning When to Trust via Selective Context Preference Optimization Purified OPSD: On-Policy Self-Distillation Without Losing How to Think

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:43.931332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:43.931332Z digest=sha256:45c311252833376f408874437f6e0d620519448f12720eb50005518ffb40acf2

Observation dce39efe-da83-462a-b9b7-a328143600b6 · outbound

This paper cites Phi-4-reasoning Technical Report.

Learning When to Trust via Selective Context Preference Optimization Phi-4-reasoning Technical Report

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:43.082824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:43.082824Z digest=sha256:6115e97ce57ddf8992c5b7e646e9ce03c9204894c03505be30c39417056f3cff

Observation ad3daa9e-4fa4-4c73-a3e0-ae48bf3b87fe · outbound

This paper cites Self-Distilled Reasoner: On-Policy Self-Distillation for Large Language Models.

Learning When to Trust via Selective Context Preference Optimization Self-Distilled Reasoner: On-Policy Self-Distillation for Large Language Models

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:44.120366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:44.120366Z digest=sha256:bed251ccfcc0c409b00d0f1e50d7511d9074a7f473cf965fa8a7dd00da958abd

Pith citing papers

No inbound Pith citation observations are available.