Pith. sign in

Paper Citation Record · LEDGER

Learning When to Trust via Selective Context Preference Optimization

As of 7 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 0 inbound Pith citation observations for arXiv:2608.06377.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.06377 v1

Coverage vector

measured 12 of 12 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:10:44.120366Z

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

12 of 12 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation de5fff26-e8b1-47a2-bcc3-f095138614c9 · outbound

This paper cites SmolLM2: When Smol Goes Big -- Data-Centric Training of a Small Language Model.

Learning When to Trust via Selective Context Preference Optimization SmolLM2: When Smol Goes Big -- Data-Centric Training of a Small Language Model

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:43.127049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:43.127049Z digest=sha256:3014eae6d4e111ddb5e49f2e02968f35dedaa54c112eec3d8760a66fcdcf7663

Observation 6e938e8c-db6b-4dc0-ac0c-5762692f62be · outbound

This paper cites Gemma 3 Technical Report.

Learning When to Trust via Selective Context Preference Optimization Gemma 3 Technical Report

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:43.503102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:43.503102Z digest=sha256:ce0ed27f8d86ef7e2c077d0aead51610e7f5c79aea9619f88b66bbaa325f7e7f

Observation ebb65bb2-77a6-48d0-b336-fc58a81573cc · outbound

This paper cites an unresolved cited work.

Learning When to Trust via Selective Context Preference Optimization Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:43.574558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:43.574558Z digest=sha256:a037ee101086ab061998f8417a63216b7aaacafba3b397b9974874c394e378a3

Observation 4d82cb14-4d33-4725-bd02-00f969e27112 · outbound

This paper cites User-Assistant Bias in LLMs.

Learning When to Trust via Selective Context Preference Optimization User-Assistant Bias in LLMs

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:43.689310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:43.689310Z digest=sha256:fe9b599fac414193f7f24098c55143fb0d65d7e1c8503947fe279f652b549dc1

Observation 89561fec-e23d-4a63-a641-293eab41b16d · outbound

This paper cites Ignore Previous Prompt: Attack Techniques For Language Models.

Learning When to Trust via Selective Context Preference Optimization Ignore Previous Prompt: Attack Techniques For Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:43.752073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:43.752073Z digest=sha256:12ed632e5f5d8fd8602fa0bb28cfc4664bbf939685df0b2e2175aa583ba18d13

Observation f6d5ff6d-3413-4f9d-87c4-f67d0d645c0f · outbound

This paper cites an unresolved cited work.

Learning When to Trust via Selective Context Preference Optimization Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:10:44.536920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:10:43.838984Z digest=sha256:4a7a73dd60c8b6175244e2fc8e40a44743d994ccdc21ce9a0ad4000311a9e142

Observation 6757a1d3-49ee-44ec-a07c-bcd9837e1cbe · outbound

This paper cites Sycophancy under Pressure: Evaluating and Mitigating Sycophantic Bias via Adversarial Dialogues in Scientific QA.

Learning When to Trust via Selective Context Preference Optimization Sycophancy under Pressure: Evaluating and Mitigating Sycophantic Bias via Adversarial Dialogues in Scientific QA

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:44.016598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:44.016598Z digest=sha256:32a0b42525fe972785e6d2d9766e5c98d48a002ea503499371a45ca3727c28d0

Observation 11c4f194-8b8c-4dc9-8b93-a3cd0d378853 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Learning When to Trust via Selective Context Preference Optimization Training Verifiers to Solve Math Word Problems

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:43.266892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:43.266892Z digest=sha256:08676a6e0f478b0a66bd5e771710d874f927571cd41826b83ef389dd418e5aa8

Observation c5bd237e-3b6e-4762-9592-bbeaf23a7bb4 · outbound

This paper cites an unresolved cited work.

Learning When to Trust via Selective Context Preference Optimization Unresolved cited work

Reference 2023

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:10:44.758253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:10:43.423326Z digest=sha256:29651f23260744defe689e89103f7179362b6ed7167c0c2fe5b281f8640acede

Observation a2ffcca9-4f3c-4ee8-9a5e-e6bd2b924f5c · outbound

This paper cites Purified OPSD: On-Policy Self-Distillation Without Losing How to Think.

Learning When to Trust via Selective Context Preference Optimization Purified OPSD: On-Policy Self-Distillation Without Losing How to Think

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:43.931332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:43.931332Z digest=sha256:991f094a2c45948ae59c13730a74437db4448e88e1861012cff3342e1a719a7e

Observation dce39efe-da83-462a-b9b7-a328143600b6 · outbound

This paper cites Phi-4-reasoning Technical Report.

Learning When to Trust via Selective Context Preference Optimization Phi-4-reasoning Technical Report

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:43.082824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:43.082824Z digest=sha256:74e90a73a8b5cae80503e190c68bc2b952da42313bc30aff2c53d12fa936a96b

Observation ad3daa9e-4fa4-4c73-a3e0-ae48bf3b87fe · outbound

This paper cites Self-Distilled Reasoner: On-Policy Self-Distillation for Large Language Models.

Learning When to Trust via Selective Context Preference Optimization Self-Distilled Reasoner: On-Policy Self-Distillation for Large Language Models

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:44.120366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:44.120366Z digest=sha256:b209470d34a577c3d8cca66c1a9d72a6047e81d04ad241fe41ab49b414441048

Pith citing papers

No inbound Pith citation observations are available.