Pith. sign in

Paper Citation Record · LEDGER

Learning When to Trust in Contextual Social Bandits

As of 19 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 0 inbound Pith citation observations for arXiv:2603.13356.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2603.13356 v2

Coverage vector

measured 12 of 12 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-15T12:59:04.484292Z

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

12 of 12 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ae9fa60c-e91c-4c2c-9bee-4bb89b3ced2a · outbound

This paper cites Concrete Problems in AI Safety.

Learning When to Trust in Contextual Social Bandits Concrete Problems in AI Safety

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-15T12:59:04.484292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T12:59:04.484292Z digest=sha256:aea1f08e0257297cafe632a231c157b8ed972d822f9a13afcf14ba6c83981687

Observation ad294c40-e4c1-4e54-9533-7e74bf8a3416 · outbound

This paper cites Constitutional AI: Harmlessness from AI Feedback.

Learning When to Trust in Contextual Social Bandits Constitutional AI: Harmlessness from AI Feedback

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-15T12:59:04.484292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T12:59:04.484292Z digest=sha256:c89e5ac3f0430eb0f7288fa63f9ed8f85902599f35e624cac0a317b47ef5748b

Observation ca93db08-baeb-405b-b3ec-e9cc1a04fb04 · outbound

This paper cites Measuring Progress on Scalable Oversight for Large Language Models.

Learning When to Trust in Contextual Social Bandits Measuring Progress on Scalable Oversight for Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-15T12:59:04.484292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T12:59:04.484292Z digest=sha256:d35af2193de5db1509600cf3f35e51210812ee7fc6f91186d4e47b2aaef5d728

Observation 42e08fb8-2846-423b-9d71-894d5f1b9948 · outbound

This paper cites Towards Understanding the Robustness of LLM-based Evaluations under Perturbations.

Learning When to Trust in Contextual Social Bandits Towards Understanding the Robustness of LLM-based Evaluations under Perturbations

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-15T12:59:04.484292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T12:59:04.484292Z digest=sha256:dd02641636c6caba4832df9fab3174c7199f60f6ee7fce679b4aeeac1d9d72c1

Observation 67e7adb4-8275-450c-b2a8-884ba2dd578d · outbound

This paper cites Objective decoupling in social reinforcement learning: Recovering ground truth from sycophantic majorities.arXiv preprint arXiv:2602.08092,.

Learning When to Trust in Contextual Social Bandits Objective decoupling in social reinforcement learning: Recovering ground truth from sycophantic majorities.arXiv preprint arXiv:2602.08092,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-15T12:59:04.484292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T12:59:04.484292Z digest=sha256:7ce47e108132ea67b044f9b7aaec3aa2490229134aa9e073f0a6cf54cbbbea98

Observation aa673148-ca26-4042-be9a-fefa213a439c · outbound

This paper cites an unresolved cited work.

Learning When to Trust in Contextual Social Bandits Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-15T12:59:04.484292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T12:59:04.484292Z digest=sha256:5fd1cd953bcde803c9179a1a1414b994c4a9cade8328639f2500ea877b203ee7

Observation af9cc9f0-4faa-4402-90e9-4a5b49d5fb5f · outbound

This paper cites Fairness and Bias in Truth Discovery Algorithms: An Experimental Analysis.

Learning When to Trust in Contextual Social Bandits Fairness and Bias in Truth Discovery Algorithms: An Experimental Analysis

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-15T12:59:04.484292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T12:59:04.484292Z digest=sha256:a4ed54717bf5a3c08ff2cea45d06dd1c3447fc4e8a4d2832e713addf637106c3

Observation 9e9c2a5f-f1a0-423b-834c-23cf77d26b75 · outbound

This paper cites Provably optimal algorithms for generalized linear contextual bandits.

Learning When to Trust in Contextual Social Bandits Provably optimal algorithms for generalized linear contextual bandits

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-15T12:59:04.484292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T12:59:04.484292Z digest=sha256:6fe8622d330443f2f347138b8b9063f54ed9584c856babac7a3562c7e1b0f06d

Observation e7d48e3e-edbd-4f3c-bc2d-3b7341c5eefb · outbound

This paper cites AI Alignment through Reinforcement Learning from Human Feedback? Contradictions and Limitations.

Learning When to Trust in Contextual Social Bandits AI Alignment through Reinforcement Learning from Human Feedback? Contradictions and Limitations

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-15T12:59:04.484292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T12:59:04.484292Z digest=sha256:ffe3390094b9ddc6be4b20830988fbd800454ae876b75539dbd7d0a11afa9e14

Observation db83a34c-cb96-493f-889b-35c0dd745e64 · outbound

This paper cites Discovering language model behaviors with model-written evaluations.

Learning When to Trust in Contextual Social Bandits Discovering language model behaviors with model-written evaluations

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-15T12:59:04.484292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T12:59:04.484292Z digest=sha256:3f5fa6e52682f090d7b2e477d8e5fc79b584f553afc383aa54d8d3dbe6c85b5e

Observation ff1ecbac-5d5c-487b-875f-5e4162f64762 · outbound

This paper cites Towards Understanding Sycophancy in Language Models.

Learning When to Trust in Contextual Social Bandits Towards Understanding Sycophancy in Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-15T12:59:04.484292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T12:59:04.484292Z digest=sha256:6281a905bf44124952c21e6538302e239a69fb7efc58b2c6bfa8c4a35d12914d

Observation 3bedcf98-d033-4bef-9643-7864b0f633fe · outbound

This paper cites Antagonising explanation and revealing bias directly through sequencing and multimodal inference.

Learning When to Trust in Contextual Social Bandits Antagonising explanation and revealing bias directly through sequencing and multimodal inference

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-15T12:59:04.484292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T12:59:04.484292Z digest=sha256:3b3cd904924f917c1f62c9ace685d4fae9b20c1a72323529ab976c68a2f082cf

Pith citing papers

No inbound Pith citation observations are available.