Pith. sign in

Paper Citation Record · LEDGER

Training Language Models to Generate Text with Citations via Fine-grained Rewards

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2402.04315.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.04315 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T20:57:07.616813Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-23T21:08:25.887635Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a27c65ae-2341-40df-bf92-4389f7cc90c1 · inbound

Trustworthiness in Retrieval-Augmented Generation Systems: A Survey cites this paper.

Trustworthiness in Retrieval-Augmented Generation Systems: A Survey Training Language Models to Generate Text with Citations via Fine-grained Rewards

Reference 135

Resolution
verified exact
arxiv_id, observed 2026-05-23T21:08:25.890876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-23T21:08:11.787013Z digest=sha256:0de72c096bbce7da15f4ce34233161d995e3d0bd3de65c2c736273a62af72ceb

Observation e327d508-df15-4fcf-b359-77f8e7798027 · inbound

SelfCite: Self-Supervised Alignment for Context Attribution in Large Language Models cites this paper.

SelfCite: Self-Supervised Alignment for Context Attribution in Large Language Models Training Language Models to Generate Text with Citations via Fine-grained Rewards

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T20:57:07.616813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T20:57:07.616813Z digest=sha256:0a271dbb47b55491d91f2bd3699da1fa2e20817a0c7a060f393b5587c5fcf3d2

Observation 45617367-ef31-4fb2-9d91-6a7148d34bca · inbound

Lessons from Training Grounded LLMs with Verifiable Rewards cites this paper.

Lessons from Training Grounded LLMs with Verifiable Rewards Training Language Models to Generate Text with Citations via Fine-grained Rewards

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T23:59:48.358493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:59:48.358493Z digest=sha256:d6072c6106f3953054de05a64b58acfa9c07845f55cc6367bd82912117fa12e2

Observation 1d3dd4eb-ea7d-4ae0-af3f-ede33479728b · inbound

Cite Pretrain: Retrieval-Free Knowledge Attribution for Large Language Models cites this paper.

Cite Pretrain: Retrieval-Free Knowledge Attribution for Large Language Models Training Language Models to Generate Text with Citations via Fine-grained Rewards

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-19T07:42:09.116389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T07:39:55.633854Z digest=sha256:06bb190ec3e1a267adef5c1de1b7a94c7ad8a4124d75aa8e9a23414e252bc186

Observation a6831e99-0c4d-4a1a-a604-6fc20f69371b · inbound

Context Attribution with Multi-Armed Bandit Optimization cites this paper.

Context Attribution with Multi-Armed Bandit Optimization Training Language Models to Generate Text with Citations via Fine-grained Rewards

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-19T07:22:09.100951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T07:21:20.893732Z digest=sha256:6117d4090662acb86a51874f771391eb34bed9caed17cf7b13acc60f13e36550