Pith. sign in

Paper Citation Record · LEDGER

Sample-Efficient Alignment for LLMs

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 3 inbound Pith citation observations for arXiv:2411.01493.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.01493 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 3 of 3 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T23:03:53.158373Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T11:54:38.401628Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation d5bd172c-6ebc-44d1-a806-25d0f200b1b2 · inbound

PILAF: Optimal Human Preference Sampling for Reward Modeling cites this paper.

PILAF: Optimal Human Preference Sampling for Reward Modeling Sample-Efficient Alignment for LLMs

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-08T23:03:53.158373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T23:03:53.158373Z digest=sha256:479599eece3938e4a5f4211a39f169fff13a0ae987935f22de2b278eabb2cc3f

Observation 14811b16-46c7-4250-ad44-9592ac2e4446 · inbound

ActiveDPO: Active Direct Preference Optimization for Sample-Efficient Alignment cites this paper.

ActiveDPO: Active Direct Preference Optimization for Sample-Efficient Alignment Sample-Efficient Alignment for LLMs

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-22T01:10:51.573542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T01:06:19.756032Z digest=sha256:bdf247c212d47a8911b18f3cf229b9919f0032bf7a6e79ae6306e981e05811a3

Observation 0310cb40-3e92-40ed-970d-e43a98e0add8 · inbound

Active Learning for Stochastic Contextual Linear Bandits cites this paper.

Active Learning for Stochastic Contextual Linear Bandits Sample-Efficient Alignment for LLMs

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-06-30T11:54:38.403214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T11:49:35.327709Z digest=sha256:eaef52c69a4f6b048dd05f434e562936f98db9536624e4a040a096130d55db35