Pith. sign in

Paper Citation Record · LEDGER

Rule Based Rewards for Language Model Safety

As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 4 inbound Pith citation observations for arXiv:2411.01111.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.01111 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 4 of 4 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T00:55:26.055928Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-21T22:10:42.078748Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation bb7de678-5cf5-4827-82b0-70da4b9c2303 · inbound

RLBFF: Binary Flexible Feedback to bridge between Human Feedback & Verifiable Rewards cites this paper.

RLBFF: Binary Flexible Feedback to bridge between Human Feedback & Verifiable Rewards Rule Based Rewards for Language Model Safety

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-21T22:10:42.080709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-21T22:09:47.649346Z digest=sha256:a11b355ce71f0e4edc76440a3d8f7f58f815c121b7ebc1bbe49e592f107e94ae

Observation 9af86ce4-e33d-4c79-8b70-148c9cabb665 · inbound

MentalThink: Shaping Thoughts in Mental SVG World cites this paper.

MentalThink: Shaping Thoughts in Mental SVG World Rule Based Rewards for Language Model Safety

Reference 251

Resolution
unresolved
no resolver link, observed 2026-07-12T01:50:59.184754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-12T01:50:59.184754Z digest=sha256:48e5bdfd71f51d12b61ce0a6325bcff4dd75c70c7d17a8271a79ec9dbf0bdf54

Observation 7deb59a2-68fe-428c-a488-23b16ef433e5 · inbound

Mach-Mind-4-Flash Technical Report cites this paper.

Mach-Mind-4-Flash Technical Report Rule Based Rewards for Language Model Safety

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-13T03:29:34.486347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T03:29:34.486347Z digest=sha256:af164d5de27993ffda78b540b4e51ac605c67eb14e646de199f47c24c3ac4b49

Observation 219d03dd-1472-40de-bdd5-1ad7d8a314ea · inbound

Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges cites this paper.

Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges Rule Based Rewards for Language Model Safety

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-03T00:55:26.055928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T00:55:26.055928Z digest=sha256:6d01f425ac782abdb10cfbb5384fed9ed72079b92b06f4d5cd69fd0fefd66351