Pith. sign in

Paper Citation Record · LEDGER

SteerLM: Attribute Conditioned SFT as an (User-Steerable) Alternative to RLHF

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2310.05344.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2310.05344 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T22:20:13.691866Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T12:43:25.917797Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 6eac74c1-acf7-456a-abaf-aaef16daf23b · inbound

BRiTE: Bootstrapping Reinforced Thinking Process to Enhance Language Model Reasoning cites this paper.

BRiTE: Bootstrapping Reinforced Thinking Process to Enhance Language Model Reasoning SteerLM: Attribute Conditioned SFT as an (User-Steerable) Alternative to RLHF

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-09T22:20:13.691866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:20:13.691866Z digest=sha256:d0c3538d1a9730fbdacd07f24a96d9a199e34cf00f7f96efca06d0534e878b3b

Observation 9026ddb6-4b52-45c5-b7a9-6d578d66c114 · inbound

AI Alignment at Your Discretion cites this paper.

AI Alignment at Your Discretion SteerLM: Attribute Conditioned SFT as an (User-Steerable) Alternative to RLHF

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-08T16:14:57.231708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:14:57.231708Z digest=sha256:da194cdac02e38184d9486ecf50ecb853894c96f2bc34808f347378222bdcdd6

Observation 1bc69fa1-00f7-492d-a817-92d8f5655d7e · inbound

CALMA: A Process for Deriving Context-aligned Axes for Language Model Alignment cites this paper.

CALMA: A Process for Deriving Context-aligned Axes for Language Model Alignment SteerLM: Attribute Conditioned SFT as an (User-Steerable) Alternative to RLHF

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T18:10:16.759470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:10:16.759470Z digest=sha256:75dc31bf0883b65f4a41492e2a9ac1b8afd7d8b13925d7c6ec235f15cb18ff6f

Observation 95ddb21d-32d7-4cd5-84cf-81fee0122848 · inbound

Boundary Suppression Asymmetry in Post-trained Assistants: Over-expansion as a Controllability Cost cites this paper.

Boundary Suppression Asymmetry in Post-trained Assistants: Over-expansion as a Controllability Cost SteerLM: Attribute Conditioned SFT as an (User-Steerable) Alternative to RLHF

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-06-29T12:43:25.919250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T12:36:18.979940Z digest=sha256:a33929d8491ebe1a724286ce6a5c20690c168e80ed33588d69873601047ea98c

Observation 3bb022fa-2cd9-4970-96fc-1dc62aa41205 · inbound

Federated Variational Preference Alignment with Gumbel-Softmax Prior for Personalized User Preferences cites this paper.

Federated Variational Preference Alignment with Gumbel-Softmax Prior for Personalized User Preferences SteerLM: Attribute Conditioned SFT as an (User-Steerable) Alternative to RLHF

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T23:22:46.645528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-28T23:21:35.338959Z digest=sha256:461b8c98e5c53e941055408e60dbfa1163fa863b508a0feb7b1c567aeb4e23cc

Observation 75b356a4-3b6e-4008-a853-ca5d470408ca · inbound

Step-Level Preference Learning for Generative Agents in Social Simulations cites this paper.

Step-Level Preference Learning for Generative Agents in Social Simulations SteerLM: Attribute Conditioned SFT as an (User-Steerable) Alternative to RLHF

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T02:00:25.187877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T02:00:25.187877Z digest=sha256:f2a864a0850ddf2674dd8093a97ed4c1bae7db41f348dffa5f418955bb3410f4