Pith. sign in

Paper Citation Record · LEDGER

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time

As of 8 August 2026, this Paper Citation Record lists 33 of 33 outbound references and 1 inbound Pith citation observation for arXiv:2505.23729.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.23729 v2

Coverage vector

measured 33 of 33 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:44:21.494870Z

measured 34 of 34 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-07T16:37:58.860183Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T23:36:36.042184Z

Reference resolution

33 of 33 outbound references displayed

  • verified exact1
  • verified fuzzy4
  • unresolved28
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 891401d7-bf5f-47c7-9988-f9ac8ed458a0 · outbound

This paper cites GPT-4 Technical Report.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:19.523685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:19.523685Z digest=sha256:caae2568f36c48f64a216bf1ab81f33f25c82a68a6fbedfd814f2a9903cf8bb6

Observation ee71dafe-1d62-4a28-8fbe-9dbfed2a7323 · outbound

This paper cites A General Language Assistant as a Laboratory for Alignment.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time A General Language Assistant as a Laboratory for Alignment

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:19.570449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:19.570449Z digest=sha256:6c684b2ba67580550ee3a866da4aeffd78260cb68bcb67477b0110ad019ed6f2

Observation 337d917b-3659-4729-a39a-759dfda77eb1 · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:19.670874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:19.670874Z digest=sha256:452eec41b2edffa3629e0dce79eb6f7f7c36dd85cbc644713a04bd9c6aca6a56

Observation 1c67c8a4-6df2-4479-8e87-fd1722545232 · outbound

This paper cites an unresolved cited work.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:19.715883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:19.715883Z digest=sha256:57fe943413728f99e7b624b5e0cad3844f2b76951dd4477f16664c69209c5c5e

Observation 3402ccd8-14fa-4630-85fb-68f2df3180cb · outbound

This paper cites Transfer Q Star: Principled Decoding for LLM Alignment.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time Transfer Q Star: Principled Decoding for LLM Alignment

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:19.777137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:19.777137Z digest=sha256:50aa918113614158506fdad4ac37a4473be878c53efa45d0fff865327fe86647

Observation c8a48e54-160c-4bad-be69-9dc30f79c723 · outbound

This paper cites Self-Play Fine-Tuning Converts Weak Language Models to Strong Language Models.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time Self-Play Fine-Tuning Converts Weak Language Models to Strong Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:19.813333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:19.813333Z digest=sha256:d86530909c420ad31afeca6003870c3219f87c4209466bbcf2f4e121c36bda2b

Observation 2b558640-201d-451f-865b-bbbced157d82 · outbound

This paper cites Safe rlhf: Safe reinforcement learning from human feedback.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time Safe rlhf: Safe reinforcement learning from human feedback

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:44:23.483038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:44:19.848413Z digest=sha256:d31698b1945d8ee453298ef24b5fba6e52fbc4a20013364f06f29b118b90335f

Observation f8a887c8-851c-46d3-9390-2239f5b8c0b1 · outbound

This paper cites RAFT: Reward rAnked FineTuning for Generative Foundation Model Alignment.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time RAFT: Reward rAnked FineTuning for Generative Foundation Model Alignment

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:19.890936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:19.890936Z digest=sha256:85bbf584a5dfee16502691b4796c33eb8f2073cbbb5e2e0393e3c935e2db0e10

Observation 54d9b4db-9fb7-4734-9ae9-fff583816ab7 · outbound

This paper cites LLMCarbon: Modeling the end-to-end Carbon Footprint of Large Language Models.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time LLMCarbon: Modeling the end-to-end Carbon Footprint of Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:19.930035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:19.930035Z digest=sha256:4c5add4bbdae2ab95c61471c39767ef9c152cc215be5eb8b69b0616f4130b1fd

Observation ffc54cc7-49e8-4bb7-aa0c-757e345d2a40 · outbound

This paper cites Improving alignment of dialogue agents via targeted human judgements.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time Improving alignment of dialogue agents via targeted human judgements

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:19.989378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:19.989378Z digest=sha256:3295de26d281a96928ca20b71edf337299e85c240967e20147d4d30ba2da266c

Observation ceadafea-0f80-4831-9f45-f568c50a5e4a · outbound

This paper cites Y., Sengupta, S., Bonadiman, D., Lai, Y.-a., Gupta, A., Pappas, N., Mansour, S., Kirchhoff, K., and Roth, D.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time Y., Sengupta, S., Bonadiman, D., Lai, Y.-a., Gupta, A., Pappas, N., Mansour, S., Kirchhoff, K., and Roth, D

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:20.011975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:20.011975Z digest=sha256:e56dda011b2661131b8f768bf267a8a77a67db57ab328e218e8db5271a3a5c5c

Observation 61a98890-281d-4596-b4b6-033b3d05fab0 · outbound

This paper cites One-Shot Safety Alignment for Large Language Models via Optimal Dualization.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time One-Shot Safety Alignment for Large Language Models via Optimal Dualization

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:44:21.936517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:44:20.062006Z digest=sha256:8050a8ef3dfb9abd91703b00ca784de2f191c8c5fb39c12abcc62153425ad997

Observation 4abcb6a1-a83f-4656-9288-eb089b4004cb · outbound

This paper cites Personalized Soups: Personalized Large Language Model Alignment via Post-hoc Parameter Merging.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time Personalized Soups: Personalized Large Language Model Alignment via Post-hoc Parameter Merging

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:20.132271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:20.132271Z digest=sha256:9e0c7a48574031b07a584b2163cc554cc13bc83d7dcee1d39350948fabc0d708

Observation 75620268-4462-4d24-b277-de0d52d9690e · outbound

This paper cites PKU-SafeRLHF: Towards Multi-Level Safety Alignment for LLMs with Human Preference.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time PKU-SafeRLHF: Towards Multi-Level Safety Alignment for LLMs with Human Preference

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:20.174221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:20.174221Z digest=sha256:76dcae895d52122b5fd840b0f02f2227dec127f39b9a34330a65b5cfa57ae3fd

Observation c3258aed-6092-456c-8e05-fe4a1ca085de · outbound

This paper cites ARGS: Alignment as Reward-Guided Search.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time ARGS: Alignment as Reward-Guided Search

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:20.211826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:20.211826Z digest=sha256:c5a79ef2033458e370f793f21b479804cbced2cbb24464fe535b32b3c528f215

Observation e7c2b313-9bb8-48d7-872c-9e44e93999a4 · outbound

This paper cites Chain of Hindsight Aligns Language Models with Feedback.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time Chain of Hindsight Aligns Language Models with Feedback

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:20.272639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:20.272639Z digest=sha256:6080bb5dda46ff2f08d932b767435b604cb90f3422446ccf8a33d9d4d6d2cbc3

Observation c4a2c54a-ce43-422a-8f00-1e73656fbb5c · outbound

This paper cites L., Daly, R.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time L., Daly, R

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:20.378678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:20.378678Z digest=sha256:ec0d396b1948c0d330c6739f96bfc6e352a763edf2b0492e06007506e3c7dd17

Observation 22dfaf66-14ba-4494-a9a8-68b624cc6a72 · outbound

This paper cites Controlled Decoding from Language Models.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time Controlled Decoding from Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:20.437302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:20.437302Z digest=sha256:a6f8e102049a0445f70111f08e361112ac1202f53e5fdf3626ca33326cf8fddd

Observation 15dd78a4-4dfc-4755-8df5-edf90a34b73e · outbound

This paper cites WebGPT: Browser-assisted question-answering with human feedback.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time WebGPT: Browser-assisted question-answering with human feedback

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:20.495851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:20.495851Z digest=sha256:d5310631b6c7e25170216b1745ea34399ec5df9c540b8e93ea71f5fda5efd626

Observation 65e6c99c-b9f3-4de9-8e2f-26867fc6a3e0 · outbound

This paper cites and Ozdaglar, A.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time and Ozdaglar, A

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:44:23.367957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:44:20.531442Z digest=sha256:1dddda16f026ac056ff8dd20cbd3d98ad6db9ee53507a5182b721f73500ea459

Observation 203f129a-4eb8-476a-89f7-a4d8295f03b5 · outbound

This paper cites Training language models to follow instructions with human feedback.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time Training language models to follow instructions with human feedback

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:20.625568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:20.625568Z digest=sha256:e47342014fadd9340fc7a95626c6ec96d8fc389073716c5a5bd3086ed0c32c63

Observation 40b10372-d403-445e-adc2-30945b7b7ae3 · outbound

This paper cites an unresolved cited work.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:20.716463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:20.716463Z digest=sha256:a8bf78a47ef279c1c389bc144212b05a44a7a7772e88397c753e6e47d5b1dffc

Observation 5ebeae0c-190a-4e3c-80ce-7337da08b8dd · outbound

This paper cites D., Ermon, S., and Finn, C.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time D., Ermon, S., and Finn, C

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:44:23.230777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:44:20.846734Z digest=sha256:1748d6f612e49a63e2cc77022f11e3c521ac552985e8e5fc17b602440bfd73e6

Observation 13965ca7-aa9b-49e0-8301-5245044ce286 · outbound

This paper cites D., Ermon, S., and Finn, C.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time D., Ermon, S., and Finn, C

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:20.920937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:20.920937Z digest=sha256:3511543a1364bfb0a0881edd1f896bb9c79542b2427099acde0299bd531ee4bb

Observation 0db0aa03-e18c-47eb-abb0-17911a8b4930 · outbound

This paper cites A., and Du, S.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time A., and Du, S

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:44:23.044947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:44:20.960306Z digest=sha256:113b105eddb4ccc831596e538e588b8103ca4619ad1d8e85c8590cc7d618554e

Observation 741c5dab-a3fa-48ca-bc1b-d8a5fa8bd3ea · outbound

This paper cites an unresolved cited work.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:44:22.879347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:44:21.004124Z digest=sha256:bb63e374ff82151a460e02f5ad672c3720f45fa1132593c446edf54cfe00cb20

Observation dd9d3a93-36f1-439e-9f1b-53c7170e4048 · outbound

This paper cites S., Tang, X., and Bogunovic, I.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time S., Tang, X., and Bogunovic, I

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:21.054411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:21.054411Z digest=sha256:7cc1f2e6a7e265a401e7c55e412f1b6d008ab36ec7770bb2bded5dc8a1ccf938

Observation 24c55602-d5f2-498e-abe4-7162cafa0752 · outbound

This paper cites an unresolved cited work.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:44:22.669996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:44:21.100780Z digest=sha256:bf299db752099aea90583b6dc5e16b2776696d170a83039b0442e9882e211906

Observation f7f74f73-2063-42e7-9102-3bc219834832 · outbound

This paper cites an unresolved cited work.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:44:22.525760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:44:21.296915Z digest=sha256:1a7f9b9e4f4543e2cacf4d32708c50f7c78ba6a10f6d5758391ea94e7797800b

Observation 51c96f80-76e3-460b-8126-d49253d2555d · outbound

This paper cites an unresolved cited work.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:44:22.365506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T12:44:21.338251Z digest=sha256:2c636d5f34daa68db180fe502e2a58a9d0c99cf17a361ba9950be138728e5307

Observation 7a22bb6e-889f-4edf-9875-f3ce74226827 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:21.380872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:21.380872Z digest=sha256:ff207e86498829e1db3a7ca5bba1a6aa06911fbb79a4e5f555417b79d51eb954

Observation c601bb67-6395-46bb-bcef-466d80a3ef90 · outbound

This paper cites M., and Wolf, T.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time M., and Wolf, T

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:21.446086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:21.446086Z digest=sha256:82f76bcc779fbff23f85bec9a67e711ab233f59f9ec2aca9eec0d5f80fa9bf45

Observation 95c77aef-c487-46b2-87fe-2c1dbbb47ae5 · outbound

This paper cites write newline.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time write newline

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:21.494870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:21.494870Z digest=sha256:0f60cbc2c3b56cd9baa56a4e5cc74a88c1671194fadded91e9a7713a89de194c

Pith citing papers

Observation 8f601549-2e01-452d-b487-9eec480b1316 · inbound

Cooperate to Compete: Strategic Coordination in Multi-Agent Conquest cites this paper.

Cooperate to Compete: Strategic Coordination in Multi-Agent Conquest Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:36:36.044970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-07T16:37:58.860183Z digest=sha256:c6a6e3770eabbfe581522c549089265ad94cbfa0cc86c0edf2c709f945c3589c