Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T17:16:36.596455Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 1 inbound Pith citation observation for arXiv:2507.11344.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T17:16:36.596455Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-03T03:04:43.641141Z
A source-named dated measurement, never combined with another source.
Source: cited_works
52 of 52 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation cc0a3585-45aa-400c-838d-eacbb62a00c1 · outbound
Guiding LLM Decision-Making with Fairness Reward Models Measuring Gender and Racial Biases in Large Language Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4590bfba-9322-4d44-873f-0f74b1f7c85a · outbound
Guiding LLM Decision-Making with Fairness Reward Models Machine bias
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6913c2f1-197a-417e-ad0d-12708f06defb · outbound
Guiding LLM Decision-Making with Fairness Reward Models Constitutional AI: Harmlessness from AI Feedback
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bef5e4b2-8799-4ec8-9c16-6493857ec85b · outbound
Guiding LLM Decision-Making with Fairness Reward Models Fairness and Machine Learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3eb81b7a-0873-4112-a139-5ff3577d6dc2 · outbound
Guiding LLM Decision-Making with Fairness Reward Models Scaling test-time compute with open models, 2024
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4fe649e-1b8b-4dc1-89ca-8e56190bfcda · outbound
Guiding LLM Decision-Making with Fairness Reward Models On the dangers of stochastic parrots: Can language models be too big? In Proceedings of the 2021 ACM conference on fairness, accountability, and transparency, 2021
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dc893af8-465b-4303-8687-e6e0658b8567 · outbound
Guiding LLM Decision-Making with Fairness Reward Models Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92d27a1c-2a61-4ff2-90cc-dfa72e7dab6e · outbound
Guiding LLM Decision-Making with Fairness Reward Models Nuanced metrics for measuring unintended bias with real data for text classification
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4262d806-a84a-47fd-b2a0-ee476739ab06 · outbound
Guiding LLM Decision-Making with Fairness Reward Models Large Language Monkeys: Scaling Inference Compute with Repeated Sampling
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 239b4c31-6913-46f3-9ac9-11e8803335b2 · outbound
Guiding LLM Decision-Making with Fairness Reward Models Alphamath almost zero: Process supervision without process
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dbf97b6b-53b1-4d69-a55f-9a8f909d4e79 · outbound
Guiding LLM Decision-Making with Fairness Reward Models Fair prediction with disparate impact: A study of bias in recidivism prediction instruments
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 343f7e75-8f04-4cd2-94e2-ecf33192fb98 · outbound
Guiding LLM Decision-Making with Fairness Reward Models Bias in bios: A case study of semantic representation bias in a high-stakes setting
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00a43689-dcc9-4192-bb08-549b22fc27af · outbound
Guiding LLM Decision-Making with Fairness Reward Models Evaluation of A frican A merican language bias in natural language generation
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f03e6d9e-b996-492b-aec9-b8026dcf5c23 · outbound
Guiding LLM Decision-Making with Fairness Reward Models The accuracy, fairness, and limits of predicting recidivism
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b10125fd-8937-4a2b-992b-c60ee033dfe7 · outbound
Guiding LLM Decision-Making with Fairness Reward Models Gaebler, Sharad Goel, Aziz Huq, and Prasanna Tambe
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0451b8b3-e928-4242-98f0-b59fb8bd9c21 · outbound
Guiding LLM Decision-Making with Fairness Reward Models Gallegos, Ryan A
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a90cb941-9963-4151-ac92-b617a720e87a · outbound
Guiding LLM Decision-Making with Fairness Reward Models Debiasing pre-trained language models via efficient fine-tuning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b79ff65-f079-4643-808a-ac6998336767 · outbound
Guiding LLM Decision-Making with Fairness Reward Models Bias in Large Language Models: Origin, Evaluation, and Mitigation
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd06d18d-8403-4b1c-ab91-237c5b3ab031 · outbound
Guiding LLM Decision-Making with Fairness Reward Models Equality of opportunity in supervised learning
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 38241025-f83b-4642-b5c0-b8ff424c9674 · outbound
Guiding LLM Decision-Making with Fairness Reward Models V- ST ar: Training verifiers for self-taught reasoners
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 947a5ebd-3ab4-4f84-bdae-3aba1c02254e · outbound
Guiding LLM Decision-Making with Fairness Reward Models Prompting Techniques for Reducing Social Bias in LLMs through System 1 and System 2 Cognitive Processes
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2332d80b-b2c0-4c99-a2ea-c78009ee6401 · outbound
Guiding LLM Decision-Making with Fairness Reward Models Evaluating Gender Bias in Large Language Models via Chain-of-Thought Prompting
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4162730-f66a-4fd6-a73f-a2ff60492451 · outbound
Guiding LLM Decision-Making with Fairness Reward Models Gender bias and stereotypes in large language models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 53e20214-70cc-4ab7-9c6c-fc3770a5e16d · outbound
Guiding LLM Decision-Making with Fairness Reward Models When do pre-training biases propagate to downstream tasks? a case study in text summarization
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be259f34-6ba6-4adb-8604-6aafe16ff023 · outbound
Guiding LLM Decision-Making with Fairness Reward Models Let's verify step by step
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 190bf357-95f7-4a39-aaa5-303b39125cf7 · outbound
Guiding LLM Decision-Making with Fairness Reward Models Debiasing large language models with structured knowledge
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2598aec9-9663-4997-98ea-eabef42f61dc · outbound
Guiding LLM Decision-Making with Fairness Reward Models Fairness-guided few-shot prompting for large language models
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9462b191-baa9-4f9c-94b9-b23826ce9b3b · outbound
Guiding LLM Decision-Making with Fairness Reward Models Evaluating Gender Bias Transfer between Pre-trained and Prompt-Adapted Language Models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 859b9ccb-1d88-421a-a44e-d564ed349bdf · outbound
Guiding LLM Decision-Making with Fairness Reward Models GPT-4 Technical Report
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 783d5837-a48e-499d-a173-cb7bbc3cbe03 · outbound
Guiding LLM Decision-Making with Fairness Reward Models Bias in word embeddings
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e44a4c7-fa98-46e4-9acf-a30085eee4c7 · outbound
Guiding LLM Decision-Making with Fairness Reward Models BBQ : A hand-built bias benchmark for question answering
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9cc90f10-3013-4f17-84e3-a0c5d69a5fa5 · outbound
Guiding LLM Decision-Making with Fairness Reward Models Divine LL a MA s: Bias, stereotypes, stigmatization, and emotion representation of religion in large language models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ffbd460-dc6e-49e7-ad19-a74ea5f316da · outbound
Guiding LLM Decision-Making with Fairness Reward Models Proximal Policy Optimization Algorithms
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e61e2c6e-e099-4736-b1a5-13274a76cb76 · outbound
Guiding LLM Decision-Making with Fairness Reward Models On second thought, let ' s not think step by step! bias and toxicity in zero-shot reasoning
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 245ea5f8-203a-4e87-bff7-c98f74010643 · outbound
Guiding LLM Decision-Making with Fairness Reward Models DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 030d9034-1f56-404f-8756-deceb6222eda · outbound
Guiding LLM Decision-Making with Fairness Reward Models Scaling LLM test-time compute optimally can be more effective than scaling parameters for reasoning
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95250af8-c97c-46be-a444-562177b8bedc · outbound
Guiding LLM Decision-Making with Fairness Reward Models Unveiling Gender Bias in Terms of Profession Across LLMs: Analyzing and Addressing Sociological Implications
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8608611d-f52e-4665-9817-335e646aebb2 · outbound
Guiding LLM Decision-Making with Fairness Reward Models Toward self-improvement of llms via imagination, searching, and criticizing
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d7875b3e-8f1c-4a5f-8efc-69c0403a9a3d · outbound
Guiding LLM Decision-Making with Fairness Reward Models Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c9db766-0caa-4e05-9fec-c551aea0eb4b · outbound
Guiding LLM Decision-Making with Fairness Reward Models Unresolved cited work
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3dc15b25-52bb-478c-806a-b86504c164a0 · outbound
Guiding LLM Decision-Making with Fairness Reward Models Solving math word problems with process- and outcome-based feedback
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b12e44ce-f502-4ebc-bb0c-56257c9e3fd6 · outbound
Guiding LLM Decision-Making with Fairness Reward Models ``kelly is a warm person, joseph is a role model'': Gender biases in LLM -generated reference letters
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29756c0c-8d77-4321-9d5a-05cec139544b · outbound
Guiding LLM Decision-Making with Fairness Reward Models Alphazero-like tree-search can guide large language model decoding and training
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e181325b-299b-43e9-bb30-c8b23d172bf4 · outbound
Guiding LLM Decision-Making with Fairness Reward Models Math-shepherd: Verify and reinforce LLM s step-by-step without human annotations
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c31743d-7da5-4686-854c-00a05a8c5382 · outbound
Guiding LLM Decision-Making with Fairness Reward Models Chi, Sharan Narang, Aakanksha Chowdhery, and Denny Zhou
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 772ded21-26f9-4b76-b258-59b652fba045 · outbound
Guiding LLM Decision-Making with Fairness Reward Models Chain-of-thought prompting elicits reasoning in large language models
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88bccc54-0090-4147-bc83-1e6642df3a25 · outbound
Guiding LLM Decision-Making with Fairness Reward Models Blueprint for an ai bill of rights: Making automated systems work for the american people, 2022
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c54b5ecf-f6f7-4f39-9b13-2616c96509ba · outbound
Guiding LLM Decision-Making with Fairness Reward Models Gender, Race, and Intersectional Bias in Resume Screening via Language Model Retrieval, page 1578–1590
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1d5547e0-35a6-46fd-99a2-8b8afe9eff28 · outbound
Guiding LLM Decision-Making with Fairness Reward Models Tree of thoughts: Deliberate problem solving with large language models
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cb3dabd6-abbe-4d12-8f41-a7466d145002 · outbound
Guiding LLM Decision-Making with Fairness Reward Models Scaling Relationship on Learning Mathematical Reasoning with Large Language Models
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b9ef331-a297-4a24-8175-a6f9d936f503 · outbound
Guiding LLM Decision-Making with Fairness Reward Models Star: Bootstrapping reasoning with reasoning
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a49fec15-69e9-48ca-ac85-03ab78d9f5f8 · outbound
Guiding LLM Decision-Making with Fairness Reward Models Towards effective discrimination testing for generative ai
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9610f636-a3ec-4be5-8de9-6de0de29ed12 · inbound
Reward Modeling for Reinforcement Learning-Based LLM Reasoning: Design, Challenges, and Evaluation Guiding LLM Decision-Making with Fairness Reward Models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.