Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-27T04:43:04.740925Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 3 inbound Pith citation observations for arXiv:2606.15127.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-27T04:43:04.740925Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T07:01:22.131040Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-09T23:16:36.698169Z
14 of 14 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 447b601f-3e95-4e48-9367-8fbd2feb2c6d · outbound
Beyond Accuracy: Measuring Bias Acknowledgment in Chain-of-Thought Reasoning for Responsible AI Evaluation Chain-of-Thought Reasoning In The Wild Is Not Always Faithful
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 60250f88-3172-43ba-bf75-dd6341c2c6f1 · outbound
Beyond Accuracy: Measuring Bias Acknowledgment in Chain-of-Thought Reasoning for Responsible AI Evaluation Reasoning Models Don't Always Say What They Think
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0c74a992-f189-46c3-9fa2-7788131d90d3 · outbound
Beyond Accuracy: Measuring Bias Acknowledgment in Chain-of-Thought Reasoning for Responsible AI Evaluation Does RAG Know When Retrieval Is Wrong? Diagnosing Context Compliance under Knowledge Conflict
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ef290be6-4202-4db5-98df-be8aaead5094 · outbound
Beyond Accuracy: Measuring Bias Acknowledgment in Chain-of-Thought Reasoning for Responsible AI Evaluation CURE:Circuit-Aware Unlearning for LLM-based Recommendation
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7fee1767-a400-4a05-b1de-dc869ce1bc54 · outbound
Beyond Accuracy: Measuring Bias Acknowledgment in Chain-of-Thought Reasoning for Responsible AI Evaluation CRAB: Codebook Rebalancing for Bias Mitigation in Generative Recommendation
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 233b53d9-2659-49c3-b08a-20235ea4514c · outbound
Beyond Accuracy: Measuring Bias Acknowledgment in Chain-of-Thought Reasoning for Responsible AI Evaluation M., Li, Z., Wu, X., Visweswaran, S., and Wang, Y
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6d52b44d-ad2f-44b9-8289-36360c2938c1 · outbound
Beyond Accuracy: Measuring Bias Acknowledgment in Chain-of-Thought Reasoning for Responsible AI Evaluation Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f66aff0b-e9f9-4bcf-93bb-4d35cf604577 · outbound
Beyond Accuracy: Measuring Bias Acknowledgment in Chain-of-Thought Reasoning for Responsible AI Evaluation Lin, J., Zhu, C., Kneuertz, P
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fc5f91cc-f1de-4cc8-9eec-600b4abba7f4 · outbound
Beyond Accuracy: Measuring Bias Acknowledgment in Chain-of-Thought Reasoning for Responsible AI Evaluation Reflect-Guard: Enhancing LLM Safeguards against Adversarial Prompts via Logical Self-Reflection
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a61174eb-32e5-4925-bd92-5d5fe7ea1178 · outbound
Beyond Accuracy: Measuring Bias Acknowledgment in Chain-of-Thought Reasoning for Responsible AI Evaluation FAIntbench: A Holistic and Precise Benchmark for Bias Evaluation in Text-to-Image Models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9f2226b5-0f54-4c49-b3a4-2a2b8140b0df · outbound
Beyond Accuracy: Measuring Bias Acknowledgment in Chain-of-Thought Reasoning for Responsible AI Evaluation AtelierEval: Agentic Evaluation of Humans & LLMs as Text-to-Image Prompters
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2207baad-4d7e-4fea-b47f-2ace7c490240 · outbound
Beyond Accuracy: Measuring Bias Acknowledgment in Chain-of-Thought Reasoning for Responsible AI Evaluation Relevant Is Not Warranted: Evidence-Force Calibration for Cited RAG
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 123b41f0-d534-434c-a872-d8e802ff5717 · outbound
Beyond Accuracy: Measuring Bias Acknowledgment in Chain-of-Thought Reasoning for Responsible AI Evaluation When Safe Skills Collide: Measuring Compositional Risk in Agent Skill Ecosystems
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 236ea5ec-cb9c-41e6-ae8a-ce87bd908d21 · outbound
Beyond Accuracy: Measuring Bias Acknowledgment in Chain-of-Thought Reasoning for Responsible AI Evaluation arXiv preprint arXiv:2508.15126 , year =
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 49ae5f4f-492d-4d08-86d6-394f5243dc46 · inbound
Evaluating LLM Robustness Under Domain-Specific Prompt Perturbations in Public Health Applications Beyond Accuracy: Measuring Bias Acknowledgment in Chain-of-Thought Reasoning for Responsible AI Evaluation
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6c8c60fb-bb4c-4cbc-bdd1-5eb025eeb94b · inbound
Phantom Guardrails: When Self-Improving Agent Harnesses Fix Failures That Never Happened Beyond Accuracy: Measuring Bias Acknowledgment in Chain-of-Thought Reasoning for Responsible AI Evaluation
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7e7b098-61eb-4a79-b54e-45e114912c23 · inbound
Toward User-Conditioned Evaluation of Personal LLM Agents under Temporal Interventions Beyond Accuracy: Measuring Bias Acknowledgment in Chain-of-Thought Reasoning for Responsible AI Evaluation
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.