Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 15 inbound Pith citation observations for arXiv:2410.01257.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:18:39.220697Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T01:37:30.542430Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 0f408b60-18da-45ec-990d-743993458d17 · inbound
LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods HelpSteer2-Preference: Complementing Ratings with Preferences
Reference 246
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 42666668-3a53-487f-97a0-02efe5af086b · inbound
From Macro to Micro: Probing Dataset Diversity in Language Model Fine-Tuning HelpSteer2-Preference: Complementing Ratings with Preferences
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16781b5f-b038-4ecf-9483-0c3d591e4a18 · inbound
RewardBench 2: Advancing Reward Model Evaluation HelpSteer2-Preference: Complementing Ratings with Preferences
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 38ed5144-7e39-40df-b130-03924ff07c98 · inbound
Well Begun is Half Done: Low-resource Preference Alignment by Weak-to-Strong Decoding HelpSteer2-Preference: Complementing Ratings with Preferences
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a39d22f1-4ff5-4198-b742-58d1984e87cb · inbound
Chasing Moving Targets with Online Self-Play Reinforcement Learning for Safer Language Models HelpSteer2-Preference: Complementing Ratings with Preferences
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 755f426c-4980-4b25-9170-c6be89aa72b4 · inbound
Textual Bayes: Quantifying Prompt Uncertainty in LLM-Based Systems HelpSteer2-Preference: Complementing Ratings with Preferences
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e9bba9b6-ce48-4fcb-83d2-d29a073e870d · inbound
OpenCodeReasoning-II: A Simple Test Time Scaling Approach via Self-Critique HelpSteer2-Preference: Complementing Ratings with Preferences
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0941c192-6f1f-491f-82f9-a7d3820bddc5 · inbound
Alignment and Safety in Large Language Models: Safety Mechanisms, Training Paradigms, and Emerging Challenges HelpSteer2-Preference: Complementing Ratings with Preferences
Reference 192
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d92521d-7af7-4991-839f-b0e355584ce3 · inbound
Multimodal LLMs as Customized Reward Models for Text-to-Image Generation HelpSteer2-Preference: Complementing Ratings with Preferences
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e57d647d-c45f-4c8d-8daf-a9bf780f2275 · inbound
Multi-Modal Requirements Data-based Acceptance Criteria Generation using LLMs HelpSteer2-Preference: Complementing Ratings with Preferences
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a379593-88ca-4d08-bb53-4ebb4da394de · inbound
HEAL: A Hypothesis-Based Preference-Aware Analysis Framework HelpSteer2-Preference: Complementing Ratings with Preferences
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4039719-d289-4f68-abe4-e1b8e027db86 · inbound
Adaptive Margin RLHF via Preference over Preferences HelpSteer2-Preference: Complementing Ratings with Preferences
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49e3ca38-f659-4cf0-8018-470ee97cc75a · inbound
Nautilus Compass: Black-box Persona Drift Detection for Production LLM Agents HelpSteer2-Preference: Complementing Ratings with Preferences
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0d6e39b8-e452-49dc-b58d-4e4be105394e · inbound
Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization HelpSteer2-Preference: Complementing Ratings with Preferences
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ce7f2b4a-8299-4273-a58c-2adf792cb670 · inbound
Small, Free, and Effective: Orchestrating Open-Weight Small Language Models to Outperform Single LLM for Malware Analysis HelpSteer2-Preference: Complementing Ratings with Preferences
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.