Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 60 inbound Pith citation observations for arXiv:2407.00215.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T04:42:07.455016Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
8
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 704ca59b-50c4-4424-a10c-a3888922ef0d · inbound
AIGS: Generating Science from AI-Powered Automated Falsification LLM Critics Help Catch LLM Bugs
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6bba815f-51b3-421d-a58c-56362845ad24 · inbound
Self-Generated Critiques Boost Reward Modeling for Language Models LLM Critics Help Catch LLM Bugs
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d63ecfa-947c-4889-b4fe-80eec2bde86e · inbound
Critic-V: VLM Critics Help Catch VLM Errors in Multimodal Reasoning LLM Critics Help Catch LLM Bugs
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4600b532-afe7-4fae-b2e0-e91787dd4a7c · inbound
ProcessBench: Identifying Process Errors in Mathematical Reasoning LLM Critics Help Catch LLM Bugs
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b09ff6f-28a8-459d-a149-4e8803cf23fd · inbound
SpearBot: Leveraging Large Language Models in a Generative-Critique Framework for Spear-Phishing Email Generation LLM Critics Help Catch LLM Bugs
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b66629ff-e189-4614-bdf9-3f3a95ab5386 · inbound
The Superalignment of Superhuman Intelligence with Large Language Models LLM Critics Help Catch LLM Bugs
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c7c6204-94bc-45af-85ed-6d0161a80de5 · inbound
Private Yet Social: How LLM Chatbots Support and Challenge Eating Disorder Recovery LLM Critics Help Catch LLM Bugs
Reference 104
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3960071c-58ae-468d-b6a3-da9ae9e73077 · inbound
Algebraic Evaluation Theorems LLM Critics Help Catch LLM Bugs
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c75c41e-3a5a-4e49-b466-6577666deebf · inbound
Teaching LLMs to Refine with Tools LLM Critics Help Catch LLM Bugs
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7776990f-41bc-480f-abc3-1f265ee59a59 · inbound
Online Preference-based Reinforcement Learning with Self-augmented Feedback from Large Language Model LLM Critics Help Catch LLM Bugs
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa4b0705-096b-4dae-951b-3826112411cc · inbound
Distilling Desired Comments for Enhanced Code Review with Large Language Models LLM Critics Help Catch LLM Bugs
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 103d1241-3dc4-4e3c-89d8-bec9e63a76df · inbound
Iterative Label Refinement Matters More than Preference Optimization under Weak Supervision LLM Critics Help Catch LLM Bugs
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bfc1b0ac-95ab-4255-8260-d903ff18887b · inbound
PairJudge RM: Perform Best-of-N Sampling with Knockout Tournament LLM Critics Help Catch LLM Bugs
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d77895bb-706e-4f44-a70e-a052df001dbc · inbound
Debate Helps Weak-to-Strong Generalization LLM Critics Help Catch LLM Bugs
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5cdd2b78-d9b6-4281-8027-413f78e3733b · inbound
BitsAI-CR: Automated Code Review via LLM in Practice LLM Critics Help Catch LLM Bugs
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d120828-c66a-41d1-9a16-1800fc0e996a · inbound
LongDPO: Unlock Better Long-form Generation Abilities for LLMs via Critique-augmented Stepwise Information LLM Critics Help Catch LLM Bugs
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9bbd0236-5e0d-4e77-9466-a2c223a80c5e · inbound
Automated Capability Discovery via Foundation Model Self-Exploration LLM Critics Help Catch LLM Bugs
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee956037-fb1a-4883-9db4-5f2edfa415a8 · inbound
InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models LLM Critics Help Catch LLM Bugs
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 16947f13-c97c-49be-be09-eb5680b7a1da · inbound
DeepCritic: Deliberate Critique with Large Language Models LLM Critics Help Catch LLM Bugs
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee70976a-dda7-4ce8-83d6-b3643658689c · inbound
J1: Exploring Simple Test-Time Scaling for LLM-as-a-Judge LLM Critics Help Catch LLM Bugs
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71262754-d894-4f4a-a2d8-3c63aa8ffb1a · inbound
Optimizing LLM-Based Multi-Agent System with Textual Feedback: A Case Study on Software Development LLM Critics Help Catch LLM Bugs
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d664ad6a-ee01-489f-9aac-0f635ddb0cc3 · inbound
Breakpoint: Scalable evaluation of system-level reasoning in LLM code agents LLM Critics Help Catch LLM Bugs
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b36489cf-d06b-4897-b3c9-75ceb9958269 · inbound
CRScore++: Reinforcement Learning with Verifiable Tool and AI Feedback for Code Review LLM Critics Help Catch LLM Bugs
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b97f700a-87c7-4fa0-8cb7-7840c538bd65 · inbound
Exchange of Perspective Prompting Enhances Reasoning in Large Language Models LLM Critics Help Catch LLM Bugs
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12932292-e91b-4517-9c01-14b43c65cf96 · inbound
A Survey of Automatic Evaluation Methods on Text, Visual and Speech Generations LLM Critics Help Catch LLM Bugs
Reference 261
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02bc1dfe-46d6-417a-bbef-fef24ef561b9 · inbound
CRITICTOOL: Evaluating Self-Critique Capabilities of Large Language Models in Tool-Calling Error Scenarios LLM Critics Help Catch LLM Bugs
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7fae0959-5729-4815-b117-045361ba19bc · inbound
BMMR: A Large-Scale Bilingual Multimodal Multi-Discipline Reasoning Dataset LLM Critics Help Catch LLM Bugs
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6868fb9f-f436-4f7d-959a-d7d92e1ba5c1 · inbound
CodeJudgeBench: Benchmarking LLM-as-a-Judge for Coding Tasks LLM Critics Help Catch LLM Bugs
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d82d28eb-19e1-4866-b18f-23d8360e1bd6 · inbound
Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety LLM Critics Help Catch LLM Bugs
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 395ff2a7-3a94-42ad-afc1-0e52bc8d9e71 · inbound
RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback LLM Critics Help Catch LLM Bugs
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9adddc11-31b6-4e55-a087-ba9a64a6eb06 · inbound
CoLD: Counterfactually-Guided Length Debiasing for Process Reward Models in Mathematical Reasoning LLM Critics Help Catch LLM Bugs
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation da4951f9-e021-4c94-b175-b9086c68171d · inbound
ViseGPT: Towards Better Alignment of LLM-generated Data Wrangling Scripts and User Prompts LLM Critics Help Catch LLM Bugs
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0539d43-abd0-461c-8d1c-864fe2e38c9e · inbound
Are Today's LLMs Ready to Explain Well-Being Concepts? LLM Critics Help Catch LLM Bugs
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48faffc0-1935-46f9-a8d2-a314fba9378f · inbound
Let's Revise Step-by-Step: A Unified Local Search Framework for Code Generation with LLMs LLM Critics Help Catch LLM Bugs
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd190c0b-b0b9-4c20-bef5-218520ea3ce5 · inbound
Mind the Generation Process: Fine-Grained Confidence Estimation During LLM Generation LLM Critics Help Catch LLM Bugs
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1347977-d5ef-4634-a2c0-d3068ce5e254 · inbound
Reinforcement Learning with Rubric Anchors LLM Critics Help Catch LLM Bugs
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a10b93e-dbd6-43be-9e52-76ceea82646f · inbound
OnGoal: Tracking and Visualizing Conversational Goals in Multi-Turn Dialogue with Large Language Models LLM Critics Help Catch LLM Bugs
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 007db2f9-132e-4e26-8ae4-18318fdfbb5d · inbound
Dream-Coder 7B: An Open Diffusion Language Model for Code LLM Critics Help Catch LLM Bugs
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 983d0e01-9175-409f-93ce-b594b6db0d16 · inbound
Human-AI Complementarity: A Goal for Amplified Oversight LLM Critics Help Catch LLM Bugs
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1730ac7-4802-48f8-9c5f-ae186195875e · inbound
No More Stale Feedback: Co-Evolving Critics for Open-World Agent Learning LLM Critics Help Catch LLM Bugs
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 28fe0286-729e-4c34-97ad-3e29aef28991 · inbound
ReCodeAgent: A Multi-agent Workflow for Language-Agnostic Translation and Validation of Large-Scale Repositories LLM Critics Help Catch LLM Bugs
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation a5806370-9914-4158-9a1a-0b5fc8847c67 · inbound
Building a Precise Video Language with Human-AI Oversight LLM Critics Help Catch LLM Bugs
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 954bd355-01ed-4240-a08c-369ca313ce8e · inbound
BenchGuard: Who Guards the Benchmarks? Automated Auditing of LLM Agent Benchmarks LLM Critics Help Catch LLM Bugs
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 0435cd55-1b72-478f-b536-62a92a510796 · inbound
AI Alignment via Incentives and Correction LLM Critics Help Catch LLM Bugs
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 647e6ab5-a8ed-487c-80aa-51a9564a395b · inbound
AI Alignment via Incentives and Correction LLM Critics Help Catch LLM Bugs
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation f47c57aa-73df-4737-95fe-b9752c8cb113 · inbound
LLM Wardens: Mitigating Adversarial Persuasion with Third-Party Conversational Oversight LLM Critics Help Catch LLM Bugs
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 21af213a-b98c-4ed4-820b-f4daf76d597f · inbound
Beyond Binary: Reframing GUI Critique as Continuous Semantic Alignment LLM Critics Help Catch LLM Bugs
Reference 99
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation fd0fc45f-6ddd-4935-a697-a8f60dc628b8 · inbound
Beyond Binary: Reframing GUI Critique as Continuous Semantic Alignment LLM Critics Help Catch LLM Bugs
Reference 99
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 580e838f-3c39-49ec-ae90-046e36b15931 · inbound
Philosophical Dispositions as Behavioral Constraints for AI-Assisted Code Review: An Empirical Study LLM Critics Help Catch LLM Bugs
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e8979a3c-728b-483d-bbb9-c65035c546bb · inbound
Weak Critics Make Strong Learners: On-Policy Critique Distillation for Scalable Oversight LLM Critics Help Catch LLM Bugs
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 75ea0ca6-6ce5-4326-9f02-775fcb3bd834 · inbound
Proof-or-Stop: Don't Trust the Agent, Trust the Evidence -- Loop Engineering for Verifiable Evidence-Gated Lifecycle Control LLM Critics Help Catch LLM Bugs
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee3d1936-9e66-4f5c-933f-b42f661aa8ea · inbound
Fantastic Adaptive Taxonomies and How to Use Them LLM Critics Help Catch LLM Bugs
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a5900ed-38ad-4d3e-b70c-10be22d5ff0d · inbound
Code Monitor Red Teaming for Public-Test-Passing Code LLM Critics Help Catch LLM Bugs
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5b8ce2a-122f-47d0-b665-1c83909c0789 · inbound
A dataset of rated conceptual arguments LLM Critics Help Catch LLM Bugs
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c37075e3-5fbb-44c3-91c3-1f907f29e722 · inbound
Beyond a Single Judge: The Evidence-Grounded, Social-Weighted Persona Panel for Generative UI Evaluation LLM Critics Help Catch LLM Bugs
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 088a8d31-2fec-43ca-9875-79c8f24ad036 · inbound
Beyond a Single Judge: The Evidence-Grounded, Social-Weighted Persona Panel for Generative UI Evaluation LLM Critics Help Catch LLM Bugs
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12947eb9-15c3-41cd-ac6d-73160fddd018 · inbound
Judging Is Not Enumerating: Silent Omissions in LLM-Authored Acceptable Sets LLM Critics Help Catch LLM Bugs
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b97cf54-a4ec-4b8a-a023-7d41cfc8b182 · inbound
Reference 112
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eae49f4d-c649-49a3-ad1a-6ff3ca62d586 · inbound
Apodex Discovery: Reality Benchmarks and Environments for Evaluating and Building Discoverative Artificial Intelligence LLM Critics Help Catch LLM Bugs
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04cf6c93-592c-4d2a-afaa-6fbdbd114091 · inbound
Benchmarking LLM Judges for Mobile Agent Evaluation LLM Critics Help Catch LLM Bugs
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.