Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 28 inbound Pith citation observations for arXiv:2206.10498.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T10:12:00.512507Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
50
pith, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 46c28c21-9af8-4463-b3f1-f41756e9b448 · inbound
LLM+P: Empowering Large Language Models with Optimal Planning Proficiency PlanBench: An Extensible Benchmark for Evaluating Large Language Models on Planning and Reasoning about Change
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation bbaa211d-562c-4d18-983e-97e9def07777 · inbound
Distilling Step-by-Step! Outperforming Larger Language Models with Less Training Data and Smaller Model Sizes PlanBench: An Extensible Benchmark for Evaluating Large Language Models on Planning and Reasoning about Change
Reference 103
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 2c3f695a-f259-42a7-aab1-b7dac58cbd74 · inbound
Reasoning with Language Model is Planning with World Model PlanBench: An Extensible Benchmark for Evaluating Large Language Models on Planning and Reasoning about Change
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a17288dc-56c7-4a4c-98f2-39360714a03e · inbound
Cognitive Architectures for Language Agents PlanBench: An Extensible Benchmark for Evaluating Large Language Models on Planning and Reasoning about Change
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 458aa1fb-88d4-41bd-ba59-2b94e36d04d1 · inbound
CodeMind: Evaluating Large Language Models for Code Reasoning PlanBench: An Extensible Benchmark for Evaluating Large Language Models on Planning and Reasoning about Change
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 72e0e72d-42d8-4193-93b0-f6eec1c4488e · inbound
GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language Models PlanBench: An Extensible Benchmark for Evaluating Large Language Models on Planning and Reasoning about Change
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 3e779026-5705-4a77-815d-9cbe0a9710eb · inbound
A Review on Generative AI Models for Synthetic Medical Text, Time Series, and Longitudinal Data PlanBench: An Extensible Benchmark for Evaluating Large Language Models on Planning and Reasoning about Change
Reference 103
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fdac3854-9208-4f7e-b8a9-0ce2b26796fe · inbound
Evaluating LLM Reasoning in the Operations Research Domain with ORQA PlanBench: An Extensible Benchmark for Evaluating Large Language Models on Planning and Reasoning about Change
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c337865-8896-49b5-8a4f-99a5a4cdd197 · inbound
A Tool for In-depth Analysis of Code Execution Reasoning of Large Language Models PlanBench: An Extensible Benchmark for Evaluating Large Language Models on Planning and Reasoning about Change
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fafff5a4-8a2a-4493-9f1c-9db9facc9e43 · inbound
Successor-Generator Planning with LLM-generated Heuristics PlanBench: An Extensible Benchmark for Evaluating Large Language Models on Planning and Reasoning about Change
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 974da3df-3e6e-49d1-a9d6-e115a0e0bc69 · inbound
Lost in Cultural Translation: Do LLMs Struggle with Math Across Cultural Contexts? PlanBench: An Extensible Benchmark for Evaluating Large Language Models on Planning and Reasoning about Change
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 96af8fdd-1d7b-41c9-be48-323a75670dcb · inbound
Toward Generalizable Evaluation in the LLM Era: A Survey Beyond Benchmarks PlanBench: An Extensible Benchmark for Evaluating Large Language Models on Planning and Reasoning about Change
Reference 107
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70585823-cd63-4605-b721-20a1a7f3b1ec · inbound
Toward Generalizable Evaluation in the LLM Era: A Survey Beyond Benchmarks PlanBench: An Extensible Benchmark for Evaluating Large Language Models on Planning and Reasoning about Change
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8588f832-6e54-4e86-a15a-c5f113590284 · inbound
The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity PlanBench: An Extensible Benchmark for Evaluating Large Language Models on Planning and Reasoning about Change
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 3f61024a-0ca6-4b6f-a3fe-7ec72afc538f · inbound
GenPlanX. Generation of Plans and Execution PlanBench: An Extensible Benchmark for Evaluating Large Language Models on Planning and Reasoning about Change
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2e22f05-25ae-4671-b40e-d5b1c81925dc · inbound
Application of LLMs to Multi-Robot Path Planning and Task Allocation PlanBench: An Extensible Benchmark for Evaluating Large Language Models on Planning and Reasoning about Change
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f25c7e5-5586-4b32-85a8-31548ce6f10d · inbound
ISO-Bench: Benchmarking Multimodal Causal Reasoning in Visual-Language Models through Procedural Plans PlanBench: An Extensible Benchmark for Evaluating Large Language Models on Planning and Reasoning about Change
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d73f5b39-ae97-4245-838c-a31c63948ff2 · inbound
Assessing Coherency and Consistency of Code Execution Reasoning by Large Language Models PlanBench: An Extensible Benchmark for Evaluating Large Language Models on Planning and Reasoning about Change
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 1dab2a04-e5c0-4f11-a94c-aa669c1db994 · inbound
SAT: Sequential Agent Tuning for Coordinator Free Plug and Play Multi-LLM Training with Monotonic Improvement Guarantees PlanBench: An Extensible Benchmark for Evaluating Large Language Models on Planning and Reasoning about Change
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 95108b2b-5ec8-4fbd-8306-a6b0d4747c0b · inbound
OPT-BENCH: Evaluating the Iterative Self-Optimization of LLM Agents in Large-Scale Search Spaces PlanBench: An Extensible Benchmark for Evaluating Large Language Models on Planning and Reasoning about Change
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 55b9f9ec-cad0-4ceb-9d3d-29102fa8297e · inbound
Consistency as a Testable Property: Statistical Methods to Evaluate AI Agent Reliability PlanBench: An Extensible Benchmark for Evaluating Large Language Models on Planning and Reasoning about Change
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation d460edfd-ebd1-4424-a809-5fcbf51d7b1e · inbound
Zero-Shot Goal Recognition with Large Language Models PlanBench: An Extensible Benchmark for Evaluating Large Language Models on Planning and Reasoning about Change
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 7eb88bfb-b7f7-4f44-b3ce-5b81605174dc · inbound
HyperGuide: Hyperbolic Guidance for Efficient Multi-Step Reasoning in Large Language Models PlanBench: An Extensible Benchmark for Evaluating Large Language Models on Planning and Reasoning about Change
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ff5f9a1-93ce-4990-9f19-baf398000339 · inbound
Managing Uncertainty in LLM-Generated Procedural Knowledge for Virtual Laboratory Planning PlanBench: An Extensible Benchmark for Evaluating Large Language Models on Planning and Reasoning about Change
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 5e40099d-66c1-4566-bf85-85ca3420a0c6 · inbound
REPOT: Recoverable Program-of-Thought via Checkpoint Repair PlanBench: An Extensible Benchmark for Evaluating Large Language Models on Planning and Reasoning about Change
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation ce0c0d53-c2c9-4b7d-b831-feb1191cd8a1 · inbound
Lost in Aggregation: A Multi-Scale Diagnostic Benchmark for LLM Spatial Navigation PlanBench: An Extensible Benchmark for Evaluating Large Language Models on Planning and Reasoning about Change
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 5d6a6729-a5ca-414c-aa49-15f52a989fe1 · inbound
Search, Fail, Recover: A Training Framework for Correction-Aware Reasoning PlanBench: An Extensible Benchmark for Evaluating Large Language Models on Planning and Reasoning about Change
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation cbeebfab-0187-479c-a309-613c04f069b1 · inbound
LatticeMind: A Conflict-Aware Memory Primitive for Multi-Agent Systems PlanBench: An Extensible Benchmark for Evaluating Large Language Models on Planning and Reasoning about Change
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.