Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T04:28:47.872578Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 0 inbound Pith citation observations for arXiv:2506.10527.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T04:28:47.872578Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
38 of 38 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 13daddeb-bf98-4bd7-9a13-2efcb2f7d51f · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs How far can transformers reason? the globality barrier and inductive scratchpad
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4a180558-2c7b-4523-96ac-45160c5a1113 · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Graph of logic: Enhancing llm reasoning with graphs and symbolic logic
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5d0623af-db05-4bc4-9540-25aa8c057937 · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Claude 3.7 sonnet
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 333256f3-6047-47a1-ae20-4dd8f38e1ec4 · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Eureka: Evaluating and Understanding Large Foundation Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cffd9cab-db11-42e7-ada0-fde6f9ff23c5 · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs The Llama 3 Herd of Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab83d070-3c04-44dc-b9e6-6c70cba195da · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs NPHardEval: Dynamic Benchmark on Reasoning Ability of Large Language Models via Complexity Classes
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52e18b69-3525-43f3-9d0c-641f6da8c5a4 · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Magentic-One: A Generalist Multi-Agent System for Solving Complex Tasks
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e381ab0e-bd16-476e-afba-2cea2d894a18 · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Gemini 2.0 pro experimental
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6a919f3c-4495-4f5c-a3f8-3db8f25737dd · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Gemini 2.0 flash thinking
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cdfb6979-b996-405f-b957-0f4421b6c7fe · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Reinforced Self-Training (ReST) for Language Modeling
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d341136-a6eb-4881-b45a-183389793d9e · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34041c10-3514-41ee-8637-d2c7787ecfc9 · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs GPT-4o System Card
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0e8ec66-628e-48b0-ab30-204a7afcd6f8 · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs OpenAI o1 System Card
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6c594e3-5fc4-4736-a8c4-6c679e502ebf · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Finding all the elementary circuits of a directed graph
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ee54dd7e-7624-49af-8097-9be19ed5a087 · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Same task, more tokens: the impact of input length on the reasoning performance of large language models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 40a4ecee-aee7-46c8-9ade-57a6469fbfae · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Large Language Models for Supply Chain Optimization
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 859e10bb-1161-4366-bb42-3df6f0a3b072 · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Llms for relational reasoning: How far are we? In Proceedings of the 1st International Workshop on Large Language Models for Code, pp.\ 119--126, 2024
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 75549691-995f-42ec-85b7-bf025908a967 · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Let's verify step by step
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0585ec8d-bfb5-4b8b-8c30-03a8a3682e69 · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1198c1e5-a59f-4cd8-a408-54ed93a49fad · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Evaluating cognitive maps and planning in large language models with cogeval
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 71007d16-5174-49f2-af57-cf863ebb13bb · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Foster, and Michael W
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8d9ed02d-68bb-4d41-9b87-63c757460c4f · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Openai o3-mini system card
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e493e0fb-5ddf-4608-b0c2-eb598dbf49d3 · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Logicbench: Towards systematic evaluation of logical reasoning ability of large language models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation eec52d23-b3ff-4bc4-a8f0-e17870037a28 · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Multi-LogiEval: Towards Evaluating Multi-Step Logical Reasoning Ability of Large Language Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9571dae9-3f7e-403e-a0e1-e884514fadfb · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Divide and translate: Compositional first-order logic translation and verification for complex logical reasoning
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation da054956-7561-4149-95b1-ce862f3536ca · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs MoreHopQA: More Than Multi-hop Reasoning
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7cfadf7a-1629-496a-aadb-5c31cfd83f56 · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Causal Language Modeling Can Elicit Search and Reasoning Capabilities on Logic Puzzles
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d708a61d-0097-4c7d-8de8-c0b73add71d0 · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Musique: Multihop questions via single-hop question composition
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8a51838-5ab7-44b6-bc01-16188e1fbe78 · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Holy Grail 2.0: From Natural Language to Constraint Models
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 21847219-af4a-477b-958a-87aa8001bb95 · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs On the planning abilities of large language models-a critical investigation
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16a779d1-6d61-4191-9907-fea88806053e · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs LLMs Still Can't Plan; Can LRMs? A Preliminary Evaluation of OpenAI's o1 on PlanBench
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a527187-c70f-4ec1-b20d-88839e383437 · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Is A Picture Worth A Thousand Words? Delving Into Spatial Reasoning for Vision Language Models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eba4517b-6b66-496a-b150-5b5060f6f192 · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs On The Planning Abilities of OpenAI's o1 Models: Feasibility, Optimality, and Generalizability
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6923359-51ac-47db-a56c-e6e919cba353 · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Math-shepherd: Verify and reinforce llms step-by-step without human annotations
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2f3f4aa8-d39d-4e1d-832d-41bd705e4542 · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs AutoGen: Enabling Next-Gen LLM Applications via Multi-Agent Conversation
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4979fd67-1488-487f-acdc-b9661780569c · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Faithful Logical Reasoning via Symbolic Chain-of-Thought
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5bae4fd2-961f-4349-bb38-18dbad1a73e3 · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f1b9139-7109-4809-907e-8bb7be6e5fb6 · outbound
LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs NATURAL PLAN: Benchmarking LLMs on Natural Language Planning
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.