Pith. sign in

Paper Citation Record · LEDGER

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs

As of 7 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 0 inbound Pith citation observations for arXiv:2506.10527.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.10527 v1

Coverage vector

measured 38 of 38 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:28:47.872578Z

measured 38 of 38 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

38 of 38 outbound references displayed

  • verified exact1
  • verified fuzzy13
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 13daddeb-bf98-4bd7-9a13-2efcb2f7d51f · outbound

This paper cites How far can transformers reason? the globality barrier and inductive scratchpad.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs How far can transformers reason? the globality barrier and inductive scratchpad

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:51.870479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T04:28:37.631385Z digest=sha256:c42f5eac868e4653600e5027dadc4d06a1ec4cdc427a18629a281e1def660844

Observation 4a180558-2c7b-4523-96ac-45160c5a1113 · outbound

This paper cites Graph of logic: Enhancing llm reasoning with graphs and symbolic logic.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Graph of logic: Enhancing llm reasoning with graphs and symbolic logic

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:51.648962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T04:28:39.992525Z digest=sha256:577cd71de8e4bb972d7896ff8bbb59d54dac4e620254f62b5dbfeb25dfbe327c

Observation 5d0623af-db05-4bc4-9540-25aa8c057937 · outbound

This paper cites Claude 3.7 sonnet.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Claude 3.7 sonnet

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:51.345069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T04:28:40.109199Z digest=sha256:5b6d80e800a07c694fbcd365923b4124e76ee44b8662958cff0a09f6ced21394

Observation 333256f3-6047-47a1-ae20-4dd8f38e1ec4 · outbound

This paper cites Eureka: Evaluating and Understanding Large Foundation Models.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Eureka: Evaluating and Understanding Large Foundation Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:40.218979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:40.218979Z digest=sha256:ba38a22e557260b2a566ceb113ed89abcbc469fa7c36bac04119089f9a48ee63

Observation cffd9cab-db11-42e7-ada0-fde6f9ff23c5 · outbound

This paper cites The Llama 3 Herd of Models.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs The Llama 3 Herd of Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:40.339630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:40.339630Z digest=sha256:3e7702ac172aaafdecca0bf38553944ce5892d16bc23a2e8eb434f2b73d0ccde

Observation ab83d070-3c04-44dc-b9e6-6c70cba195da · outbound

This paper cites NPHardEval: Dynamic Benchmark on Reasoning Ability of Large Language Models via Complexity Classes.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs NPHardEval: Dynamic Benchmark on Reasoning Ability of Large Language Models via Complexity Classes

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:40.881171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:40.881171Z digest=sha256:63e44a169737a352a8af9373b8198bf1ce3e8f256f2d3817de9bfc225937091d

Observation 52e18b69-3525-43f3-9d0c-641f6da8c5a4 · outbound

This paper cites Magentic-One: A Generalist Multi-Agent System for Solving Complex Tasks.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Magentic-One: A Generalist Multi-Agent System for Solving Complex Tasks

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:43.800298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:43.800298Z digest=sha256:424e139682e84302d6168cf0bf35ad70a02a7137556a69ef637875b532748b49

Observation e381ab0e-bd16-476e-afba-2cea2d894a18 · outbound

This paper cites Gemini 2.0 pro experimental.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Gemini 2.0 pro experimental

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:51.060732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T04:28:43.913151Z digest=sha256:982fc87fe1cf70aa8f933c2945cb49d495d25a08f5b7626bee686cbf144701f5

Observation 6a919f3c-4495-4f5c-a3f8-3db8f25737dd · outbound

This paper cites Gemini 2.0 flash thinking.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Gemini 2.0 flash thinking

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:50.806928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T04:28:44.022394Z digest=sha256:2e4873c542d8e332e984b9c44d654154ed710264df4cbb2ca5ea80b123d7ac05

Observation cdfb6979-b996-405f-b957-0f4421b6c7fe · outbound

This paper cites Reinforced Self-Training (ReST) for Language Modeling.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Reinforced Self-Training (ReST) for Language Modeling

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:44.141478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:44.141478Z digest=sha256:c20f291e897b9756dfa03f75782a11d118be55f2cee842d05b76bceeed8507cf

Observation 1d341136-a6eb-4881-b45a-183389793d9e · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:44.268329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:44.268329Z digest=sha256:a7041e6f76ed53c8b74a3af764d0eba11d9ab8b41f912cebab3aa54bf977ec43

Observation 34041c10-3514-41ee-8637-d2c7787ecfc9 · outbound

This paper cites GPT-4o System Card.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs GPT-4o System Card

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:44.369152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:44.369152Z digest=sha256:77ca251bc5189e51b41cd63e514055ccbe8288f5e2b593755045fecedc513836

Observation e0e8ec66-628e-48b0-ab30-204a7afcd6f8 · outbound

This paper cites OpenAI o1 System Card.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs OpenAI o1 System Card

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:44.472582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:44.472582Z digest=sha256:ea5c85a11f6329a983e9b97550219378f16a541fb43fa17508bfaa851b53ea9a

Observation e6c594e3-5fc4-4736-a8c4-6c679e502ebf · outbound

This paper cites Finding all the elementary circuits of a directed graph.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Finding all the elementary circuits of a directed graph

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:50.551608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T04:28:44.662847Z digest=sha256:ff1dc4901fd703ba8047cb0dd50cb08eaf9f3ea67fe57442a07cb7745859e61d

Observation ee54dd7e-7624-49af-8097-9be19ed5a087 · outbound

This paper cites Same task, more tokens: the impact of input length on the reasoning performance of large language models.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Same task, more tokens: the impact of input length on the reasoning performance of large language models

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:50.304512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T04:28:44.779665Z digest=sha256:0bd9b134cdf612186098d2b03b3413f3d0e0a38efea4289723f110629574fa80

Observation 40a4ecee-aee7-46c8-9ade-57a6469fbfae · outbound

This paper cites Large Language Models for Supply Chain Optimization.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Large Language Models for Supply Chain Optimization

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:44.907012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:44.907012Z digest=sha256:03640a97ebf24e3e3e48a6dc4a9f8c7bdd947ddda2a980b3d1b38c066c152d66

Observation 859e10bb-1161-4366-bb42-3df6f0a3b072 · outbound

This paper cites Llms for relational reasoning: How far are we? In Proceedings of the 1st International Workshop on Large Language Models for Code, pp.\ 119--126, 2024.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Llms for relational reasoning: How far are we? In Proceedings of the 1st International Workshop on Large Language Models for Code, pp.\ 119--126, 2024

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:50.020611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T04:28:45.061476Z digest=sha256:0a1ab1fbfe058b10f0747c3a12e89fd96cf2c9fad657b5fae6597ab20936e0e1

Observation 75549691-995f-42ec-85b7-bf025908a967 · outbound

This paper cites Let's verify step by step.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Let's verify step by step

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:45.169542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:45.169542Z digest=sha256:7a9fd549fa190e47d1e26f4dd12f1d59e21c5c74f78a6f6c36af046954d0b3fa

Observation 0585ec8d-bfb5-4b8b-8c30-03a8a3682e69 · outbound

This paper cites ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:45.311682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:45.311682Z digest=sha256:4e4c2e6242214c67f1ee58cee815c6093d71f1b138f2f3f57c5e76f2d7c67b67

Observation 1198c1e5-a59f-4cd8-a408-54ed93a49fad · outbound

This paper cites Evaluating cognitive maps and planning in large language models with cogeval.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Evaluating cognitive maps and planning in large language models with cogeval

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:49.777154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T04:28:45.438885Z digest=sha256:2e59bea5226a4503713dd6928e944d1f968286c14888713cbccb2e34a2ce988d

Observation 71007d16-5174-49f2-af57-cf863ebb13bb · outbound

This paper cites Foster, and Michael W.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Foster, and Michael W

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:49.524885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T04:28:45.587781Z digest=sha256:56f1ec8f41af7e2bcce1a6ea926236c86590e6a5febadf780f41aa5f49c5b956

Observation 8d9ed02d-68bb-4d41-9b87-63c757460c4f · outbound

This paper cites Openai o3-mini system card.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Openai o3-mini system card

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:45.681138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:45.681138Z digest=sha256:484001e75899d62e124ea1988a492d8d0b53092c7bc8b53d71a2b4e2d53307c2

Observation e493e0fb-5ddf-4608-b0c2-eb598dbf49d3 · outbound

This paper cites Logicbench: Towards systematic evaluation of logical reasoning ability of large language models.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Logicbench: Towards systematic evaluation of logical reasoning ability of large language models

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:49.303330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T04:28:45.808086Z digest=sha256:802d786925c7dd3adc1c8d73c7e2a9f771811ab5986fe3a092267545637fe41c

Observation eec52d23-b3ff-4bc4-a8f0-e17870037a28 · outbound

This paper cites Multi-LogiEval: Towards Evaluating Multi-Step Logical Reasoning Ability of Large Language Models.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Multi-LogiEval: Towards Evaluating Multi-Step Logical Reasoning Ability of Large Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:45.947400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:45.947400Z digest=sha256:5f0372569bb7e89a8b98900dfaa2d767de270912dbac905d32f590518215d8e6

Observation 9571dae9-3f7e-403e-a0e1-e884514fadfb · outbound

This paper cites Divide and translate: Compositional first-order logic translation and verification for complex logical reasoning.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Divide and translate: Compositional first-order logic translation and verification for complex logical reasoning

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:49.020000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T04:28:46.143857Z digest=sha256:a29ddc37bdd1f216c125bea8f9cb47725bd3cb5a407bf4eb22e5945ba79252e3

Observation da054956-7561-4149-95b1-ce862f3536ca · outbound

This paper cites MoreHopQA: More Than Multi-hop Reasoning.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs MoreHopQA: More Than Multi-hop Reasoning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:46.289978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:46.289978Z digest=sha256:1f92bdfe636283cca3eb3e84e3fa7a7559b05618352e55f6428f19adfebc177a

Observation 7cfadf7a-1629-496a-aadb-5c31cfd83f56 · outbound

This paper cites Causal Language Modeling Can Elicit Search and Reasoning Capabilities on Logic Puzzles.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Causal Language Modeling Can Elicit Search and Reasoning Capabilities on Logic Puzzles

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:46.408842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:46.408842Z digest=sha256:e80b8d6f975b017e6bce602b38d2c78c78e9850aa9522b581bb93e564a1e66e2

Observation d708a61d-0097-4c7d-8de8-c0b73add71d0 · outbound

This paper cites Musique: Multihop questions via single-hop question composition.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Musique: Multihop questions via single-hop question composition

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:46.557922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:46.557922Z digest=sha256:4f7380c317e63359326d31f6353d81a323fec6ae035fee87a1702c1069bed7ed

Observation b8a51838-5ab7-44b6-bc01-16188e1fbe78 · outbound

This paper cites Holy Grail 2.0: From Natural Language to Constraint Models.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Holy Grail 2.0: From Natural Language to Constraint Models

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-07T04:28:48.259984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T04:28:46.653749Z digest=sha256:272ca2171f72e3f74b3a4f841327b1c12a0740b974d7137b27a45a9938f31cf2

Observation 21847219-af4a-477b-958a-87aa8001bb95 · outbound

This paper cites On the planning abilities of large language models-a critical investigation.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs On the planning abilities of large language models-a critical investigation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:46.787264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:46.787264Z digest=sha256:107acde5bc46a4e851b0198e8fbd903a24474a221ab40aa5fdab93b40aff9d62

Observation 16a779d1-6d61-4191-9907-fea88806053e · outbound

This paper cites LLMs Still Can't Plan; Can LRMs? A Preliminary Evaluation of OpenAI's o1 on PlanBench.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs LLMs Still Can't Plan; Can LRMs? A Preliminary Evaluation of OpenAI's o1 on PlanBench

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:46.939636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:46.939636Z digest=sha256:82b3065129972d969f02a34d5af540154b6c34d2c0c203bc5a73e54de2af8dbb

Observation 0a527187-c70f-4ec1-b20d-88839e383437 · outbound

This paper cites Is A Picture Worth A Thousand Words? Delving Into Spatial Reasoning for Vision Language Models.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Is A Picture Worth A Thousand Words? Delving Into Spatial Reasoning for Vision Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:47.081061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:47.081061Z digest=sha256:07c544bba3405884704b242f9f92a656a20056ec7b34366e1c6fac3353cab5cc

Observation eba4517b-6b66-496a-b150-5b5060f6f192 · outbound

This paper cites On The Planning Abilities of OpenAI's o1 Models: Feasibility, Optimality, and Generalizability.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs On The Planning Abilities of OpenAI's o1 Models: Feasibility, Optimality, and Generalizability

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:47.210808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:47.210808Z digest=sha256:de2f8efbf9538e00a1a8f0634a795ad8a3fa8b9acbb5602c640014f0c4d1263e

Observation d6923359-51ac-47db-a56c-e6e919cba353 · outbound

This paper cites Math-shepherd: Verify and reinforce llms step-by-step without human annotations.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Math-shepherd: Verify and reinforce llms step-by-step without human annotations

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:28:48.740642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T04:28:47.330230Z digest=sha256:04f55942e58d695193e644a29d6aef3f954724e46475f5b553c7e3a38c992115

Observation 2f3f4aa8-d39d-4e1d-832d-41bd705e4542 · outbound

This paper cites AutoGen: Enabling Next-Gen LLM Applications via Multi-Agent Conversation.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs AutoGen: Enabling Next-Gen LLM Applications via Multi-Agent Conversation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:47.511050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:47.511050Z digest=sha256:41747723557fffb4db239aade297eefe488194cb8e7d7d438a3dbf94e446ef43

Observation 4979fd67-1488-487f-acdc-b9661780569c · outbound

This paper cites Faithful Logical Reasoning via Symbolic Chain-of-Thought.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Faithful Logical Reasoning via Symbolic Chain-of-Thought

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:47.609701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:47.609701Z digest=sha256:d55247fe1df8d7bc0c145e769d855d06549695a9d4d0012539d5b43dc751a14a

Observation 5bae4fd2-961f-4349-bb38-18dbad1a73e3 · outbound

This paper cites HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:47.750383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:47.750383Z digest=sha256:41461a50c73961f4c0f803254654cbce8cf5c96e7645b86450192e3352c839f5

Observation 3f1b9139-7109-4809-907e-8bb7be6e5fb6 · outbound

This paper cites NATURAL PLAN: Benchmarking LLMs on Natural Language Planning.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs NATURAL PLAN: Benchmarking LLMs on Natural Language Planning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:47.872578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:47.872578Z digest=sha256:89d8d0385ba71120c1abea5f445bc47d4ccf2c03a16b9e4ca0a36bbc31e568b5

Pith citing papers

No inbound Pith citation observations are available.