Pith. sign in

Paper Citation Record · LEDGER

LogicBench: Towards Systematic Evaluation of Logical Reasoning Ability of Large Language Models

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2404.15522.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.15522 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T05:24:24.117256Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T18:31:44.623520Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5ab0e704-cf2f-4298-af86-26d990931283 · inbound

Enhancing Transformers for Generalizable First-Order Logical Entailment cites this paper.

Enhancing Transformers for Generalizable First-Order Logical Entailment LogicBench: Towards Systematic Evaluation of Logical Reasoning Ability of Large Language Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-10T22:48:53.560481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:48:53.560481Z digest=sha256:b157c130bac0f66c444b54f4ba17b11a96833d50aeb1f8d6229d89ffdfd07175

Observation 55ea891f-6fc9-4345-a66f-146e99f7de15 · inbound

Boosting Self-Efficacy and Performance of Large Language Models via Verbal Efficacy Stimulations cites this paper.

Boosting Self-Efficacy and Performance of Large Language Models via Verbal Efficacy Stimulations LogicBench: Towards Systematic Evaluation of Logical Reasoning Ability of Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T14:46:41.032275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T14:46:41.032275Z digest=sha256:feddafe5f87abb42d3a5de2a0dae59db0d8af528b2dc1c623841758899b1c4a4

Observation 5a67629b-de4b-4984-99ee-a77bf9081f80 · inbound

Computational Reasoning of Large Language Models cites this paper.

Computational Reasoning of Large Language Models LogicBench: Towards Systematic Evaluation of Logical Reasoning Ability of Large Language Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-16T05:24:24.117256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:24:24.117256Z digest=sha256:ddf2ca9f45061a4c92587103c11e305bea80d56cfcb4a97060f70d11da3c0edc

Observation 4201fc29-0509-4782-83bb-4ef2f3b2b8bb · inbound

Evaluation of LLMs for mathematical problem solving cites this paper.

Evaluation of LLMs for mathematical problem solving LogicBench: Towards Systematic Evaluation of Logical Reasoning Ability of Large Language Models

Reference 133

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:13.215948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:11:13.215948Z digest=sha256:6ad74f691ba0be1e9d28346b7aa68b3a3a899c03e8720f4e25adc2f4751739ac

Observation 53813fa4-b8c5-430f-86d1-f321aa31da0f · inbound

BAR Conjecture: the Feasibility of Inference Budget-Constrained LLM Services with Authenticity and Reasoning cites this paper.

BAR Conjecture: the Feasibility of Inference Budget-Constrained LLM Services with Authenticity and Reasoning LogicBench: Towards Systematic Evaluation of Logical Reasoning Ability of Large Language Models

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-06T11:04:47.237610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:04:47.237610Z digest=sha256:f29807395110bf904fc59e31a131c81bf4984b2d3afeb5911febd92e705ba3f1

Observation a5dbacfa-40e5-4177-9235-3f517a90ed75 · inbound

HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation cites this paper.

HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation LogicBench: Towards Systematic Evaluation of Logical Reasoning Ability of Large Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T05:41:42.117693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:41:42.117693Z digest=sha256:eff9194ed4d3a6b0bada07e159b4b37ee233ff6f084845cec409bdbb079e865d

Observation f43b4129-4a13-48a5-b28b-1a658df5655a · inbound

PuzzleClone: A DSL-Powered Framework for Synthesizing Verifiable Data cites this paper.

PuzzleClone: A DSL-Powered Framework for Synthesizing Verifiable Data LogicBench: Towards Systematic Evaluation of Logical Reasoning Ability of Large Language Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-05T18:08:43.092102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T18:08:43.092102Z digest=sha256:467264b987d4cced087f3f2ac8b7febb66fcf89d85d774e3d99bfde278f19605

Observation face0bee-84d8-40ce-a37e-b282b40a2e71 · inbound

Self-Aligned Reward: Towards Effective and Efficient Reasoners cites this paper.

Self-Aligned Reward: Towards Effective and Efficient Reasoners LogicBench: Towards Systematic Evaluation of Logical Reasoning Ability of Large Language Models

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-18T18:31:44.625901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-18T18:27:23.076544Z digest=sha256:45680a4f31a8685c77960b67a8f9bf41720c5857ac4e19997d3fefc4c48abc15

Observation f9abc7f1-d0a3-4ad0-a0c1-f462ebf1ec15 · inbound

UniCode: Augmenting Evaluation for Code Reasoning cites this paper.

UniCode: Augmenting Evaluation for Code Reasoning LogicBench: Towards Systematic Evaluation of Logical Reasoning Ability of Large Language Models

Reference 1999

Resolution
unresolved
no resolver link, observed 2026-08-04T09:39:47.686180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:39:47.686180Z digest=sha256:aa67751709a7015545a58c5862f87f29b9b2b08fda9e961551496b730f652e47

Observation 982d34ca-6019-45b2-ad98-773a2f9aa53a · inbound

Author-in-the-Loop Response Generation and Evaluation: Integrating Author Expertise and Intent in Responses to Peer Review cites this paper.

Author-in-the-Loop Response Generation and Evaluation: Integrating Author Expertise and Intent in Responses to Peer Review LogicBench: Towards Systematic Evaluation of Logical Reasoning Ability of Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T09:42:58.223989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T09:42:58.223989Z digest=sha256:8b0fa44ce0f9aa4187ad41e1ba0d6e763c01e8a5a3a355271dcb82426112efff

Observation 34bd2d4e-085c-4dc3-90ce-571afc952d0d · inbound

HyperLens: Quantifying Cognitive Effort in LLMs with Fine-grained Confidence Trajectory cites this paper.

HyperLens: Quantifying Cognitive Effort in LLMs with Fine-grained Confidence Trajectory LogicBench: Towards Systematic Evaluation of Logical Reasoning Ability of Large Language Models

Reference 41

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T19:36:09.585732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-08T11:38:49.630171Z digest=sha256:be4b28ac3fa89c7cc07b784dc4171d921dad1f4649242fc7d03df50f90622be1

Observation 6b3a175f-ba4e-4b38-ae9e-fe6c940a92f0 · inbound

SAE-StatSteer: Statistical Consensus Feature Selection for Optimization-Free Activation Steering of Large Language Models cites this paper.

SAE-StatSteer: Statistical Consensus Feature Selection for Optimization-Free Activation Steering of Large Language Models LogicBench: Towards Systematic Evaluation of Logical Reasoning Ability of Large Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-02T12:14:23.445389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:14:23.445389Z digest=sha256:38e8b8fb1a1381c7cc042757d650e6c255108f7f29f080cc80ceebbc38bde711