Pith. sign in

Paper Citation Record · LEDGER

Multi-LogiEval: Towards Evaluating Multi-Step Logical Reasoning Ability of Large Language Models

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2406.17169.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.17169 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:50:27.973380Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T22:17:25.555027Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 46850007-b9cd-4d20-b553-92a5208124ac · inbound

Meta-Thinking in LLMs via Multi-Agent Reinforcement Learning: A Survey cites this paper.

Meta-Thinking in LLMs via Multi-Agent Reinforcement Learning: A Survey Multi-LogiEval: Towards Evaluating Multi-Step Logical Reasoning Ability of Large Language Models

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:27.973380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:27.973380Z digest=sha256:8cda1a0901b94bca8aa9406dc494c99ec31d9cf6ea7168aa8ef66e0e37592d93

Observation 21030ce5-2cc6-421a-bfce-18ab298eb529 · inbound

Computational Reasoning of Large Language Models cites this paper.

Computational Reasoning of Large Language Models Multi-LogiEval: Towards Evaluating Multi-Step Logical Reasoning Ability of Large Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-16T05:24:24.112681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:24:24.112681Z digest=sha256:d4c2d36286607f2f1e072b91a8fbb4b43f553b1838b7e2bcebfca9fe5de6e39a

Observation 33def663-f585-4cd1-a77b-8c762ccdecdd · inbound

DeepMath-Creative: A Benchmark for Evaluating Mathematical Creativity of Large Language Models cites this paper.

DeepMath-Creative: A Benchmark for Evaluating Mathematical Creativity of Large Language Models Multi-LogiEval: Towards Evaluating Multi-Step Logical Reasoning Ability of Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T21:51:21.327994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:51:21.327994Z digest=sha256:740ca261c121a0be76e0d3cb21dbe1405495f64a8d6e8ea4612365e5288e0492

Observation 6a1a912d-e0cc-43de-b472-3dce54265012 · inbound

ALAS: A Stateful Multi-LLM Agent Framework for Disruption-Aware Planning cites this paper.

ALAS: A Stateful Multi-LLM Agent Framework for Disruption-Aware Planning Multi-LogiEval: Towards Evaluating Multi-Step Logical Reasoning Ability of Large Language Models

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-15T20:37:37.420595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:37:37.420595Z digest=sha256:a4b61fa39d71b9c84cfe0a7f2fef51a0c3fe989ef950e73d72128ef3f39c0f95

Observation eec52d23-b3ff-4bc4-a8f0-e17870037a28 · inbound

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs cites this paper.

LogiPlan: A Structured Benchmark for Logical Planning and Relational Reasoning in LLMs Multi-LogiEval: Towards Evaluating Multi-Step Logical Reasoning Ability of Large Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T04:28:45.947400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:28:45.947400Z digest=sha256:056030c45d5884b47d20a472bdaad1dd4550076d9f82fe78b77a68faee9b6965

Observation 288015c9-5e86-4673-a88d-77973173f6a6 · inbound

KnowRL: Exploring Knowledgeable Reinforcement Learning for Factuality cites this paper.

KnowRL: Exploring Knowledgeable Reinforcement Learning for Factuality Multi-LogiEval: Towards Evaluating Multi-Step Logical Reasoning Ability of Large Language Models

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-19T07:37:08.980700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-19T07:33:08.719028Z digest=sha256:28325a286fbd27dbc851bb2b18f478afdd480a74e2345a2e98c8aedb8cc0bf32

Observation 8a41a6b2-7454-4f51-a7d3-22c42bd207b2 · inbound

Throttling Web Agents Using Reasoning Gates cites this paper.

Throttling Web Agents Using Reasoning Gates Multi-LogiEval: Towards Evaluating Multi-Step Logical Reasoning Ability of Large Language Models

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:04.522619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:04.522619Z digest=sha256:895effe85963cde02c8537ae74227e863153f8950207d1f3f57c55b27076619f

Observation 3478f068-ac84-4797-b25b-1ff9a9413c58 · inbound

Semantic-Aware Logical Reasoning via a Semiotic Framework cites this paper.

Semantic-Aware Logical Reasoning via a Semiotic Framework Multi-LogiEval: Towards Evaluating Multi-Step Logical Reasoning Ability of Large Language Models

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T13:01:23.919675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T12:57:45.584017Z digest=sha256:27a847b079d5ac3668c7c5643de2bd115396fb3fc65408afaf75ba55828cdb6c

Observation fdc9579a-9d33-4b69-83e7-c4149bb22a22 · inbound

Ling and Ring 2.6 Technical Report: Efficient and Instant Agentic Intelligence at Trillion-Parameter Scale cites this paper.

Ling and Ring 2.6 Technical Report: Efficient and Instant Agentic Intelligence at Trillion-Parameter Scale Multi-LogiEval: Towards Evaluating Multi-Step Logical Reasoning Ability of Large Language Models

Reference 131

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T22:17:25.556581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-07-02T22:10:59.568675Z digest=sha256:a583dfa0e29137036c8ad459fdce8d810bdaf45f3a6bb0a03bc669f7a18a247f

Observation d62b1c3c-9bd6-4f42-879d-7444999a3c03 · inbound

The Effect of Perceived Race and Gender on Police Language Use: Experimental Evidence from VR Simulations cites this paper.

The Effect of Perceived Race and Gender on Police Language Use: Experimental Evidence from VR Simulations Multi-LogiEval: Towards Evaluating Multi-Step Logical Reasoning Ability of Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T10:34:08.625645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:34:08.625645Z digest=sha256:e56297370e0683a54edebafadf6d928c219b4025ea47f7ca91798d42d90fc5d7

Observation d67ff240-788d-4a0b-be18-c6ac63cf2fad · inbound

Surfacing the Unsaid: CUE-Bench for Affective Stance in Chinese Discourse cites this paper.

Surfacing the Unsaid: CUE-Bench for Affective Stance in Chinese Discourse Multi-LogiEval: Towards Evaluating Multi-Step Logical Reasoning Ability of Large Language Models

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-12T16:59:19.351541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T16:59:19.351541Z digest=sha256:d4a07af1e12e697a78ee64fc197baf296f4b63f2eb5e84059f4d589fce2e5a91

Observation 735f023c-3c43-49b8-a49c-78b21bbd32a2 · inbound

Surfacing the Unsaid: CUE-Bench for Affective Stance in Chinese Discourse cites this paper.

Surfacing the Unsaid: CUE-Bench for Affective Stance in Chinese Discourse Multi-LogiEval: Towards Evaluating Multi-Step Logical Reasoning Ability of Large Language Models

Reference 145

Resolution
unresolved
no resolver link, observed 2026-08-15T14:19:34.811705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:19:34.811705Z digest=sha256:2574800ecf32a19fed15fc385b601e357673431c5baaea073f4bec70fa797d97