Pith. sign in

Paper Citation Record · LEDGER

ReCEval: Evaluating Reasoning Chains via Correctness and Informativeness

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2304.10703.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2304.10703 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T04:40:19.139964Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T04:37:36.781188Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3024974b-e2c8-4826-ac94-5d257b215237 · inbound

MINERVA: Evaluating Complex Video Reasoning cites this paper.

MINERVA: Evaluating Complex Video Reasoning ReCEval: Evaluating Reasoning Chains via Correctness and Informativeness

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.139964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.139964Z digest=sha256:fc068df73620f84052ab9e5d862a52c5c52002a7625f840887e8eaba12464ef6

Observation b69102e0-1c11-49d3-ae9f-dcaab0c0559f · inbound

Detection and Mitigation of Hallucination in Large Reasoning Models: A Mechanistic Perspective cites this paper.

Detection and Mitigation of Hallucination in Large Reasoning Models: A Mechanistic Perspective ReCEval: Evaluating Reasoning Chains via Correctness and Informativeness

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T20:28:19.061679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:28:19.061679Z digest=sha256:cb97ac99a4bf381ae3900aec78c08b3bc710d507753c9d4b6dec26580d63d23c

Observation 818a946b-3f6c-4f31-a07f-6200ea3d67e2 · inbound

Retrieval Augmented Decision-Making: A Requirements-Driven, Multi-Criteria Framework for Structured Decision Support cites this paper.

Retrieval Augmented Decision-Making: A Requirements-Driven, Multi-Criteria Framework for Structured Decision Support ReCEval: Evaluating Reasoning Chains via Correctness and Informativeness

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:33:47.710576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:33:47.710576Z digest=sha256:b9677a18b615509c034c998215dafbbd4428bc35d4e3153f7a0be41695450ea4

Observation 9e83b7d7-971b-4220-8f1f-5ce5d603d714 · inbound

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM cites this paper.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM ReCEval: Evaluating Reasoning Chains via Correctness and Informativeness

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:10.405095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:10.405095Z digest=sha256:2d46c1ea300f68dcd5b992d4750ebc8d72a6bba84ccb6e28f6542e07417c035d

Observation 4d15d7d0-eef1-42aa-a693-9fbaf3263aff · inbound

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains cites this paper.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains ReCEval: Evaluating Reasoning Chains via Correctness and Informativeness

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:32.114341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:32.114341Z digest=sha256:02358551146959ed21eff507645a583a1debaa1899a10fec0ea33b50e37b9cf3

Observation 7d6bf28a-8846-4d2c-a4c9-afa79093e908 · inbound

Chain-of-Code Collapse: Reasoning Failures in LLMs via Adversarial Prompting in Code Generation cites this paper.

Chain-of-Code Collapse: Reasoning Failures in LLMs via Adversarial Prompting in Code Generation ReCEval: Evaluating Reasoning Chains via Correctness and Informativeness

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T05:51:08.254881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:51:08.254881Z digest=sha256:f1db520c01b55fc8293615450ccb2fe5344aa478ac66eee07e696bceaf7f5baf

Observation b8e998bf-7eae-4b59-bb1d-43c97d681ccf · inbound

Rethinking Human Preference Evaluation of LLM Rationales cites this paper.

Rethinking Human Preference Evaluation of LLM Rationales ReCEval: Evaluating Reasoning Chains via Correctness and Informativeness

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.887155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.887155Z digest=sha256:2984ebbbc4edb4bd1bcea413d75e701b510b992ceb682bc7c95ed387d51a9b1e

Observation 9fea0482-9d0f-4543-996a-c99c71027027 · inbound

A Survey on Evaluating Quality and Trustworthiness in LLM-Generated Data cites this paper.

A Survey on Evaluating Quality and Trustworthiness in LLM-Generated Data ReCEval: Evaluating Reasoning Chains via Correctness and Informativeness

Reference 173

Resolution
unresolved
no resolver link, observed 2026-08-03T08:15:28.060444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T08:15:28.060444Z digest=sha256:1aee1df461d678d11711e5872909cc7935e54771ce8190e231c3f2e54a5b9acf

Observation c123eb9d-781b-49eb-9ee1-5674ca531a17 · inbound

Fragile Thoughts: How Large Language Models Handle Chain-of-Thought Perturbations cites this paper.

Fragile Thoughts: How Large Language Models Handle Chain-of-Thought Perturbations ReCEval: Evaluating Reasoning Chains via Correctness and Informativeness

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-16T03:47:15.332782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-16T03:43:18.987241Z digest=sha256:6c6b04e0e0177b302c484835b8fbc75f1ab684ba07749b9d8debf59340d54a6a

Observation 94b106a3-1bc7-4b3d-8944-042e81ff2d1f · inbound

Beyond Output Correctness: Benchmarking and Evaluating Large Language Model Reasoning in Coding Tasks cites this paper.

Beyond Output Correctness: Benchmarking and Evaluating Large Language Model Reasoning in Coding Tasks ReCEval: Evaluating Reasoning Chains via Correctness and Informativeness

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T09:36:02.566440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T15:56:03.196091Z digest=sha256:d6391594e8d4f4327ea0d74c86ce4203c2d0c21c4b6d2cabc0c53891247a58cc

Observation cce65705-223c-4626-b753-e9dab7a3b6e1 · inbound

Supervised Fine-tuning with Synthetic Rationale Data Hurts Real-World Disease Prediction cites this paper.

Supervised Fine-tuning with Synthetic Rationale Data Hurts Real-World Disease Prediction ReCEval: Evaluating Reasoning Chains via Correctness and Informativeness

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T04:37:36.786700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-27T13:48:39.832057Z digest=sha256:9d85c67f0106214b4634e7b171573187eea70925a2e7eae3a1777d27950ce203