Pith. sign in

Paper Citation Record · LEDGER

MultiNRC: A Challenging and Native Multilingual Reasoning Evaluation Benchmark for LLMs

As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2507.17476.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.17476 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T01:32:18.540133Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T05:57:41.648856Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation abd9ad50-4753-4ed2-964c-1c53e6da9354 · inbound

CulturALL: Benchmarking Multilingual and Multicultural Competence of LLMs on Grounded Tasks cites this paper.

CulturALL: Benchmarking Multilingual and Multicultural Competence of LLMs on Grounded Tasks MultiNRC: A Challenging and Native Multilingual Reasoning Evaluation Benchmark for LLMs

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:16:06.172934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T02:04:53.929158Z digest=sha256:f91338526e8a40d76024035e7e025a75eb8ebdfbaf589e529ecb8dddcaa9974e

Observation f365a8c4-7e82-44c6-9556-0729c8f201e3 · inbound

CroCo: Cross-Lingual Contrastive Preference Tuning on Self-Generations cites this paper.

CroCo: Cross-Lingual Contrastive Preference Tuning on Self-Generations MultiNRC: A Challenging and Native Multilingual Reasoning Evaluation Benchmark for LLMs

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T21:33:59.068091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-29T21:30:39.993873Z digest=sha256:12ef07135765165af973debafe5c10383092db033b9dcc3bbda8632e495cbb8d

Observation 5d82aa61-7089-4649-8ab7-ebe59097f8f3 · inbound

CultureForest: Understanding and Evaluating Cultural Norm Grounded Reasoning in LLMs cites this paper.

CultureForest: Understanding and Evaluating Cultural Norm Grounded Reasoning in LLMs MultiNRC: A Challenging and Native Multilingual Reasoning Evaluation Benchmark for LLMs

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-01T23:16:23.741698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T14:34:52.644249Z digest=sha256:b77546d9370c88cb68b02e00d3663213dfa6cb94b0d0ce527942fbe1c7cc713f

Observation 21aee8e3-8550-4bd9-872a-8d7b0ebc04d9 · inbound

The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes cites this paper.

The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes MultiNRC: A Challenging and Native Multilingual Reasoning Evaluation Benchmark for LLMs

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-07-03T05:57:41.650298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T12:59:51.091008Z digest=sha256:e5f8615513fb72c438f0d9d27cfaa6cceff88fb8720aad321883a1c106dfcb5b

Observation 912c0fec-b5af-4690-8d14-bf9d701d022e · inbound

SFBench: The SciFy Scientific Feasibility Benchmark cites this paper.

SFBench: The SciFy Scientific Feasibility Benchmark MultiNRC: A Challenging and Native Multilingual Reasoning Evaluation Benchmark for LLMs

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T07:04:21.319521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T06:59:19.857066Z digest=sha256:e3256145efe43c9925362b59759a18aa081630be675938ce64fca75ae486d67b

Observation 39b8cba8-2720-4852-8948-b72dbf75b1ce · inbound

Reasoning Consensus: Structural Ensembling of LLM Reasoning via Weighted DAG Aggregation cites this paper.

Reasoning Consensus: Structural Ensembling of LLM Reasoning via Weighted DAG Aggregation MultiNRC: A Challenging and Native Multilingual Reasoning Evaluation Benchmark for LLMs

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T01:32:18.540133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T01:32:18.540133Z digest=sha256:0ce2823299b24b2d72ff43b49dd67f905696556c9a9ea8ca39516ef3cce4bd12