Pith. sign in

Paper Citation Record · LEDGER

mHumanEval -- A Multilingual Benchmark to Evaluate Large Language Models for Code Generation

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 3 inbound Pith citation observations for arXiv:2410.15037.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.15037 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 3 of 3 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:29:19.036173Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T12:11:08.258913Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 420e91c7-9ceb-44a7-802d-b81cf01f07ec · inbound

SwiftEval: Developing a Language-Specific Benchmark for LLM-generated Code Evaluation cites this paper.

SwiftEval: Developing a Language-Specific Benchmark for LLM-generated Code Evaluation mHumanEval -- A Multilingual Benchmark to Evaluate Large Language Models for Code Generation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:19.036173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:19.036173Z digest=sha256:fa81f9c5360d4c8aba258b8b0dc267eab6983bb2edf31897d7702038bbb67bd5

Observation b9f315b6-09bb-4839-8937-5233d01e3212 · inbound

SIMCODE: A Benchmark for Natural Language to ns-3 Network Simulation Code Generation cites this paper.

SIMCODE: A Benchmark for Natural Language to ns-3 Network Simulation Code Generation mHumanEval -- A Multilingual Benchmark to Evaluate Large Language Models for Code Generation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T17:24:05.899426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:24:05.899426Z digest=sha256:618ff39e39642892641ed0b8bcee46831460ebe9b97147a8cd4c46c0632b8262

Observation fd395e90-bbf8-426b-80fa-84ca13aafef1 · inbound

MORPHOGEN: A Multilingual Benchmark for Evaluating Gender-Aware Morphological Generation cites this paper.

MORPHOGEN: A Multilingual Benchmark for Evaluating Gender-Aware Morphological Generation mHumanEval -- A Multilingual Benchmark to Evaluate Large Language Models for Code Generation

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:11:08.267658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T04:04:12.610168Z digest=sha256:fa18e9c0b47be15f9017e507d7fe287f5e3093f8e2e16a4a01aaf6ad2fc6be72