Pith. sign in

Paper Citation Record · LEDGER

The NordDRG AI Benchmark for Large Language Models

As of 10 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 0 inbound Pith citation observations for arXiv:2506.13790.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.13790 v3

Coverage vector

measured 32 of 32 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:47:18.089814Z

measured 32 of 32 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

32 of 32 outbound references displayed

  • verified exact3
  • verified fuzzy2
  • unresolved26
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 515b9af0-e0cb-483d-a53b-8e470be481d7 · outbound

This paper cites Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models.

The NordDRG AI Benchmark for Large Language Models Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:17.997316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:47:17.997316Z digest=sha256:4b892fb767af74c41df1f820c48ead591078554b25fb6bd4a4ac2cc5c2acae9f

Observation c94917ea-8bb7-453c-8480-42d323205069 · outbound

This paper cites Large Language Model in Medical Informatics: Direct Classification and Enhanced Text Representations for Automatic ICD Coding.

The NordDRG AI Benchmark for Large Language Models Large Language Model in Medical Informatics: Direct Classification and Enhanced Text Representations for Automatic ICD Coding

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-07T04:47:18.380851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T04:47:18.001490Z digest=sha256:3fd997916fb1962dc930785956cfcaf3cfe57e2bd656f4df0fd1e27c3f66fca4

Observation 25ca01e7-cd56-4f69-b036-04bc8ee9c5b0 · outbound

This paper cites Automated clinical coding using off-the-shelf large language models.

The NordDRG AI Benchmark for Large Language Models Automated clinical coding using off-the-shelf large language models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:18.004914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:47:18.004914Z digest=sha256:da6ab03043b3e9f47918c9e01e108dc4166d604cbab77064f1749a81290a3300

Observation 98ff3bf9-d929-4ca1-b338-843240bfabd2 · outbound

This paper cites D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al.

The NordDRG AI Benchmark for Large Language Models D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:18.008910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:47:18.008910Z digest=sha256:a99868a4d45fee95b0a35582086d727380ccd24604dd36ed5946ab0463f6ad06

Observation f0a069d9-389c-4f9c-bfc5-ef28c79e0ca8 · outbound

This paper cites A Comprehensive Survey of AI-Generated Content (AIGC): A History of Generative AI from GAN to ChatGPT.

The NordDRG AI Benchmark for Large Language Models A Comprehensive Survey of AI-Generated Content (AIGC): A History of Generative AI from GAN to ChatGPT

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:18.011925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:47:18.011925Z digest=sha256:773000b7558cd5d71da140578442c3b37bd58df37425ed0c5dd33f7ad2ce9c25

Observation 5c20ff77-eb65-4dfd-8876-1ceb460c72b2 · outbound

This paper cites PaLM: Scaling Language Modeling with Pathways.

The NordDRG AI Benchmark for Large Language Models PaLM: Scaling Language Modeling with Pathways

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:18.015342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:47:18.015342Z digest=sha256:10ac2619ab3407d5c0254474e6e7c2284182204285661394d8cec46f0a0fe4b1

Observation 76a817a4-4fc0-49a5-8343-ddf76abd9208 · outbound

This paper cites an unresolved cited work.

The NordDRG AI Benchmark for Large Language Models Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:47:18.534353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T04:47:18.018921Z digest=sha256:67035cf464a54cba021ffdac8a35f1de5b2047d554449fe8c12e59b3d99638e7

Observation 8181321a-3273-4a7e-9703-6124463aac58 · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

The NordDRG AI Benchmark for Large Language Models BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:18.021758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:47:18.021758Z digest=sha256:eaf6c965784449dea1869a2fdedd82a9dd5a66b817e39a7467e39a7a46bb3617

Observation 706dadd2-271f-4412-b247-894223f4cb74 · outbound

This paper cites an unresolved cited work.

The NordDRG AI Benchmark for Large Language Models Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:47:18.525022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T04:47:18.024777Z digest=sha256:7292dffe141db13de34588a03d003bb78099e46239521675be70e66427f9d938

Observation ad01dd9b-1a8b-4236-9ad1-15515ca6102e · outbound

This paper cites KG-MTT-BERT: Knowledge Graph Enhanced BERT for Multi-Type Medical Text Classification.

The NordDRG AI Benchmark for Large Language Models KG-MTT-BERT: Knowledge Graph Enhanced BERT for Multi-Type Medical Text Classification

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-07T04:47:18.320567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T04:47:18.027497Z digest=sha256:ba15b254fc597905c1de18bb1c9882059bc6167f305e90ae6f7c9fe78b17ad21

Observation d7ec89bf-6c0b-43a8-989c-31ea1e04be74 · outbound

This paper cites Measuring Massive Multitask Language Understanding.

The NordDRG AI Benchmark for Large Language Models Measuring Massive Multitask Language Understanding

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:18.030278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:47:18.030278Z digest=sha256:5260f0a20f4d63db25cab46faced3437704c51cb1ce7fd2579dde5edc0019e6e

Observation a593e703-acda-4513-8cde-781110510a29 · outbound

This paper cites Training Compute-Optimal Large Language Models.

The NordDRG AI Benchmark for Large Language Models Training Compute-Optimal Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:18.034739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:47:18.034739Z digest=sha256:9151ebf4536bdc91d2ea5e8d7712d0f894fd7609c27c099fe29bcfa49e11a140

Observation 724d74b9-b986-4d39-93fd-1e761a2c6918 · outbound

This paper cites an unresolved cited work.

The NordDRG AI Benchmark for Large Language Models Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:47:18.514570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T04:47:18.037464Z digest=sha256:1e08da339de07df14fa6c897a3a2e2d1b29233557536e8da3b2ec00d675668cf

Observation 858e0aec-dfe5-44d8-95dc-a5b9fd582e3f · outbound

This paper cites Large language models are good medical coders, if provided with tools.

The NordDRG AI Benchmark for Large Language Models Large language models are good medical coders, if provided with tools

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-08-07T04:47:18.287099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T04:47:18.040148Z digest=sha256:3e51de7c2044f8bfe25a79526d5775cf54612811a88ecdac5deffb72d8b6622b

Observation 51ec63c1-4f0b-4b96-94f9-a00835c174b3 · outbound

This paper cites Holistic Evaluation of Language Models.

The NordDRG AI Benchmark for Large Language Models Holistic Evaluation of Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:18.042981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:47:18.042981Z digest=sha256:30c8e913839e05c6fb65e0c6343d0f10b3a412e4bf2f9313e1089ed813d75fb8

Observation 64835f72-9593-43a5-ac45-635a35268ecc · outbound

This paper cites A Comprehensive Overview of Large Language Models.

The NordDRG AI Benchmark for Large Language Models A Comprehensive Overview of Large Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:18.045941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:47:18.045941Z digest=sha256:ba3ca380c0ec64c1166203d33d2896793412b9a523abf8aeb323637ecdd9e43b

Observation eca5baa8-3d39-4a87-85bc-9a48fb5e34e0 · outbound

This paper cites Norddrg specification, url: https://nordcase.org/.

The NordDRG AI Benchmark for Large Language Models Norddrg specification, url: https://nordcase.org/

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:18.505630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T04:47:18.048743Z digest=sha256:e5f59be62afe2be363b3530f626c26ac019d86d6662a115ba1d9a264e07c90f0

Observation 93518a54-1235-4272-b398-3379e7673b57 · outbound

This paper cites E., Rossi, M., Hui, W., Virtanen, V., and Bragge, J.

The NordDRG AI Benchmark for Large Language Models E., Rossi, M., Hui, W., Virtanen, V., and Bragge, J

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:18.496622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T04:47:18.051398Z digest=sha256:dc35f9b66594ef9fb56c8d97ab97e07166e0bbbd9f0868698a3c813abca70df0

Observation 3f2a5257-e0c2-4c52-b1a6-cbdaa92f45c8 · outbound

This paper cites an unresolved cited work.

The NordDRG AI Benchmark for Large Language Models Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:47:18.487933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T04:47:18.054163Z digest=sha256:cc2bd4acb338bd1ca3e4fcaad34f9659605141c33e1d302c62414b3d0422db37

Observation 8a0ecc40-b408-465a-8fad-a03c444d4215 · outbound

This paper cites an unresolved cited work.

The NordDRG AI Benchmark for Large Language Models Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:47:18.478964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T04:47:18.056970Z digest=sha256:dcae5e26cc6ddf34d752df21e7a0b019677b4a8b1ba47af04d313ca833da5ec0

Observation f54c6b6f-d567-46ce-9f93-6d0cca99cd40 · outbound

This paper cites an unresolved cited work.

The NordDRG AI Benchmark for Large Language Models Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:18.059547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:47:18.059547Z digest=sha256:6f767ea09a1cbca7f48776dc639c65188edff67e6f4711964266488cd692f74e

Observation 2084bc1f-c49c-426d-9219-66f44dad12ec · outbound

This paper cites Zero Shot Health Trajectory Prediction Using Transformer.

The NordDRG AI Benchmark for Large Language Models Zero Shot Health Trajectory Prediction Using Transformer

Reference 22

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T04:47:18.254806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T04:47:18.062146Z digest=sha256:3ccc910128c57abb1e63200d56875e93c60fa2666127dddea8eebceff870809f

Observation de4be81f-1f9a-4f3c-b049-72aff053aa76 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

The NordDRG AI Benchmark for Large Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:18.064980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:47:18.064980Z digest=sha256:c3aa36d36e5ce66b04c7dac1190ebb0ad7fdef6436da4f1c5a9e653678324665

Observation 80c47429-3c67-45f5-bb50-1609f0c6ebfd · outbound

This paper cites N., Kaiser, ., and Polosukhin, I.

The NordDRG AI Benchmark for Large Language Models N., Kaiser, ., and Polosukhin, I

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:18.067520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:47:18.067520Z digest=sha256:88f2693b56b031cceb3e8b556c225bd7ef1b30959703c8e0b9b517ec495aaff2

Observation 1b1fcbd5-2791-42c1-9d97-b10dae9525fe · outbound

This paper cites an unresolved cited work.

The NordDRG AI Benchmark for Large Language Models Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:47:18.458299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T04:47:18.070258Z digest=sha256:c96b6afca6829b46a17897f5e52444af43cd39b67bc5f3fd6073b38f7c6f4fde

Observation a9ad48ce-647e-4790-8e7e-9617b7c9b383 · outbound

This paper cites an unresolved cited work.

The NordDRG AI Benchmark for Large Language Models Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:47:18.446912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T04:47:18.072883Z digest=sha256:653bca8d9ea4eb884ec9958667e08aac462a1f5241a0771d1f859475b0740bdc

Observation 1016ed56-20b4-42e5-9443-c65c4dd29ce4 · outbound

This paper cites an unresolved cited work.

The NordDRG AI Benchmark for Large Language Models Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:18.075336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:47:18.075336Z digest=sha256:c32e6e9199c24959c09d31d915bbed7d47df9257ebb2d0140724c80e9e38c8b0

Observation a633e670-b955-464e-bba5-9e2ba0737daf · outbound

This paper cites Back-Propagation Optimization and Multi-Valued Artificial Neural Networks for Highly Vivid Structural Color Filter Metasurfaces.

The NordDRG AI Benchmark for Large Language Models Back-Propagation Optimization and Multi-Valued Artificial Neural Networks for Highly Vivid Structural Color Filter Metasurfaces

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:18.078257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:47:18.078257Z digest=sha256:18ef8ec809a0810652d7aba8682d13f0eba91229a285ad7b6465f1bc3aff340f

Observation 9cd76605-2022-42e1-9752-dd2cd1669153 · outbound

This paper cites an unresolved cited work.

The NordDRG AI Benchmark for Large Language Models Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:47:18.427000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T04:47:18.081450Z digest=sha256:a7dbfc5bee95b29c2fddf6b8cd3e4f8067ef409c22859476db3d711694cbfa69

Observation aaaadb01-e9f4-4254-8684-2c78d9e7e90f · outbound

This paper cites an unresolved cited work.

The NordDRG AI Benchmark for Large Language Models Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:47:18.412518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T04:47:18.084183Z digest=sha256:7da73e88765270842f20143d88bbf5d0e21225812238d795b027e63722f1445b

Observation 313537b3-6259-43e4-a80b-c68f4cbabfe5 · outbound

This paper cites OPT: Open Pre-trained Transformer Language Models.

The NordDRG AI Benchmark for Large Language Models OPT: Open Pre-trained Transformer Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:18.087129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:47:18.087129Z digest=sha256:f865069c80e53ad3786bec5eb08977034d3e84e2050d7298409b0f35d220fa49

Observation 54b5d437-ff84-44ff-b5d4-724fb0e255d3 · outbound

This paper cites Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena.

The NordDRG AI Benchmark for Large Language Models Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:18.089814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:47:18.089814Z digest=sha256:3ff1a7298527369e7d215139376039792315d939d5aeca8174d92dfdceffcc72

Pith citing papers

No inbound Pith citation observations are available.