Pith. sign in

Paper Citation Record · LEDGER

The NordDRG AI Benchmark for Large Language Models

As of 8 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 0 inbound Pith citation observations for arXiv:2506.13790.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.13790 v3

Coverage vector

measured 32 of 32 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:47:18.089814Z

measured 32 of 32 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

32 of 32 outbound references displayed

  • verified exact3
  • verified fuzzy2
  • unresolved26
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 515b9af0-e0cb-483d-a53b-8e470be481d7 · outbound

This paper cites Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models.

The NordDRG AI Benchmark for Large Language Models Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:17.997316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:47:17.997316Z digest=sha256:1b7a30af9662b7c45e43cedf4839fcf73806c5b051af7e5a11e4a5b014986da7

Observation c94917ea-8bb7-453c-8480-42d323205069 · outbound

This paper cites Large Language Model in Medical Informatics: Direct Classification and Enhanced Text Representations for Automatic ICD Coding.

The NordDRG AI Benchmark for Large Language Models Large Language Model in Medical Informatics: Direct Classification and Enhanced Text Representations for Automatic ICD Coding

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-07T04:47:18.380851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:47:18.001490Z digest=sha256:9e35d55d948f36e02b7d56f345302fc761a290ea983361eeac68ba72cfc1b4f6

Observation 25ca01e7-cd56-4f69-b036-04bc8ee9c5b0 · outbound

This paper cites Automated clinical coding using off-the-shelf large language models.

The NordDRG AI Benchmark for Large Language Models Automated clinical coding using off-the-shelf large language models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:18.004914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:47:18.004914Z digest=sha256:fd4c84267db5e7de64017e4aacb08ae3f759d533fbf0c0472a62d419eac3b185

Observation 98ff3bf9-d929-4ca1-b338-843240bfabd2 · outbound

This paper cites D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al.

The NordDRG AI Benchmark for Large Language Models D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:18.008910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:47:18.008910Z digest=sha256:25fdeed56f3436f60b34b055665ab273541513d5e3578fb02306a14b650eb31d

Observation f0a069d9-389c-4f9c-bfc5-ef28c79e0ca8 · outbound

This paper cites A Comprehensive Survey of AI-Generated Content (AIGC): A History of Generative AI from GAN to ChatGPT.

The NordDRG AI Benchmark for Large Language Models A Comprehensive Survey of AI-Generated Content (AIGC): A History of Generative AI from GAN to ChatGPT

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:18.011925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:47:18.011925Z digest=sha256:bfc89422e3cfccb7c676a52ff2ce90fb51bd0b97449e7a3d0dc661bb4041360b

Observation 5c20ff77-eb65-4dfd-8876-1ceb460c72b2 · outbound

This paper cites PaLM: Scaling Language Modeling with Pathways.

The NordDRG AI Benchmark for Large Language Models PaLM: Scaling Language Modeling with Pathways

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:18.015342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:47:18.015342Z digest=sha256:d6325f74eedf16c86db6bac608300d404a9d6b1f7ba45f3b6a2898ca8d596565

Observation 76a817a4-4fc0-49a5-8343-ddf76abd9208 · outbound

This paper cites an unresolved cited work.

The NordDRG AI Benchmark for Large Language Models Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:47:18.534353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:47:18.018921Z digest=sha256:115501bf759c124f57854e8dd04713042d38a08695277662189bddeff93009bd

Observation 8181321a-3273-4a7e-9703-6124463aac58 · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

The NordDRG AI Benchmark for Large Language Models BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:18.021758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:47:18.021758Z digest=sha256:5b9b0705668f4262347b51eb3043a7721034e8be8f7b6053696c6af962040ed4

Observation 706dadd2-271f-4412-b247-894223f4cb74 · outbound

This paper cites an unresolved cited work.

The NordDRG AI Benchmark for Large Language Models Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:47:18.525022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:47:18.024777Z digest=sha256:43f7c5b4e2e0427d6417985b8a76d98b786b44237e6d1b914cee07cf34dc1ca6

Observation ad01dd9b-1a8b-4236-9ad1-15515ca6102e · outbound

This paper cites KG-MTT-BERT: Knowledge Graph Enhanced BERT for Multi-Type Medical Text Classification.

The NordDRG AI Benchmark for Large Language Models KG-MTT-BERT: Knowledge Graph Enhanced BERT for Multi-Type Medical Text Classification

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-07T04:47:18.320567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:47:18.027497Z digest=sha256:4c282db6980655bc70e2023c513af2ceff7b2452c771c7ff82820976ac45e4bc

Observation d7ec89bf-6c0b-43a8-989c-31ea1e04be74 · outbound

This paper cites Measuring Massive Multitask Language Understanding.

The NordDRG AI Benchmark for Large Language Models Measuring Massive Multitask Language Understanding

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:18.030278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:47:18.030278Z digest=sha256:7e14eaf902599b4a266b2da44cf963340a7acab23f139edb7004c48678d3e502

Observation a593e703-acda-4513-8cde-781110510a29 · outbound

This paper cites Training Compute-Optimal Large Language Models.

The NordDRG AI Benchmark for Large Language Models Training Compute-Optimal Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:18.034739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:47:18.034739Z digest=sha256:51fa1152690176c700817bbb1dbd8e84d590abfbea2dafeb1638ab1521e47676

Observation 724d74b9-b986-4d39-93fd-1e761a2c6918 · outbound

This paper cites an unresolved cited work.

The NordDRG AI Benchmark for Large Language Models Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:47:18.514570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:47:18.037464Z digest=sha256:1d2d881a1fbdf799b93d3643afe593ce9c9e0ab773a3fc725570ef690f129e6e

Observation 858e0aec-dfe5-44d8-95dc-a5b9fd582e3f · outbound

This paper cites Large language models are good medical coders, if provided with tools.

The NordDRG AI Benchmark for Large Language Models Large language models are good medical coders, if provided with tools

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-08-07T04:47:18.287099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:47:18.040148Z digest=sha256:3931d7279d79948e15271a1209f39a77297dd7984208d7735de79260d513db2d

Observation 51ec63c1-4f0b-4b96-94f9-a00835c174b3 · outbound

This paper cites Holistic Evaluation of Language Models.

The NordDRG AI Benchmark for Large Language Models Holistic Evaluation of Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:18.042981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:47:18.042981Z digest=sha256:7a73bc580e322ad3843eee509c561e7fa35229c3ba6f902afe28683aa8eb1259

Observation 64835f72-9593-43a5-ac45-635a35268ecc · outbound

This paper cites A Comprehensive Overview of Large Language Models.

The NordDRG AI Benchmark for Large Language Models A Comprehensive Overview of Large Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:18.045941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:47:18.045941Z digest=sha256:e3e56fcb56e3418c871eeece58e7650ed80d4de42a06acdf6fee88c55e273707

Observation eca5baa8-3d39-4a87-85bc-9a48fb5e34e0 · outbound

This paper cites Norddrg specification, url: https://nordcase.org/.

The NordDRG AI Benchmark for Large Language Models Norddrg specification, url: https://nordcase.org/

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:18.505630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:47:18.048743Z digest=sha256:1a8b3ccd2bc9389c19126f08e26409eb1304114404b4c9ac7e299985a4a4f355

Observation 93518a54-1235-4272-b398-3379e7673b57 · outbound

This paper cites E., Rossi, M., Hui, W., Virtanen, V., and Bragge, J.

The NordDRG AI Benchmark for Large Language Models E., Rossi, M., Hui, W., Virtanen, V., and Bragge, J

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:18.496622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:47:18.051398Z digest=sha256:94a56c835bae9e119d8c367be267e338a990646662b2e3034f3922191d0eea3b

Observation 3f2a5257-e0c2-4c52-b1a6-cbdaa92f45c8 · outbound

This paper cites an unresolved cited work.

The NordDRG AI Benchmark for Large Language Models Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:47:18.487933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:47:18.054163Z digest=sha256:47667381cc4195b2aa0769540cf3e210d38b45c28a31eb3c73b17be5ef64ea0f

Observation 8a0ecc40-b408-465a-8fad-a03c444d4215 · outbound

This paper cites an unresolved cited work.

The NordDRG AI Benchmark for Large Language Models Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:47:18.478964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:47:18.056970Z digest=sha256:336637e7d6bd339d008eeabf8e004815db0cca0e8c6b35e498b6aba001e488ad

Observation f54c6b6f-d567-46ce-9f93-6d0cca99cd40 · outbound

This paper cites an unresolved cited work.

The NordDRG AI Benchmark for Large Language Models Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:18.059547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:47:18.059547Z digest=sha256:a9dc69f921c4742477d4d501c198cbbfb3729a5f91b0b64e62baf88a8413d841

Observation 2084bc1f-c49c-426d-9219-66f44dad12ec · outbound

This paper cites Zero Shot Health Trajectory Prediction Using Transformer.

The NordDRG AI Benchmark for Large Language Models Zero Shot Health Trajectory Prediction Using Transformer

Reference 22

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T04:47:18.254806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:47:18.062146Z digest=sha256:efac2e8edb7b43dd117827b0aa968cc41bfebe8e82489a07a6c4179e07a16d1e

Observation de4be81f-1f9a-4f3c-b049-72aff053aa76 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

The NordDRG AI Benchmark for Large Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:18.064980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:47:18.064980Z digest=sha256:34ff956a19131e64ed4d99f243608588d9a694929c23785cc1db748c85ead2f1

Observation 80c47429-3c67-45f5-bb50-1609f0c6ebfd · outbound

This paper cites N., Kaiser, ., and Polosukhin, I.

The NordDRG AI Benchmark for Large Language Models N., Kaiser, ., and Polosukhin, I

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:18.067520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:47:18.067520Z digest=sha256:94c2cd73bac218b6269855316abd4b519ded974d2408c08eb8cee083913ed0a5

Observation 1b1fcbd5-2791-42c1-9d97-b10dae9525fe · outbound

This paper cites an unresolved cited work.

The NordDRG AI Benchmark for Large Language Models Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:47:18.458299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:47:18.070258Z digest=sha256:342470ea525b1ff201504c9b95ef05d4e543de688f79051a3b191d6d6328696e

Observation a9ad48ce-647e-4790-8e7e-9617b7c9b383 · outbound

This paper cites an unresolved cited work.

The NordDRG AI Benchmark for Large Language Models Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:47:18.446912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:47:18.072883Z digest=sha256:f327acadac09ea72ca6397523fb297ba39634b55a1220b57f211b492620d04be

Observation 1016ed56-20b4-42e5-9443-c65c4dd29ce4 · outbound

This paper cites an unresolved cited work.

The NordDRG AI Benchmark for Large Language Models Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:18.075336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:47:18.075336Z digest=sha256:ad594cf7a498b711b8da2e8093e0e3b61b497b1ccd9c80a3c514d01ce045535b

Observation a633e670-b955-464e-bba5-9e2ba0737daf · outbound

This paper cites Back-Propagation Optimization and Multi-Valued Artificial Neural Networks for Highly Vivid Structural Color Filter Metasurfaces.

The NordDRG AI Benchmark for Large Language Models Back-Propagation Optimization and Multi-Valued Artificial Neural Networks for Highly Vivid Structural Color Filter Metasurfaces

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:18.078257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:47:18.078257Z digest=sha256:2dc8fb04a5353dc882e88cb16aeb4959b92f70517c9945d032f596550f50e462

Observation 9cd76605-2022-42e1-9752-dd2cd1669153 · outbound

This paper cites an unresolved cited work.

The NordDRG AI Benchmark for Large Language Models Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:47:18.427000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:47:18.081450Z digest=sha256:dab50fd9fde3e118dfae36ef63f2ce744432cacd753c308ab7ab7d10d528f507

Observation aaaadb01-e9f4-4254-8684-2c78d9e7e90f · outbound

This paper cites an unresolved cited work.

The NordDRG AI Benchmark for Large Language Models Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:47:18.412518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T04:47:18.084183Z digest=sha256:550a42507ea4e5416be34933e6033430b5252fba4b76d6d60a2c4fc4efde5d41

Observation 313537b3-6259-43e4-a80b-c68f4cbabfe5 · outbound

This paper cites OPT: Open Pre-trained Transformer Language Models.

The NordDRG AI Benchmark for Large Language Models OPT: Open Pre-trained Transformer Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:18.087129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:47:18.087129Z digest=sha256:a50a00e2735118618a3bd4ec7aff954df0797e9cf5efd2dc63f66f62843a94eb

Observation 54b5d437-ff84-44ff-b5d4-724fb0e255d3 · outbound

This paper cites Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena.

The NordDRG AI Benchmark for Large Language Models Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:18.089814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:47:18.089814Z digest=sha256:dd13c88030e0db8b4b27d2f86382d87d57f702db80844924afcbb65daa7e8ab5

Pith citing papers

No inbound Pith citation observations are available.