Pith. sign in

Paper Citation Record · LEDGER

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks

As of 7 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 2 inbound Pith citation observations for arXiv:2507.08538.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.08538 v1

Coverage vector

measured 45 of 45 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:19:59.488074Z

measured 47 of 47 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-27T22:16:42.836058Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T08:09:41.063430Z

Reference resolution

45 of 45 outbound references displayed

  • verified exact2
  • verified fuzzy1
  • unresolved42
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ea5dc9a5-023e-4979-9cef-feceaabe33a4 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:20:00.116357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T18:19:59.286503Z digest=sha256:02ac9881cbafb06505c74b7acd7629c2eaa89c50262f62de71f200cde62d141e

Observation a3f2fbe0-9388-49c3-8b8f-62c6a6591335 · outbound

This paper cites SIB-200: A Simple, Inclusive, and Big Evaluation Dataset for Topic Classification in 200+ Languages and Dialects.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks SIB-200: A Simple, Inclusive, and Big Evaluation Dataset for Topic Classification in 200+ Languages and Dialects

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.297703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.297703Z digest=sha256:64cc8c6c7c1b987a83e544f21fa551c4766eef8ee109b577500103a535191f6a

Observation 74157662-1675-4467-8bcc-c0cddf292b77 · outbound

This paper cites IrokoBench: A New Benchmark for African Languages in the Age of Large Language Models.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks IrokoBench: A New Benchmark for African Languages in the Age of Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.301933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.301933Z digest=sha256:e6ce6d51af3c40f0cfcebb88d696b031f83daa32daed8194b0e2d8eb7c52a35b

Observation c5d39465-3f03-4cb6-9ff7-818308f676a7 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.313039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.313039Z digest=sha256:2d1cbc61db5430d97fb70d021289af5d9de879678a0070e4504cde103f878541

Observation 73d888ff-7b96-4575-bfc2-f15485c3c338 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.316902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.316902Z digest=sha256:414359225051ece9d4398f54c36842c3092a3c626c2634dd7d6c82a38497261d

Observation 12900b22-1ea5-432d-9e9a-7d5c63184090 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:20:00.103872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T18:19:59.320491Z digest=sha256:eccb4bf196f02227ee703a96c3114fc6ef24764184d16cad674286f501e353ec

Observation f50d2a13-f26b-4a51-ab1f-a6eab9fb5177 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.324622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.324622Z digest=sha256:25913d5dc3d94c15c910cb086518254bc0937d9deb2d09bc330f53feef16855a

Observation 9495b29a-229b-43fe-b043-f51a8d8d39c5 · outbound

This paper cites Uhura: A Benchmark for Evaluating Scientific Question Answering and Truthfulness in Low-Resource African Languages.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Uhura: A Benchmark for Evaluating Scientific Question Answering and Truthfulness in Low-Resource African Languages

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.327369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.327369Z digest=sha256:efb5429ccec5999d8b2bda6d8d4c78d3e3476c1f26d561591cb066e76aaca5d5

Observation b13a21ec-78e8-4094-977a-93b2acdc80f9 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.331114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.331114Z digest=sha256:e0bdafb41ae2b0e3ef5c09a6a6d6f7ff4798b22f9fc46e6da2411c2d2e8b0f6c

Observation 316ba190-ca4d-4be0-804d-3cb9a2a4f783 · outbound

This paper cites Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.345260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.345260Z digest=sha256:9e99219e58e6c558da2abd70c429b94b726eb8283877169d6fb3f91751d02502

Observation 2dadaf57-a4d2-4e71-9310-6510fe1dfcf5 · outbound

This paper cites Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.348657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.348657Z digest=sha256:fe3e60daf73a0215a07ce28ff57966ac9539db931cbf49063d28a3ba0702e364

Observation 2ccd1855-5476-4a80-a4e0-891fa333c7cc · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Training Verifiers to Solve Math Word Problems

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.352217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.352217Z digest=sha256:149ef5a73132151a132b94ceec1cc158b164adb8c91bb0106cbeb038b9d78205

Observation 40d3f81d-f183-46b7-9f57-013f7a0f15e9 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Training Verifiers to Solve Math Word Problems

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.358937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.358937Z digest=sha256:a04ae7c67b064f145868b1d9c6f8e30468222be3519302a1465224763f4b7663

Observation 5b815c6d-6f48-449e-a8d6-69bbf1ee79f4 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.362269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.362269Z digest=sha256:0957bd4ff40bc3adce3ae33bc64f70fcfb335fe456686a81695df0d4b9deb5b9

Observation 68de6cc5-cffb-4362-b86b-306e8654b6e7 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.370724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.370724Z digest=sha256:b29b5b3b0cc9bddae2eccaa07b99c08019671a70aa9611e01126bc4b72155992

Observation a3ebffed-3383-44b8-b7c6-1c8850839acb · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.373566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.373566Z digest=sha256:4aef1018edec3f42d7c546ba56a26c8ec436cb52998ddd0d25aa0df1eeecc864

Observation 40f01db5-8738-4e51-a15a-31e2aabac34a · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.376093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.376093Z digest=sha256:477122349fd5a825c8ca482bb7b722a42ec6a6f17a650b5c047a3a0d231a3ac6

Observation 2d44566c-ee72-4965-bd9b-62fbb6cefabf · outbound

This paper cites Eberhard, Gary F.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Eberhard, Gary F

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:20:00.084843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T18:19:59.380102Z digest=sha256:c877f2d74c019213f02f58948b76e0ac0acece4722b320f607abc9d4b2a70535

Observation 371f5463-c200-41bf-a438-ba27692daa18 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:20:00.065614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T18:19:59.387423Z digest=sha256:a7aa33ed8e2060855c3bcfb2b7a08c3c208681087c34d40b64dee5fb3b663cc1

Observation 7fa8c939-2e20-4dbd-a464-da82de0f37d9 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 21

Resolution
verified exact
doi, observed 2026-08-06T18:19:59.597764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T18:19:59.397423Z digest=sha256:39fda117e1dd8d8aac556ef01ca333e278c32f07121c56c41e7fa415a829bca8

Observation 1d22288d-4ec5-4ebf-8937-6ec2001dda73 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.401528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.401528Z digest=sha256:f03268cb81b600be38e12d09ecd9415aa43cb32d24ff7365c67aa6de4816179c

Observation d646cb52-a33e-4b21-9279-fee469173916 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.404568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.404568Z digest=sha256:e80762031da1f9cf1fd64f7b117094056a4a0c40c9e5b786be68043ec44bf52a

Observation 2abf76eb-59de-4ed3-b413-675fdb541f32 · outbound

This paper cites The Llama 3 Herd of Models.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks The Llama 3 Herd of Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.420042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.420042Z digest=sha256:054a9b541ae21adf6acae5ff83a98a222f14a3268f4bb4c752edcfeb10bd08d0

Observation 2a3720dd-1d5f-4fa7-a27b-a521650b0d69 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 25

Resolution
verified exact
doi, observed 2026-08-06T18:19:59.572001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T18:19:59.422684Z digest=sha256:f1b8b5ab0b496878ae10192d5c12ec3efbf0df5aaa116c290a7246afe6282c9a

Observation 3f6dd766-7819-4598-9c5c-3e5fee4c167e · outbound

This paper cites Measuring Massive Multitask Language Understanding.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Measuring Massive Multitask Language Understanding

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.425715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.425715Z digest=sha256:3a95bcc7584fa5bb859d5bad232a1283083d76d571397bf3ca0b44ab897c17a0

Observation b0ccd95b-80db-4b7d-ae55-6aa4eaa3fd47 · outbound

This paper cites Measuring Massive Multitask Language Understanding.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Measuring Massive Multitask Language Understanding

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.429610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.429610Z digest=sha256:41544eb62edc7003b5bce272bf6fbf277d4540f9ffa334ad763e01132b9c276e

Observation 9f4be4c7-7d94-47bf-b8bf-a289c9cc45c2 · outbound

This paper cites Okapi: Instruction-tuned Large Language Models in Multiple Languages with Reinforcement Learning from Human Feedback.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Okapi: Instruction-tuned Large Language Models in Multiple Languages with Reinforcement Learning from Human Feedback

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.432414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.432414Z digest=sha256:0e5067e397fd9b146154b20581ae7dec2e05f77ec36a01281e080b23f681013d

Observation 6abaf680-c7f8-44c5-9d50-82fecfaff93b · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.435453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.435453Z digest=sha256:6a0ab1a80414bb6763289d634ba03417e93704932db13756e7437189ad389fcd

Observation edc97634-f5c4-44aa-b52d-5005161a89a1 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:20:00.049239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T18:19:59.438133Z digest=sha256:b940e5b97b32b9b6c5013349f1f586741c66b37a8687863b373905e1b5a7f019

Observation f2fa17cc-eb8c-4d46-b7b0-47104a8c9d1c · outbound

This paper cites AfroBench: How Good are Large Language Models on African Languages?.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks AfroBench: How Good are Large Language Models on African Languages?

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.443625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.443625Z digest=sha256:9e09616bf92aa536d26f62c4efb49259210ad57123d3f422e9f2c093ea9725fd

Observation 0d0cf63a-55ed-4e2c-a021-74a74d668996 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.446447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.446447Z digest=sha256:2571c3129235ebadd1c43eb58b8bc9e2a1231178ac9b9372b8e2c4e73dc43814

Observation 09b840e1-b651-4402-bccc-354a17892cfa · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:20:00.031481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T18:19:59.449243Z digest=sha256:02f642da51571c2d615bb3aa93dd48f8aa050f92c5a15050a86d51fb81fcad70

Observation ea9e2e71-2600-448a-b84c-7cbb86697667 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:20:00.018207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T18:19:59.451844Z digest=sha256:69a990813806618b3e3f539648f3fc0a3a053f35bc5aa1cbad19f8df13ba81a2

Observation ec21d80d-d331-42aa-8f16-d7cbebd7a210 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.454603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.454603Z digest=sha256:d8f22c62c869d5d7b17d60e54e317c4210b42341a5ed2ab3707cca60971eeabe

Observation 2fa77bbf-7f2e-4af4-9480-bbf1bc74644e · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:20:00.000793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T18:19:59.457204Z digest=sha256:f054e4831c7e15212eb472d1bbafa379b6d6def67d2fe0877d6a95378e2f95d7

Observation f4dbf787-c595-4360-92b2-1435ec378768 · outbound

This paper cites Language Models are Multilingual Chain-of-Thought Reasoners.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Language Models are Multilingual Chain-of-Thought Reasoners

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.459779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.459779Z digest=sha256:8ebab2987b45f8eab6f11ec8543bebf7a08cbdbd645b0bd4eb41c39e2e7fc854

Observation 75c7a333-5511-4883-9966-214c158b7954 · outbound

This paper cites Global MMLU: Understanding and Addressing Cultural and Linguistic Biases in Multilingual Evaluation.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Global MMLU: Understanding and Addressing Cultural and Linguistic Biases in Multilingual Evaluation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.462845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.462845Z digest=sha256:1b109df8320a44bc9ea01e2c57c85e358a360cb4ca3037a9c1478215853c4e86

Observation c39204d1-abe3-4fc3-a480-15c1b4b0e12c · outbound

This paper cites Gemma 3 Technical Report.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Gemma 3 Technical Report

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.466179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.466179Z digest=sha256:f2f886614f8a7cdb86bcc1334c0f2e978ef26d652d8008c3175903b94e4cc994

Observation 9e9bfe70-e9d4-449f-82b4-ed9ee62f7b75 · outbound

This paper cites Towards Multilingual LLM Evaluation for European Languages.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Towards Multilingual LLM Evaluation for European Languages

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.469267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.469267Z digest=sha256:0b2296a3593372deaf784371453a4c3b0d4b704bc40299ece6b8573232483bd2

Observation b138c725-3ee2-4ae8-94b3-f9fa36f6f61f · outbound

This paper cites Towards Multilingual LLM Evaluation for European Languages.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Towards Multilingual LLM Evaluation for European Languages

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.472849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.472849Z digest=sha256:1c2b29545258455bf8d298e3ebb715bd0a54d27ffcd9e2829b903bae09b4e03e

Observation 492a34b0-5344-4c66-b1ad-0411e78bc728 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.475927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.475927Z digest=sha256:589b838d5c740f01dbf37c5aea1bcc4f39ee1d2085265f6de6875fdd4371fba3

Observation 47f9d69e-db9d-46d3-8415-ca814bdfb838 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.478444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.478444Z digest=sha256:a1b0dbfe3604349bd9b42b7f31f370b917eed766a2a7b4a1ecc9e00872f76319

Observation 92444df7-09a0-46c4-bbfb-5e41e53ca32a · outbound

This paper cites BERTScore: Evaluating Text Generation with BERT.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks BERTScore: Evaluating Text Generation with BERT

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.481633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.481633Z digest=sha256:c1eec428ea45ebb91802c685fdb15af27d567798c91d143fefa3fd2c09930ebe

Observation a693c908-41b3-4107-baa7-f147dc88334d · outbound

This paper cites online" 'onlinestring :=.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks online" 'onlinestring :=

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.484979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.484979Z digest=sha256:b72c727105994b6f1bb93a67b70df33c813522fb947af316cbc653499affb1e9

Observation 15b845e3-1c0e-4bc5-a390-2e997e86499a · outbound

This paper cites write newline.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks write newline

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.488074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.488074Z digest=sha256:08ed47ed50f98731f0d07e0905bb3ec38810dd875eb5d981d4b60882054d6ed5

Pith citing papers

Observation d6b8e740-e173-4c1f-be05-b7e98a48b732 · inbound

MADE: Beyond Scoring via a Multilingual Agentic Diagnosing Engine for Fine-Grained Evaluation Insights cites this paper.

MADE: Beyond Scoring via a Multilingual Agentic Diagnosing Engine for Fine-Grained Evaluation Insights The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks

Reference 65

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T16:57:09.878647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-27T22:16:42.836058Z digest=sha256:4f403a05dd6c0ce52aa3a5f4311d26a150c995ba88ca83beb6ff67dde26c8933

Observation fe845ad2-462c-4b79-9366-31d37fc2fe62 · inbound

The Language-Energy Divide: Measuring Energy Costs of Multilingual LLM Inference cites this paper.

The Language-Energy Divide: Measuring Energy Costs of Multilingual LLM Inference The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-04T08:09:41.065578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T12:12:45.853559Z digest=sha256:5ed8c97f33f7731dd603ab7d058721ab2babb1dcea2fd9f2d6d360791ba582ce