Pith. sign in

Paper Citation Record · LEDGER

The Belebele Benchmark: a Parallel Reading Comprehension Dataset in 122 Language Variants

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2308.16884.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2308.16884 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:28:01.729249Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-23T06:25:27.841513Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 425af3ab-0780-4a70-9640-6849a3e6b153 · inbound

SailCompass: Towards Reproducible and Robust Evaluation for Southeast Asian Languages cites this paper.

SailCompass: Towards Reproducible and Robust Evaluation for Southeast Asian Languages The Belebele Benchmark: a Parallel Reading Comprehension Dataset in 122 Language Variants

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T04:40:15.381678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:40:15.381678Z digest=sha256:176f40360e2e6193d5340267e87bf5b3c9e6e395e6c31bfcdc030ebfe6bab0f8

Observation f14223e4-a499-4149-98f8-e35f2ed1763b · inbound

2M-BELEBELE: Highly Multilingual Speech and American Sign Language Comprehension Dataset cites this paper.

2M-BELEBELE: Highly Multilingual Speech and American Sign Language Comprehension Dataset The Belebele Benchmark: a Parallel Reading Comprehension Dataset in 122 Language Variants

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-11T18:04:14.067390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T18:04:14.067390Z digest=sha256:318cfe1c3867f8b477aca0ce943960a4632f532a314fb4c2977d9e7b1e37ab7b

Observation 8148b7bd-ee9d-41f2-99fe-1c375ceb790b · inbound

BgGPT 1.0: Extending English-centric LLMs to other languages cites this paper.

BgGPT 1.0: Extending English-centric LLMs to other languages The Belebele Benchmark: a Parallel Reading Comprehension Dataset in 122 Language Variants

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T15:35:55.266907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:35:55.266907Z digest=sha256:5fed6b8cb932ab3b2d0f1c92ea973a23d431baf83647780f45ed91c59ce41889

Observation 148813b0-329c-4d88-a404-310f1eb67be9 · inbound

Qwen2.5 Technical Report cites this paper.

Qwen2.5 Technical Report The Belebele Benchmark: a Parallel Reading Comprehension Dataset in 122 Language Variants

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-23T06:25:27.844262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-23T06:25:00.376073Z digest=sha256:5a18122177506a2ee75373d31bc229ceda7e0241538eddb38292e3f4a9bfd039

Observation 00d773f3-783b-40a6-9d5e-4c81af4e6cae · inbound

Fleurs-SLU: A Massively Multilingual Benchmark for Spoken Language Understanding cites this paper.

Fleurs-SLU: A Massively Multilingual Benchmark for Spoken Language Understanding The Belebele Benchmark: a Parallel Reading Comprehension Dataset in 122 Language Variants

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T21:10:04.081234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:10:04.081234Z digest=sha256:78ca6f98473207ed6973f1861f69f0d380389e19fc9db7fe73ee5675fc18425d

Observation 1bac3e2e-1e6b-44b0-9804-bb872246d603 · inbound

Compass-V2 Technical Report cites this paper.

Compass-V2 Technical Report The Belebele Benchmark: a Parallel Reading Comprehension Dataset in 122 Language Variants

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-16T11:28:01.729249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:28:01.729249Z digest=sha256:1116cea6bc8593a8793774a3f5814f54b6bf749eae1b39039ef2c5e55621a6ab

Observation ae538013-3ffe-476a-807a-4989d90432be · inbound

Qwen3 Technical Report cites this paper.

Qwen3 Technical Report The Belebele Benchmark: a Parallel Reading Comprehension Dataset in 122 Language Variants

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:35:28.449337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-09T06:35:27.813995Z digest=sha256:b5740128764a0498b8e6129af48a14febde8099f4a1007dee70e12f6b92b640f

Observation c6af4b12-d765-47ac-b21d-9c1593108e11 · inbound

Multilingual Prompt Engineering in Large Language Models: A Survey Across NLP Tasks cites this paper.

Multilingual Prompt Engineering in Large Language Models: A Survey Across NLP Tasks The Belebele Benchmark: a Parallel Reading Comprehension Dataset in 122 Language Variants

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T20:53:03.228815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:53:03.228815Z digest=sha256:d7f7b5c088048bb2259747a668a6f322bef962ed8398a3b2d4b97feebfb41b1a

Observation 5d41d539-f2b3-404b-92c1-b6a48ae8adda · inbound

Cross-Lingual Optimization for Language Transfer in Large Language Models cites this paper.

Cross-Lingual Optimization for Language Transfer in Large Language Models The Belebele Benchmark: a Parallel Reading Comprehension Dataset in 122 Language Variants

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:07.811099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:07.811099Z digest=sha256:d922609f8e2cb4c3d78e1fc4f44040ecc31f73e647dcaf195c55ca959d2ce667

Observation 1d8a31b4-d56b-4e9e-9abc-43af9b76335d · inbound

Lost in the Mix: Evaluating LLM Understanding of Code-Switched Text cites this paper.

Lost in the Mix: Evaluating LLM Understanding of Code-Switched Text The Belebele Benchmark: a Parallel Reading Comprehension Dataset in 122 Language Variants

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:09.509245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:09.509245Z digest=sha256:6a8a0975bd0b22f31749e88e88c9b6e60006ddcb5387993f793ec1a85b722a0e

Observation 3d170110-4d0f-4283-a254-f16331871022 · inbound

A Vietnamese Dataset for Text Segmentation and Multiple Choices Reading Comprehension cites this paper.

A Vietnamese Dataset for Text Segmentation and Multiple Choices Reading Comprehension The Belebele Benchmark: a Parallel Reading Comprehension Dataset in 122 Language Variants

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:24.093318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:24.093318Z digest=sha256:d3f4c3feb062b59951b820b255e522e055ed07a5b2a21b0c97f0e20529edf567

Observation 6cdb0b51-9bfc-47a8-8ffb-fd49eeacacdd · inbound

From Ambiguity to Accuracy: The Transformative Effect of Coreference Resolution on Retrieval-Augmented Generation systems cites this paper.

From Ambiguity to Accuracy: The Transformative Effect of Coreference Resolution on Retrieval-Augmented Generation systems The Belebele Benchmark: a Parallel Reading Comprehension Dataset in 122 Language Variants

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T05:32:05.794157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-19T05:30:33.121799Z digest=sha256:ed254d1d990c911e9775318a078eff2a7f985de315681453ae09abe02a7717cc

Observation e2c362b1-a10c-423e-82e7-6427b600c90d · inbound

Improving Korean-English Cross-Lingual Retrieval: A Data-Centric Study of Language Composition and Model Merging cites this paper.

Improving Korean-English Cross-Lingual Retrieval: A Data-Centric Study of Language Composition and Model Merging The Belebele Benchmark: a Parallel Reading Comprehension Dataset in 122 Language Variants

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-22T00:40:51.024145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-22T00:39:24.748381Z digest=sha256:48233538d5184e0070f7a4c72f187e39e8036f18a4c72c593d12b2b1e9f40622

Observation 1b0a24ad-03f2-4b3a-995b-dd68e180a1cd · inbound

AgentScope 1.0: A Developer-Centric Framework for Building Agentic Applications cites this paper.

AgentScope 1.0: A Developer-Centric Framework for Building Agentic Applications The Belebele Benchmark: a Parallel Reading Comprehension Dataset in 122 Language Variants

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T17:26:28.791709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:26:28.791709Z digest=sha256:23902156f0ab899892849fe394f7d32ff92637ea8a7adce24b73bd64ac5aca7d

Observation 2db290fc-bd00-4632-983f-50814d13c33d · inbound

A Data-Efficient Path to Multilingual LLMs: Language Expansion via Post-training PARAM$\Delta$ Integration into Upcycled MoE cites this paper.

A Data-Efficient Path to Multilingual LLMs: Language Expansion via Post-training PARAM$\Delta$ Integration into Upcycled MoE The Belebele Benchmark: a Parallel Reading Comprehension Dataset in 122 Language Variants

Reference 27

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T11:13:13.604415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-20T11:09:22.027588Z digest=sha256:6e2a6a47ac61d7a25ee7f54e59e362ad2bf9834685426efde4c7aed6463344fd

Observation f372923f-1309-4463-829a-c640d0853e9f · inbound

Fairness Pruning: Locating Demographic Bias in GLU-MLP Layers via Differential Activations cites this paper.

Fairness Pruning: Locating Demographic Bias in GLU-MLP Layers via Differential Activations The Belebele Benchmark: a Parallel Reading Comprehension Dataset in 122 Language Variants

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-31T11:25:39.532395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T11:25:39.532395Z digest=sha256:531c94f32668d496a9d60fd75bb4950ecf0ad21e8297951ea2e054b51dd58a3b