Pith. sign in

Paper Citation Record · LEDGER

BALSAM: A Platform for Benchmarking Arabic Large Language Models

As of 10 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 1 inbound Pith citation observation for arXiv:2507.22603.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.22603 v1

Coverage vector

measured 12 of 12 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T11:36:00.280571Z

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-10T19:03:03.350371Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-10T23:30:52.664737Z

Reference resolution

12 of 12 outbound references displayed

  • verified exact0
  • verified fuzzy5
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 01da45ee-27e7-4952-83a4-46425ac1fb33 · outbound

This paper cites an unresolved cited work.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:36:00.480710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T11:36:00.259136Z digest=sha256:5eb384160f00f348e4f2a0922ec29ee0228c57fab990a4175d82b3267f31ef6b

Observation e06ab83d-1c55-4735-8fea-d57a18341fac · outbound

This paper cites an unresolved cited work.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:36:00.463511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T11:36:00.264374Z digest=sha256:de1ffbff0780cd73a5007414cd8a1f5a76c5ba6ebbda002d7354e45a04bb2876

Observation ac4e2a38-4a84-45c0-83f9-18e31d69a0b2 · outbound

This paper cites Dan Hendrycks, Collin Burns, Steven Basart, Andy Zou, Mantas Mazeika, Dawn Song, and Jacob Steinhardt.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Dan Hendrycks, Collin Burns, Steven Basart, Andy Zou, Mantas Mazeika, Dawn Song, and Jacob Steinhardt

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T11:36:00.539955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T11:36:00.228330Z digest=sha256:c51cc3fe09d4d80829cc2634c8961178e377c1ba3341165e1196a254b88262c6

Observation 157a8e67-c6f6-40ac-8138-da5596c037cc · outbound

This paper cites Arena Learning: Build Data Flywheel for LLMs Post-training via Simulated Chatbot Arena.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Arena Learning: Build Data Flywheel for LLMs Post-training via Simulated Chatbot Arena

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T11:36:00.239773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:36:00.239773Z digest=sha256:f1e691c6174f018886eb2fe616f3a9eefb1658ad7f20db771ef91ffb8f3dbcde

Observation 1afaefda-ccae-4945-b94e-a2e33ff066bd · outbound

This paper cites Evaluate the generated output by comparing it to the ground truth, considering how well it addresses the original prompt.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Evaluate the generated output by comparing it to the ground truth, considering how well it addresses the original prompt

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T11:36:00.445382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T11:36:00.269899Z digest=sha256:8fc607d1661db412577a813ac4c3c3552b60c319209df0002e9559c800bb463f

Observation 2e23c6ce-d51d-4c64-8042-f66350784495 · outbound

This paper cites an unresolved cited work.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:36:00.427852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T11:36:00.275656Z digest=sha256:de98ceeb2e0ab220f3f47a172948cd00e7c9c971bbfdc59b1836d0aca782a800

Observation a1be0072-2efd-47be-8c2d-b37fd6a6c9bc · outbound

This paper cites score": 3,.

BALSAM: A Platform for Benchmarking Arabic Large Language Models score": 3,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T11:36:00.410565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T11:36:00.280571Z digest=sha256:0e7368aa6034d02f0298fd81a2fa936a0b3fd34f81903363e96052ae4be4f07f

Observation 32a637dc-4852-47fd-a38b-a61e37a53dee · outbound

This paper cites In International Conference on Learning Representations.

BALSAM: A Platform for Benchmarking Arabic Large Language Models In International Conference on Learning Representations

Reference 2021

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T11:36:00.521004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T11:36:00.233976Z digest=sha256:3711155b86e6e84f3753d3f9a88c355185b5c1012a61616d56e556da5d2e1246

Observation 9c55c0a3-d89c-49d2-82b3-628f3da1c94d · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-06T11:36:00.215303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:36:00.215303Z digest=sha256:f06cba479a7b44d23dae5cee196e0c3547ce7bbc96b7c42741498a4d135dcddf

Observation bd0921df-52e8-440c-92f8-00ef16da9de4 · outbound

This paper cites Jais and Jais-chat: Arabic-Centric Foundation and Instruction-Tuned Open Generative Large Language Models.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Jais and Jais-chat: Arabic-Centric Foundation and Instruction-Tuned Open Generative Large Language Models

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T11:36:00.250729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:36:00.250729Z digest=sha256:4195cd51efc7e80a9d207669fe878d76522b2cbb7e7972d2cc3b0ef7285bd340

Observation d3f97aaa-4e76-433a-8eda-725ae6fffd57 · outbound

This paper cites Creating Arabic LLM Prompts at Scale.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Creating Arabic LLM Prompts at Scale

Reference 2024

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T11:36:00.366212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T11:36:00.222651Z digest=sha256:a9a8da72bd1402f816e517f4524216981ddf429861206443ee306df32e03a8f5

Observation 06ffc0bd-6a50-4701-a2e3-efbf85ba8cfe · outbound

This paper cites In Proceedings of the 31st International Conference on Computational Lin- guistics, pages 4186–4218, Abu Dhabi, UAE.

BALSAM: A Platform for Benchmarking Arabic Large Language Models In Proceedings of the 31st International Conference on Computational Lin- guistics, pages 4186–4218, Abu Dhabi, UAE

Reference 2025

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T11:36:00.500643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T11:36:00.245539Z digest=sha256:df874fa4718b4a55b767ce0a4d24d1dc5eb0393e687ec33010fa6fadbb2c46ef

Pith citing papers

Observation d292673a-38ce-4876-b9e2-315e12e1b824 · inbound

Context-Aware Dialectal Arabic Machine Translation with Interactive Region and Register Selection cites this paper.

Context-Aware Dialectal Arabic Machine Translation with Interactive Region and Register Selection BALSAM: A Platform for Benchmarking Arabic Large Language Models

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:30:52.671051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T19:03:03.350371Z digest=sha256:997a5db2ea9b2ece6f9e5936c1f165949e36494debde869d0516a49068e5f246