Pith. sign in

Paper Citation Record · LEDGER

BALSAM: A Platform for Benchmarking Arabic Large Language Models

As of 7 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 1 inbound Pith citation observation for arXiv:2507.22603.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.22603 v1

Coverage vector

measured 12 of 12 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T11:36:00.280571Z

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-10T19:03:03.350371Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-10T23:30:52.664737Z

Reference resolution

12 of 12 outbound references displayed

  • verified exact0
  • verified fuzzy5
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 01da45ee-27e7-4952-83a4-46425ac1fb33 · outbound

This paper cites an unresolved cited work.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:36:00.480710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:36:00.259136Z digest=sha256:a491ecef9fd76001a481094673872bdb041232f4b2f8e3577cfd9a6194ff670a

Observation e06ab83d-1c55-4735-8fea-d57a18341fac · outbound

This paper cites an unresolved cited work.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:36:00.463511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:36:00.264374Z digest=sha256:cf4de7bcdc8ab7e543d08bb66260f08e35b89bfe7808101a185f71587736ea98

Observation ac4e2a38-4a84-45c0-83f9-18e31d69a0b2 · outbound

This paper cites Dan Hendrycks, Collin Burns, Steven Basart, Andy Zou, Mantas Mazeika, Dawn Song, and Jacob Steinhardt.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Dan Hendrycks, Collin Burns, Steven Basart, Andy Zou, Mantas Mazeika, Dawn Song, and Jacob Steinhardt

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T11:36:00.539955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:36:00.228330Z digest=sha256:1597f48450734b910e809b9aaf03ebd59597fd1e2fa53cb70ad1d6a40f029d1a

Observation 157a8e67-c6f6-40ac-8138-da5596c037cc · outbound

This paper cites Arena Learning: Build Data Flywheel for LLMs Post-training via Simulated Chatbot Arena.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Arena Learning: Build Data Flywheel for LLMs Post-training via Simulated Chatbot Arena

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T11:36:00.239773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:36:00.239773Z digest=sha256:bac8d0adf891f74aef3f1b7b2570b4f5ea7de455aaa5063b714999bb56647e1c

Observation 1afaefda-ccae-4945-b94e-a2e33ff066bd · outbound

This paper cites Evaluate the generated output by comparing it to the ground truth, considering how well it addresses the original prompt.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Evaluate the generated output by comparing it to the ground truth, considering how well it addresses the original prompt

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T11:36:00.445382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:36:00.269899Z digest=sha256:10f835459481d22bc20c1bb6ad39807b5f9cf29b61d97c80eb2725cdde08577d

Observation 2e23c6ce-d51d-4c64-8042-f66350784495 · outbound

This paper cites an unresolved cited work.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:36:00.427852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:36:00.275656Z digest=sha256:46bde7510aa947623d2ed087aa435c701a50b76789539b989971570f76eceb9f

Observation a1be0072-2efd-47be-8c2d-b37fd6a6c9bc · outbound

This paper cites score": 3,.

BALSAM: A Platform for Benchmarking Arabic Large Language Models score": 3,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T11:36:00.410565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:36:00.280571Z digest=sha256:2e3514a6670dcd8e93a503022f7d2557b1915a057215d04398675acdcc5f6dbe

Observation 32a637dc-4852-47fd-a38b-a61e37a53dee · outbound

This paper cites In International Conference on Learning Representations.

BALSAM: A Platform for Benchmarking Arabic Large Language Models In International Conference on Learning Representations

Reference 2021

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T11:36:00.521004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:36:00.233976Z digest=sha256:7b338a799c52657a6b30ff48f9553cf4d9a212a37f00193903ff5fd3baf7cbc0

Observation 9c55c0a3-d89c-49d2-82b3-628f3da1c94d · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-06T11:36:00.215303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:36:00.215303Z digest=sha256:0d91c3ee8bca99a33c6040fb93b8b894ea1fd24301bc20817cb947de48699636

Observation bd0921df-52e8-440c-92f8-00ef16da9de4 · outbound

This paper cites Jais and Jais-chat: Arabic-Centric Foundation and Instruction-Tuned Open Generative Large Language Models.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Jais and Jais-chat: Arabic-Centric Foundation and Instruction-Tuned Open Generative Large Language Models

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T11:36:00.250729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:36:00.250729Z digest=sha256:65aeb28bad7b2d8ae7e22442c123b32962439d748236b8878514a93f6311c6b6

Observation d3f97aaa-4e76-433a-8eda-725ae6fffd57 · outbound

This paper cites Creating Arabic LLM Prompts at Scale.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Creating Arabic LLM Prompts at Scale

Reference 2024

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T11:36:00.366212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:36:00.222651Z digest=sha256:960d8c41bb22d96249907937f90610c8d7faea3380420b6ee528612587a9f0ab

Observation 06ffc0bd-6a50-4701-a2e3-efbf85ba8cfe · outbound

This paper cites In Proceedings of the 31st International Conference on Computational Lin- guistics, pages 4186–4218, Abu Dhabi, UAE.

BALSAM: A Platform for Benchmarking Arabic Large Language Models In Proceedings of the 31st International Conference on Computational Lin- guistics, pages 4186–4218, Abu Dhabi, UAE

Reference 2025

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T11:36:00.500643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:36:00.245539Z digest=sha256:b4de85c4e09279b30ce7b671ee94e3f20fb7f6537b91ddf48fd5991bb4064bb3

Pith citing papers

Observation d292673a-38ce-4876-b9e2-315e12e1b824 · inbound

Context-Aware Dialectal Arabic Machine Translation with Interactive Region and Register Selection cites this paper.

Context-Aware Dialectal Arabic Machine Translation with Interactive Region and Register Selection BALSAM: A Platform for Benchmarking Arabic Large Language Models

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:30:52.671051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T19:03:03.350371Z digest=sha256:40a243967dc05afb549dedcdbe31bd0dcfbbc3c98f5d3b3f6b33a369bc2a1c25