Pith. sign in

Paper Citation Record · LEDGER

PL-Guard: Benchmarking Language Model Safety for Polish

As of 7 August 2026, this Paper Citation Record lists 27 of 27 outbound references and 0 inbound Pith citation observations for arXiv:2506.16322.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.16322 v1

Coverage vector

measured 27 of 27 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:49:54.958475Z

measured 27 of 27 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

27 of 27 outbound references displayed

  • verified exact0
  • verified fuzzy8
  • unresolved14
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch5

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d563ce13-f171-458d-8fca-4797387eb0e7 · outbound

This paper cites AEGIS: Online Adaptive AI Content Safety Moderation with Ensemble of LLM Experts.

PL-Guard: Benchmarking Language Model Safety for Polish AEGIS: Online Adaptive AI Content Safety Moderation with Ensemble of LLM Experts

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:52.956697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:52.956697Z digest=sha256:ceb6bc49ec35dec32f2bcd51222a9c07fcaa79c6cd31fa5c03d535fd1b79be7b

Observation f6f3ef22-2a04-45be-898f-1198ea6710b9 · outbound

This paper cites Evaluation of Few-Shot Learning for Classification Tasks in the Polish Language.

PL-Guard: Benchmarking Language Model Safety for Polish Evaluation of Few-Shot Learning for Classification Tasks in the Polish Language

Reference 6

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T23:49:55.996013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:49:53.054733Z digest=sha256:903caf41580874c4d696819eaf0d046cded6c6325a067a54f1d4a0feb900672f

Observation d7f0966d-2071-457b-b502-d5dff6ebc6b8 · outbound

This paper cites WildGuard: Open One-Stop Moderation Tools for Safety Risks, Jailbreaks, and Refusals of LLMs.

PL-Guard: Benchmarking Language Model Safety for Polish WildGuard: Open One-Stop Moderation Tools for Safety Risks, Jailbreaks, and Refusals of LLMs

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:53.145378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:53.145378Z digest=sha256:2b924ff30c339366c7ae5e17e3c05688f9f24ba1f874dcc91a9aed60852cc1c1

Observation 1683a507-e3d5-4e7f-b7a0-19a007254198 · outbound

This paper cites Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations.

PL-Guard: Benchmarking Language Model Safety for Polish Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:53.313221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:53.313221Z digest=sha256:09a92fe23629ef4b4aa68e4698d1cf2ec91b1da26827e0c5f24fd6a9640343b5

Observation 2822e7cd-a2f2-4041-8e20-9aedce5dc085 · outbound

This paper cites LLMzSz{\L}: a comprehensive LLM benchmark for Polish.

PL-Guard: Benchmarking Language Model Safety for Polish LLMzSz{\L}: a comprehensive LLM benchmark for Polish

Reference 10

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T23:49:55.736284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:49:53.402804Z digest=sha256:8d85a42fc86052e0b501a39a94614df1a3ff5becd7e92153188a4dd8a126eb52

Observation e75931c0-7e22-401a-b751-997bf2ec2ae2 · outbound

This paper cites Mixtral of Experts.

PL-Guard: Benchmarking Language Model Safety for Polish Mixtral of Experts

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:53.469273Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:53.469273Z digest=sha256:b8f4e4ea7fd0607fad7a5b3466bd385fff2db2d1e5e1c15c3c68237e803e24b5

Observation fd0ab551-9a12-4305-9a11-2835994141ca · outbound

This paper cites Towards Safe Multilingual Frontier AI.

PL-Guard: Benchmarking Language Model Safety for Polish Towards Safe Multilingual Frontier AI

Reference 12

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T23:49:55.502792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:49:53.562541Z digest=sha256:7b83df60c5f98e871852aaea1c560af57922b3a5d567a3f8896c4dc54a59d061

Observation 862b81b8-59a1-4b90-a768-bc3031e8afc7 · outbound

This paper cites pl web service.

PL-Guard: Benchmarking Language Model Safety for Polish pl web service

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:49:57.534778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:49:53.735948Z digest=sha256:eea68ac98980a45145d9926d8aa54b7606b347b3f7cd3be92b83a860ccaa1a0b

Observation 2a6911cd-3f3e-4831-868b-664a65d8abd8 · outbound

This paper cites MultiSlav: Using Cross-Lingual Knowledge Transfer to Combat the Curse of Multilinguality.

PL-Guard: Benchmarking Language Model Safety for Polish MultiSlav: Using Cross-Lingual Knowledge Transfer to Combat the Curse of Multilinguality

Reference 15

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T23:49:55.310297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:49:53.804783Z digest=sha256:1e5ca66c9f42e09f320d7e266f25f01441fc702d79ba92aa335845e881957e74

Observation cb8ca907-6a43-48bb-be6f-871615814c1e · outbound

This paper cites In Proceedings of the 5th Workshop on Trustworthy NLP (TrustNLP 2025), pages 155–165, Albuquerque, New Mexico.

PL-Guard: Benchmarking Language Model Safety for Polish In Proceedings of the 5th Workshop on Trustworthy NLP (TrustNLP 2025), pages 155–165, Albuquerque, New Mexico

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:49:57.182539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:49:53.894659Z digest=sha256:21d53713de8be2bc2e9f20ca6038abb85c4557dbbe9aa5c0f897eed9ec1d0944

Observation 37e15ba6-6626-4d86-a81c-8bdbc38eb222 · outbound

This paper cites PolyGuard: A Multilingual Safety Moderation Tool for 17 Languages.

PL-Guard: Benchmarking Language Model Safety for Polish PolyGuard: A Multilingual Safety Moderation Tool for 17 Languages

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:53.979957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:53.979957Z digest=sha256:f97f7fb7473404eddbfe7209427c931ea95ab202bd320c46fe7002b30203f102

Observation 2fea2721-5fb1-469d-9151-7ea7b3a47127 · outbound

This paper cites The Llama 3 Herd of Models.

PL-Guard: Benchmarking Language Model Safety for Polish The Llama 3 Herd of Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:54.053004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:54.053004Z digest=sha256:98f868a70a264320294dc16783618f85026fa5e37e18a6e94b71655136c493c3

Observation 4d20f9c6-99ba-41cf-ab72-fa054af8f63b · outbound

This paper cites arXiv preprint arXiv:2410.18565.

PL-Guard: Benchmarking Language Model Safety for Polish arXiv preprint arXiv:2410.18565

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:54.124021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:54.124021Z digest=sha256:6230c46092aa6e5d55565040b22a9867eb8b73abc50084eeda5b35c7f29e1cfa

Observation 74c45ba4-24c6-4356-9b84-50d65f7b0c21 · outbound

This paper cites In Proceed- ings of the 2022 Conference on Empirical Methods in Natural Language Processing, pages 3419–3448, Abu Dhabi, United Arab Emirates.

PL-Guard: Benchmarking Language Model Safety for Polish In Proceed- ings of the 2022 Conference on Empirical Methods in Natural Language Processing, pages 3419–3448, Abu Dhabi, United Arab Emirates

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:49:56.913132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:49:54.237682Z digest=sha256:775385712bcdd101494d6a0f677e99522f3f076df88bff8559c911b993a373a9

Observation bbdbab36-2dbc-476a-96c5-d739514d512a · outbound

This paper cites PL-MTEB: Polish Massive Text Embedding Benchmark.

PL-Guard: Benchmarking Language Model Safety for Polish PL-MTEB: Polish Massive Text Embedding Benchmark

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:54.320650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:54.320650Z digest=sha256:fa2fbbd9a538062583b6787d71de41cd85455e394f45686890dd13c1f542e02d

Observation fc5139f5-97b7-46bd-bf94-7d162cdc300e · outbound

This paper cites KLEJ: Comprehensive Benchmark for Polish Language Understanding.

PL-Guard: Benchmarking Language Model Safety for Polish KLEJ: Comprehensive Benchmark for Polish Language Understanding

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:54.408495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:54.408495Z digest=sha256:2ec43b0f28e50b7d53d6415d9ba600902c0d23cf2d842f74e480f9e0f7c66c0a

Observation 1cb4159f-6176-4025-95d0-185f464b418e · outbound

This paper cites Rainbow Teaming: Open-Ended Generation of Diverse Adversarial Prompts.

PL-Guard: Benchmarking Language Model Safety for Polish Rainbow Teaming: Open-Ended Generation of Diverse Adversarial Prompts

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:54.533216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:54.533216Z digest=sha256:f5474fd1cd9a40ed7cc0bd16a8d937b3eeefc9e425bbdd7850cd54ebb402b006

Observation 230949b0-5aea-4608-86cc-662ff0aed381 · outbound

This paper cites Accessed: 2024-12-15.

PL-Guard: Benchmarking Language Model Safety for Polish Accessed: 2024-12-15

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:49:56.678025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:49:54.638486Z digest=sha256:55ee4f37967d7169a67dc5641fb17c1f65b94fe5389fe719e7070fe573f60fd5

Observation 5a4dfae1-1e72-4b6d-9d1e-f4eaccb58eee · outbound

This paper cites In Pro- ceedings of the 2024 Joint International Conference on Computational Linguistics, Language Resources and Evaluation (LREC-COLING 2024), pages 12574– 12584.

PL-Guard: Benchmarking Language Model Safety for Polish In Pro- ceedings of the 2024 Joint International Conference on Computational Linguistics, Language Resources and Evaluation (LREC-COLING 2024), pages 12574– 12584

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:49:56.446772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:49:54.744608Z digest=sha256:30229b9af1a1b798b509dfc8c2daa8d4f993f5b1b3fb59b9478e77f169eb3e9d

Observation 939b3006-0f26-4958-8805-fe088265ce9d · outbound

This paper cites ShieldGemma: Generative AI Content Moderation Based on Gemma.

PL-Guard: Benchmarking Language Model Safety for Polish ShieldGemma: Generative AI Content Moderation Based on Gemma

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:54.836720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:54.836720Z digest=sha256:86f303fed61fc28ea4361f2f7c696abb87ba0483d11535bf8f54fd68b5aec919

Observation 5152cbc0-bd5b-4048-866a-a5f5ef1b9b59 · outbound

This paper cites SafetyBench: Evaluating the Safety of Large Language Models.

PL-Guard: Benchmarking Language Model Safety for Polish SafetyBench: Evaluating the Safety of Large Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:54.958475Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:54.958475Z digest=sha256:b66538eb30ac6a23f741de611608904cc947147ee0a21456ad0295614c25a573

Observation b6f9d19e-2201-4150-a598-12790ce4e260 · outbound

This paper cites Anna Kolos, Inez Okulska, Kinga Gł ˛ abi´nska, Agnieszka Karli´nska, Emilia Wi´snios, Paweł Ellerik, and An- drzej Prałat.

PL-Guard: Benchmarking Language Model Safety for Polish Anna Kolos, Inez Okulska, Kinga Gł ˛ abi´nska, Agnieszka Karli´nska, Emilia Wi´snios, Paweł Ellerik, and An- drzej Prałat

Reference 2019

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:49:57.752433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:49:53.629472Z digest=sha256:7889343cc1b4b5cc5382323625fa3894cffbef364f24d618325c8c8c0b4aa1ea

Observation ae4d56f3-acaa-416a-b29e-d1be1d4f0b0d · outbound

This paper cites In Findings of the Association for Computational Linguistics: EMNLP 2020 , pages 3356–3369, Online.

PL-Guard: Benchmarking Language Model Safety for Polish In Findings of the Association for Computational Linguistics: EMNLP 2020 , pages 3356–3369, Online

Reference 2020

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:49:58.034194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:49:52.915554Z digest=sha256:11aabf57d83293db2a125781cc364cd5facbeb33b35de12d4d972afa58ab134d

Observation 75904f59-fdc4-4187-9fda-2f3695cacbb4 · outbound

This paper cites ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection.

PL-Guard: Benchmarking Language Model Safety for Polish ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:53.215936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:53.215936Z digest=sha256:1e362d6ad7674b6ccf290c6c28935ff48b6f4d500ad47eb0b504b9e452e62ff9

Observation 2e1182d2-c42d-48c9-8d81-b7028e4a16c1 · outbound

This paper cites Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment.

PL-Guard: Benchmarking Language Model Safety for Polish Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:52.753276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:52.753276Z digest=sha256:6af3ba2901e249b3d49f8b4e34fda54a3c18f46893fab82d0e006d9ab5c072b4

Observation 53cd5c61-d416-4d6e-9678-b0252c4839bb · outbound

This paper cites Accessed: 2024-12-15.

PL-Guard: Benchmarking Language Model Safety for Polish Accessed: 2024-12-15

Reference 2024

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:49:58.275088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:49:52.654311Z digest=sha256:2f1bc2fc05b3d88fb0fa59f8d13dc783a7d54167dc11ccbc5bb9a983081b3c1d

Observation 7daf2794-dfba-4877-a328-33cb345afc63 · outbound

This paper cites Evaluating LLMs Robustness in Less Resourced Languages with Proxy Models.

PL-Guard: Benchmarking Language Model Safety for Polish Evaluating LLMs Robustness in Less Resourced Languages with Proxy Models

Reference 2025

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T23:49:56.218135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:49:52.821536Z digest=sha256:20af2d130cb3945fe24826590965155e2c8de319bbb7ab8c132723aaadae9c40

Pith citing papers

No inbound Pith citation observations are available.