Pith. sign in

Paper Citation Record · LEDGER

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers

As of 21 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 0 inbound Pith citation observations for arXiv:2607.21010.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.21010 v1

Coverage vector

measured 54 of 54 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T08:46:17.594637Z

measured 54 of 54 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

54 of 54 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved54
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ba862811-772a-4f4e-b97d-8401b36248bd · outbound

This paper cites A comprehensive survey on legal summarization: Challenges and future directions.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers A comprehensive survey on legal summarization: Challenges and future directions

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:12.710693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:12.710693Z digest=sha256:6233289578ee53f9ec28dc8a1bbff1e9965bfd0a20acc9e4ae631b25a37dd6ce

Observation 4517b5ee-256d-4809-98e2-046cbb1cefde · outbound

This paper cites Large language models robustness against perturbation: S.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Large language models robustness against perturbation: S

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:12.775600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:12.775600Z digest=sha256:4e96a96a2fed4619fb41cbe10c4b843fac224a1889ddf4adbe22801047e9f4d6

Observation 79f977dd-d613-41a7-929e-e0a1eac5cea5 · outbound

This paper cites Using llm (large language model) to improve efficiency in literature review for undergraduate research.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Using llm (large language model) to improve efficiency in literature review for undergraduate research

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:12.857316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:12.857316Z digest=sha256:0f1af5a52956c83ec69ab91e8f90164e5369f1efaea6e16cf0c9e45a8bfcc8b4

Observation 4c16b665-e324-4179-a1d2-9e266f785d28 · outbound

This paper cites Robust tests for the equality of variances.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Robust tests for the equality of variances

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:12.918614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:12.918614Z digest=sha256:bcf02dd2de0493c44d6456c80f8162812a01e4f81d391539e9ab104a2b198695

Observation 7d7d2ada-1917-4f51-a68e-cdf0e061cb7b · outbound

This paper cites Efficient inference for noisy llm-as-a-judge evaluation.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Efficient inference for noisy llm-as-a-judge evaluation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:13.001406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:13.001406Z digest=sha256:7c6d9c105b1da8d5d284b31c79a4b88c194b7a75985d34c11179b6cbf76d6e58

Observation 5d7adfc0-191e-41c0-84f9-cc58b96d029e · outbound

This paper cites an unresolved cited work.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:13.056939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:13.056939Z digest=sha256:1e1e7508e9cc00d502ebdce3530488b680020068bf23ccf2257fe8f0b4ba8589

Observation 419922fc-7cf8-43b4-83f8-35130c2586e0 · outbound

This paper cites Statistical comparisons of classifiers over multiple data sets.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Statistical comparisons of classifiers over multiple data sets

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:13.137024Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:13.137024Z digest=sha256:f9b3a949f58856c518bf511160ef9a535b960697069c65ebf1d9371351eb78c7

Observation 35c44920-4697-49da-99a8-8f8c807b6887 · outbound

This paper cites Applicability of large language models and generative models for legal case judgement summarization.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Applicability of large language models and generative models for legal case judgement summarization

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:13.242928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:13.242928Z digest=sha256:58ccfcd3f6285b118263a2a98567e06374a9520daa99cd0db1bd05a0dfd65255

Observation 34c92790-8a86-4ac3-94c8-19b2a16735ec · outbound

This paper cites Explainability meets text summarization: A survey, in: Proceedings of the 17th International Natural Language Generation Conference, pp.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Explainability meets text summarization: A survey, in: Proceedings of the 17th International Natural Language Generation Conference, pp

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:13.328803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:13.328803Z digest=sha256:3007bfc0a5bd352343b1375f666a062f58629f465dbe9540ef366fe6de46d924

Observation cabed90e-9dca-40e7-9d8c-2ea4f23e8236 · outbound

This paper cites Consistency Evaluation of News Article Summaries Generated by Large (and Small) Language Models.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Consistency Evaluation of News Article Summaries Generated by Large (and Small) Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:13.384716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:13.384716Z digest=sha256:d3ef038bca58b7c18e7ffbde44659a53eccb0f020936b625db890105775d4490

Observation 809600eb-c6dd-4c0b-8520-469eb2391571 · outbound

This paper cites an unresolved cited work.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:13.464978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:13.464978Z digest=sha256:104c26feb7ebf7fa8b6e23250ec558fcd3fa97513fd6cc91b7a69fbeff15a655

Observation 5fb3496f-d1c1-427d-8a82-5f6890f80ebc · outbound

This paper cites an unresolved cited work.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:13.526027Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:13.526027Z digest=sha256:03bc6adfa856564a98df549b26919241089be73c71fa84cce0c0a264b59af800

Observation aba51b5b-1f25-4649-91a3-294785f454d0 · outbound

This paper cites TrustLLM: Trustworthiness in Large Language Models.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers TrustLLM: Trustworthiness in Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:13.605521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:13.605521Z digest=sha256:4ba8046384cd63daeb8296ca3ae870c5ddb3848878e21ce1cbcae0adfb73c733

Observation d060040c-c398-43a2-bf95-9d03be857d02 · outbound

This paper cites an unresolved cited work.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:13.655817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:13.655817Z digest=sha256:08adbc3781f48bdf2669a81eb1e0fc277a49a028ce3de2965d8a5ced2c10aceb

Observation 95f74450-44a6-4caf-a246-cc60b71ae1ab · outbound

This paper cites Consistency analysis of chatgpt, in: Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing, pp.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Consistency analysis of chatgpt, in: Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing, pp

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:13.743863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:13.743863Z digest=sha256:ea06f35dd9cc28d091ac0890f3cc5f6c73945ac610ceb79469d11feba33afcca

Observation c3386394-0a16-4ba4-85ab-2f6af86a7628 · outbound

This paper cites Context-Aware Sports Highlight Generation Leveraging Large Language Models.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Context-Aware Sports Highlight Generation Leveraging Large Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:13.826062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:13.826062Z digest=sha256:50a180ea590f86bca473d09e272214b3738d97078e523c733d1b295c98a6b40f

Observation 7bd3b05b-7ab4-400f-8a84-06ac5f5d09a7 · outbound

This paper cites an unresolved cited work.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:13.908595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:13.908595Z digest=sha256:f8deaa974c2bfd89016c96521ad2798b41480f2026f56dba0fa40da62a451cd6

Observation 126b27e8-6410-4740-b003-af20fc2c6b6e · outbound

This paper cites Llms cannot reliably judge (yet?): A comprehensive assessment on the robustness of llm-as-a-judge.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Llms cannot reliably judge (yet?): A comprehensive assessment on the robustness of llm-as-a-judge

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:13.964911Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:13.964911Z digest=sha256:7baeea0bf23efad07894b01502c5d6cc1b7f6bc1a64f1e823a42ff02f20d78f4

Observation e9df35f9-1f64-44ac-add0-315d9a909051 · outbound

This paper cites Do not abstain! identify and solve the uncertainty, in: Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (V olume 1: Long Papers), pp.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Do not abstain! identify and solve the uncertainty, in: Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (V olume 1: Long Papers), pp

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:14.045044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:14.045044Z digest=sha256:6eb56adf00b0a74353d23fac72fa62a5d91c28b270e75de398bed4a2cf9790e5

Observation 4b6ac1fc-ceec-4945-83a1-0f0c3c5671c6 · outbound

This paper cites Sumsurvey: An abstractive dataset of scientific survey papers for long document summarization, in: Findings of the ACL 2024, pp.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Sumsurvey: An abstractive dataset of scientific survey papers for long document summarization, in: Findings of the ACL 2024, pp

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:14.129595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:14.129595Z digest=sha256:4809126bd37e48066e5f141dda93f428eacab317168412194514733a03629236

Observation 343ff456-85ca-420d-abde-9d4959a2f7d4 · outbound

This paper cites Low-resource court judgment summarization for common law systems.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Low-resource court judgment summarization for common law systems

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:14.214472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:14.214472Z digest=sha256:5bc5ecb1e9619f46e6a635d02bebc007f2a55a11c12d1bd898b6dba378d09ea7

Observation e3b93c37-3db2-4232-a37d-e3c3a644c50d · outbound

This paper cites an unresolved cited work.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:14.297541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:14.297541Z digest=sha256:aa743cf45674a7bd148437a0debd37257088e1287ce9e7ce8e60ee6ad72e30fd

Observation 10f62dba-8f8e-4273-a3e8-f9a326ee90f8 · outbound

This paper cites Tools in the Loop: Quantifying Uncertainty of LLM Question Answering Systems That Use Tools.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Tools in the Loop: Quantifying Uncertainty of LLM Question Answering Systems That Use Tools

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:14.378320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:14.378320Z digest=sha256:04bed782b9517d527aa7295c5c31f91ce88d4c65466370fbbbe2f373ddd68007

Observation f824fb54-7766-4edd-aa89-b918c519da33 · outbound

This paper cites Effectiveness in retrieving legal precedents: exploring text summarization and cutting-edge language models toward a cost-efficient approach.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Effectiveness in retrieving legal precedents: exploring text summarization and cutting-edge language models toward a cost-efficient approach

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:14.438296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:14.438296Z digest=sha256:119e4b3311fb1b74218ff5dcc0c844076f12fd56c5c8de69cba241dac35a6c33

Observation 06d89a72-c36d-4ab3-972a-98e481ec4533 · outbound

This paper cites State of what art? a call for multi-prompt llm evaluation.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers State of what art? a call for multi-prompt llm evaluation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:14.496388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:14.496388Z digest=sha256:600a9374c416b91ace96414a4e404f1f9e4efe80810e4b811610ac962d839423

Observation b0193b8d-9fdb-48fd-922e-c76541e7a415 · outbound

This paper cites Abstractive text summarization using sequence-to-sequence rnns and beyond, in: Proceedings of the 20th SIGNLL conference on computational natural language learning, pp.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Abstractive text summarization using sequence-to-sequence rnns and beyond, in: Proceedings of the 20th SIGNLL conference on computational natural language learning, pp

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:14.547936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:14.547936Z digest=sha256:cf9ea0d61d3252abcba6cfcc5f2ffb2a14a08cb2fa811b8fd7c6094e2e4c9ac5

Observation 202d3509-ebfc-412a-9779-072afd236e84 · outbound

This paper cites Evaluating Variance in Visual Question Answering Benchmarks.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Evaluating Variance in Visual Question Answering Benchmarks

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:14.599132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:14.599132Z digest=sha256:df6b8138cc229f47bdc328ef294b3137932942713fdfe6cff5ee29a76689ae5a

Observation 974951f5-4266-4556-ab26-7c495e28da2b · outbound

This paper cites an unresolved cited work.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:14.681971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:14.681971Z digest=sha256:1383729f581eb29424157cc9d507cfe46820540ce98820bd30990fdb273fa8b6

Observation 2f8529f4-8b82-46e4-a21f-bd02679036bd · outbound

This paper cites Efficient multi-prompt evaluation of llms, in: Proceedings of the 38th International Conference on Neural Information Processing Systems, pp.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Efficient multi-prompt evaluation of llms, in: Proceedings of the 38th International Conference on Neural Information Processing Systems, pp

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:14.757339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:14.757339Z digest=sha256:84cab48b7b5042aa1f726c02b86726cce468c0a365e858e47c041c3e4a1ef1c8

Observation 7cc65411-dfe0-4a9b-ba71-48ab1b18503e · outbound

This paper cites an unresolved cited work.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:14.842281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:14.842281Z digest=sha256:3ced6672b40672ba2f1cede58a37c77c141f465413b8ecffab972c9013d65da5

Observation e60734a2-dc61-4c68-952b-2471bbe58cd7 · outbound

This paper cites Leveraging large language models on the traditional scientific writing workflow, in: 2024 Conference on AI, Science, Engineering, and Technology (AIxSET), IEEE.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Leveraging large language models on the traditional scientific writing workflow, in: 2024 Conference on AI, Science, Engineering, and Technology (AIxSET), IEEE

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:14.903488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:14.903488Z digest=sha256:7f4d0c4507ed6e9b259b9936aa2b8bab3ad8b0a1d092753a3b0639eb3fedc0b5

Observation 7ae50f9e-85e8-4705-b520-83a86618b120 · outbound

This paper cites an unresolved cited work.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:14.953010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:14.953010Z digest=sha256:fdcf877be97df18036edb7d40d85712992334f9cd9bfe5c783d7b3b752c0569f

Observation 16d23be8-631b-4a6b-8395-42dc8c138b97 · outbound

This paper cites How resilient are language models to text perturbations?, in: International Conference on Intelligent Data Engineering and Automated Learning, Springer.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers How resilient are language models to text perturbations?, in: International Conference on Intelligent Data Engineering and Automated Learning, Springer

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:15.030965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:15.030965Z digest=sha256:08389f1ac18cbdfbf6b6da9e00fbcf6383dcd31ec3fff49d20b7a119bbf12c75

Observation 7de55c08-2825-4751-9ff4-6e4fcd5c413a · outbound

This paper cites Can You Trust LLM Judgments? Reliability of LLM-as-a-Judge.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Can You Trust LLM Judgments? Reliability of LLM-as-a-Judge

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:15.119758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:15.119758Z digest=sha256:927b1b9744a4285f5cef76a5a12257d11be0851d3a043c3a350953bad9af9d39

Observation 16003fb3-6a20-4e28-a272-1381f5281c10 · outbound

This paper cites A coin flip for safety: Llm judges fail to reliably measure adversarial robustness.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers A coin flip for safety: Llm judges fail to reliably measure adversarial robustness

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:15.195403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:15.195403Z digest=sha256:9490749945978635fd8cd6417ffcb28f0c0bf80aa4168df2e840d27fae24e4d6

Observation 2d827d14-b1c1-4bd3-928a-5498c3224a70 · outbound

This paper cites Quantifying Language Models' Sensitivity to Spurious Features in Prompt Design or: How I learned to start worrying about prompt formatting.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Quantifying Language Models' Sensitivity to Spurious Features in Prompt Design or: How I learned to start worrying about prompt formatting

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:15.357112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:15.357112Z digest=sha256:25904b850651f47e1211e4c77a143294f9ec47f8c08519d1d7e1b291a9a3c077

Observation b067281a-a2b1-40b2-8680-38783516dc41 · outbound

This paper cites an unresolved cited work.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:15.442442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:15.442442Z digest=sha256:15ea4eda2fe61d4e18f9e3d8236839cb82f192d21ef3ec397e5f746f275d80d7

Observation 03ef4978-01b3-4b3d-bb78-18fa5c883a29 · outbound

This paper cites Robustness of large language models to perturbations in text.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Robustness of large language models to perturbations in text

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:15.611011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:15.611011Z digest=sha256:d6ed5d598e6f9eb4b8dc31f88eea6173481984d963b969214fbd8301bee52bc3

Observation edb6d993-40e4-4486-88d7-22a441d7efb7 · outbound

This paper cites Legal text summarization via judicial syllogism with large language models.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Legal text summarization via judicial syllogism with large language models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:15.768217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:15.768217Z digest=sha256:857456fef4ebf3c73ec05bc163217d536cae2f07bf788824d194c1c469cd298f

Observation c3afb610-bf4a-44da-995a-d0367277be74 · outbound

This paper cites Artificial intelligence risk management framework (ai rmf 1.0).(2023).

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Artificial intelligence risk management framework (ai rmf 1.0).(2023)

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:15.873375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:15.873375Z digest=sha256:f46a254e19c93808fefc11b0614ca503902dcb922b4d80ac3e69c95be769ba24

Observation 3fe07b83-615a-413d-a7cb-e1c110969d97 · outbound

This paper cites Evaluating the factual consistency of large language models through news summarization, in: Findings of the Association for Computational Linguistics: ACL 2023, pp.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Evaluating the factual consistency of large language models through news summarization, in: Findings of the Association for Computational Linguistics: ACL 2023, pp

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:16.037489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:16.037489Z digest=sha256:6cbfbc1d5d914599860ef44e37be9372ab3eab85fab5c8e07c781838cc6f7d69

Observation aaf3ebbc-c2ec-49e6-a647-898ea236ded5 · outbound

This paper cites an unresolved cited work.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:16.098103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:16.098103Z digest=sha256:6d102a4e6c952524e1a75570072c364a9f9acbd2b5bafee3e2c3e43be91ef45c

Observation 43e5511e-d97d-4d8b-89df-5a7c1429c8dd · outbound

This paper cites Prompt engineering in consistency and reliability with the evidence-based guideline for llms.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Prompt engineering in consistency and reliability with the evidence-based guideline for llms

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:16.188424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:16.188424Z digest=sha256:caa2f038cbdf96310fa25dd70e33a43816fe14fb54d47188544b1a68c954add2

Observation a1ea5236-7a7f-46ce-9374-8b81b6d1da4c · outbound

This paper cites Using llm-supported lecture summarization system to improve knowledge recall and student satisfaction.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Using llm-supported lecture summarization system to improve knowledge recall and student satisfaction

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:16.271309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:16.271309Z digest=sha256:9e05fb9de70ad8af3e72f0b8ba1814c144d958dbc2f4a6295f9f1724472fd4ee

Observation ed5ebf97-8d16-4988-b3a8-d53c42bc02a2 · outbound

This paper cites Comparisons of various types of normality tests.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Comparisons of various types of normality tests

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:16.352441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:16.352441Z digest=sha256:1fe66c6edefc9357405ff18072cdc28897cdfb6412b2592a4f14befbda6a3a8f

Observation 4e76288c-eb55-4ac2-a35f-b22e8972999a · outbound

This paper cites Event-based evaluation of abstractive news summarization, in: Proceedings of the Fourth Workshop on Generation, Evaluation and Metrics (GEM2), pp.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Event-based evaluation of abstractive news summarization, in: Proceedings of the Fourth Workshop on Generation, Evaluation and Metrics (GEM2), pp

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:16.413745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:16.413745Z digest=sha256:e77d359f97e2e813859c8f66673645623515844dc13e8c6fcbe0926218169652

Observation 9e5fd6cd-414d-4bc3-bcf3-e01497129cb9 · outbound

This paper cites Alignscore: Evaluating factual consistency with a unified alignment function, in: Proceedings of the 61st Annual Meeting of the ACL (V olume 1: Long Papers), pp.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Alignscore: Evaluating factual consistency with a unified alignment function, in: Proceedings of the 61st Annual Meeting of the ACL (V olume 1: Long Papers), pp

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:16.565036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:16.565036Z digest=sha256:47abd4ef82dbe420d59a4d47b8044b9e08a42fb28a22785128e262338b105e5b

Observation d92b3102-aed1-4c81-9747-fbfc74587ea9 · outbound

This paper cites A systematic survey of text summarization: From statistical methods to large language models.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers A systematic survey of text summarization: From statistical methods to large language models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:16.726842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:16.726842Z digest=sha256:bd6f81233add4122f489e440af20641a3320c1421fd75c59315b2c5d6e05d41b

Observation 84908382-9703-49db-9ac0-85d915a637cf · outbound

This paper cites BERTScore: Evaluating Text Generation with BERT.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers BERTScore: Evaluating Text Generation with BERT

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:16.840476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:16.840476Z digest=sha256:ce426dbf6274801661892ef86dd6603fd59c687d0e4f8deacdfea4711923d718

Observation dc0fea92-c0bd-4775-8685-7f08ff666f63 · outbound

This paper cites Benchmarking large language models for news summarization.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Benchmarking large language models for news summarization

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:17.000413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:17.000413Z digest=sha256:90034f6d9d1511e07a1e7ed6e8871ac7f5acbe7bf6e6b8a3e4b3d09a1ea29536

Observation 8dbd9767-522a-4ddd-89e7-84bfdc68063f · outbound

This paper cites Trustworthy evaluation of large language models.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Trustworthy evaluation of large language models

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:17.115073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:17.115073Z digest=sha256:78526067d4d14ac12dbc7ebda47f839a5537fb2c5e17380707d77c6d678d9a51

Observation 5146d519-3232-4e64-a9ca-5721081d5061 · outbound

This paper cites A comprehensive survey on process-oriented automatic text summarization with exploration of llm-based methods.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers A comprehensive survey on process-oriented automatic text summarization with exploration of llm-based methods

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:17.278674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:17.278674Z digest=sha256:fca1e86967e729c5186da1715a418972974cea6968e7a9e88b8c9e2e26418a9b

Observation 451aec2d-0551-4685-a02b-cbde61db5322 · outbound

This paper cites Assessing the accuracy of artificial intelligence-generated clinical summaries from ambulatory glaucoma subspecialty clinical encounters.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Assessing the accuracy of artificial intelligence-generated clinical summaries from ambulatory glaucoma subspecialty clinical encounters

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:17.438254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:17.438254Z digest=sha256:65634e75f127961ded7371d92b76b5d8e8b966af44e91ed2be89af5f9a09a940

Observation b4f7f642-9430-4782-96af-eb8d3d58dd67 · outbound

This paper cites an unresolved cited work.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers Unresolved cited work

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:17.594637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:17.594637Z digest=sha256:eaf1c3c2ee7b6d64c15618360ff58a87524e2a43903e8ede2d822829166fe758

Pith citing papers

No inbound Pith citation observations are available.