Pith. sign in

Paper Citation Record · LEDGER

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives

As of 22 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 2 inbound Pith citation observations for arXiv:2412.10220.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.10220 v1

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T16:17:07.599252Z

measured 49 of 49 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-10T15:25:46.925378Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T10:36:02.611276Z

Reference resolution

47 of 47 outbound references displayed

  • verified exact3
  • verified fuzzy15
  • unresolved22
  • parse uncertain0
  • malformed identifier7
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 66e8265a-312d-4d50-bcfb-756138f8f99a · outbound

This paper cites an unresolved cited work.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-11T16:17:08.108551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T16:17:07.458997Z digest=sha256:f833ad72b7738f108c75f7afc0aad48694fcd9fe32020e1b80ae97495ece9597

Observation 9369f04f-3c22-4cf1-a56b-7868e9b8e011 · outbound

This paper cites an unresolved cited work.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-11T16:17:08.100788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T16:17:07.463646Z digest=sha256:1cd3853053b5be7120b26278eec1d821d0575f5a42f8b07625b13bd8a296ab62

Observation 100c07a3-073b-4e03-bfc9-2f220f0236f0 · outbound

This paper cites Flemish AI Research Program.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives Flemish AI Research Program

Reference 3

Resolution
malformed identifier
raw_fallback, observed 2026-08-11T16:17:07.953985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T16:17:07.467290Z digest=sha256:539ac448471851bce61636975889fb29830dc794d78f0b96e88cca5b88cf5645

Observation cd76681c-7d66-48e2-87c5-d59b7ea23faf · outbound

This paper cites Lundberg and Su-In Lee.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives Lundberg and Su-In Lee

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:17:08.093224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T16:17:07.471642Z digest=sha256:7cde03740898d87310e17c12465ead0ccceb6d9c7e2dcf4418da01cbfefeec8d

Observation 066475d5-bf48-49a3-b03e-2192637f8faf · outbound

This paper cites ”why should i trust you?”: Explain- ing the predictions of any classifier.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives ”why should i trust you?”: Explain- ing the predictions of any classifier

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T16:17:07.475368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:17:07.475368Z digest=sha256:5fdb17020f95f1eb008da36cdc186a1a310f699be36ce8c9740ef956ba7d41b2

Observation b73734ac-5de6-45a1-b424-00d5f4ae1e04 · outbound

This paper cites A value for n-person games.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives A value for n-person games

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T16:17:07.478794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:17:07.478794Z digest=sha256:3faa03e3e807f839ec3ec02b43865decb6177a00106561b729fdf20c02f599cf

Observation 9ced43c6-22ad-4380-9140-5644f9125d11 · outbound

This paper cites The inadequacy of shapley values for explainabil- ity, 2023.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives The inadequacy of shapley values for explainabil- ity, 2023

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:17:08.081653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T16:17:07.481445Z digest=sha256:946209edbb74e449c1496c33e2147251ad2a072c71fab3162190b6eb1d33f1fc

Observation 9ad79fb9-c267-417e-85d8-edbd5a8324eb · outbound

This paper cites Ex- plainability is not a game.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives Ex- plainability is not a game

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T16:17:07.484973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:17:07.484973Z digest=sha256:9bc8a0f8256b94af33dada37bc5d50c8c85acc4ed24f07ea068c10f84f75c32a

Observation 49413a5f-6702-4cb8-91a9-8198cb441cc1 · outbound

This paper cites Natural language explanations for machine learning classification decisions.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives Natural language explanations for machine learning classification decisions

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T16:17:07.487993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:17:07.487993Z digest=sha256:ef6182cd1d280b4401eee355f4e716977a5ab003eedff23d83d469709f99d963

Observation deceb318-f251-4882-a7e4-c8a7c4ff4d12 · outbound

This paper cites Tell Me a Story! Narrative-Driven XAI with Large Language Models.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives Tell Me a Story! Narrative-Driven XAI with Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T16:17:07.492051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:17:07.492051Z digest=sha256:67369c065d9c2678224f6f5700fdf55c0f76402fe3349a17389015d18f183a2c

Observation 524364df-b89a-4dfc-ac09-9230b5223d68 · outbound

This paper cites LLMs for XAI: Future Directions for Explaining Explanations.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives LLMs for XAI: Future Directions for Explaining Explanations

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T16:17:07.494861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:17:07.494861Z digest=sha256:71b0f2645449c8ae82f4cb8c3fe3c6d90c84f5b7eb26ee935319a8da8b72eb20

Observation 50f871d1-740f-4857-ad95-1f1bcdb6c489 · outbound

This paper cites Natural Language Counterfactual Explanations for Graphs Using Large Language Models.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives Natural Language Counterfactual Explanations for Graphs Using Large Language Models

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-08-11T16:17:07.743925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T16:17:07.498907Z digest=sha256:8211995199e485da3edd20ba045c92d8bdaef86f8ac2f6ab67190b7eb78ea359

Observation 8add35a7-d315-4c68-a1c4-4c6cc4f2ab2d · outbound

This paper cites GraphNarrator: Generating Textual Explanations for Graph Neural Networks.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives GraphNarrator: Generating Textual Explanations for Graph Neural Networks

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T16:17:07.502476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:17:07.502476Z digest=sha256:21cdaeb6728a0b504011ac1ab5ffb3a65353268bb7218329f8a1bc40abd9c338

Observation bcc5ea74-d2d2-4da0-9243-819377d8c44b · outbound

This paper cites GraphXAIN: Narratives to Explain Graph Neural Networks.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives GraphXAIN: Narratives to Explain Graph Neural Networks

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T16:17:07.506101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:17:07.506101Z digest=sha256:cb8fb9f400a9a4d4e5ebe729005bf6ff63c92908023a07b84feb7970c7eefe42

Observation 8d0861c0-51d3-4439-ac24-10a72a6582af · outbound

This paper cites Faithful and plausible natu- ral language explanations for image classifi- cation: A pipeline approach.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives Faithful and plausible natu- ral language explanations for image classifi- cation: A pipeline approach

Reference 15

Resolution
malformed identifier
no resolver link, observed 2026-08-11T16:17:07.508986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:17:07.508986Z digest=sha256:d9c4924b1c4fe6dc6767475acc223b1586fc89c5a48aabd3fdc59837c8fb2c32

Observation f65c85bd-728f-4e63-a4f9-bae8820dd64e · outbound

This paper cites In-Context Explainers: Harnessing LLMs for Explaining Black Box Models.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives In-Context Explainers: Harnessing LLMs for Explaining Black Box Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T16:17:07.512632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:17:07.512632Z digest=sha256:6af4c58b73d65429bd83575ce37ba03a76294306342553d6e6a9c3f63d935637

Observation 00ae3619-fd37-4e18-8a74-c1197510ebd9 · outbound

This paper cites Explaining ma- chine learning models with interactive natural language conversations using talktomodel.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives Explaining ma- chine learning models with interactive natural language conversations using talktomodel

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:17:08.074462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T16:17:07.516461Z digest=sha256:33599bd1e7138dd06c344805246a1d4eb304a6f8f672967e146fb29dbdd6d053

Observation 80d6ab0b-8b0c-4aef-875f-12ceea63da2d · outbound

This paper cites ME- TEOR: An automatic metric for MT evalu- ation with improved correlation with human judgments.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives ME- TEOR: An automatic metric for MT evalu- ation with improved correlation with human judgments

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:17:08.058600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T16:17:07.530835Z digest=sha256:a9d079b4afa1b859140d9981c877466c4d55c713906e92553f65bf2d0aaf8f8f

Observation 3460f186-b813-4886-88c3-85778c99a5fa · outbound

This paper cites Keane, Eoin M.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives Keane, Eoin M

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:17:08.066492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T16:17:07.521210Z digest=sha256:6a33f71c2741d0218c8952e3de5ac21ad074f8ee7e5766d08dbad31d5496e331

Observation 2fbded58-85d5-4a66-bd6a-bc8075889d01 · outbound

This paper cites Do models explain them- selves? Counterfactual simulatability of natural language explanations.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives Do models explain them- selves? Counterfactual simulatability of natural language explanations

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:17:08.050942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T16:17:07.536547Z digest=sha256:3db324b6c4572190745529d53a985560ecf9cc32e0ffc56358a78fa1b1beb8f1

Observation 70fdefc5-6b79-40cd-ac59-42c48798b14c · outbound

This paper cites an unresolved cited work.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives Unresolved cited work

Reference 21

Resolution
malformed identifier
no resolver link, observed 2026-08-11T16:17:07.525849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:17:07.525849Z digest=sha256:71d19d7bd9569d49cfa71a0f216099d881c46dcade3ecaa47abf56d0830a7040

Observation 88041425-3202-4a4e-b038-41cd32b37a42 · outbound

This paper cites BLEURT: Learning robust metrics for text generation.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives BLEURT: Learning robust metrics for text generation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T16:17:07.528002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:17:07.528002Z digest=sha256:bcaa77c09df1f01fa1a0725f75f6ddb13bb084df67ae7fe50e0c800690c65ae2

Observation 15b8aa88-629c-4543-ae6c-bbf0700e4d07 · outbound

This paper cites The disagreement problem in ex- plainable machine learning: A practitioner’s perspective.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives The disagreement problem in ex- plainable machine learning: A practitioner’s perspective

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:17:08.033311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T16:17:07.550045Z digest=sha256:fa1605ce1002913badaf5d4efd14aff76a61fef2f146ff28bf2b5485c425adb3

Observation 1403399b-4080-476a-9e3f-8fcbddee5ac0 · outbound

This paper cites F ActScore: Fine-grained atomic evaluation of factual precision in long form text generation.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives F ActScore: Fine-grained atomic evaluation of factual precision in long form text generation

Reference 24

Resolution
malformed identifier
no resolver link, observed 2026-08-11T16:17:07.533968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:17:07.533968Z digest=sha256:01b26351526b1b82d0c8143290eed2b50d47916578dfe3e26a3ec210dd6fb47e

Observation 6c835f72-4275-4d2e-b9b4-426295af1f3a · outbound

This paper cites PRobELM: Plausibility ranking evaluation for language models.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives PRobELM: Plausibility ranking evaluation for language models

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:17:08.022375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T16:17:07.556439Z digest=sha256:eab0f02fcdc45b0df604e715e3f938303b4e2a766fbf175a504e6457f7b2a3db

Observation 24136dbe-fe04-4b8d-a2a1-a58fbf74c0d6 · outbound

This paper cites A Survey on Natural Language Counterfactual Generation.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives A Survey on Natural Language Counterfactual Generation

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-08-11T16:17:07.693937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T16:17:07.558746Z digest=sha256:294494b3d9a9bb0f14e86e9a6ba8c43d13f53a7333e3a4430fa5f18d93948d50

Observation 60bd9c9f-41c1-4e4d-898b-e54c2fdf6162 · outbound

This paper cites You Can Generate It Again: Data-to-Text Generation with Verification and Correction Prompting.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives You Can Generate It Again: Data-to-Text Generation with Verification and Correction Prompting

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-08-11T16:17:07.709487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T16:17:07.543352Z digest=sha256:e89ad7886c58810f60c9d772632ed12481c404516b0c6f8efe6cf0e161ef3573

Observation e249ad58-702f-42e6-a1f1-85fb671a1de3 · outbound

This paper cites The Landscape of Emerging AI Agent Architectures for Reasoning, Planning, and Tool Calling: A Survey.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives The Landscape of Emerging AI Agent Architectures for Reasoning, Planning, and Tool Calling: A Survey

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T16:17:07.546973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:17:07.546973Z digest=sha256:d9af88c9d4ffe85a10be4953896d07001140f62a953148c3cfe322b2a5a5ba38

Observation 2e1eb895-4274-40c0-b4b4-964cfe3e99eb · outbound

This paper cites Sentence- bert: Sentence embeddings using siamese bert-networks.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives Sentence- bert: Sentence embeddings using siamese bert-networks

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:17:08.014238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T16:17:07.568624Z digest=sha256:43d386df5c476653cd3d17c669e62ee2688fb94623c10ce08e1225aa43643779

Observation d1972177-6ae1-4774-bb6c-5862a0a41aeb · outbound

This paper cites Towards few-shot fact-checking via perplexity.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives Towards few-shot fact-checking via perplexity

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T16:17:07.553636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:17:07.553636Z digest=sha256:1a746f9be9419bd55078f02c2084374611f0cefb22e6f6311acebd714ddfc791

Observation 72b38990-3eef-4e12-81ad-7f6e730e9a4a · outbound

This paper cites GPT-4 Technical Report.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives GPT-4 Technical Report

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T16:17:07.576150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:17:07.576150Z digest=sha256:d17ab9c80d6a7135183953fc79e48e9b11849fe7c7b3e6d21a1c1154180c60f3

Observation 34835850-e086-40d9-aeb6-43a911d25f3a · outbound

This paper cites Claude sonnet 3.5.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives Claude sonnet 3.5

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:17:07.999887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T16:17:07.580124Z digest=sha256:15d439e87d6feb9d37b50968b61d61c7ade3499bb28f78090fdb21e0358856aa

Observation df8d44f5-21e0-4902-95f6-bcf83eb50d43 · outbound

This paper cites Perplexity from PLM Is Unreliable for Evaluating Text Quality.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives Perplexity from PLM Is Unreliable for Evaluating Text Quality

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T16:17:07.562462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:17:07.562462Z digest=sha256:5c7c00982445c608b0c924fb0ee5eb89d04724de63817b274d7c1864b4147512

Observation da3b3d94-ac13-40f7-8456-cc48b14adc3f · outbound

This paper cites Efficient Estimation of Word Representations in Vector Space.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives Efficient Estimation of Word Representations in Vector Space

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-11T16:17:07.565772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:17:07.565772Z digest=sha256:8d03bfb85d26089ea374a396361d06752ad93b5bb37b6c24c77ece396bb06699

Observation 592a9244-b580-4898-aba6-188aca8c656f · outbound

This paper cites Mistral large 2.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives Mistral large 2

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:17:07.979102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T16:17:07.586417Z digest=sha256:3d7f11f87d32024517f9d01219a0149c7c452ac0d51aae4b2aaf85a6dad2bd31

Observation e04ac323-90c7-487e-94fc-275ee54e9ca6 · outbound

This paper cites Zhang, Mark Har- man, and Meng Wang.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives Zhang, Mark Har- man, and Meng Wang

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-11T16:17:07.592213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:17:07.592213Z digest=sha256:9380c77528075861ead09a1203e73dcdf1f8669ea31384b6f0471ed73bb913d7

Observation fe4b8619-b339-47a3-b923-d630d7b01e23 · outbound

This paper cites NV-Embed: Improved Techniques for Training LLMs as Generalist Embedding Models.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives NV-Embed: Improved Techniques for Training LLMs as Generalist Embedding Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T16:17:07.573421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:17:07.573421Z digest=sha256:aa72fb7e52e5b64fd815aa0ec8e5648301753233be96561564a37579334b8631

Observation 48258627-7921-4712-bed7-593597329f0d · outbound

This paper cites Adaptive chameleon or stubborn sloth: Revealing the behavior of large language models in knowledge conflicts.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives Adaptive chameleon or stubborn sloth: Revealing the behavior of large language models in knowledge conflicts

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:17:07.962454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T16:17:07.596922Z digest=sha256:4b492e71704052abd81f981eae7f1d9382f6933d83cd4f4d24300d35a987de70

Observation e4af75fb-0b5a-4fe1-a1ed-d181f3b854b7 · outbound

This paper cites Context-faithful prompting for large language models.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives Context-faithful prompting for large language models

Reference 39

Resolution
malformed identifier
no resolver link, observed 2026-08-11T16:17:07.599252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:17:07.599252Z digest=sha256:8b7b73a0e206fb1677892590279982ce11bbe02f56473e42a2f2450cc5f931e4

Observation 82bafdfe-e3be-47da-a95d-ef1fc45eebde · outbound

This paper cites Llama 3 model card.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives Llama 3 model card

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T16:17:07.582600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:17:07.582600Z digest=sha256:d5c959597c296e599cfa9a110a4861c620d4ab19bb3787d81c6db03c768ee874

Observation 8c428be9-4abb-4c61-a14a-ae9af0b4199e · outbound

This paper cites The llama 3 herd of mod- els, 2024.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives The llama 3 herd of mod- els, 2024

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:17:07.987051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T16:17:07.584489Z digest=sha256:67b54fadaced51178c25e0b46478e303ad3573079bc5ae962d8ec4817f50ea1f

Observation b975b56b-0f68-4a67-a197-23a7941a6e87 · outbound

This paper cites Entity-based knowledge con- flicts in question answering.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives Entity-based knowledge con- flicts in question answering

Reference 45

Resolution
malformed identifier
no resolver link, observed 2026-08-11T16:17:07.594559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:17:07.594559Z digest=sha256:7df2fbe3de31f6c083775d1227b48a078a51e37840ceb8665410057362c853db

Observation 532d0c77-1923-4217-964a-a52dad5963ae · outbound

This paper cites 19 org/CorpusID:201646309.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives 19 org/CorpusID:201646309

Reference 2019

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:17:08.006627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T16:17:07.571095Z digest=sha256:fc6cf3414559ae3a44116055ebfe894ade1f39eba4b0ffb142514a6a41e32cec

Observation 0c52d410-b766-4a0c-bc5c-1417b54f10fa · outbound

This paper cites URL https: //doi.org/10.24963/ijcai.2021/609.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives URL https: //doi.org/10.24963/ijcai.2021/609

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-11T16:17:07.523664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:17:07.523664Z digest=sha256:dfe983c92874563bcd5dd3203a1866d264799a65b336490a5cc0212443a86df8

Observation a7508e05-4ea2-42db-9c91-289addbcbf2d · outbound

This paper cites doi:10.1038/s42256- 023-00692-8.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives doi:10.1038/s42256- 023-00692-8

Reference 2023

Resolution
malformed identifier
no resolver link, observed 2026-08-11T16:17:07.518635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:17:07.518635Z digest=sha256:1ae5afcaecd61a54d86bcdb292c7e23d689e39003cfcd61abd7fd662407d9627

Observation 9a678ebe-0199-40d4-a588-da6812b0cc67 · outbound

This paper cites an unresolved cited work.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives Unresolved cited work

Reference 2024

Resolution
unresolved
raw_fallback, observed 2026-08-11T16:17:08.041654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T16:17:07.539365Z digest=sha256:994107d103ff74486ae62b1625bb4c639aa682fffe8bda9862236aeaeaff78f4

Observation ca98e4fd-328e-43dc-bcd2-a8cfd1cdd7b3 · outbound

This paper cites URL https://huggingface.co/ mistralai/Mistral-Large-Instruct-2407.

How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives URL https://huggingface.co/ mistralai/Mistral-Large-Instruct-2407

Reference 2407

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:17:07.971441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T16:17:07.589377Z digest=sha256:a991c8e3d1e2ae82b7dbad9c52775be2d2a51fd6551497a43f7b63f71de6e3ae

Pith citing papers

Observation 3ecd87ac-c046-4939-811b-e477401c76f1 · inbound

A Two-Stage LLM Framework for Accessible and Verified XAI Explanations cites this paper.

A Two-Stage LLM Framework for Accessible and Verified XAI Explanations How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:36:02.614307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T15:25:46.925378Z digest=sha256:94c6f7d12e2f287ad9e82845096eafd0d9d4d826bafd3a00f38336bff5d61c69

Observation 94a37fcc-18ec-4d54-b785-7b9f871ab5e1 · inbound

On the Importance and Evaluation of Narrativity in Natural Language AI Explanations cites this paper.

On the Importance and Evaluation of Narrativity in Natural Language AI Explanations How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:23:37.985155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T05:21:12.989059Z digest=sha256:39d4c20b02686aead1a7a93b6ebd1b869fdc8d3360de355a6914c7e560f7362f