Pith. sign in

Paper Citation Record · LEDGER

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop

As of 13 August 2026, this Paper Citation Record lists 100 of 118 outbound references and 0 inbound Pith citation observations for arXiv:2608.11171.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.11171 v1

Coverage vector

measured 100 of 118 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T04:51:17.259117Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 118 outbound references displayed

  • verified exact6
  • verified fuzzy12
  • unresolved82
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8d9ca6f8-ec40-4cd2-9a0e-f359d768f7d3 · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.630599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.630599Z digest=sha256:ac5cff9a563b66ddd3edd167da15fac0d122de5114cb88d0881f92b43354f449

Observation de088d98-e2f9-4e2d-af73-5700f1053308 · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.648681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.648681Z digest=sha256:a9cb5c7b99f7294b2d88f07bc6ec5eb5d2b27d25647be001ffb564d768f709b5

Observation bf5ef2da-7092-4ecf-aada-16d42922fd86 · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.667018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.667018Z digest=sha256:fbb82a5b45d28507e90e8706a9abeb40c6599eb7da413af370b950377428c2da

Observation c011e2b4-367e-4497-9fa6-fa23c58f332f · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.671317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.671317Z digest=sha256:beed08bba0c604002953b7aae64d18cb92f5c042c9a905cc04c079c7be3a8a38

Observation c8e8bca1-0a47-4526-81e7-a4ff31fb9f57 · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.685034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.685034Z digest=sha256:d9c58cae6188af26104f294268f84fd41ff72fa54470317b0a29c1a197202e91

Observation baac49ac-a54a-4be1-b053-273459137e61 · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.690286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.690286Z digest=sha256:46f0efa338fa9f16750ab66c3e66ee3b8353c36e7da9e790a566653a6e72e8ba

Observation c4ffb726-c5f2-42d2-b97f-597a473e3724 · outbound

This paper cites don't forget the teachers.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop don't forget the teachers

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.737630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.737630Z digest=sha256:5b1570c2597652b76d1354670dcfd547d7e66f102a25de4d0a71ba2d366785e8

Observation 6b16f8c3-3af4-417c-bddc-72aef843c6b5 · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.745842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.745842Z digest=sha256:a65c2632a2a89cbb2532b1eba40958120485e50f267c044a3ef1ec7b2a06e815

Observation 9870e611-b528-4ab9-9e31-792d137b9324 · outbound

This paper cites TrustLLM: Trustworthiness in Large Language Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop TrustLLM: Trustworthiness in Large Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.754194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.754194Z digest=sha256:88f876bdf1deb4071e350c8e6b3696b85732db3161007f690cef63dab8191ad0

Observation 46401178-93bd-4429-a47e-4e174c89416b · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.758736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.758736Z digest=sha256:6e0a7a4f3c7c36c9c4ab5d91c4e92df36f1ad54df6f9b31c90a65ce6403ed166

Observation 831c3195-f441-40e5-a3e8-22ca521a3feb · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.767502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.767502Z digest=sha256:ded59b108d2566c5ce123488d250005ff4de035ac8b0666f7020ad2922dded17

Observation 163ffdfc-8028-4fbd-a4a9-71c18d0b6783 · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.771478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.771478Z digest=sha256:57dfc3c6d2dab757786ef2db8b88cda7debba4ccff2959d114656c0c85501599

Observation 1a998e40-6912-4942-baef-7f9c23cbe17e · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.775652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.775652Z digest=sha256:8b5d95ba7538ba99098dc1d44aeb9c146bf10c1d2c95021c13823917d70a023e

Observation 645cca4e-9f50-4fb8-93ce-98e5cd658204 · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.811315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.811315Z digest=sha256:ffc624e28524fe0d87b1277fb404883b90dca147da7a534017f41e0225268e73

Observation f74c9f77-3b20-4494-981d-04f072dc2f58 · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.815380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.815380Z digest=sha256:20411473cf3f4e9924c21dde3d082d47d2d5bec5d4630dae441f2c3a14b786af

Observation b1e16acf-8f1e-42f5-a2a3-a5e70d8948fa · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.827962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.827962Z digest=sha256:408626c778d7d5cb6a9b1a7e15763bf1e9dec53694578335a5b00bdbe6167c89

Observation 6041a784-e8cb-42c8-aa8a-ad32cb84e29a · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.857656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.857656Z digest=sha256:588fcb27244c7cf40fe69ccfc239cb753e51d6ef310dd9d6fe7bb768f4520ee4

Observation 7cddb846-25b5-4780-9965-db36f7c79c4b · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.880222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.880222Z digest=sha256:890d0308b64e5bd2986220195bc2011425a5c28a6f0b4070afde41bccccd0832

Observation 68a32cee-7b52-4259-8f3a-8fe60d1019e6 · outbound

This paper cites DecodingTrust: A Comprehensive Assessment of Trustworthiness in GPT Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop DecodingTrust: A Comprehensive Assessment of Trustworthiness in GPT Models

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.888702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.888702Z digest=sha256:881b8c8b893fccec63bf2c0f1747988bd583e26daf462e7f25f0080428dfbdcb

Observation 02bafdd9-53fb-4a9a-a152-1238249fb910 · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.901804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.901804Z digest=sha256:82ed4537248687b56d4b4e466b25cec10080df139fa31ae9cf245da5fbcaf9e8

Observation 3c7f868c-804e-474f-8090-2f9f3b95bf1c · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.919064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.919064Z digest=sha256:60ba99280a1ebad119f6b3c5b63ae624ef9380852619bf84be4da66e4a3f4d53

Observation 0c8f1fc8-d6fe-4850-8274-79fed12bacdf · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.923408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.923408Z digest=sha256:4b660c205e2e94b13d900bf5f1c3a15dd6ece192599820bcf7a4b5a828d5c2cc

Observation db943417-ff45-4671-91b0-ed7794a7416a · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.927400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.927400Z digest=sha256:e806c68a1b398731217ffd1aca5ca0af30d35d831ca0c8f2c497dadefa4d062b

Observation c3b4fb2d-528f-4e0a-89e8-0382adcebc0b · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.931692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.931692Z digest=sha256:e27b3caf0c5fb254044a322ece79db1fcf5e2487512a264ab7bff8ac4b763263

Observation c479cfcd-9596-4125-8055-ad17f421356b · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.935849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.935849Z digest=sha256:049fb85ccb8f43ef191a8ed44cc0884dbc849768f6613efcfd55112f34aaca7d

Observation 84c9f417-3d2e-4107-bed1-25062866139e · outbound

This paper cites 2026 , howpublished =.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop 2026 , howpublished =

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.939878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.939878Z digest=sha256:d6307481611a34b93567fe8167e5d6014c2531a8310b7f74713c7d4a0ff2455e

Observation 26f9553c-e3d4-432a-a788-02e54cf8dc1e · outbound

This paper cites Human-Centered Explainable AI : Towards a Reflective Sociotechnical Approach.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Human-Centered Explainable AI : Towards a Reflective Sociotechnical Approach

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.944004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.944004Z digest=sha256:85c507c8a596fd2526e1b4136cd4ed02cc978986d07156ecd5c84a13255d8fcc

Observation 107975d5-4d78-4da4-8757-05817c127b83 · outbound

This paper cites and Wintersberger, Philipp and Manger, Carina and Hubig, Nina and Savage, Saiph and Weisz, Justin D.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop and Wintersberger, Philipp and Manger, Carina and Hubig, Nina and Savage, Saiph and Weisz, Justin D

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.948539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.948539Z digest=sha256:177b183f1a3b0125dc5e0e805861258f34efd6b334942a72333188e5d552614d

Observation 0e2ea7ae-a4f9-46c1-9066-b8b0cf35c35a · outbound

This paper cites A Survey on Medical Large Language Models: Technology, Application, Trustworthiness, and Future Directions.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop A Survey on Medical Large Language Models: Technology, Application, Trustworthiness, and Future Directions

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.952670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.952670Z digest=sha256:f9b35b59dedf95dc8dc7eb55fabb0365b9e87baf9ae68f3ff124079b7a694993

Observation da851883-7758-46cc-820b-29c1860d8b97 · outbound

This paper cites Standard Benchmarks Fail -- Auditing LLM Agents in Finance Must Prioritize Risk.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Standard Benchmarks Fail -- Auditing LLM Agents in Finance Must Prioritize Risk

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.956798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.956798Z digest=sha256:16bd73eb8220d6a0dc0946a003aac58fbdd9b6187eb05fe49492a0bd98e1618f

Observation 91bc9c9d-60ba-4cbf-9cf8-667c86e04d82 · outbound

This paper cites Don't Forget the Teachers.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Don't Forget the Teachers

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.960742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.960742Z digest=sha256:c9cb9eba02cf5538a6feda3fb89c87e3d25a35c6443ebd9d4e5a8eacd52991c5

Observation 913fa84b-6be8-4080-8c9f-6d436179cc43 · outbound

This paper cites 2025 Silicon Valley Cybersecurity Conference (SVCC) , pages=.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop 2025 Silicon Valley Cybersecurity Conference (SVCC) , pages=

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.965221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.965221Z digest=sha256:67085ff0923b4473884f6d0e9cb18dba7dcfdf42a6e016e49f2a3544a7454a24

Observation 3bd9c0e2-d80d-419f-a421-d48a1d9b36af · outbound

This paper cites 2023 , url =.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop 2023 , url =

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.969245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.969245Z digest=sha256:672bed75f18bd4f5438fd8271261859194c3d28b2177fc89c47a5a05123e47a1

Observation 4cf52372-36ab-4930-b5d2-9427461bb01c · outbound

This paper cites 2024 , url=.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop 2024 , url=

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.973582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.973582Z digest=sha256:1048ca1546e58d6b5a02f65fc4fa141dedbb5aa831fe2d4bfd731300d7f53786

Observation 428c3343-a31c-4b1a-a90f-2fc70ef11344 · outbound

This paper cites Trustworthy LLMs: a Survey and Guideline for Evaluating Large Language Models' Alignment.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Trustworthy LLMs: a Survey and Guideline for Evaluating Large Language Models' Alignment

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.977677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.977677Z digest=sha256:ea9e001e1fc80f2c9d70f7018bcab3ce9a5a70b25ba14526d537224248a46490

Observation 5efb6439-584c-4411-a97b-fb0b4b0a62ce · outbound

This paper cites 2025 , publisher=.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop 2025 , publisher=

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.981843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.981843Z digest=sha256:e1cb4045a67319d1ae44c4a01855de1a52a9182139a72216031e1d999ec57e81

Observation 498ceeaf-2846-448f-bc20-7fc603cd67ae · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.986141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.986141Z digest=sha256:6f9ee416377ea589d199d421ef0d75066556421a68a35bd7f12fed415b21802a

Observation 37a8edc6-c58f-404c-b0a7-82c03f5dcfb3 · outbound

This paper cites Interpretability Rules: Jointly Bootstrapping a Neural Relation Extractor with an Explanation Decoder.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Interpretability Rules: Jointly Bootstrapping a Neural Relation Extractor with an Explanation Decoder

Reference 85

Resolution
verified exact
doi, observed 2026-08-12T04:51:17.579945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:16.990589Z digest=sha256:ae198bd5c036565d72cbe97f863560812e9ba7fd0e7e456b62fc6b3fb4210cc2

Observation fa99828f-2299-4b74-8df4-932c0a7707fd · outbound

This paper cites Measuring Biases of Word Embeddings: What Similarity Measures and Descriptive Statistics to Use?.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Measuring Biases of Word Embeddings: What Similarity Measures and Descriptive Statistics to Use?

Reference 86

Resolution
verified exact
doi, observed 2026-08-12T04:51:17.897013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:16.994632Z digest=sha256:19120ba4defba46ecc93ce39c8ea3d07c7842f9e9b90e7ce24b785c02f472dc1

Observation 0d500bcb-aa3f-47bf-870e-b15566842501 · outbound

This paper cites and Kiritchenko, Svetlana and Balkir, Esma.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop and Kiritchenko, Svetlana and Balkir, Esma

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.999053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.999053Z digest=sha256:719a31dae85146d3b87c38a09cd1f01fb48e168524848331b412d1c227af349a

Observation 977ce2f1-f298-4342-b774-2ceae3df2885 · outbound

This paper cites GPT s Don ' t Keep Secrets: Searching for Backdoor Watermark Triggers in Autoregressive Language Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop GPT s Don ' t Keep Secrets: Searching for Backdoor Watermark Triggers in Autoregressive Language Models

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.003186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.003186Z digest=sha256:b8372fdc3161f9a3da4e9cc4bdd7536e0f0a6123810136daa209aafc72699343

Observation 855102a8-6674-4972-b483-eab017c5bbfb · outbound

This paper cites Reliability Check: An Analysis of GPT -3's Response to Sensitive Topics and Prompt Wording.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Reliability Check: An Analysis of GPT -3's Response to Sensitive Topics and Prompt Wording

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.007660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.007660Z digest=sha256:ab55530e8e71fb32718f13907bfddde5111cc8d92b29778c1c8b485f4d4f1817

Observation a2e7835e-f913-4a22-b938-6112a387e6b0 · outbound

This paper cites Driving Context into Text-to-Text Privatization.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Driving Context into Text-to-Text Privatization

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.011852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.011852Z digest=sha256:1e4d3735fe06853253c49dbe637f5649169dc15f02d36a6e703e87995ef9512b

Observation a569e7ad-2d73-44ca-8d2e-99b7c9e94b1e · outbound

This paper cites Expanding Scope: Adapting E nglish Adversarial Attacks to C hinese.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Expanding Scope: Adapting E nglish Adversarial Attacks to C hinese

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.016058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.016058Z digest=sha256:5b1707e19f6ee06f2e9afe466d0b9365acfaaea951efbfeb36238ac53ec5c17f

Observation 0b3994e2-3d6e-410b-8a35-77b802cde3ba · outbound

This paper cites Flatness-Aware Gradient Descent for Safe Conversational AI.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Flatness-Aware Gradient Descent for Safe Conversational AI

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.020464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.020464Z digest=sha256:2dc0ae3b5795040100f83b538a4ea228f268034b29fb014968e5b31edf137c75

Observation 8a824ce0-ecf5-4508-b1c1-5e2749353908 · outbound

This paper cites PBI -Attack: Prior-Guided Bimodal Interactive Black-Box Jailbreak Attack for Toxicity Maximization.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop PBI -Attack: Prior-Guided Bimodal Interactive Black-Box Jailbreak Attack for Toxicity Maximization

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.024785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.024785Z digest=sha256:cbf691cfc1341e81caa9fd1a9bd7496afa5365c369e73e8ec326971ecc41a749

Observation 70156afd-43f2-4b56-9937-21f1a44382fc · outbound

This paper cites Beyond Text-to- SQL for IoT Defense: A Comprehensive Framework for Querying and Classifying IoT Threats.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Beyond Text-to- SQL for IoT Defense: A Comprehensive Framework for Querying and Classifying IoT Threats

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.028942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.028942Z digest=sha256:42d77822077b64c14e6bff4c5ceae57e967e3f0cce7dca5eecd6629559c1d031

Observation 43c6f142-7da7-4d57-95c6-a3c650c01f21 · outbound

This paper cites Minimal Evidence Group Identification for Claim Verification.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Minimal Evidence Group Identification for Claim Verification

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.033333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.033333Z digest=sha256:1bd026526b13cab9d81b247ccc5da0c864e132767510954feb15620ca25aa8a9

Observation cd32d46d-f950-4093-ad1d-ba9016f142d7 · outbound

This paper cites Estimating Knowledge in Large Language Models Without Generating a Single Token.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Estimating Knowledge in Large Language Models Without Generating a Single Token

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.037344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.037344Z digest=sha256:e6115b87f7219a14d46b6f5b4abdf162ca2067f52a69b339eedc9f49bbfbd62b

Observation e0f4c745-e219-4e72-80c3-ed429bbe5883 · outbound

This paper cites Intrinsic Test of Unlearning Using Parametric Knowledge Traces.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Intrinsic Test of Unlearning Using Parametric Knowledge Traces

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.041415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.041415Z digest=sha256:952b425adca68ef7051386be9aea1414e6d07e0c5b18c997f54a2d91b4b77add

Observation fde013c0-e009-4fed-9000-b955d5bd1970 · outbound

This paper cites Can we trust the evaluation on C hat GPT ?.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Can we trust the evaluation on C hat GPT ?

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.045477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.045477Z digest=sha256:f05b8f76765812fa5f5e69256e047ce57dd8190897c4653c39788779e1d23ae1

Observation bb79b921-8c81-4537-bcb9-a161c973fd60 · outbound

This paper cites Improving Factuality of Abstractive Summarization via Contrastive Reward Learning.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Improving Factuality of Abstractive Summarization via Contrastive Reward Learning

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.049608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.049608Z digest=sha256:182d3b7f787a57ade1c5cac37b9a95776c943e9f7b0b284f0803f57c4928bfb2

Observation 81f29426-d024-4297-bba4-f31f8c9d42e3 · outbound

This paper cites Exploring Causal Mechanisms for Machine Text Detection Methods.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Exploring Causal Mechanisms for Machine Text Detection Methods

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.053733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.053733Z digest=sha256:fae0ad1b7ec1455a5de9143df389f7d2f394a4999fd7375768032802f52d2900

Observation 29cfc788-99b7-49f6-9547-553936eebf07 · outbound

This paper cites On the Robustness of Agentic Function Calling.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop On the Robustness of Agentic Function Calling

Reference 101

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.057835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.057835Z digest=sha256:359ac6d2acc5989a36281bd3a2325f47d0ee854d83a47c5ee188f6efd90d7ceb

Observation e27a5b00-a18a-48c3-8fc1-db57c08867f8 · outbound

This paper cites Cross-Task Defense: Instruction-Tuning LLM s for Content Safety.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Cross-Task Defense: Instruction-Tuning LLM s for Content Safety

Reference 102

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.061990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.061990Z digest=sha256:0e21d4de05c1e7e3b1d7d88c61b128a0235e85891411c448c2077489b47507f6

Observation d9e1eb20-7381-4116-8447-fdb5428c4c1e · outbound

This paper cites Gender Bias in Natural Language Processing Across Human Languages.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Gender Bias in Natural Language Processing Across Human Languages

Reference 103

Resolution
verified exact
doi, observed 2026-08-12T04:51:17.714511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.066494Z digest=sha256:0e85321e03cb5aba913b3a8429d381b8a1f1530a5a696f71e615c5a0d14886e8

Observation c4df4eb7-a254-42fa-a742-7af291b2e8df · outbound

This paper cites Into the Gap between What Language Models Say and What They Know.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Into the Gap between What Language Models Say and What They Know

Reference 104

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.071469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.071469Z digest=sha256:904a0c70e4c981c8310d330ce63b1eb0bb1d7a8b0de8d440d6c0613f0a7e71c3

Observation 1dd145f3-52cc-4863-b58e-f57ea78259a9 · outbound

This paper cites The False Sense of Privacy in LLM s: Non-Verbatim Memorization and Semantic Leakage.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop The False Sense of Privacy in LLM s: Non-Verbatim Memorization and Semantic Leakage

Reference 105

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.075909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.075909Z digest=sha256:5ad8c9b3b34f3c20d9b8458d42ac240901ce0818903603536d576e5806320a5f

Observation 665d41c0-4741-4d55-bfd6-a3f69d137367 · outbound

This paper cites and Raimundo, Marcos M.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop and Raimundo, Marcos M

Reference 106

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.081130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.081130Z digest=sha256:94412f1d96ce3efcaebe7f0d07fd1218388faff1928b12cbcbcd903f207dabbf

Observation df6bb622-cc75-4d30-a7c2-7f3f0cf08de8 · outbound

This paper cites 2023 , howpublished =.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop 2023 , howpublished =

Reference 107

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.085233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.085233Z digest=sha256:d1364c29cca3afd19b719e9bde87544cf86d72403e05f38f24067dccfe74822c

Observation b8febde7-8661-42d9-a1ad-9d74840a86c0 · outbound

This paper cites Nature Machine Intelligence , volume=.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Nature Machine Intelligence , volume=

Reference 108

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.089472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.089472Z digest=sha256:773b1dc80ecb5c5df0908a351a8dbfe1dcf8f6fdc60b974432afee2596767c8e

Observation 86f3659e-2cca-43b0-90de-d40d432325ad · outbound

This paper cites Formalizing Trust in Artificial Intelligence: Prerequisites, Causes and Goals of Human Trust in AI , year =.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Formalizing Trust in Artificial Intelligence: Prerequisites, Causes and Goals of Human Trust in AI , year =

Reference 109

Resolution
verified exact
arxiv_id_nonexistent, observed 2026-08-12T04:51:18.118140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.093468Z digest=sha256:809f302fad80394204dafd0fe230200ba0922efcc713e32ddb6462938f6c2a95

Observation 4ae1e4f8-d7f1-4291-852c-520c7d802d47 · outbound

This paper cites FAccT 2022 , year =.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop FAccT 2022 , year =

Reference 110

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.097633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.097633Z digest=sha256:d2bc6f52b5c5087cacea9e5995ecc4b6e8a9b9ac352d1b410b06f1d94ab9da6c

Observation 0147a468-2552-4eda-a574-a7dee4c5c3ae · outbound

This paper cites SODAPOP : Open-Ended Discovery of Social Biases in Social Commonsense Reasoning Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop SODAPOP : Open-Ended Discovery of Social Biases in Social Commonsense Reasoning Models

Reference 111

Resolution
verified exact
doi, observed 2026-08-12T04:51:17.453501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.102185Z digest=sha256:4238d6cbef5e0bd1e6480477c9c081ff6d3f9ca04d40c6266eeb4195c7e527df

Observation a4a719c5-83ac-4805-bb0e-d9e500235ea1 · outbound

This paper cites F air B elief - Assessing Harmful Beliefs in Language Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop F air B elief - Assessing Harmful Beliefs in Language Models

Reference 112

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.106463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.106463Z digest=sha256:bf1601b8f35b2c4590634184e96929854a74d074f9fb8879f84fe87965c9d27f

Observation 6f1b58e1-2ae7-4f1e-b79b-900131422633 · outbound

This paper cites Investigating and Addressing Hallucinations of LLM s in Tasks Involving Negation.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Investigating and Addressing Hallucinations of LLM s in Tasks Involving Negation

Reference 113

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.110769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.110769Z digest=sha256:14099d017ffcfa9bf040e6de0a3278f3802bdbcb4fc8a6222fa9b759a67ff9e1

Observation b6c78b3f-8159-465f-af7a-ee3361fa9b89 · outbound

This paper cites Introducing G en C eption for Multimodal LLM Benchmarking: You May Bypass Annotations.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Introducing G en C eption for Multimodal LLM Benchmarking: You May Bypass Annotations

Reference 114

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.114972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.114972Z digest=sha256:864156d491e7ea3de81a2c81ecf660d61277b79c828c796772c8eb90ece570c3

Observation 1ae8427c-ff08-47d1-a8a4-5c26bd86b8ef · outbound

This paper cites Tell Me Why: Explainable Public Health Fact-Checking with Large Language Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Tell Me Why: Explainable Public Health Fact-Checking with Large Language Models

Reference 115

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.119168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.119168Z digest=sha256:d68618cd5abb39e43506cca5a8adf4a00f8a600c063e4602128bcbf1287acc60

Observation 006a2779-9e8f-4f0a-bcfe-076eb44a3cef · outbound

This paper cites Disentangling Linguistic Features with Dimension-Wise Analysis of Vector Embeddings.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Disentangling Linguistic Features with Dimension-Wise Analysis of Vector Embeddings

Reference 116

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.123464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.123464Z digest=sha256:66a26672e4c9644cf7ef32f95049003acae785dd1c2be362dde470d0b2168b11

Observation bb1c49fa-76ae-401e-b19e-adb654a3615f · outbound

This paper cites On The Real-world Performance of Machine Translation: Exploring Social Media Post-authors' Perspectives.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop On The Real-world Performance of Machine Translation: Exploring Social Media Post-authors' Perspectives

Reference 117

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.127721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.127721Z digest=sha256:fe3687329151084097d78e98994ef987f75343bb90325dd1bc8c12b28a1e3d8f

Observation 3e020f1c-d738-4c26-8411-4ed746e324a8 · outbound

This paper cites V i B e: A Text-to-Video Benchmark for Evaluating Hallucination in Large Multimodal Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop V i B e: A Text-to-Video Benchmark for Evaluating Hallucination in Large Multimodal Models

Reference 118

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.131935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.131935Z digest=sha256:278ca8399d69496ddc1a40112e6f603a5397e98db0e1c4a4561ac4220b3bde38

Observation 525ff90f-5571-4d9d-8c82-0760ff712954 · outbound

This paper cites FACTOID : FAC tual en T ailment f O r halluc I nation Detection.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop FACTOID : FAC tual en T ailment f O r halluc I nation Detection

Reference 119

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.136528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.136528Z digest=sha256:290d4179427b3ca978a8b0157342572d3978dd8488a37dfca4e6f2fcfa2971cb

Observation 90472fb8-a17a-4039-83a5-b2ab7fd06287 · outbound

This paper cites Private Release of Text Embedding Vectors.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Private Release of Text Embedding Vectors

Reference 120

Resolution
verified exact
doi, observed 2026-08-12T04:51:17.384662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.141076Z digest=sha256:18e0d1fe879dda763740f7ed844e810d017a6058a2d6fef0ce96066d4f879b68

Observation 74282c53-747a-475f-98ee-858b036da7a9 · outbound

This paper cites Challenges in Applying Explainability Methods to Improve the Fairness of NLP Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Challenges in Applying Explainability Methods to Improve the Fairness of NLP Models

Reference 121

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.145503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.145503Z digest=sha256:b50866903eec375da0ffa30eb4948dc5bb67a1125889ae80b4138b8a05f28e35

Observation 2bb8e51e-980a-4099-8f0a-675e9c86307c · outbound

This paper cites An Encoder Attribution Analysis for Dense Passage Retriever in Open-Domain Question Answering.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop An Encoder Attribution Analysis for Dense Passage Retriever in Open-Domain Question Answering

Reference 122

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.149773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.149773Z digest=sha256:19523b4533dba1fff3e3c5d6e79a4455cc31d134e5f9fbf00e2ff549d9fce69a

Observation b90ff77c-5bbe-4c31-85ed-7a757a7e202a · outbound

This paper cites A Keyword Based Approach to Understanding the Overpenalization of Marginalized Groups by E nglish Marginal Abuse Models on T witter.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop A Keyword Based Approach to Understanding the Overpenalization of Marginalized Groups by E nglish Marginal Abuse Models on T witter

Reference 123

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.154226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.154226Z digest=sha256:7de8be3f0aca31ea63a148ec48313f1cc823f3c3e359f97dea4ecd37cc2de141

Observation adf32ab9-b7de-43be-b04c-19f465a8f73d · outbound

This paper cites Examining the Causal Impact of First Names on Language Models: The Case of Social Commonsense Reasoning.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Examining the Causal Impact of First Names on Language Models: The Case of Social Commonsense Reasoning

Reference 124

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.541575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.158335Z digest=sha256:155392c148b80470b7d34cff6a8cd893c46d2d084a6bce4e118c104280d86c92

Observation 9fa36f13-191e-494a-b98d-78ff65c182fa · outbound

This paper cites An Empirical Study of Metrics to Measure Representational Harms in Pre-Trained Language Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop An Empirical Study of Metrics to Measure Representational Harms in Pre-Trained Language Models

Reference 125

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.528229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.162510Z digest=sha256:0987025e4e369183bc1f5f1b10216a539a727afdb3cd175eb87ce842d1143dc7

Observation 341e203f-54ec-4c36-8eb9-9c88008c453a · outbound

This paper cites Beyond T uring: A Comparative Analysis of Approaches for Detecting Machine-Generated Text.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Beyond T uring: A Comparative Analysis of Approaches for Detecting Machine-Generated Text

Reference 126

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.514337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.166956Z digest=sha256:0a0cf6d7125d829db3c6f9d0e5f9068f57d0893bf28ddd340af0e5ee941e96b8

Observation 005475b9-56ca-431d-848e-697fccbee092 · outbound

This paper cites Automated Adversarial Discovery for Safety Classifiers.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Automated Adversarial Discovery for Safety Classifiers

Reference 127

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.500532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.171286Z digest=sha256:6142aa63637423e2daa8fc299bcb186fcf8f682a9a8105cf451f23e1f0fb48c1

Observation 7ea20418-09e4-482a-9562-f91ede53ab8f · outbound

This paper cites The Trade-off between Performance, Efficiency, and Fairness in Adapter Modules for Text Classification.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop The Trade-off between Performance, Efficiency, and Fairness in Adapter Modules for Text Classification

Reference 128

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.487078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.175925Z digest=sha256:dd36b3546e6787fa6c2de4df025eca59fd2651f765c33d5fd1f3641051f9ce6d

Observation 190fa7b7-a3c4-4d80-876e-5395f7ffa0f7 · outbound

This paper cites On the Interplay between Fairness and Explainability.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop On the Interplay between Fairness and Explainability

Reference 129

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.473132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.180480Z digest=sha256:818b7dd498b729fedb9c34cd3508e5a6331b22c143219f2f37f3069565f8ed6c

Observation 55ba8d40-09b8-46fd-96da-4a6b551689e9 · outbound

This paper cites F act A lign: Fact-Level Hallucination Detection and Classification Through Knowledge Graph Alignment.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop F act A lign: Fact-Level Hallucination Detection and Classification Through Knowledge Graph Alignment

Reference 130

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.458476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.185252Z digest=sha256:e12bbffe7a08d947bd4a2d6d11c4b18cfa825e8b644c97a40c311ca4c605e112

Observation 8c75aa3a-5794-4c63-88b4-404d2bc63ee8 · outbound

This paper cites Break the Breakout: Reinventing LM Defense Against Jailbreak Attacks with Self-Refine.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Break the Breakout: Reinventing LM Defense Against Jailbreak Attacks with Self-Refine

Reference 131

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.443147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.190099Z digest=sha256:d76d00dcd700806d0b8dd30968af0a864f95637a2c579ed1fc920e2ad386d6f3

Observation 6f66e172-2292-46bb-b6a1-953db8cc11f4 · outbound

This paper cites Ambiguity Detection and Uncertainty Calibration for Question Answering with Large Language Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Ambiguity Detection and Uncertainty Calibration for Question Answering with Large Language Models

Reference 132

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.428494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.194565Z digest=sha256:5ae917a3e1e74438d9b8985d95e04ecb54123f147f16804e4d7a286994ce137d

Observation 29146040-7016-4cff-8c6c-ea063beffc48 · outbound

This paper cites Error Detection for Multimodal Classification.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Error Detection for Multimodal Classification

Reference 133

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.198931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.198931Z digest=sha256:da2e66f5c0497e7ebc352410b98b091f5ffd081b864643cf006a3ab0fd947628

Observation 91c8a12d-937d-4e08-ba69-3eec055fa2a8 · outbound

This paper cites Know What You do Not Know: Verbalized Uncertainty Estimation Robustness on Corrupted Images in Vision-Language Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Know What You do Not Know: Verbalized Uncertainty Estimation Robustness on Corrupted Images in Vision-Language Models

Reference 134

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.414831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.203373Z digest=sha256:c2acbf641ae4ed9494b48740351bf1df208e41f733b97f118ed7571ab115edcb

Observation a7212c6c-36cd-4ffb-b00b-7e7f9fc293b7 · outbound

This paper cites Multi-lingual Multi-turn Automated Red Teaming for LLM s.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Multi-lingual Multi-turn Automated Red Teaming for LLM s

Reference 135

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.400454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.207859Z digest=sha256:3a8fb08a91a10fcc430378b09c00ac53b06901c77f19d3e492c4de53ec32f487

Observation 5be57346-eaa7-4f20-8844-ef8f0db1a5bf · outbound

This paper cites Line of Duty: Evaluating LLM Self-Knowledge via Consistency in Feasibility Boundaries.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Line of Duty: Evaluating LLM Self-Knowledge via Consistency in Feasibility Boundaries

Reference 136

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.385972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.212116Z digest=sha256:8ea13b82b3db0fb9d0cf62c445099f3d8b3b4e022007ac885b200eb534523faa

Observation 4fb8417f-d540-425e-9264-e7743f4588b2 · outbound

This paper cites MoNaCo: More Natural and Complex Questions for Reasoning Across Dozens of Documents.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop MoNaCo: More Natural and Complex Questions for Reasoning Across Dozens of Documents

Reference 137

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.216216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.216216Z digest=sha256:6462ef92dca3f6572da746422a68130a309f7d01095b8088edebe5138373af3b

Observation 54f6a97a-3cd1-4e93-afd0-4e4e809588cf · outbound

This paper cites and Aletras, Nikolaos and Ma, Ning.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop and Aletras, Nikolaos and Ma, Ning

Reference 138

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.220598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.220598Z digest=sha256:743d1cd548515e586e6b4e4227c7673318888c076193f1e459fa085d4694cffd

Observation 7046bf1a-afc2-43c3-b9cf-0cafdae14f5b · outbound

This paper cites A Survey on Gender Bias in Natural Language Processing.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop A Survey on Gender Bias in Natural Language Processing

Reference 139

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.224833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.224833Z digest=sha256:32a40ba6f6ebe2bb37888341b3df0500f73d5633e9861d8eb36dcd7fbd8f8346

Observation 443dba07-8f5f-40a6-8d92-ae7ad2edc7fa · outbound

This paper cites Inducing Positive Perspectives with Text Reframing.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Inducing Positive Perspectives with Text Reframing

Reference 140

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.229554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.229554Z digest=sha256:767169a18e37643d9d97bc448312bd59641d8eb9dd69d7915fbbe7f516e97f69

Observation a71ace5c-74d1-4d18-9a9d-2fe24ea8f402 · outbound

This paper cites The Importance of Modeling Social Factors of Language: Theory and Practice.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop The Importance of Modeling Social Factors of Language: Theory and Practice

Reference 141

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.233451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.233451Z digest=sha256:c2c729be6275ece44b08f48398f3c8f9e247c5d29e6b9c4a3dbe000f386cd207

Observation 1e9ce44b-5ca1-4dc6-b711-3998d2ef9c6d · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 142

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.237236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.237236Z digest=sha256:e211f0156ac26d1d530394a20edb29ad67c5c0bd21df098fb0cfb94bb28d14da

Observation 258c868f-edfa-44fd-8a50-a06277381181 · outbound

This paper cites 2023 , howpublished =.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop 2023 , howpublished =

Reference 143

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.241762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.241762Z digest=sha256:988a8aa1c0710d3e92929b2b769975276cb8ac611acf6bdbfa0bc347daa0ba9e

Observation 7650e705-1670-4fde-aa0c-ef1a3c49304a · outbound

This paper cites GPT-4 Technical Report.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop GPT-4 Technical Report

Reference 144

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.246222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.246222Z digest=sha256:2a5b84af1bcbfdfb30f508ec46607226b72e24a0e2997e94b3326c9e2cad095e

Observation 4a4c4ba2-69bc-47c8-92bb-119dbf2d7bf4 · outbound

This paper cites Strength in Numbers: Estimating Confidence of Large Language Models by Prompt Agreement.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Strength in Numbers: Estimating Confidence of Large Language Models by Prompt Agreement

Reference 145

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.250929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.250929Z digest=sha256:49f9d31e877958418ed9fe205dc5508d2616b68a142960239e72569be0776109

Observation 4617b8f3-ce41-4840-a759-dd30de6748b9 · outbound

This paper cites On the Intrinsic and Extrinsic Fairness Evaluation Metrics for Contextualized Language Representations.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop On the Intrinsic and Extrinsic Fairness Evaluation Metrics for Contextualized Language Representations

Reference 146

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.255122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.255122Z digest=sha256:00e3be4ed9730d601613523251d8821c91c2772a84093b53692275c834c59064

Observation 9b6dd65d-2dd3-4dc9-b107-a6e7c1e3b72c · outbound

This paper cites Pay Attention to the Robustness of C hinese Minority Language Models! Syllable-level Textual Adversarial Attack on T ibetan Script.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Pay Attention to the Robustness of C hinese Minority Language Models! Syllable-level Textual Adversarial Attack on T ibetan Script

Reference 147

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.259117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.259117Z digest=sha256:e6a78b93495bb09890753a5f39edf0f49c9fe52045f02b6306a1e8be94b397b6

Pith citing papers

No inbound Pith citation observations are available.