Pith. sign in

Paper Citation Record · LEDGER

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop

As of 12 August 2026, this Paper Citation Record lists 100 of 118 outbound references and 0 inbound Pith citation observations for arXiv:2608.11171.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.11171 v1

Coverage vector

measured 100 of 118 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T04:51:17.259117Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 118 outbound references displayed

  • verified exact6
  • verified fuzzy12
  • unresolved82
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8d9ca6f8-ec40-4cd2-9a0e-f359d768f7d3 · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.630599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.630599Z digest=sha256:896eb659f845b1162fea52e77439ed61f8b6ade05ad7e99c1643f2f3ff80e3e6

Observation de088d98-e2f9-4e2d-af73-5700f1053308 · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.648681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.648681Z digest=sha256:4489f3329cc49fdf4fbffae39d3ab4240cb82d9a50b017b6740bb8645dc70406

Observation bf5ef2da-7092-4ecf-aada-16d42922fd86 · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.667018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.667018Z digest=sha256:1e609e0d9ab403051a5c51bd24a62411ee4af65f5cf333ea9d50590f97b8f543

Observation c011e2b4-367e-4497-9fa6-fa23c58f332f · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.671317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.671317Z digest=sha256:8da993a1879e5b64f84b888ec0b122614fce22cbb321c2797c8db854939f9827

Observation c8e8bca1-0a47-4526-81e7-a4ff31fb9f57 · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.685034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.685034Z digest=sha256:082bde99321d62f4387fa54879c13ffc4e054aca70d8637a10350b02e9f82206

Observation baac49ac-a54a-4be1-b053-273459137e61 · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.690286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.690286Z digest=sha256:a910655f168313ac7aa4af0c70c4c87ca973b8a4c3f4aa0737270829f3f0143f

Observation c4ffb726-c5f2-42d2-b97f-597a473e3724 · outbound

This paper cites don't forget the teachers.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop don't forget the teachers

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.737630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.737630Z digest=sha256:87c61df5334dcdcac1d701933bb602f35de1ee3725038ec929aeccade7e7ef02

Observation 6b16f8c3-3af4-417c-bddc-72aef843c6b5 · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.745842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.745842Z digest=sha256:b7d2d8a37b0186148f9b0a779f7744f564568d7046ae21b61a354e5cecb079c8

Observation 9870e611-b528-4ab9-9e31-792d137b9324 · outbound

This paper cites TrustLLM: Trustworthiness in Large Language Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop TrustLLM: Trustworthiness in Large Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.754194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.754194Z digest=sha256:6ef63fcbfdd01adc356cced7e3a0d30318061a0223cf6e3721f4d10541dfd252

Observation 46401178-93bd-4429-a47e-4e174c89416b · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.758736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.758736Z digest=sha256:68c550663abc08e1e9ce897e6c4cb9062717190ad6694e9a91deec1156ad89b4

Observation 831c3195-f441-40e5-a3e8-22ca521a3feb · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.767502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.767502Z digest=sha256:ffa21ed28ffd8259fea6967c7890af0c7e1f211d7e27dfd3d089fb823df8bbf4

Observation 163ffdfc-8028-4fbd-a4a9-71c18d0b6783 · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.771478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.771478Z digest=sha256:3ff6053bcff85cf8ede210826cc788a9917fe7005b0a3bcd2111e7c3fa048f37

Observation 1a998e40-6912-4942-baef-7f9c23cbe17e · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.775652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.775652Z digest=sha256:3d9269259c2f176c04189001e8d5e5510b57726fa7fd06bd2be0c51db2c550ba

Observation 645cca4e-9f50-4fb8-93ce-98e5cd658204 · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.811315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.811315Z digest=sha256:b6c1f0db27bd4964cbc9a11d2339572e408082a448653ea0f62bceaacfe137b9

Observation f74c9f77-3b20-4494-981d-04f072dc2f58 · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.815380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.815380Z digest=sha256:3a43d1e1bea4bf548345b52d5ff6fdcb713948d6bdc496edb7f0ffc9b1cafed9

Observation b1e16acf-8f1e-42f5-a2a3-a5e70d8948fa · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.827962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.827962Z digest=sha256:a6b6c7a48c15ab0513246a13643788465f86b45957eca83db831a3d8ea8a514b

Observation 6041a784-e8cb-42c8-aa8a-ad32cb84e29a · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.857656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.857656Z digest=sha256:78d3f96d78de9388a07267095e09c89653b92fa7fdd315431a9ce2addda35f2e

Observation 7cddb846-25b5-4780-9965-db36f7c79c4b · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.880222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.880222Z digest=sha256:d3f2036ed98dcc8f5bf775fda6c81056c5a0abfa2adb06357203a8c54040dc7b

Observation 68a32cee-7b52-4259-8f3a-8fe60d1019e6 · outbound

This paper cites DecodingTrust: A Comprehensive Assessment of Trustworthiness in GPT Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop DecodingTrust: A Comprehensive Assessment of Trustworthiness in GPT Models

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.888702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.888702Z digest=sha256:4fcef0252485f4381ca92d253abcd6141719b4e0a1ce696f835ecf73cefff458

Observation 02bafdd9-53fb-4a9a-a152-1238249fb910 · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.901804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.901804Z digest=sha256:0f9f891ec180f122ca3cca9922591689606f5c8cffa5c00c5023eb2387deb178

Observation 3c7f868c-804e-474f-8090-2f9f3b95bf1c · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.919064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.919064Z digest=sha256:b1ba6046ea773bceb06adc521efba295535ad9e9d1d40e313aad9bbc7a21a030

Observation 0c8f1fc8-d6fe-4850-8274-79fed12bacdf · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.923408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.923408Z digest=sha256:615d999b0d651175b76561e16e4db030da94c6828b8e3db065ccf5a7ce85d4fd

Observation db943417-ff45-4671-91b0-ed7794a7416a · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.927400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.927400Z digest=sha256:4412073eb5044ce0e510356311f7bc354a417847de77504c0d1cc16677b70340

Observation c3b4fb2d-528f-4e0a-89e8-0382adcebc0b · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.931692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.931692Z digest=sha256:f313699afd76f1e0682c19b47aec166572ddaa6b2de2a5fd65cc7d2cebbe5820

Observation c479cfcd-9596-4125-8055-ad17f421356b · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.935849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.935849Z digest=sha256:6f759a0f3780e608cd46f6bc0f5a8ce97c1dcc86b9252449d0c01ba667c1adac

Observation 84c9f417-3d2e-4107-bed1-25062866139e · outbound

This paper cites 2026 , howpublished =.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop 2026 , howpublished =

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.939878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.939878Z digest=sha256:81884c787e534f37991b83623a043de67ac22e9315ca950aa0043253757f9a36

Observation 26f9553c-e3d4-432a-a788-02e54cf8dc1e · outbound

This paper cites Human-Centered Explainable AI : Towards a Reflective Sociotechnical Approach.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Human-Centered Explainable AI : Towards a Reflective Sociotechnical Approach

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.944004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.944004Z digest=sha256:901440277ac69be41df3239d2237feedf2dbb65c2b78093e3ff27f239674f823

Observation 107975d5-4d78-4da4-8757-05817c127b83 · outbound

This paper cites and Wintersberger, Philipp and Manger, Carina and Hubig, Nina and Savage, Saiph and Weisz, Justin D.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop and Wintersberger, Philipp and Manger, Carina and Hubig, Nina and Savage, Saiph and Weisz, Justin D

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.948539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.948539Z digest=sha256:19910f3738eae7764ce50f973cdbaae2927a9e29b19240576bb6650b55fd67f9

Observation 0e2ea7ae-a4f9-46c1-9066-b8b0cf35c35a · outbound

This paper cites A Survey on Medical Large Language Models: Technology, Application, Trustworthiness, and Future Directions.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop A Survey on Medical Large Language Models: Technology, Application, Trustworthiness, and Future Directions

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.952670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.952670Z digest=sha256:30a9b8c3dab3762dba69b3c2270576cc1910a429c43d9169c390eb9446c7bf5d

Observation da851883-7758-46cc-820b-29c1860d8b97 · outbound

This paper cites Standard Benchmarks Fail -- Auditing LLM Agents in Finance Must Prioritize Risk.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Standard Benchmarks Fail -- Auditing LLM Agents in Finance Must Prioritize Risk

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.956798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.956798Z digest=sha256:1e524715253d99db9f2d3fbc11e4dd7680477f8d14d03d41c70191eaca4eae28

Observation 91bc9c9d-60ba-4cbf-9cf8-667c86e04d82 · outbound

This paper cites Don't Forget the Teachers.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Don't Forget the Teachers

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.960742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.960742Z digest=sha256:f6a3575a1d77776ce222f18cca3d66668cc70a9697de2afe4c7455163e123407

Observation 913fa84b-6be8-4080-8c9f-6d436179cc43 · outbound

This paper cites 2025 Silicon Valley Cybersecurity Conference (SVCC) , pages=.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop 2025 Silicon Valley Cybersecurity Conference (SVCC) , pages=

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.965221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.965221Z digest=sha256:ec80055ae4e0ca734b0c3d7ed31ee69410ecd445bbfef87221b28b465e93a18f

Observation 3bd9c0e2-d80d-419f-a421-d48a1d9b36af · outbound

This paper cites 2023 , url =.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop 2023 , url =

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.969245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.969245Z digest=sha256:3f8d2e1e8f5108f71e4c922f3f6534b45736886c9a18e9496fc5752dceadb5e7

Observation 4cf52372-36ab-4930-b5d2-9427461bb01c · outbound

This paper cites 2024 , url=.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop 2024 , url=

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.973582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.973582Z digest=sha256:3af53152a9d57014a7084c25dbd2cd5eea2263c48d6ce111e8c3630f85d7bd65

Observation 428c3343-a31c-4b1a-a90f-2fc70ef11344 · outbound

This paper cites Trustworthy LLMs: a Survey and Guideline for Evaluating Large Language Models' Alignment.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Trustworthy LLMs: a Survey and Guideline for Evaluating Large Language Models' Alignment

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.977677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.977677Z digest=sha256:7715c99116aad6439193622f29ff70f8e8fad050dccb8d73bbea4d6d15bdadfb

Observation 5efb6439-584c-4411-a97b-fb0b4b0a62ce · outbound

This paper cites 2025 , publisher=.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop 2025 , publisher=

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.981843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.981843Z digest=sha256:237059b9253212af97936c3e08f314b8e3e998210097b72accb5929a388e2a5d

Observation 498ceeaf-2846-448f-bc20-7fc603cd67ae · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.986141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.986141Z digest=sha256:a2f81d8f32f00de4b24729b0c12014fa76fd4aa8e65aae69592498abc4a9ac0c

Observation 37a8edc6-c58f-404c-b0a7-82c03f5dcfb3 · outbound

This paper cites Interpretability Rules: Jointly Bootstrapping a Neural Relation Extractor with an Explanation Decoder.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Interpretability Rules: Jointly Bootstrapping a Neural Relation Extractor with an Explanation Decoder

Reference 85

Resolution
verified exact
doi, observed 2026-08-12T04:51:17.579945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:16.990589Z digest=sha256:38ac35cd52f6cd0578a9b3244f7f1a8bd2d47576feaee1ed9b38ba46565aee16

Observation fa99828f-2299-4b74-8df4-932c0a7707fd · outbound

This paper cites Measuring Biases of Word Embeddings: What Similarity Measures and Descriptive Statistics to Use?.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Measuring Biases of Word Embeddings: What Similarity Measures and Descriptive Statistics to Use?

Reference 86

Resolution
verified exact
doi, observed 2026-08-12T04:51:17.897013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:16.994632Z digest=sha256:1f44b557a9b99d37c189be409d84ee77983cdc9727683b2f44b465fe058963ca

Observation 0d500bcb-aa3f-47bf-870e-b15566842501 · outbound

This paper cites and Kiritchenko, Svetlana and Balkir, Esma.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop and Kiritchenko, Svetlana and Balkir, Esma

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.999053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.999053Z digest=sha256:f7f9f413c2f1ca2dd05ea4038bcd6d9a078798665277560d600cff27a33f6117

Observation 977ce2f1-f298-4342-b774-2ceae3df2885 · outbound

This paper cites GPT s Don ' t Keep Secrets: Searching for Backdoor Watermark Triggers in Autoregressive Language Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop GPT s Don ' t Keep Secrets: Searching for Backdoor Watermark Triggers in Autoregressive Language Models

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.003186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.003186Z digest=sha256:5866fe5cb9cfb21c3013d9d07e0d0c4a0652972aac42f67b36b4aa4407bfdc6c

Observation 855102a8-6674-4972-b483-eab017c5bbfb · outbound

This paper cites Reliability Check: An Analysis of GPT -3's Response to Sensitive Topics and Prompt Wording.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Reliability Check: An Analysis of GPT -3's Response to Sensitive Topics and Prompt Wording

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.007660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.007660Z digest=sha256:e0146e75b9802988885e9edcc1d71cbf37d74098475832a45e2d444929949c16

Observation a2e7835e-f913-4a22-b938-6112a387e6b0 · outbound

This paper cites Driving Context into Text-to-Text Privatization.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Driving Context into Text-to-Text Privatization

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.011852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.011852Z digest=sha256:881a0b51cb7f90ba4ed20b77c5363b8c014986c48d0a38296161959bd3d6226a

Observation a569e7ad-2d73-44ca-8d2e-99b7c9e94b1e · outbound

This paper cites Expanding Scope: Adapting E nglish Adversarial Attacks to C hinese.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Expanding Scope: Adapting E nglish Adversarial Attacks to C hinese

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.016058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.016058Z digest=sha256:bfa88f505c0d3495c3a160794ed4b82d636633bfbfd91eaa7c247960d8222291

Observation 0b3994e2-3d6e-410b-8a35-77b802cde3ba · outbound

This paper cites Flatness-Aware Gradient Descent for Safe Conversational AI.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Flatness-Aware Gradient Descent for Safe Conversational AI

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.020464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.020464Z digest=sha256:4f4ad5c353bb0c15482fd5464a948fdda545c819e63f3dceb434ce5a5271831e

Observation 8a824ce0-ecf5-4508-b1c1-5e2749353908 · outbound

This paper cites PBI -Attack: Prior-Guided Bimodal Interactive Black-Box Jailbreak Attack for Toxicity Maximization.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop PBI -Attack: Prior-Guided Bimodal Interactive Black-Box Jailbreak Attack for Toxicity Maximization

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.024785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.024785Z digest=sha256:ca054651330f39b580cf36462499cb8601c1eb94f8e2605b26dc8a6dff919877

Observation 70156afd-43f2-4b56-9937-21f1a44382fc · outbound

This paper cites Beyond Text-to- SQL for IoT Defense: A Comprehensive Framework for Querying and Classifying IoT Threats.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Beyond Text-to- SQL for IoT Defense: A Comprehensive Framework for Querying and Classifying IoT Threats

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.028942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.028942Z digest=sha256:d2351355878d111ef08d670226c015aae3dd543f4cbc01fdc3fbd0f912ac7a9b

Observation 43c6f142-7da7-4d57-95c6-a3c650c01f21 · outbound

This paper cites Minimal Evidence Group Identification for Claim Verification.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Minimal Evidence Group Identification for Claim Verification

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.033333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.033333Z digest=sha256:b84cf3086d1ea7fe8498c2f8dd72380916f2acec676997e74b33690651a8d94b

Observation cd32d46d-f950-4093-ad1d-ba9016f142d7 · outbound

This paper cites Estimating Knowledge in Large Language Models Without Generating a Single Token.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Estimating Knowledge in Large Language Models Without Generating a Single Token

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.037344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.037344Z digest=sha256:06724b25dcedc9359d8b89d6e3670d4614dff6c45b9ffbd930c95e53d6cde0bf

Observation e0f4c745-e219-4e72-80c3-ed429bbe5883 · outbound

This paper cites Intrinsic Test of Unlearning Using Parametric Knowledge Traces.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Intrinsic Test of Unlearning Using Parametric Knowledge Traces

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.041415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.041415Z digest=sha256:2b2f8215ac9ad0839cff000cdebd7a767a4876bcbdd15c711002f506f28026e7

Observation fde013c0-e009-4fed-9000-b955d5bd1970 · outbound

This paper cites Can we trust the evaluation on C hat GPT ?.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Can we trust the evaluation on C hat GPT ?

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.045477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.045477Z digest=sha256:bb131b50efe73a4d051fa70f40346568da636e387c39dca368197f6d013e0887

Observation bb79b921-8c81-4537-bcb9-a161c973fd60 · outbound

This paper cites Improving Factuality of Abstractive Summarization via Contrastive Reward Learning.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Improving Factuality of Abstractive Summarization via Contrastive Reward Learning

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.049608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.049608Z digest=sha256:1008ef6a1f34b591693dc3091add91a1c094d83cc276083fcb732ba3c75c12bd

Observation 81f29426-d024-4297-bba4-f31f8c9d42e3 · outbound

This paper cites Exploring Causal Mechanisms for Machine Text Detection Methods.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Exploring Causal Mechanisms for Machine Text Detection Methods

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.053733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.053733Z digest=sha256:bdaed2ad861dde6e1ea65676beb6dd896a8271c6c76d312afab39558944e915f

Observation 29cfc788-99b7-49f6-9547-553936eebf07 · outbound

This paper cites On the Robustness of Agentic Function Calling.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop On the Robustness of Agentic Function Calling

Reference 101

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.057835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.057835Z digest=sha256:a3e1798a4a040ca2c2d1b27cc1cde5ab618af8ca7811041ff1c25a515420176b

Observation e27a5b00-a18a-48c3-8fc1-db57c08867f8 · outbound

This paper cites Cross-Task Defense: Instruction-Tuning LLM s for Content Safety.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Cross-Task Defense: Instruction-Tuning LLM s for Content Safety

Reference 102

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.061990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.061990Z digest=sha256:2cb8d0a2e87252ef8dd3544667ddea5a66251acfd2ca04d800b89bde85266bc0

Observation d9e1eb20-7381-4116-8447-fdb5428c4c1e · outbound

This paper cites Gender Bias in Natural Language Processing Across Human Languages.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Gender Bias in Natural Language Processing Across Human Languages

Reference 103

Resolution
verified exact
doi, observed 2026-08-12T04:51:17.714511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.066494Z digest=sha256:820127466c4f30c543b3296717732e5364ca6423c32fbc16f8b945b0cb8d4254

Observation c4df4eb7-a254-42fa-a742-7af291b2e8df · outbound

This paper cites Into the Gap between What Language Models Say and What They Know.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Into the Gap between What Language Models Say and What They Know

Reference 104

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.071469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.071469Z digest=sha256:e4a2b12e5c4a0e6cfe1c407a56abc24cb9b740d7ecc484daba1f4fd878a6ff85

Observation 1dd145f3-52cc-4863-b58e-f57ea78259a9 · outbound

This paper cites The False Sense of Privacy in LLM s: Non-Verbatim Memorization and Semantic Leakage.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop The False Sense of Privacy in LLM s: Non-Verbatim Memorization and Semantic Leakage

Reference 105

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.075909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.075909Z digest=sha256:db8d3387d093e5df208562b8ed20be4e506fc4256c89b3725e0cb840f3294e81

Observation 665d41c0-4741-4d55-bfd6-a3f69d137367 · outbound

This paper cites and Raimundo, Marcos M.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop and Raimundo, Marcos M

Reference 106

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.081130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.081130Z digest=sha256:b59fe09ebc0b78663320223c7e6428c427b5f1db1dbfe152ce895ddbd681b374

Observation df6bb622-cc75-4d30-a7c2-7f3f0cf08de8 · outbound

This paper cites 2023 , howpublished =.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop 2023 , howpublished =

Reference 107

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.085233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.085233Z digest=sha256:fe1582411f5f7207e9a783931aae2bc73bf4953636a4960a3023dfa15ebbc828

Observation b8febde7-8661-42d9-a1ad-9d74840a86c0 · outbound

This paper cites Nature Machine Intelligence , volume=.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Nature Machine Intelligence , volume=

Reference 108

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.089472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.089472Z digest=sha256:877616b323b07109a3c6ca441f1ffb2dd24442450679888feccd278a55e820b6

Observation 86f3659e-2cca-43b0-90de-d40d432325ad · outbound

This paper cites Formalizing Trust in Artificial Intelligence: Prerequisites, Causes and Goals of Human Trust in AI , year =.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Formalizing Trust in Artificial Intelligence: Prerequisites, Causes and Goals of Human Trust in AI , year =

Reference 109

Resolution
verified exact
arxiv_id_nonexistent, observed 2026-08-12T04:51:18.118140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.093468Z digest=sha256:6d64accffd653fd1e325267b766d119a7a5eb225c3944e3de01945a4c1bb324c

Observation 4ae1e4f8-d7f1-4291-852c-520c7d802d47 · outbound

This paper cites FAccT 2022 , year =.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop FAccT 2022 , year =

Reference 110

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.097633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.097633Z digest=sha256:7782ecca1ca0bf56e1c550f42efd0a62ec79cca720630dfe76e9494d4137889b

Observation 0147a468-2552-4eda-a574-a7dee4c5c3ae · outbound

This paper cites SODAPOP : Open-Ended Discovery of Social Biases in Social Commonsense Reasoning Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop SODAPOP : Open-Ended Discovery of Social Biases in Social Commonsense Reasoning Models

Reference 111

Resolution
verified exact
doi, observed 2026-08-12T04:51:17.453501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.102185Z digest=sha256:364ffd6a094bb41c9790bc9993e31974c56423cd00dcaa020e83e1b4b0b4f2dc

Observation a4a719c5-83ac-4805-bb0e-d9e500235ea1 · outbound

This paper cites F air B elief - Assessing Harmful Beliefs in Language Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop F air B elief - Assessing Harmful Beliefs in Language Models

Reference 112

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.106463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.106463Z digest=sha256:9c7a1cd428d64b3977ea1d0f59dc49226238b4d24c52377bf919b6142ba63d95

Observation 6f1b58e1-2ae7-4f1e-b79b-900131422633 · outbound

This paper cites Investigating and Addressing Hallucinations of LLM s in Tasks Involving Negation.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Investigating and Addressing Hallucinations of LLM s in Tasks Involving Negation

Reference 113

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.110769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.110769Z digest=sha256:c1459dabce59efe0c4b8e96ff192ab0b42a00948bdbd22595c4c6a371858eb44

Observation b6c78b3f-8159-465f-af7a-ee3361fa9b89 · outbound

This paper cites Introducing G en C eption for Multimodal LLM Benchmarking: You May Bypass Annotations.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Introducing G en C eption for Multimodal LLM Benchmarking: You May Bypass Annotations

Reference 114

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.114972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.114972Z digest=sha256:c96875b33014e7b5d35e74b3f2e77cec3f2aec3e92167ff296870a4bad3d0a6b

Observation 1ae8427c-ff08-47d1-a8a4-5c26bd86b8ef · outbound

This paper cites Tell Me Why: Explainable Public Health Fact-Checking with Large Language Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Tell Me Why: Explainable Public Health Fact-Checking with Large Language Models

Reference 115

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.119168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.119168Z digest=sha256:6d07b92bf1901144c65ddbdff90c0275f23dda8c5fcd5ecb956f5c7a3c3ae1d3

Observation 006a2779-9e8f-4f0a-bcfe-076eb44a3cef · outbound

This paper cites Disentangling Linguistic Features with Dimension-Wise Analysis of Vector Embeddings.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Disentangling Linguistic Features with Dimension-Wise Analysis of Vector Embeddings

Reference 116

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.123464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.123464Z digest=sha256:6104fccb3a2b1cb3ddca1bc796d99e0fce3e1348ff73f68e4dd790f9b5c8b974

Observation bb1c49fa-76ae-401e-b19e-adb654a3615f · outbound

This paper cites On The Real-world Performance of Machine Translation: Exploring Social Media Post-authors' Perspectives.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop On The Real-world Performance of Machine Translation: Exploring Social Media Post-authors' Perspectives

Reference 117

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.127721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.127721Z digest=sha256:1cd2c6ec488b92881b3874f34d6b10e1436316f1c0c620fa99f512d8fd49ebd9

Observation 3e020f1c-d738-4c26-8411-4ed746e324a8 · outbound

This paper cites V i B e: A Text-to-Video Benchmark for Evaluating Hallucination in Large Multimodal Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop V i B e: A Text-to-Video Benchmark for Evaluating Hallucination in Large Multimodal Models

Reference 118

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.131935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.131935Z digest=sha256:74babb3cf523da9d5eeb8007ce467bf08306244182e47c77da5998b1cbfc7dac

Observation 525ff90f-5571-4d9d-8c82-0760ff712954 · outbound

This paper cites FACTOID : FAC tual en T ailment f O r halluc I nation Detection.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop FACTOID : FAC tual en T ailment f O r halluc I nation Detection

Reference 119

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.136528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.136528Z digest=sha256:b879004198746f85e64711bfcfff77652b16d68e3abcf7c12f4be942ec94eb46

Observation 90472fb8-a17a-4039-83a5-b2ab7fd06287 · outbound

This paper cites Private Release of Text Embedding Vectors.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Private Release of Text Embedding Vectors

Reference 120

Resolution
verified exact
doi, observed 2026-08-12T04:51:17.384662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.141076Z digest=sha256:26a6b5fe06159d153a943632d5029ca4c736d125a03dca3cf76093197c56582d

Observation 74282c53-747a-475f-98ee-858b036da7a9 · outbound

This paper cites Challenges in Applying Explainability Methods to Improve the Fairness of NLP Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Challenges in Applying Explainability Methods to Improve the Fairness of NLP Models

Reference 121

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.145503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.145503Z digest=sha256:62a170cba7ba5854824c41c6d1e8ca5b4cb831353d1a5a6ebef1a750ca4ceb25

Observation 2bb8e51e-980a-4099-8f0a-675e9c86307c · outbound

This paper cites An Encoder Attribution Analysis for Dense Passage Retriever in Open-Domain Question Answering.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop An Encoder Attribution Analysis for Dense Passage Retriever in Open-Domain Question Answering

Reference 122

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.149773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.149773Z digest=sha256:374379140550e3842e70bfd088276f4bc35f0cfb716ce4c8b6d3ec09afcd2468

Observation b90ff77c-5bbe-4c31-85ed-7a757a7e202a · outbound

This paper cites A Keyword Based Approach to Understanding the Overpenalization of Marginalized Groups by E nglish Marginal Abuse Models on T witter.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop A Keyword Based Approach to Understanding the Overpenalization of Marginalized Groups by E nglish Marginal Abuse Models on T witter

Reference 123

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.154226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.154226Z digest=sha256:3e5b5519928cada7aab695d145016588194444184927985a394dbab4b010eb43

Observation adf32ab9-b7de-43be-b04c-19f465a8f73d · outbound

This paper cites Examining the Causal Impact of First Names on Language Models: The Case of Social Commonsense Reasoning.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Examining the Causal Impact of First Names on Language Models: The Case of Social Commonsense Reasoning

Reference 124

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.541575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.158335Z digest=sha256:efe21a1d656a0f248669ae72091df36b92a772fe0adc80b3af6d2b8fb443769b

Observation 9fa36f13-191e-494a-b98d-78ff65c182fa · outbound

This paper cites An Empirical Study of Metrics to Measure Representational Harms in Pre-Trained Language Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop An Empirical Study of Metrics to Measure Representational Harms in Pre-Trained Language Models

Reference 125

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.528229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.162510Z digest=sha256:5c09f0c674864b6972ce6f43178bfe8ec08290c755a4520631714e1f5d609541

Observation 341e203f-54ec-4c36-8eb9-9c88008c453a · outbound

This paper cites Beyond T uring: A Comparative Analysis of Approaches for Detecting Machine-Generated Text.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Beyond T uring: A Comparative Analysis of Approaches for Detecting Machine-Generated Text

Reference 126

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.514337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.166956Z digest=sha256:1c0c2b431ee2483d27b2a2d7cd9086e3c994708f19fafce9377a2b3cd08b1de4

Observation 005475b9-56ca-431d-848e-697fccbee092 · outbound

This paper cites Automated Adversarial Discovery for Safety Classifiers.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Automated Adversarial Discovery for Safety Classifiers

Reference 127

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.500532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.171286Z digest=sha256:7b3c0d5acf8c160295c95657cbb8099ede84f4530a84e475421765b743a57500

Observation 7ea20418-09e4-482a-9562-f91ede53ab8f · outbound

This paper cites The Trade-off between Performance, Efficiency, and Fairness in Adapter Modules for Text Classification.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop The Trade-off between Performance, Efficiency, and Fairness in Adapter Modules for Text Classification

Reference 128

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.487078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.175925Z digest=sha256:dfa41a09f5ae5d22d8811f50ae717fde95b0e288583cd13f103d28303aed717f

Observation 190fa7b7-a3c4-4d80-876e-5395f7ffa0f7 · outbound

This paper cites On the Interplay between Fairness and Explainability.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop On the Interplay between Fairness and Explainability

Reference 129

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.473132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.180480Z digest=sha256:19fbe500354d1cae5d79e35985e24d6f19fb64053656aac9c47961614fd5a767

Observation 55ba8d40-09b8-46fd-96da-4a6b551689e9 · outbound

This paper cites F act A lign: Fact-Level Hallucination Detection and Classification Through Knowledge Graph Alignment.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop F act A lign: Fact-Level Hallucination Detection and Classification Through Knowledge Graph Alignment

Reference 130

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.458476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.185252Z digest=sha256:86e5876975642227712eff80ccbf66623e0a2b5d0b5d98ed4bb9754d32546974

Observation 8c75aa3a-5794-4c63-88b4-404d2bc63ee8 · outbound

This paper cites Break the Breakout: Reinventing LM Defense Against Jailbreak Attacks with Self-Refine.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Break the Breakout: Reinventing LM Defense Against Jailbreak Attacks with Self-Refine

Reference 131

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.443147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.190099Z digest=sha256:f8f8d3874fffbcfcd17d10f747b944afb3c41dd95d9c60ddb148ccd816a99b7b

Observation 6f66e172-2292-46bb-b6a1-953db8cc11f4 · outbound

This paper cites Ambiguity Detection and Uncertainty Calibration for Question Answering with Large Language Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Ambiguity Detection and Uncertainty Calibration for Question Answering with Large Language Models

Reference 132

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.428494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.194565Z digest=sha256:d9ac6e005bbb47ee1b995927d1abdcb1a69190fb94a98a256e04c6ef00b5f2e7

Observation 29146040-7016-4cff-8c6c-ea063beffc48 · outbound

This paper cites Error Detection for Multimodal Classification.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Error Detection for Multimodal Classification

Reference 133

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.198931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.198931Z digest=sha256:05bb71979ed585f3f9016b32c80ceefbea27137483ac8a395a9d52af20e5cad1

Observation 91c8a12d-937d-4e08-ba69-3eec055fa2a8 · outbound

This paper cites Know What You do Not Know: Verbalized Uncertainty Estimation Robustness on Corrupted Images in Vision-Language Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Know What You do Not Know: Verbalized Uncertainty Estimation Robustness on Corrupted Images in Vision-Language Models

Reference 134

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.414831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.203373Z digest=sha256:6c8c7310e2d44d7d1c30f69c49ec9e9e27e17f2f8aede05b45995ffd5a333dbd

Observation a7212c6c-36cd-4ffb-b00b-7e7f9fc293b7 · outbound

This paper cites Multi-lingual Multi-turn Automated Red Teaming for LLM s.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Multi-lingual Multi-turn Automated Red Teaming for LLM s

Reference 135

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.400454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.207859Z digest=sha256:54c16435405f8f9e232f9e02f1d431617bb56774040684f138101c4c9286858a

Observation 5be57346-eaa7-4f20-8844-ef8f0db1a5bf · outbound

This paper cites Line of Duty: Evaluating LLM Self-Knowledge via Consistency in Feasibility Boundaries.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Line of Duty: Evaluating LLM Self-Knowledge via Consistency in Feasibility Boundaries

Reference 136

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.385972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.212116Z digest=sha256:c5abb819a88f4678dfcda656e409b8f7f6f739de360cf884b54e9af2036227d4

Observation 4fb8417f-d540-425e-9264-e7743f4588b2 · outbound

This paper cites MoNaCo: More Natural and Complex Questions for Reasoning Across Dozens of Documents.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop MoNaCo: More Natural and Complex Questions for Reasoning Across Dozens of Documents

Reference 137

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.216216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.216216Z digest=sha256:7d99e13df8d854d6be21b9f1fe5c86b882f1a3df071c5d41f662f788a6532e66

Observation 54f6a97a-3cd1-4e93-afd0-4e4e809588cf · outbound

This paper cites and Aletras, Nikolaos and Ma, Ning.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop and Aletras, Nikolaos and Ma, Ning

Reference 138

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.220598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.220598Z digest=sha256:105e1a14c4f8c1e3832c23d75a5b62ed8828a33799972e1a5194d7f130e611ee

Observation 7046bf1a-afc2-43c3-b9cf-0cafdae14f5b · outbound

This paper cites A Survey on Gender Bias in Natural Language Processing.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop A Survey on Gender Bias in Natural Language Processing

Reference 139

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.224833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.224833Z digest=sha256:e9f64d1e7d86145bb6d73c7c1bba8ff9114b270918d812b687e4ffac32c3a0bd

Observation 443dba07-8f5f-40a6-8d92-ae7ad2edc7fa · outbound

This paper cites Inducing Positive Perspectives with Text Reframing.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Inducing Positive Perspectives with Text Reframing

Reference 140

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.229554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.229554Z digest=sha256:252d6c14130bdf217beb85cd2eca7ab3f4ce66c37071110a2c7de2e5c35cfbae

Observation a71ace5c-74d1-4d18-9a9d-2fe24ea8f402 · outbound

This paper cites The Importance of Modeling Social Factors of Language: Theory and Practice.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop The Importance of Modeling Social Factors of Language: Theory and Practice

Reference 141

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.233451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.233451Z digest=sha256:1f2b9976e6a8a00b319d5701db68b098c29393e80a7f231590cd59b87b248f8e

Observation 1e9ce44b-5ca1-4dc6-b711-3998d2ef9c6d · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 142

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.237236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.237236Z digest=sha256:e6a8d82e0c4fad451bb5ebb13d00e895f6bc7224dc899ec50eb32fefb467ba97

Observation 258c868f-edfa-44fd-8a50-a06277381181 · outbound

This paper cites 2023 , howpublished =.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop 2023 , howpublished =

Reference 143

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.241762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.241762Z digest=sha256:5c74d49104bb0615654a739ba86443385a4d028da41ddf9c864a68a3c33d5ecc

Observation 7650e705-1670-4fde-aa0c-ef1a3c49304a · outbound

This paper cites GPT-4 Technical Report.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop GPT-4 Technical Report

Reference 144

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.246222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.246222Z digest=sha256:b94e6c63d189561692409c2250337279faa6c1380e40ef313927c07d1c39adec

Observation 4a4c4ba2-69bc-47c8-92bb-119dbf2d7bf4 · outbound

This paper cites Strength in Numbers: Estimating Confidence of Large Language Models by Prompt Agreement.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Strength in Numbers: Estimating Confidence of Large Language Models by Prompt Agreement

Reference 145

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.250929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.250929Z digest=sha256:e4f43fdd6541a9aee2a92a690743fb5718bb4cb014f71eb5c036354f940d7067

Observation 4617b8f3-ce41-4840-a759-dd30de6748b9 · outbound

This paper cites On the Intrinsic and Extrinsic Fairness Evaluation Metrics for Contextualized Language Representations.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop On the Intrinsic and Extrinsic Fairness Evaluation Metrics for Contextualized Language Representations

Reference 146

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.255122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.255122Z digest=sha256:1fad777d5e66eca9944724ddd337aff7ec50691949370586f9830e078441b880

Observation 9b6dd65d-2dd3-4dc9-b107-a6e7c1e3b72c · outbound

This paper cites Pay Attention to the Robustness of C hinese Minority Language Models! Syllable-level Textual Adversarial Attack on T ibetan Script.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Pay Attention to the Robustness of C hinese Minority Language Models! Syllable-level Textual Adversarial Attack on T ibetan Script

Reference 147

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.259117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.259117Z digest=sha256:734d917c77916dbfcdd40764b73bf2d64400d733e204e99bdde33e61379427a2

Pith citing papers

No inbound Pith citation observations are available.