Pith. sign in

Paper Citation Record · LEDGER

On Identifiability in Transformers

As of 20 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 7 inbound Pith citation observations for arXiv:1908.04211.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1908.04211 v4

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T13:58:53.311191Z

measured 56 of 56 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T19:00:37.098270Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-22T23:42:16.124037Z

Reference resolution

49 of 49 outbound references displayed

  • verified exact2
  • verified fuzzy22
  • unresolved24
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f74aad54-9e40-43cc-ada8-6b5ce92a081f · outbound

This paper cites A BERT Baseline for the Natural Questions.

On Identifiability in Transformers A BERT Baseline for the Natural Questions

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-14T13:58:53.133516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T13:58:53.133516Z digest=sha256:e2dd9b1053c10e398f9f5510e3b249cd7ee2cbffef8aa39d3fe319b0e0f6e90d

Observation cd3a8c9e-17fa-440d-b038-522b43084f5c · outbound

This paper cites Do Transformer Attention Heads Provide Transparency in Abstractive Summarization?.

On Identifiability in Transformers Do Transformer Attention Heads Provide Transparency in Abstractive Summarization?

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-14T13:58:53.630069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T13:58:53.138076Z digest=sha256:50198d654641be936dab970287c5ae12737536adce2be87f8a356014200ff708

Observation b7bc76cd-955f-41e3-9e61-8f8a0c4bcfcd · outbound

This paper cites Neural Machine Translation by Jointly Learning to Align and Translate.

On Identifiability in Transformers Neural Machine Translation by Jointly Learning to Align and Translate

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-14T13:58:53.142812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T13:58:53.142812Z digest=sha256:0e9e054c32935a25738675e93e1b9720ce8c87df0e31270356ba39837afaa0d2

Observation b0f1f77b-aa15-42b2-abd9-3c69d839bb78 · outbound

This paper cites Bellman and Karl Johan str \"o m.

On Identifiability in Transformers Bellman and Karl Johan str \"o m

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:58:53.931375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T13:58:53.146950Z digest=sha256:aa9cb07d150c8199fc072188a8867ac9c007c53f9f7c7a9a43f2500905ab8e1c

Observation 1f8506f2-d3f2-4c2a-9beb-9fdc8f5fc5d1 · outbound

This paper cites Brown, Stephen A.

On Identifiability in Transformers Brown, Stephen A

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:58:53.920963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T13:58:53.150942Z digest=sha256:24693e201817e2ccb8061d7446943fefa4fae636e8b091d16c88037da9c07b33

Observation 4b0ee127-7aa5-422b-8455-e178b26a6d88 · outbound

This paper cites What Does BERT Look At? An Analysis of BERT's Attention.

On Identifiability in Transformers What Does BERT Look At? An Analysis of BERT's Attention

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-14T13:58:53.154824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T13:58:53.154824Z digest=sha256:84f0bcecf8e4c6fe52ca2152736fd9910db9595e6a19f7cf8cedb042450d596f

Observation 35acef93-5806-4b1a-9433-ddd044ce590f · outbound

This paper cites Visualizing and Measuring the Geometry of BERT.

On Identifiability in Transformers Visualizing and Measuring the Geometry of BERT

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-14T13:58:53.159141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T13:58:53.159141Z digest=sha256:ffa3fe1769456aff8499aa2b4d971fa8126af686e1ef609094a0b9189e7d591a

Observation e41f26ad-837d-4f72-bfef-0b5ca326c92a · outbound

This paper cites Universal transformers.

On Identifiability in Transformers Universal transformers

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:58:53.911104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T13:58:53.163309Z digest=sha256:3633474e0d5ec9d816aa9272088c74950b2e7d649ed7659cd495ef57627d7dae

Observation 73bc7a39-c9db-487b-9fce-35ab09ce1d16 · outbound

This paper cites BERT: pre-training of deep bidirectional transformers for language understanding.

On Identifiability in Transformers BERT: pre-training of deep bidirectional transformers for language understanding

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:58:53.901087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T13:58:53.166831Z digest=sha256:ebf2bddb933649256802776092983fa56a071f1314b1ad31d717ae441d74ee45

Observation 59a37f41-52bb-4dc6-97ec-4d9f483c3164 · outbound

This paper cites Dolan and Chris Brockett.

On Identifiability in Transformers Dolan and Chris Brockett

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:58:53.890896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T13:58:53.170097Z digest=sha256:b7fd436be6700a759447b68e97a5d6b2f4e57b15fc2cfac29936c56c40d4e6fe

Observation 94bf0c2b-7380-44db-a562-c2f2e83786ed · outbound

This paper cites Understanding the difficulty of training deep feedforward neural networks.

On Identifiability in Transformers Understanding the difficulty of training deep feedforward neural networks

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:58:53.879633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T13:58:53.173791Z digest=sha256:b0a64d04ddbdd530086aa08835a4600c86e31c02341bd08890bf12a805d59738

Observation 75dee84f-6cb6-47d8-9c08-9a1ccf262815 · outbound

This paper cites Gaussian Error Linear Units (GELUs).

On Identifiability in Transformers Gaussian Error Linear Units (GELUs)

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-14T13:58:53.177287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T13:58:53.177287Z digest=sha256:a410989fd61788cf3d3c17a44cf52882a546e17f65729d371cd7310f39370bc5

Observation ae400b3d-f2b3-4065-baea-058580f47dff · outbound

This paper cites an unresolved cited work.

On Identifiability in Transformers Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-14T13:58:53.868772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T13:58:53.181149Z digest=sha256:485e07a5b4b8e570d0bea860a8603fbcea23b9627fa8a2008f330e0f46f36c96

Observation 1687f85c-0f52-4f23-92df-9cabe680fc84 · outbound

This paper cites an unresolved cited work.

On Identifiability in Transformers Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-14T13:58:53.858146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T13:58:53.184493Z digest=sha256:aeb0145201311593a85d378165d74c253339e7ceaf8341c86beb9b1f2100f804

Observation d346aa0b-67a1-4b1e-b654-5a45d9998ef5 · outbound

This paper cites Microsoft translator at wmt 2019: Towards large-scale document-level neural machine translation.

On Identifiability in Transformers Microsoft translator at wmt 2019: Towards large-scale document-level neural machine translation

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:58:53.847358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T13:58:53.187860Z digest=sha256:e9d6b08eb123a291d25efd947fe90d9038387c5bc49935fb0ae9786d0c987374

Observation 8a29dfe3-6757-43f6-8d38-7acdff17f984 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

On Identifiability in Transformers Adam: A Method for Stochastic Optimization

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-14T13:58:53.191205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T13:58:53.191205Z digest=sha256:5de3eb386b64ab083b2005deb274bcdba67c71bc9610bd00775e1a0279347848

Observation eeae91fa-05c7-4f21-a4bf-796d73a722aa · outbound

This paper cites Attention is (not) all you need for commonsense reasoning.

On Identifiability in Transformers Attention is (not) all you need for commonsense reasoning

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:58:53.836629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T13:58:53.194762Z digest=sha256:737809a9044a76aaade2ff01b4ba5bbc4406825b7463fcec64149e893c199fb6

Observation ec4fd3ad-5e73-4d0a-b51f-78a67866d002 · outbound

This paper cites Albert: A lite bert for self-supervised learning of language representations.

On Identifiability in Transformers Albert: A lite bert for self-supervised learning of language representations

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-14T13:58:53.198150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T13:58:53.198150Z digest=sha256:a542c42412a4e99bc3416769e6c809a889e3e9b3a4a9f70024a7de83e042caac

Observation ce65e208-cac8-497c-a133-7b2d52bdcc5f · outbound

This paper cites Open Sesame: Getting Inside BERT's Linguistic Knowledge.

On Identifiability in Transformers Open Sesame: Getting Inside BERT's Linguistic Knowledge

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-14T13:58:53.201568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T13:58:53.201568Z digest=sha256:31d6655c11540008028f06c8d1a125c41d03283afa90339853c3197dfeaab1b0

Observation d12daa22-ff2d-4e85-9343-03f53a22e7ff · outbound

This paper cites RoBERTa: A Robustly Optimized BERT Pretraining Approach.

On Identifiability in Transformers RoBERTa: A Robustly Optimized BERT Pretraining Approach

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-14T13:58:53.205312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T13:58:53.205312Z digest=sha256:a3650444457805879ed7ffb77b91e403e2b4c7a892970213fc6629599ca465af

Observation 1726a06c-9233-46c6-931a-7b1bd7d16523 · outbound

This paper cites Extracting syntactic trees from transformer encoder self-attentions.

On Identifiability in Transformers Extracting syntactic trees from transformer encoder self-attentions

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:58:53.819541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T13:58:53.208802Z digest=sha256:30bcc8be4ecb31e272cd53a2d51de59ff585621f107167fae833491a233426fb

Observation 83a23a39-9870-4c73-8adf-e808b850c159 · outbound

This paper cites Are Sixteen Heads Really Better than One?.

On Identifiability in Transformers Are Sixteen Heads Really Better than One?

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-14T13:58:53.211989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T13:58:53.211989Z digest=sha256:ce9bfa9d8338d6352254253e25cab13e0ea8ec8cd4f7f0105fc29d52df66d07e

Observation c9097d92-760f-4672-86f2-a8aa80daf38a · outbound

This paper cites Investigating the Successes and Failures of BERT for Passage Re-Ranking.

On Identifiability in Transformers Investigating the Successes and Failures of BERT for Passage Re-Ranking

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-14T13:58:53.215593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T13:58:53.215593Z digest=sha256:2f08917cd25ddb851e645599e0ebfb05a402b4f78d160d503d472d1e9f865ccc

Observation 7f1d55dc-4a08-4c2d-9638-926ef3b9955f · outbound

This paper cites Peters, Mark Neumann, Luke Zettlemoyer, and Wen - tau Yih.

On Identifiability in Transformers Peters, Mark Neumann, Luke Zettlemoyer, and Wen - tau Yih

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:58:53.809048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T13:58:53.219278Z digest=sha256:d04a71b5c8e9ddc597336698f24ba4ebc16b918955f5f1c0fedf6ff87809c1d7

Observation 983e18b8-ac7a-4d37-aafb-cf9dfd4c1b9c · outbound

This paper cites o rner, Hinrich Sch \.

On Identifiability in Transformers o rner, Hinrich Sch \

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:58:53.799188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T13:58:53.222736Z digest=sha256:7e1f44b7bb76a249367c8bc4ff7176e4e9db258e5b81764e29977d10e5459df7

Observation 1b488205-59dd-4b03-9bbd-983fdfcacb18 · outbound

This paper cites Learning to Deceive with Attention-Based Explanations.

On Identifiability in Transformers Learning to Deceive with Attention-Based Explanations

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-14T13:58:53.226145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T13:58:53.226145Z digest=sha256:d0c859f4c8c5a6b1ffd0cbcde19486d7f844d96048980e48bb89e80a81ff4975

Observation 2aba314c-3256-401a-bd92-0d6ae8152117 · outbound

This paper cites Improving language understanding by generative pre-training.

On Identifiability in Transformers Improving language understanding by generative pre-training

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:58:53.788562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T13:58:53.230365Z digest=sha256:7a1f9808730082b2c56e71cfa1e10d6efa6b19ec5a307945cde7ca18347af0d0

Observation 8c303378-3295-4f0d-8712-1c01503568b5 · outbound

This paper cites Language models are unsupervised multitask learners.

On Identifiability in Transformers Language models are unsupervised multitask learners

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-14T13:58:53.233924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T13:58:53.233924Z digest=sha256:1d0e8913435119ea2546ec438468415913405be923a348b5e97e966c5ef3f862

Observation b0dd69b5-9d9c-4a34-8635-3545333f2194 · outbound

This paper cites An analysis of encoder representations in transformer-based machine translation.

On Identifiability in Transformers An analysis of encoder representations in transformer-based machine translation

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:58:53.772535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T13:58:53.237303Z digest=sha256:a414ceece0f85c9ed41e75d672d6bf1d8de00dc1117a07d2fbefa1ba7a26b9a0

Observation cbdcabc2-146d-405d-9a84-621836929db3 · outbound

This paper cites an unresolved cited work.

On Identifiability in Transformers Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-14T13:58:53.762475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T13:58:53.240733Z digest=sha256:8f4c6eab8713c8899adbe8e3a0e5bd400fce92e994ac4f3b710372dd9f1de18c

Observation f3055383-139a-4d7b-865a-a6b9ed834aa9 · outbound

This paper cites Deep inside convolutional networks: Visualising image classification models and saliency maps.

On Identifiability in Transformers Deep inside convolutional networks: Visualising image classification models and saliency maps

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:58:53.752493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T13:58:53.243984Z digest=sha256:04abbfb02c54e47df4852b199a331b162dd4d162066b8f813cd4002fe8d59f79

Observation 51e51d98-2b42-4208-a2cc-40c50052d828 · outbound

This paper cites An analysis of attention mechanisms: The case of word sense disambiguation in neural machine translation.

On Identifiability in Transformers An analysis of attention mechanisms: The case of word sense disambiguation in neural machine translation

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:58:53.741539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T13:58:53.247752Z digest=sha256:9e60ae770681de00fff7d6da8f759263a143a125adac55bdf94349d57884ebe6

Observation fc54cb90-280b-4bab-97bc-fa315fd64163 · outbound

This paper cites BERT rediscovers the classical NLP pipeline.

On Identifiability in Transformers BERT rediscovers the classical NLP pipeline

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:58:53.731259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T13:58:53.251204Z digest=sha256:5c16f3bc6b3e8e090775a314e9eee211015ff803d452373686d5017a4b280b02

Observation 8ebab20b-a393-4044-b301-1abe91e82fac · outbound

This paper cites Manning, and Yoram Singer.

On Identifiability in Transformers Manning, and Yoram Singer

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:58:53.720737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T13:58:53.254900Z digest=sha256:a53830792346291d6dd71695353238c5b584945dc69841fa5540a10b1577f102

Observation 8ee805d1-e827-4277-a37d-ae2c76c33ad3 · outbound

This paper cites Gomez, Lukasz Kaiser, and Illia Polosukhin.

On Identifiability in Transformers Gomez, Lukasz Kaiser, and Illia Polosukhin

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:58:53.709245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T13:58:53.258734Z digest=sha256:49ef41a5edc288747727b2f72ee6748f2b01593b2d86d1e98d810a93afa475f9

Observation 8cc4927c-8c0a-4bff-b4f7-a3040913da9f · outbound

This paper cites Visualizing Attention in Transformer-Based Language Representation Models.

On Identifiability in Transformers Visualizing Attention in Transformer-Based Language Representation Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-14T13:58:53.263300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T13:58:53.263300Z digest=sha256:081a9242e5c7f403f6c3b787e15f56dc5be5ff684d56fd29431f238da1af26d2

Observation 59fe64bb-d723-4d37-abf4-51697592ae06 · outbound

This paper cites Analyzing multi-head self-attention: Specialized heads do the heavy lifting, the rest can be pruned.

On Identifiability in Transformers Analyzing multi-head self-attention: Specialized heads do the heavy lifting, the rest can be pruned

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:58:53.699096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T13:58:53.267102Z digest=sha256:e6c46d5e8a2d8ca048047d78f6b96f795e200212f523c2f7c93beebc02e6aabf

Observation 911b3616-1320-4eee-8282-ced3bb6b8995 · outbound

This paper cites Attending to Mathematical Language with Transformers.

On Identifiability in Transformers Attending to Mathematical Language with Transformers

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-08-14T13:58:53.516256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T13:58:53.270806Z digest=sha256:905a1d89f37f9ab733945f9ede03065b5ba7982a6ade3c8db9c68e52cbb52dee

Observation 6927b9a2-8151-4cfb-8842-5a86d5c179b8 · outbound

This paper cites Neural Network Acceptability Judgments.

On Identifiability in Transformers Neural Network Acceptability Judgments

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-14T13:58:53.274383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T13:58:53.274383Z digest=sha256:60eb3e7210df3ba682db9f1b874435b6bb324da3a0f8a20c872efd8075faf349

Observation 15d4a6e4-ff5e-4b4b-8fb6-8f0d3d46a8e3 · outbound

This paper cites Attention is not not Explanation.

On Identifiability in Transformers Attention is not not Explanation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-14T13:58:53.278844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T13:58:53.278844Z digest=sha256:2351bec21791308879661c3f1faf3bce6f0df28cd2045eff4b95ca3ab32f6b25

Observation 7a016438-3966-4496-bc6c-6986f7c23240 · outbound

This paper cites A broad-coverage challenge corpus for sentence understanding through inference.

On Identifiability in Transformers A broad-coverage challenge corpus for sentence understanding through inference

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:58:53.688166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T13:58:53.282274Z digest=sha256:b1bdd35f7227985148eb42edf79eae24a0b2c8a075dd474a2994017cf48aa1b3

Observation 00d863bf-d16e-4862-9feb-c0e0cd2de30e · outbound

This paper cites Wong, Fandong Meng, Lidia S.

On Identifiability in Transformers Wong, Fandong Meng, Lidia S

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:58:53.677922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T13:58:53.285788Z digest=sha256:86c031103c5e48480ed45d9f571f0fa28379a16140e2e1aa6a28d36978dce760

Observation b826a9c6-c782-4432-98aa-dde59e02080e · outbound

This paper cites XLNet: Generalized Autoregressive Pretraining for Language Understanding.

On Identifiability in Transformers XLNet: Generalized Autoregressive Pretraining for Language Understanding

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-14T13:58:53.289063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T13:58:53.289063Z digest=sha256:23fcd4e4c243d8575dc71bea7a12f9449232370da0d626d0488c9e6bee7a6fbd

Observation 90a6031b-4205-498c-87b8-53be3df748cf · outbound

This paper cites Xlnet: Generalized autoregressive pretraining for language understanding.

On Identifiability in Transformers Xlnet: Generalized autoregressive pretraining for language understanding

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:58:53.667814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-14T13:58:53.292666Z digest=sha256:ab409460c06bbdb6e583a7d293403f93c2f722bec40fa917a345e9e93d322280

Observation 45cae2cc-f9d7-46ed-a4b8-7fbbda3de17c · outbound

This paper cites Adding Interpretable Attention to Neural Translation Models Improves Word Alignment.

On Identifiability in Transformers Adding Interpretable Attention to Neural Translation Models Improves Word Alignment

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-14T13:58:53.296228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T13:58:53.296228Z digest=sha256:953d25553829470bf2052ad5de94ce504497643043031077bc55993d8a56af69

Observation 34e9c8d9-06e9-4ad6-b6bc-c94a23fc232a · outbound

This paper cites write newline.

On Identifiability in Transformers write newline

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-14T13:58:53.299911Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T13:58:53.299911Z digest=sha256:0284cb43019804caac7a90b75d642975674a93780ba7ac74e963202dc5e2c7e2

Observation 83812952-4fdd-469a-bedf-6906bfbd2bbc · outbound

This paper cites @esa (Ref.

On Identifiability in Transformers @esa (Ref

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-14T13:58:53.304061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T13:58:53.304061Z digest=sha256:c055aef5c5d2d89b68d515ca4b184e9c434edf11a0a759cdc1a0846f2bd4853a

Observation dc48f631-c6f1-4467-8871-49683fdaa021 · outbound

This paper cites an unresolved cited work.

On Identifiability in Transformers Unresolved cited work

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-14T13:58:53.307636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T13:58:53.307636Z digest=sha256:e06cd5ff2097233a553bee8184689132e2f8b4226ffa7e44b73d2bc78f035ffa

Observation d5419d0f-fed3-48eb-ae8b-9712cd5eb8c9 · outbound

This paper cites 妤 ' x P·uy n gʅ8Oj D q ^hژ.

On Identifiability in Transformers 妤 ' x P·uy n gʅ8Oj D q ^hژ

Reference 49

Resolution
malformed identifier
no resolver link, observed 2026-08-14T13:58:53.311191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T13:58:53.311191Z digest=sha256:7fd250eceb1a5b902d3e3ea939d7442afb02f1dfe7c506de572300262aefd7aa

Pith citing papers

Observation 175b10e9-8a08-4553-8fdc-777b0e4bd12a · inbound

TabTransformer: Tabular Data Modeling Using Contextual Embeddings cites this paper.

TabTransformer: Tabular Data Modeling Using Contextual Embeddings On Identifiability in Transformers

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-16T21:32:31.296446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-16T21:32:31.241434Z digest=sha256:59b7894d7da71ff0a8ba5573d1f590a5486d0d7c3d7f64dd912785c866ffffe8

Observation 888e2002-2eb8-4f1a-9065-5696c05af852 · inbound

Enabling Global, Human-Centered Explanations for LLMs:From Tokens to Interpretable Code and Test Generation cites this paper.

Enabling Global, Human-Centered Explanations for LLMs:From Tokens to Interpretable Code and Test Generation On Identifiability in Transformers

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-22T23:42:16.127211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-22T23:41:22.017848Z digest=sha256:cf0422069e29cb4719e64043cc7f93fba1e5b2770ca7b0befb5a1df083500a23

Observation 6080de41-71b7-40f1-87dc-b490649218c7 · inbound

Probing the Embedding Space of Transformers via Minimal Token Perturbations cites this paper.

Probing the Embedding Space of Transformers via Minimal Token Perturbations On Identifiability in Transformers

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-15T19:00:37.098270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:00:37.098270Z digest=sha256:fc746f8ccefad770767dbf6fc57487ad82737e262a353f6585ec57b66013390c

Observation 5bd43240-299f-460b-b5ac-184373dd68f9 · inbound

Towards Transparent AI: A Survey on Explainable Large Language Models cites this paper.

Towards Transparent AI: A Survey on Explainable Large Language Models On Identifiability in Transformers

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T22:22:38.869579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:22:38.869579Z digest=sha256:384cc8ba37e1eec73de4eff17dcfe62134bd2f8a31a311fb3cca776ef39a19b5

Observation 1554c404-456a-4417-bf34-18971ceceec8 · inbound

PLEX: Perturbation-free Local Explanations for LLM-Based Text Classification cites this paper.

PLEX: Perturbation-free Local Explanations for LLM-Based Text Classification On Identifiability in Transformers

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T18:08:20.728496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:08:20.728496Z digest=sha256:b5a45193758d290755a340ac2272dda4b5769e43fe792f4302b85622bb4c06e5

Observation 8d1c3292-16c8-45b7-b35b-d0d3fd1a8c07 · inbound

Provably Learning Multi-Head Attention with Queries cites this paper.

Provably Learning Multi-Head Attention with Queries On Identifiability in Transformers

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T21:45:27.642304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:45:27.642304Z digest=sha256:b9010cc65da56e6cdbee3c6e8eb7a26f46419e29d5b49406e92054c30017d830

Observation 6fc581b4-b572-4423-82a8-e8121b416603 · inbound

Provably Learning Multi-Head Attention with Queries cites this paper.

Provably Learning Multi-Head Attention with Queries On Identifiability in Transformers

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-10T04:35:40.561992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T04:35:40.561992Z digest=sha256:91afd426972a7b461da4bda590bab954f2dcf244cc4fd469737be42c84e20d8e