Pith. sign in

Paper Citation Record · LEDGER

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling

As of 10 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 0 inbound Pith citation observations for arXiv:2502.02919.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.02919 v1

Coverage vector

measured 36 of 36 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T10:41:01.420244Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

36 of 36 outbound references displayed

  • verified exact0
  • verified fuzzy2
  • unresolved34
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e661855d-07eb-4557-b828-2c81dc1a4a5c · outbound

This paper cites Layer Normalization.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Layer Normalization

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.239337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.239337Z digest=sha256:bbf1793f46a3ea80a07d7edfc9e2d5d22d653d3d1359b93885e3dd73dfec0192

Observation ec137b26-5003-4bd1-b0fd-29606f6bfc7d · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-09T10:41:02.009344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T10:41:01.246337Z digest=sha256:0902142a8336e601d5bc86404b801f2808bde57d4b1193a9223930c49d96015e

Observation ec6f87cd-03b7-44fe-8263-539446193a3d · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.251977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.251977Z digest=sha256:cf46bc6cb928d4915d1056a6591088bccbfe80ff01b1ec3229347680165fbeb2

Observation 16c3790a-fb14-424e-8ea9-bc6217ab2907 · outbound

This paper cites J.; Jin, R.; and Shou, M.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling J.; Jin, R.; and Shou, M

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:41:01.982712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T10:41:01.257588Z digest=sha256:e6a6d0e808974a9289f03641f82b1b82e7affb5cff00959d6c95e60a3433a5d4

Observation 6ee06156-ff64-425f-b760-4c1971901098 · outbound

This paper cites MMDetection: Open MMLab Detection Toolbox and Benchmark.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling MMDetection: Open MMLab Detection Toolbox and Benchmark

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.262696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.262696Z digest=sha256:7f7ffb02dee2b588359510a83239d949a7711b798261b5366d7471b2849b992f

Observation 963676bc-1edb-48fd-922d-7aef6b7701bd · outbound

This paper cites Vision Transformer Adapter for Dense Predictions.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Vision Transformer Adapter for Dense Predictions

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.268204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.268204Z digest=sha256:98fa988e4a858312a75944eab88190d973a1ea6d6090d9aaedbd7398423a2960

Observation eefb3eef-0b7d-4a45-8348-3a7a6d85e486 · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-09T10:41:01.966546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T10:41:01.274027Z digest=sha256:8f68c94de0415d67e7f946db634b5ff81a93afb0c1442a0954f2849a19705430

Observation 2a89c1d7-1ebc-4665-b812-c104c6b8916a · outbound

This paper cites Conditional Positional Encodings for Vision Transformers.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Conditional Positional Encodings for Vision Transformers

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.278794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.278794Z digest=sha256:e8ad00754bd5be61c638ceea78e2190cf07378bf10aeac54eb95b6e0b5ed0f6d

Observation 0e20669a-ab27-43aa-bf4a-b58d34e2d938 · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-09T10:41:01.949151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T10:41:01.284003Z digest=sha256:5cd6dfa00436c6b1743705bd8bf5fd648df21411769088e9143717dcf8a9cafa

Observation f6cab2d6-8f57-4e7f-9bce-75c1ea147823 · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.288513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.288513Z digest=sha256:8c5ac7372cdc37eec0ed7c86c0f2261b148b7b25c107152b763f9008f9a7ae8c

Observation 591e8f8a-cc38-47e5-9a9f-73e7ee6af321 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.293518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.293518Z digest=sha256:9e476c617f3646c6600a91559e1c7453fb33c8943676c1d6aab31fe739911b3d

Observation 717870e9-5ab3-4364-a5f6-c01970772fe0 · outbound

This paper cites Escaping the Big Data Paradigm with Compact Transformers.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Escaping the Big Data Paradigm with Compact Transformers

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.298875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.298875Z digest=sha256:6135cfddb8332060547ba068681adef5c1cb591029c73f7eff3dbf2f2bc43b21

Observation 2068eef1-69cf-4add-b4fb-0f2d8febc9c1 · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.303810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.303810Z digest=sha256:d23cc889f9575d9981da74ae7c0d213675f348e9f219b4f9fce8d492cbcb77de

Observation f6034912-b7f0-4ccc-b819-7b146fb90ef1 · outbound

This paper cites Rotary Position Embedding for Vision Transformer.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Rotary Position Embedding for Vision Transformer

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.308538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.308538Z digest=sha256:4091fd83206ee1f3d8ece041e1f107b20890d0bc970a25bd1c6dc0740da5fef9

Observation 0df307c7-248e-4f89-b52d-40080797030a · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.313341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.313341Z digest=sha256:d6c54eff2f3d46e01af07f10606524328be22728508bd4a80d67223731bc1baa

Observation de472f31-358d-4b32-9ad3-a6e256d5b830 · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.318568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.318568Z digest=sha256:1fac1416691ed8fa1dd560355f914e61f3dbcabb8f835de08fbd0afa3099bf80

Observation 21c0907a-0ba6-42d3-8c64-ad4e73036f87 · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.323421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.323421Z digest=sha256:546e6f71fa1e9d33774683b6445ce26bb40cbe1123cc65b2ff74072cfa3bfe61

Observation 5a1140ae-6822-40c5-b119-1a01ee5eb784 · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.328114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.328114Z digest=sha256:20ee0b9c7e78b2c1c4db5a6dab5ce3f5cbd49aeda686413274359e3f008c0127

Observation 47b62d92-d59e-4398-9df3-b27b436c0e61 · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.332753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.332753Z digest=sha256:f8fa98673597fd6424d2670063bc9917efdfbd6bbd34de3d073538a3d4b852db

Observation 14eaa9f1-5f2c-4801-9767-c27a832e012b · outbound

This paper cites Self-Attention with Relative Position Representations.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Self-Attention with Relative Position Representations

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.337970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.337970Z digest=sha256:e2e2f7850394279f474acc67292e8e64c2b74d42a9d830c410fb8916a9762fc0

Observation f0fedcb6-1420-47ea-96b3-e9e676f782b9 · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.343091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.343091Z digest=sha256:b3c5936c7cee371893e9154368a208c496f92c93d3f3c438e55f139077a0f05b

Observation 487896cd-b4f6-4c1d-b708-342e75481ac1 · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.347656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.347656Z digest=sha256:7b9a0a5d995410a7f2fbe37b68519e5a48612e23dc9a8cb2aaf2a85f2ab1a092

Observation e051b9ba-25d3-496d-825d-f5af358bf046 · outbound

This paper cites N.; Kaiser, .; and Polosukhin, I.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling N.; Kaiser, .; and Polosukhin, I

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.355182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.355182Z digest=sha256:ec85b5988a7965ddd259f2d60655fab62a3e82e5ed143eb25902ccd33eb5b300

Observation 85715c03-046f-4f1c-a2f1-ebb7ab802e61 · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-09T10:41:01.815849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T10:41:01.360381Z digest=sha256:87cb897f5e616be9df649bd5121dc6cf1ef9e153f9b51f725d939d663ee241d3

Observation f91d2907-7005-45c4-af58-0edc020bdebb · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-09T10:41:01.797329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T10:41:01.365788Z digest=sha256:7a74448160b8d4b91b570b60e3aeb27d0dca0f3fe0ec206ed4f72789874fc0f8

Observation 81bbc29a-c45d-4554-9e30-7dffb78d04f7 · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-09T10:41:01.779959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T10:41:01.370892Z digest=sha256:9c8ce7037d81e7ff7a6c14c26db8aa2b69c631d681a037988c57da2a94c4442c

Observation d8969e2a-c39c-4bc5-a0a5-1f28b1d91e35 · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-09T10:41:01.762467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T10:41:01.375825Z digest=sha256:b9c260091bbc208f7ca16d4807bcc03bdf1c4de01ff7db16625d5647dade3e7a

Observation ee22e8d4-7482-4705-8de2-fd414c093775 · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-09T10:41:01.742833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T10:41:01.380740Z digest=sha256:f41b02a668118004f40709dce4f96138865ac7717165401c98f63a06c690ecca

Observation 7903436a-3741-458a-9701-761b25010063 · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-09T10:41:01.724563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T10:41:01.385331Z digest=sha256:5c7d0f283767cd746ac95d3223d25fc4477765992022fa579df6d10e52ecb5eb

Observation 4ecedc5b-3b8d-46ea-bf31-0f6c09762449 · outbound

This paper cites E.; Feng, J.; and Yan, S.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling E.; Feng, J.; and Yan, S

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T10:41:01.708349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T10:41:01.390265Z digest=sha256:32ff78a1bc22ca647eb7beb970b6c7ebbf3aa4950d1cd94a0b03113c7118922f

Observation 74b25d0e-f68f-4fb9-9f03-765cd1e38a1b · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-09T10:41:01.692463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T10:41:01.394845Z digest=sha256:96237649826ba11d7073701e076be1fefda0a1befb2e871df552d72a40cedbcb

Observation b5f0ee3c-3f15-45f1-a2e3-e3ca2ab611fa · outbound

This paper cites H.; et al.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling H.; et al

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.399466Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.399466Z digest=sha256:d4db1a064527d293e03d0b9f563688048d7dadab295640cfc8b0963af0455bac

Observation de705d0b-0541-444a-ac28-c90a3343f01d · outbound

This paper cites an unresolved cited work.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-09T10:41:01.664832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-09T10:41:01.404419Z digest=sha256:e3e5c1e144d3d4307320d6f67e0f3d664a5d51ae3cfee6e1f128a439b6da87f4

Observation f502618b-00cb-4b64-a0d0-66b3b71f479f · outbound

This paper cites Deformable DETR: Deformable Transformers for End-to-End Object Detection.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling Deformable DETR: Deformable Transformers for End-to-End Object Detection

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.409443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.409443Z digest=sha256:fa04d5e4a2f32d336023cecd859bcdfeca6170c428c1b6887796138ea1713c33

Observation 6cdc534d-677f-4fea-80a0-f0de778b9d9a · outbound

This paper cites , " * write output.state after.block = add.period write newline.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling , " * write output.state after.block = add.period write newline

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.414853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.414853Z digest=sha256:0d66394b634c3f598bd95b8a2cdda32c1305154234d56207997ecd3e6ada1cd6

Observation 4efd5ca3-de0c-45f2-9cb1-559c22e23405 · outbound

This paper cites write newline.

Maximizing the Position Embedding for Vision Transformers with Global Average Pooling write newline

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-09T10:41:01.420244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:41:01.420244Z digest=sha256:ba77c95f232ec06c1915c4afc5b2122b7293f6f9525df87a3d10063ace3e0f52

Pith citing papers

No inbound Pith citation observations are available.