Pith. sign in

Paper Citation Record · LEDGER

Benchmarking Detection Transfer Learning with Vision Transformers

As of 17 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2111.11429.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2111.11429 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T10:45:00.631913Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T14:28:32.127255Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 035853c0-aa10-45dc-8345-d22ba0cda705 · inbound

Adding Conditional Control to Text-to-Image Diffusion Models cites this paper.

Adding Conditional Control to Text-to-Image Diffusion Models Benchmarking Detection Transfer Learning with Vision Transformers

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-16T22:43:10.971064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-16T22:43:10.880338Z digest=sha256:572c029e23a42f0cd7fe7f4ebdac638a206bb0ba4268455cead88853fc35018f

Observation 0408340d-2dd4-42e7-9e2b-3a84518b0e99 · inbound

Robust Adaptation of Foundation Models with Black-Box Visual Prompting cites this paper.

Robust Adaptation of Foundation Models with Black-Box Visual Prompting Benchmarking Detection Transfer Learning with Vision Transformers

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-23T23:23:36.307944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-23T23:23:03.562550Z digest=sha256:bac09a08424b8dcdb7299b20c1b5ba7e57f30bebe36898e39633536994ba8a72

Observation ffe2ee62-2558-4fe2-9a91-6e2b07c7cf38 · inbound

Self-Supervised Learning for Real-World Object Detection: a Survey cites this paper.

Self-Supervised Learning for Real-World Object Detection: a Survey Benchmarking Detection Transfer Learning with Vision Transformers

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-23T19:03:21.313635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-23T19:02:11.415329Z digest=sha256:2507522d33c0c770f8e2c57f480bb5f52643f14ce565955eb19a8fd849ab6239

Observation 20b108e2-7e44-4e05-abee-833ed77af9d6 · inbound

Perception of Visual Content: Differences Between Humans and Foundation Models cites this paper.

Perception of Visual Content: Differences Between Humans and Foundation Models Benchmarking Detection Transfer Learning with Vision Transformers

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T10:45:00.631913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T10:45:00.631913Z digest=sha256:957fbe2f40acc962846fdbc6f59f1decaae9c244de0006e093297f1246ef3b71

Observation c072f7f4-658f-4ecb-99b7-25cf3fda2706 · inbound

Plancraft: an evaluation dataset for planning with LLM agents cites this paper.

Plancraft: an evaluation dataset for planning with LLM agents Benchmarking Detection Transfer Learning with Vision Transformers

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T23:21:35.055671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T23:21:35.055671Z digest=sha256:d7f86c25ff9925bb3c179d2543ba917f260ee490d68a36b531b779d11f157210

Observation 4b295d3c-d319-4c7e-9fed-ae83103af353 · inbound

A Separable Self-attention Inspired by the State Space Model for Computer Vision cites this paper.

A Separable Self-attention Inspired by the State Space Model for Computer Vision Benchmarking Detection Transfer Learning with Vision Transformers

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T22:26:10.615551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:26:10.615551Z digest=sha256:8393907bf3476ca6f9339155713b4d50a8d9e781957b10200412314afd0b4468

Observation aec7fc9d-356d-447b-897a-d1fd448a8163 · inbound

Parameter-Inverted Image Pyramid Networks for Visual Perception and Multimodal Understanding cites this paper.

Parameter-Inverted Image Pyramid Networks for Visual Perception and Multimodal Understanding Benchmarking Detection Transfer Learning with Vision Transformers

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-10T20:39:35.542024Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:39:35.542024Z digest=sha256:26a92d36d8e66bf9c2cbc02c9a5138f4da9683b333dd173b05e3994a3d98e89f

Observation f5da79d6-9ca5-4b34-b89b-d243c4f3392a · inbound

Towards more transferable adversarial attack in black-box manner cites this paper.

Towards more transferable adversarial attack in black-box manner Benchmarking Detection Transfer Learning with Vision Transformers

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:11.779207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:41:11.779207Z digest=sha256:b4fa7ca65d159a6dd30f6f81b471e09e9773d799161191ff94b4becd519a3d0d

Observation 23514c34-1a8e-4d4a-9c7e-d94883a6abcd · inbound

ViT-Split: Unleashing the Power of Vision Foundation Models via Efficient Splitting Heads cites this paper.

ViT-Split: Unleashing the Power of Vision Foundation Models via Efficient Splitting Heads Benchmarking Detection Transfer Learning with Vision Transformers

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T11:09:29.983561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:09:29.983561Z digest=sha256:77f19e9e2f38f8fe820374e3d0fb60f677efe1bac24cc0e9bc8ae1aa20936643

Observation ec110b89-46f8-47f9-8eb3-24092659f2d9 · inbound

MPT: Motion Prompt Tuning for Micro-Expression Recognition cites this paper.

MPT: Motion Prompt Tuning for Micro-Expression Recognition Benchmarking Detection Transfer Learning with Vision Transformers

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T21:05:12.990008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:05:12.990008Z digest=sha256:0a0f6629618c11cec1d7bd55b780bfedc181e0aa86179611491f624ef09f6c0e

Observation 8db358b3-08e6-4917-8f57-f906807425a8 · inbound

High-Speed Full-Color HDR Imaging via Unwrapping Modulo-Encoded Spike Streams cites this paper.

High-Speed Full-Color HDR Imaging via Unwrapping Modulo-Encoded Spike Streams Benchmarking Detection Transfer Learning with Vision Transformers

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-10T12:30:23.828688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-10T12:26:23.330051Z digest=sha256:5d4509549986f168555b1aa1dad2eb241035b1fdfbe19f4ac3c116bcab017b00

Observation edc35e0a-9445-4d58-bf6f-9bc4420c5c4b · inbound

HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers cites this paper.

HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Benchmarking Detection Transfer Learning with Vision Transformers

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T14:28:32.128646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-06-27T07:01:07.362430Z digest=sha256:5d6b2a06553584aa911f76daa23d7450aa012a3a04900ffd630490c3e2359642

Observation ab45c2d1-414b-41c5-9bfd-4eb025eb2c1e · inbound

Twins: Learn to Predict Unified Representations with Focal Loss cites this paper.

Twins: Learn to Predict Unified Representations with Focal Loss Benchmarking Detection Transfer Learning with Vision Transformers

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-01T04:29:45.886889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T04:29:45.886889Z digest=sha256:041c1624f7519e98fd0b140816c24e10f2311c44159fbfd23061b59de74fd6ad