Pith. sign in

Paper Citation Record · LEDGER

Deeper Inside Deep ViT

As of 15 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 0 inbound Pith citation observations for arXiv:2508.04181.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.04181 v1

Coverage vector

measured 32 of 32 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T00:55:21.245290Z

measured 32 of 32 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

32 of 32 outbound references displayed

  • verified exact3
  • verified fuzzy12
  • unresolved17
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ce9b271e-cd5b-4d9d-940e-c775af7487cb · outbound

This paper cites Layer Normalization.

Deeper Inside Deep ViT Layer Normalization

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T00:55:18.323666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:55:18.323666Z digest=sha256:1670ebe219dd7f4ff88c915da3e1126d18b49c3309fc3473a339afd499bce81e

Observation d537ff11-9f77-447a-afdb-510a52e6cf7a · outbound

This paper cites Relational inductive biases, deep learning, and graph networks.

Deeper Inside Deep ViT Relational inductive biases, deep learning, and graph networks

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T00:55:18.394404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:55:18.394404Z digest=sha256:5a6e7f98aa8e89b1e82148801043d20c647105dc35f46ba6aa86d31213a86468

Observation 4f6f39e6-7e90-44b9-ad55-4f8a6e01e562 · outbound

This paper cites On the Opportunities and Risks of Foundation Models.

Deeper Inside Deep ViT On the Opportunities and Risks of Foundation Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T00:55:18.484513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:55:18.484513Z digest=sha256:ee1e4793b8ee8f825b0165a4ab1a8425b2b05fb106a8f48b9072045f2c66b985

Observation aa14b0b7-5eeb-4708-b860-50af89a7d2fb · outbound

This paper cites Generative Adversarial U-Net for Domain-free Medical Image Augmentation.

Deeper Inside Deep ViT Generative Adversarial U-Net for Domain-free Medical Image Augmentation

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-08-06T00:55:21.739413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T00:55:18.623767Z digest=sha256:1bbd685365665d2205865b7f5038e9ee25081876c5e3c0e4bee09d87b2ee0f15

Observation a5a602c9-a8e9-49a9-9edf-510b727b5714 · outbound

This paper cites Palm: Scaling language modeling with pathways.

Deeper Inside Deep ViT Palm: Scaling language modeling with pathways

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T00:55:18.690708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:55:18.690708Z digest=sha256:691ff1015936699fdd102f10007a4c37c20b5e3b5a2721c3e29eeef6a6c12976

Observation 01f3bddd-061e-4613-a21c-1a5084f7a61b · outbound

This paper cites Scaling vision transformers to 22 billion parameters.

Deeper Inside Deep ViT Scaling vision transformers to 22 billion parameters

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T00:55:18.792634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:55:18.792634Z digest=sha256:1bdd5f9570390f43afdfae541ee0e6e9840aa78003022c1c17af590ec843312c

Observation 83ed8d73-6c15-47e0-b8a7-f963ee453f4a · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Deeper Inside Deep ViT An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T00:55:18.861097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:55:18.861097Z digest=sha256:5ff4602ec7440c93ae3029a6fb9509fe63024874a51a8f8fccb0255b5c83acc7

Observation cf9adae1-dc56-48e0-8c76-3ba137bf0695 · outbound

This paper cites Convit: Improving vision transformers with soft convolutional inductive biases.

Deeper Inside Deep ViT Convit: Improving vision transformers with soft convolutional inductive biases

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:55:23.676642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T00:55:18.934779Z digest=sha256:539e0923e0fd31614fb239906c689eed54392ac1fd314660d6516c50a7a83899

Observation c4cc6f83-8ac1-4b2c-8a5d-0ae99cbf394a · outbound

This paper cites Taming transformers for high-resolution image synthesis.

Deeper Inside Deep ViT Taming transformers for high-resolution image synthesis

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T00:55:19.028566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:55:19.028566Z digest=sha256:b41ca2671b2558d9f32396400569358af9eaa2a2f83dc165b77b427ade233cdb

Observation 7f89c464-df8a-4d16-abb7-8ad516a64c4e · outbound

This paper cites Se (3)-transformers: 3d roto-translation equivariant attention networks.

Deeper Inside Deep ViT Se (3)-transformers: 3d roto-translation equivariant attention networks

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:55:23.530908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T00:55:19.101598Z digest=sha256:720ab584ae36fdf392c26b3b98886cd4f20dab5c09cb4f2ba18f30763a1906c7

Observation 5198433f-80c1-47f4-95ca-d5f9a2595d80 · outbound

This paper cites Intriguing properties of transformer training instabilities, 2023.

Deeper Inside Deep ViT Intriguing properties of transformer training instabilities, 2023

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:55:23.317587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T00:55:19.212858Z digest=sha256:50c84227abf1413a4b68c9eeb175f12e923dcbc877dc1cd4d19749dd3153d8f4

Observation 8f7ce89c-cd89-4e00-a211-f16bec62d09e · outbound

This paper cites Image-to-image translation with conditional adversarial networks.

Deeper Inside Deep ViT Image-to-image translation with conditional adversarial networks

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T00:55:19.276806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:55:19.276806Z digest=sha256:0da6ec6878383d2c3cd74bd05b076c21551909f2f31a07cb368f5683b1e99213

Observation 3e82dea6-48d0-4aef-9c87-6c0f13d4f520 · outbound

This paper cites TransGAN: Two Pure Transformers Can Make One Strong GAN, and That Can Scale Up.

Deeper Inside Deep ViT TransGAN: Two Pure Transformers Can Make One Strong GAN, and That Can Scale Up

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-08-06T00:55:21.608433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T00:55:19.349820Z digest=sha256:450a2e65070f7ec65553857d11339fefa86c579ccf9014472fd2df71a6c570c9

Observation 726ddafc-e74b-4468-be06-2de580ab8233 · outbound

This paper cites A style-based generator architecture for generative adversarial networks.

Deeper Inside Deep ViT A style-based generator architecture for generative adversarial networks

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T00:55:19.394156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:55:19.394156Z digest=sha256:e1a6184a7cae71401615144140dc96f49f28d6916e29ebf8332bb48b5cc79115

Observation 169df59b-0623-41ff-bc80-a22425c70811 · outbound

This paper cites ViTGAN: Training GANs with Vision Transformers.

Deeper Inside Deep ViT ViTGAN: Training GANs with Vision Transformers

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T00:55:19.445684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:55:19.445684Z digest=sha256:32bb8fdb7465c2e33a53a7cf0b9520b84794326990940154f7aa57d93d16d4f6

Observation beac8d87-91cb-4fa4-9755-d9e8af67c31e · outbound

This paper cites Blendgan: Implicitly gan blending for arbitrary stylized face generation.

Deeper Inside Deep ViT Blendgan: Implicitly gan blending for arbitrary stylized face generation

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:55:23.121610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T00:55:19.527041Z digest=sha256:640d28d0a1dacf84ca73f251d192abbf30651395897c5fe78eb8047c46a25995

Observation ce9e0072-cd6b-4b34-ac2d-fdcc3c05310f · outbound

This paper cites Swin transformer: Hierarchical vision transformer using shifted windows.

Deeper Inside Deep ViT Swin transformer: Hierarchical vision transformer using shifted windows

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:55:22.956431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T00:55:19.604222Z digest=sha256:36f98d931cd0186a49cb71b290a2f98e1654a62bf6b1d838a1a0c2a031f88239

Observation 7ad5847d-0512-4625-8276-800f621c1cb1 · outbound

This paper cites Decoupled Weight Decay Regularization.

Deeper Inside Deep ViT Decoupled Weight Decay Regularization

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T00:55:19.692799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:55:19.692799Z digest=sha256:7d5fe7c32305ed9b6c1d6de3dc70b69602ea1847ada540dbf2ead1d34ffdcd51

Observation fbbc5d8e-e7ea-42c4-a938-fc6ba77c85c9 · outbound

This paper cites Large Scale Transfer Learning for Differentially Private Image Classification.

Deeper Inside Deep ViT Large Scale Transfer Learning for Differentially Private Image Classification

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-08-06T00:55:21.455902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T00:55:19.827204Z digest=sha256:9ad64186c7412e88a431a3672e2f9a737834769923d1d2d78d8cd81d1173f509

Observation cc19d6e6-265f-4cb9-b6af-4b1e2b37b197 · outbound

This paper cites MobileViT: Light-weight, General-purpose, and Mobile-friendly Vision Transformer.

Deeper Inside Deep ViT MobileViT: Light-weight, General-purpose, and Mobile-friendly Vision Transformer

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T00:55:19.954028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:55:19.954028Z digest=sha256:9aa725a2881920009fe9b4538285cb259a5855e024aeaffda6b44cf502238115

Observation 1a83d089-c7e7-4221-bfd1-2f50f44f79ff · outbound

This paper cites Mixed Precision Training.

Deeper Inside Deep ViT Mixed Precision Training

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T00:55:20.049340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:55:20.049340Z digest=sha256:6d2098144d66ba943773fa5e83262eb41a478231e57002c71050f00870e52538

Observation 0c51c60c-4f8b-4363-b115-da709a5188fb · outbound

This paper cites Foundation models for generalist medical artificial intelligence.

Deeper Inside Deep ViT Foundation models for generalist medical artificial intelligence

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:55:22.769106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T00:55:20.165788Z digest=sha256:45025b55fd7f431a673183c3dcc41d4521d8c32db2225319c910143d44d6f857

Observation bb34ee32-9dc7-46d1-88f1-877909059180 · outbound

This paper cites Do vision transformers see like convolutional neural networks? Advances in Neural Information Processing Systems, 34:12116–12128, 2021.

Deeper Inside Deep ViT Do vision transformers see like convolutional neural networks? Advances in Neural Information Processing Systems, 34:12116–12128, 2021

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:55:22.639380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T00:55:20.226965Z digest=sha256:2d29d90f48d57af73ce5e1e89ddcdf4aec1f35a80adae3a3230bb2dab2d4483b

Observation 2b6bbd18-ad90-4c93-b1aa-46c05470fa30 · outbound

This paper cites Zero-shot text-to-image generation.

Deeper Inside Deep ViT Zero-shot text-to-image generation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T00:55:20.320394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:55:20.320394Z digest=sha256:43c7728206ef4bd3563f230da002896bf64ddcc87340802c41b9474cd97dbb75

Observation 4c35ea16-3187-47c9-b043-fd16e546e1b9 · outbound

This paper cites U-net: Convolutional networks for biomedical image segmentation.

Deeper Inside Deep ViT U-net: Convolutional networks for biomedical image segmentation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T00:55:20.463427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:55:20.463427Z digest=sha256:c7e38bd1edc96df31fd00011e457a885d029cfcde1463aa2ca9508e512fafafb

Observation b0782635-139e-4750-a0c0-20320b789cdf · outbound

This paper cites Revisiting unreasonable effectiveness of data in deep learning era.

Deeper Inside Deep ViT Revisiting unreasonable effectiveness of data in deep learning era

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:55:22.504731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T00:55:20.565510Z digest=sha256:a3a57bc043b3f0505f261151334db7769200c58a25bf3af3e896536a07940405

Observation a1214035-30c1-4599-bb67-5c1117dd3ba1 · outbound

This paper cites LaMDA: Language Models for Dialog Applications.

Deeper Inside Deep ViT LaMDA: Language Models for Dialog Applications

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T00:55:20.670264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:55:20.670264Z digest=sha256:8832f3a3e571edafa21335e5daba505197ca22d0ab33f77697b3413e26908108

Observation 06a9e207-c09c-46d8-b923-d7710f77a39c · outbound

This paper cites Attention is all you need.

Deeper Inside Deep ViT Attention is all you need

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T00:55:20.791141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:55:20.791141Z digest=sha256:264ff9b716c98512dd6fd65ee2c83dd39b7084c6aa25424f4ee3203a57c01283

Observation 1cccc6d1-1277-40ba-b87f-2668fc5630f5 · outbound

This paper cites Internimage: Exploring large-scale vision foundation models with deformable convolutions.

Deeper Inside Deep ViT Internimage: Exploring large-scale vision foundation models with deformable convolutions

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:55:22.415097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T00:55:20.898458Z digest=sha256:2b26ffed8906c753aeb397557c8f6cfcb5c6b301db1461fe77b545da0559d7bc

Observation f2b1d65f-8e30-4f13-a132-e64a23688d9f · outbound

This paper cites Generative adversarial network in medical imaging: A review.

Deeper Inside Deep ViT Generative adversarial network in medical imaging: A review

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:55:22.252362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T00:55:21.015991Z digest=sha256:7fabb850249e2f20cb58325327bd9bd924ab03b96486796e9dbd10e0d2266583

Observation cf77a910-d2e1-4b42-92c6-fc9dc3cfe06d · outbound

This paper cites Scaling vision transform- ers.

Deeper Inside Deep ViT Scaling vision transform- ers

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:55:22.106097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T00:55:21.149799Z digest=sha256:ec5764e296fe55bb57ad99e15c4f2e81304309424ab2bb56214ae385a926b828

Observation b8e52cf8-5959-459d-84a8-1ad36ba0e74c · outbound

This paper cites Unpaired image-to-image translation using cycle-consistent adversarial networks.

Deeper Inside Deep ViT Unpaired image-to-image translation using cycle-consistent adversarial networks

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:55:21.909012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T00:55:21.245290Z digest=sha256:d2391d741059f64ab35f435c43ebfe0c8d11ddf22fd59b05dc3884b19a433722

Pith citing papers

No inbound Pith citation observations are available.