Pith. sign in

Paper Citation Record · LEDGER

Deeper Inside Deep ViT

As of 22 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 0 inbound Pith citation observations for arXiv:2508.04181.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.04181 v1

Coverage vector

measured 32 of 32 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T00:55:21.245290Z

measured 32 of 32 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

32 of 32 outbound references displayed

  • verified exact3
  • verified fuzzy12
  • unresolved17
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ce9b271e-cd5b-4d9d-940e-c775af7487cb · outbound

This paper cites Layer Normalization.

Deeper Inside Deep ViT Layer Normalization

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T00:55:18.323666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:55:18.323666Z digest=sha256:e202c46037c979c97d2b0274ca3571638373d0560d4f25e6742c73952447fdb6

Observation d537ff11-9f77-447a-afdb-510a52e6cf7a · outbound

This paper cites Relational inductive biases, deep learning, and graph networks.

Deeper Inside Deep ViT Relational inductive biases, deep learning, and graph networks

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T00:55:18.394404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:55:18.394404Z digest=sha256:a66ed03c866b7e68e564bbfaee75760367f993b8760f3411bfd8681a951e940d

Observation 4f6f39e6-7e90-44b9-ad55-4f8a6e01e562 · outbound

This paper cites On the Opportunities and Risks of Foundation Models.

Deeper Inside Deep ViT On the Opportunities and Risks of Foundation Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T00:55:18.484513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:55:18.484513Z digest=sha256:2acc52bb12cb806275917d3a2c53c3116be19dc14d43c63491fa2f517c8f9220

Observation aa14b0b7-5eeb-4708-b860-50af89a7d2fb · outbound

This paper cites Generative Adversarial U-Net for Domain-free Medical Image Augmentation.

Deeper Inside Deep ViT Generative Adversarial U-Net for Domain-free Medical Image Augmentation

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-08-06T00:55:21.739413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T00:55:18.623767Z digest=sha256:1d68f259ef345fc71f32fb1b9167a732e78af66107c9a39405fc42563f9d6b09

Observation a5a602c9-a8e9-49a9-9edf-510b727b5714 · outbound

This paper cites Palm: Scaling language modeling with pathways.

Deeper Inside Deep ViT Palm: Scaling language modeling with pathways

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T00:55:18.690708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:55:18.690708Z digest=sha256:d7dff466b601b88ab553da34b8d2cdcd8cfe0f8c1153e9b688a4f898802a8b71

Observation 01f3bddd-061e-4613-a21c-1a5084f7a61b · outbound

This paper cites Scaling vision transformers to 22 billion parameters.

Deeper Inside Deep ViT Scaling vision transformers to 22 billion parameters

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T00:55:18.792634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:55:18.792634Z digest=sha256:d93435f7e34e0ea6b9507b39c7513435c3ab0726e5f1702ed148c10e2c4d52a3

Observation 83ed8d73-6c15-47e0-b8a7-f963ee453f4a · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Deeper Inside Deep ViT An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T00:55:18.861097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:55:18.861097Z digest=sha256:0d583d8b037ea1b232e15b246f286612b5604042cdb8b29d0528cb2f938fe341

Observation cf9adae1-dc56-48e0-8c76-3ba137bf0695 · outbound

This paper cites Convit: Improving vision transformers with soft convolutional inductive biases.

Deeper Inside Deep ViT Convit: Improving vision transformers with soft convolutional inductive biases

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:55:23.676642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T00:55:18.934779Z digest=sha256:903a4d318d9a403658675edbf30d0b46888d176405043a7f0b81cf4814d4895e

Observation c4cc6f83-8ac1-4b2c-8a5d-0ae99cbf394a · outbound

This paper cites Taming transformers for high-resolution image synthesis.

Deeper Inside Deep ViT Taming transformers for high-resolution image synthesis

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T00:55:19.028566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:55:19.028566Z digest=sha256:ea068cafe94e8688a3470d34855b1c75e8ff43da57ef80c208017153c28dcd08

Observation 7f89c464-df8a-4d16-abb7-8ad516a64c4e · outbound

This paper cites Se (3)-transformers: 3d roto-translation equivariant attention networks.

Deeper Inside Deep ViT Se (3)-transformers: 3d roto-translation equivariant attention networks

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:55:23.530908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T00:55:19.101598Z digest=sha256:56a10b51d26453ec8268ac442e1dba4f2b8797179dd18960632083e7a2837194

Observation 5198433f-80c1-47f4-95ca-d5f9a2595d80 · outbound

This paper cites Intriguing properties of transformer training instabilities, 2023.

Deeper Inside Deep ViT Intriguing properties of transformer training instabilities, 2023

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:55:23.317587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T00:55:19.212858Z digest=sha256:09828c0c390789715eca50d729a5a0937bf56ddb5b60c45d143e1dd6f80a93b7

Observation 8f7ce89c-cd89-4e00-a211-f16bec62d09e · outbound

This paper cites Image-to-image translation with conditional adversarial networks.

Deeper Inside Deep ViT Image-to-image translation with conditional adversarial networks

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T00:55:19.276806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:55:19.276806Z digest=sha256:e7b368b8042a55687d772b4c9aebc88ba38882da15aab6f95be5e0b246be2b84

Observation 3e82dea6-48d0-4aef-9c87-6c0f13d4f520 · outbound

This paper cites TransGAN: Two Pure Transformers Can Make One Strong GAN, and That Can Scale Up.

Deeper Inside Deep ViT TransGAN: Two Pure Transformers Can Make One Strong GAN, and That Can Scale Up

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-08-06T00:55:21.608433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T00:55:19.349820Z digest=sha256:c3cd949ac0832ea6bb6fe8572b821772c9583b091ad21f7b2065cc25b0ea6689

Observation 726ddafc-e74b-4468-be06-2de580ab8233 · outbound

This paper cites A style-based generator architecture for generative adversarial networks.

Deeper Inside Deep ViT A style-based generator architecture for generative adversarial networks

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T00:55:19.394156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:55:19.394156Z digest=sha256:e99b928b1bbd8cc26b3fdc16a7f6aa1ec53746f5c3258ddd6a5ad7648aec582c

Observation 169df59b-0623-41ff-bc80-a22425c70811 · outbound

This paper cites ViTGAN: Training GANs with Vision Transformers.

Deeper Inside Deep ViT ViTGAN: Training GANs with Vision Transformers

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T00:55:19.445684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:55:19.445684Z digest=sha256:fe8c179ff23c71f9fff82e61754077e8cce34ba78bfbd886213979adb1928080

Observation beac8d87-91cb-4fa4-9755-d9e8af67c31e · outbound

This paper cites Blendgan: Implicitly gan blending for arbitrary stylized face generation.

Deeper Inside Deep ViT Blendgan: Implicitly gan blending for arbitrary stylized face generation

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:55:23.121610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T00:55:19.527041Z digest=sha256:89feef9ec4bc5fbeb12b8bb87abb8cb36f2085aecd0d7ef49e3825b0f0d1a1eb

Observation ce9e0072-cd6b-4b34-ac2d-fdcc3c05310f · outbound

This paper cites Swin transformer: Hierarchical vision transformer using shifted windows.

Deeper Inside Deep ViT Swin transformer: Hierarchical vision transformer using shifted windows

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:55:22.956431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T00:55:19.604222Z digest=sha256:4409d7b70507515e7a354fbfdbdcd61d28fb2dd8839dac7b9a36d206064d4cee

Observation 7ad5847d-0512-4625-8276-800f621c1cb1 · outbound

This paper cites Decoupled Weight Decay Regularization.

Deeper Inside Deep ViT Decoupled Weight Decay Regularization

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T00:55:19.692799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:55:19.692799Z digest=sha256:47283326a7978fd8160ca592d5d2f84c9746e7be131052f36a49ee0cc2f3f200

Observation fbbc5d8e-e7ea-42c4-a938-fc6ba77c85c9 · outbound

This paper cites Large Scale Transfer Learning for Differentially Private Image Classification.

Deeper Inside Deep ViT Large Scale Transfer Learning for Differentially Private Image Classification

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-08-06T00:55:21.455902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T00:55:19.827204Z digest=sha256:c1265aa1a21419b78f4d1ef71e6581d3f94091f00745090f646ba520e3819b85

Observation cc19d6e6-265f-4cb9-b6af-4b1e2b37b197 · outbound

This paper cites MobileViT: Light-weight, General-purpose, and Mobile-friendly Vision Transformer.

Deeper Inside Deep ViT MobileViT: Light-weight, General-purpose, and Mobile-friendly Vision Transformer

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T00:55:19.954028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:55:19.954028Z digest=sha256:3a5cc3672b984e32513cea5fa2746f1f289cdeef5f3291e6001513fa1cb6c85b

Observation 1a83d089-c7e7-4221-bfd1-2f50f44f79ff · outbound

This paper cites Mixed Precision Training.

Deeper Inside Deep ViT Mixed Precision Training

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T00:55:20.049340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:55:20.049340Z digest=sha256:6b1ed7b8780e080ad2b3b76c7ae80c9b53ef2b2d04f8fd6dd5a5681c000fc5d6

Observation 0c51c60c-4f8b-4363-b115-da709a5188fb · outbound

This paper cites Foundation models for generalist medical artificial intelligence.

Deeper Inside Deep ViT Foundation models for generalist medical artificial intelligence

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:55:22.769106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T00:55:20.165788Z digest=sha256:d0120776f290ae5312c73d6203fa7ffe21a0639c559bb0b4557da84532b2e7e0

Observation bb34ee32-9dc7-46d1-88f1-877909059180 · outbound

This paper cites Do vision transformers see like convolutional neural networks? Advances in Neural Information Processing Systems, 34:12116–12128, 2021.

Deeper Inside Deep ViT Do vision transformers see like convolutional neural networks? Advances in Neural Information Processing Systems, 34:12116–12128, 2021

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:55:22.639380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T00:55:20.226965Z digest=sha256:fe40384f2a85f559dd472e003f5f17464797008d8c80e481c89ed68548ec993b

Observation 2b6bbd18-ad90-4c93-b1aa-46c05470fa30 · outbound

This paper cites Zero-shot text-to-image generation.

Deeper Inside Deep ViT Zero-shot text-to-image generation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T00:55:20.320394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:55:20.320394Z digest=sha256:df2c0afb93f670820b451cc72f7bb6badc259e2929272e80feaca2c4c5743611

Observation 4c35ea16-3187-47c9-b043-fd16e546e1b9 · outbound

This paper cites U-net: Convolutional networks for biomedical image segmentation.

Deeper Inside Deep ViT U-net: Convolutional networks for biomedical image segmentation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T00:55:20.463427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:55:20.463427Z digest=sha256:d43cea8814084286a01549d8df70d7009fc2fcb4c580c3fb46cbeabbe49a44d4

Observation b0782635-139e-4750-a0c0-20320b789cdf · outbound

This paper cites Revisiting unreasonable effectiveness of data in deep learning era.

Deeper Inside Deep ViT Revisiting unreasonable effectiveness of data in deep learning era

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:55:22.504731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T00:55:20.565510Z digest=sha256:afc2cbb666c282b81d562de3e9e4aa716f2304708b20639755e52a7032e0caaa

Observation a1214035-30c1-4599-bb67-5c1117dd3ba1 · outbound

This paper cites LaMDA: Language Models for Dialog Applications.

Deeper Inside Deep ViT LaMDA: Language Models for Dialog Applications

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T00:55:20.670264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:55:20.670264Z digest=sha256:74edd7f17dfb01a2f84e4fd6df5c13d18f97d2476bbe5ac230e8e06411fefd01

Observation 06a9e207-c09c-46d8-b923-d7710f77a39c · outbound

This paper cites Attention is all you need.

Deeper Inside Deep ViT Attention is all you need

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T00:55:20.791141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:55:20.791141Z digest=sha256:4870f5307e9c47f4b3f9f696cb7482927229b13387f080fd6b766d5647210fa1

Observation 1cccc6d1-1277-40ba-b87f-2668fc5630f5 · outbound

This paper cites Internimage: Exploring large-scale vision foundation models with deformable convolutions.

Deeper Inside Deep ViT Internimage: Exploring large-scale vision foundation models with deformable convolutions

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:55:22.415097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T00:55:20.898458Z digest=sha256:a88bdd1fa9030b36ed10689228401a82d9739caafbd03347bfa53e9a0d286677

Observation f2b1d65f-8e30-4f13-a132-e64a23688d9f · outbound

This paper cites Generative adversarial network in medical imaging: A review.

Deeper Inside Deep ViT Generative adversarial network in medical imaging: A review

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:55:22.252362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T00:55:21.015991Z digest=sha256:86309230163dc2770d0cd8cb76ac7743fb3097e5953c55ccf27e921ecc538e09

Observation cf77a910-d2e1-4b42-92c6-fc9dc3cfe06d · outbound

This paper cites Scaling vision transform- ers.

Deeper Inside Deep ViT Scaling vision transform- ers

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:55:22.106097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T00:55:21.149799Z digest=sha256:37f742f3341ad921547cd605beecf1aa29f2fadcce18de70578a7c2a3a5b0ef0

Observation b8e52cf8-5959-459d-84a8-1ad36ba0e74c · outbound

This paper cites Unpaired image-to-image translation using cycle-consistent adversarial networks.

Deeper Inside Deep ViT Unpaired image-to-image translation using cycle-consistent adversarial networks

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:55:21.909012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T00:55:21.245290Z digest=sha256:b01bd192694bb0d9dcddedb78cbbe8d7d8a8aef340ce3215a35454be6767049a

Pith citing papers

No inbound Pith citation observations are available.