Pith. sign in

Paper Citation Record · LEDGER

Sensitive Image Classification by Vision Transformers

As of 12 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 0 inbound Pith citation observations for arXiv:2412.16446.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.16446 v1

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T10:37:19.977230Z

measured 35 of 35 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

35 of 35 outbound references displayed

  • verified exact2
  • verified fuzzy23
  • unresolved10
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a3247c8f-959b-49a7-8e3b-b584f6c791e6 · outbound

This paper cites Findings from WeProtect global alliance/ technology coalition survey of technology companies.

Sensitive Image Classification by Vision Transformers Findings from WeProtect global alliance/ technology coalition survey of technology companies

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.590035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T10:37:19.824734Z digest=sha256:c90c7b7451b790283784d0f6da37cd4950f7ddfea1272e3d998edb5dda245219

Observation 93612713-5a75-499e-9b9c-f6c04ed05ba9 · outbound

This paper cites The tale of Telegram governance: When the rule of thumb fails,.

Sensitive Image Classification by Vision Transformers The tale of Telegram governance: When the rule of thumb fails,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.576665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T10:37:19.829751Z digest=sha256:c2dc209556e0ac2c57569778f1bd9c25003140573dac33e8f301a71fd5bd766c

Observation 3d139b80-6b47-4932-a132-105c0127f3f6 · outbound

This paper cites Detecting child sexual abuse material: A comprehensive survey,.

Sensitive Image Classification by Vision Transformers Detecting child sexual abuse material: A comprehensive survey,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.562766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T10:37:19.834465Z digest=sha256:0d23c6eb589af217bd953e1b9ee28de7d9c0f5fd032f436fed5a1b0dc98208a1

Observation dedefd7a-479c-4c57-abae-50b524c24806 · outbound

This paper cites Smart content recognition from images using a mixture of convolutional neural networks,.

Sensitive Image Classification by Vision Transformers Smart content recognition from images using a mixture of convolutional neural networks,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.548476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T10:37:19.839109Z digest=sha256:aba9775b012404d5298656a9e489660f12930c41e2dcd11a5b82519192c7eb04

Observation 07ebffce-6460-4b8a-a787-f4e18e7ebd20 · outbound

This paper cites 20k nudity dataset,.

Sensitive Image Classification by Vision Transformers 20k nudity dataset,

Reference 5

Resolution
verified exact
raw_fallback, observed 2026-08-11T10:37:20.180497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T10:37:19.843632Z digest=sha256:63a29e1c8f8dc9d4f379b47d2fc91468a81c27082e2ebe5261705593ecb3594f

Observation 1e6a80df-73c7-4ad2-85c9-5c17aec191f8 · outbound

This paper cites AttM- CNN: Attention and metric learning based CNN for pornography, age and child sexual abuse (CSA) detection in images,.

Sensitive Image Classification by Vision Transformers AttM- CNN: Attention and metric learning based CNN for pornography, age and child sexual abuse (CSA) detection in images,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.533728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T10:37:19.848013Z digest=sha256:dab0602963c9f081a3df27cbbf8fda6b773baf227d33b333107258b7b5f3a263

Observation e823b342-7dda-4768-a42f-4d5fb91e7bcd · outbound

This paper cites Description of the neural network based on AB/DL pictures. Possible implications for forensic sexology,.

Sensitive Image Classification by Vision Transformers Description of the neural network based on AB/DL pictures. Possible implications for forensic sexology,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.519350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T10:37:19.852580Z digest=sha256:4d4113b94c12d28274e8352e1afac6a7b821a11d36303160eb57bbba8f02e90c

Observation 725679e3-c2f1-4203-ac36-cf55d46cd546 · outbound

This paper cites LSPD: A large-scale pornographic dataset for detection and classification,.

Sensitive Image Classification by Vision Transformers LSPD: A large-scale pornographic dataset for detection and classification,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.503920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T10:37:19.857381Z digest=sha256:a834d0528aae16ec978b091b67cd15da260d970ebe7eb53d5898e4f29fecf0bd

Observation b0e23209-dd61-431b-81cf-380d73dd6919 · outbound

This paper cites an unresolved cited work.

Sensitive Image Classification by Vision Transformers Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-11T10:37:20.487588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T10:37:19.861365Z digest=sha256:ec9d5f95b178f71be6238ce6f6049914ab998e5ad4453a4f0d3bb63fb34ec668

Observation 275d7f96-4e65-4209-adeb-05c41a44f927 · outbound

This paper cites Detecting pornographic images by localizing skin rois,.

Sensitive Image Classification by Vision Transformers Detecting pornographic images by localizing skin rois,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.472012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T10:37:19.865681Z digest=sha256:324f892035ccfbe3a761edd0740f7ebac2635491ec8cc5dbb8ad7972a7d4fc27

Observation fee19391-3a80-4351-8615-e122190b463f · outbound

This paper cites NuDetective: A Forensic Tool to Help Combat Child Pornography through Automatic Nudity Detection,.

Sensitive Image Classification by Vision Transformers NuDetective: A Forensic Tool to Help Combat Child Pornography through Automatic Nudity Detection,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.456656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T10:37:19.869443Z digest=sha256:2af65dbd00612eb53e59ca3ac652fff77907f1094b1b4551786e05f1a399aa53

Observation 061454a2-7421-44bd-a022-7fcacb5e3c8e · outbound

This paper cites Open nsfw model,.

Sensitive Image Classification by Vision Transformers Open nsfw model,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.443407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T10:37:19.873444Z digest=sha256:3bb2491911438cd1271850a009acfcdfceeb6e2efd961d6c3e85df8aef3f9835

Observation d5f4db3f-303e-4547-b344-a0b95be9bf8c · outbound

This paper cites Laying foundations for effective machine learning in law enforce- ment. Majura – A labelling schema for child exploitation materials,.

Sensitive Image Classification by Vision Transformers Laying foundations for effective machine learning in law enforce- ment. Majura – A labelling schema for child exploitation materials,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.429531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T10:37:19.877368Z digest=sha256:1ff161d8bb33a1b93a44b0ad9c3ee8a636b4590c54e6901aac81806128d0777f

Observation 2738a141-98ee-4173-aad6-281e53a34d06 · outbound

This paper cites Neural Machine Translation by Jointly Learning to Align and Translate.

Sensitive Image Classification by Vision Transformers Neural Machine Translation by Jointly Learning to Align and Translate

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T10:37:19.881081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:37:19.881081Z digest=sha256:7c3e10d26774d49c9a5f0e913b3ee269319cd9be00f7304c4197884b41d2e132

Observation ca7ea071-2fa1-4600-ac11-9a01e9b04636 · outbound

This paper cites Survey on the attention based RNN model and its applications in computer vision.

Sensitive Image Classification by Vision Transformers Survey on the attention based RNN model and its applications in computer vision

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-11T10:37:20.088055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T10:37:19.885023Z digest=sha256:a01859fda58fddfe023d813fc7e82043cf564fbef893c5db7955868e008ed725

Observation 5db4cb8a-494b-46fc-ad12-514e57ff39f4 · outbound

This paper cites The Kinetics Human Action Video Dataset.

Sensitive Image Classification by Vision Transformers The Kinetics Human Action Video Dataset

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T10:37:19.888909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:37:19.888909Z digest=sha256:0ae389ef2b847d0eaab9da806cccac437ad38c6a3926f5a2cec5b6f2d694f602

Observation 755b89b8-1194-42f0-b51e-774c19727d60 · outbound

This paper cites The “something something.

Sensitive Image Classification by Vision Transformers The “something something

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.413802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T10:37:19.893297Z digest=sha256:b21984216b86075179acb61062c171c087f7b60617faacf2a53e1bb5eabe620a

Observation 6477ee2c-bc56-4763-92a7-6c1aa4fca996 · outbound

This paper cites Multimodal learning with trans- formers: A survey,.

Sensitive Image Classification by Vision Transformers Multimodal learning with trans- formers: A survey,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T10:37:19.897311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:37:19.897311Z digest=sha256:24d62cbb4780d68ec7bd13c9a8585249178a6f91c1069e9558cd23813d411895

Observation b75aecde-cf7e-4111-be61-bfa67dfaf42d · outbound

This paper cites Frozen CLIP models are efficient video learners,.

Sensitive Image Classification by Vision Transformers Frozen CLIP models are efficient video learners,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.388758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T10:37:19.901198Z digest=sha256:aa34a1af422212f3ac7d700df14fb2f6d58613f99626a38e40a702d08b393474

Observation bf12f670-a356-47c9-bd56-039b168bb1f6 · outbound

This paper cites VideoMAE: Masked autoencoders are data-efficient learners for self-supervised video pre- training,.

Sensitive Image Classification by Vision Transformers VideoMAE: Masked autoencoders are data-efficient learners for self-supervised video pre- training,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.374744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T10:37:19.905164Z digest=sha256:929a47c8ef349ec3c7ce2879c152de7d47cc6ef94aa21cb38769178b74cf81af

Observation d766ce6d-4c5f-40be-8d5d-1431e86b02eb · outbound

This paper cites Is Space-Time Attention All You Need for Video Understanding?.

Sensitive Image Classification by Vision Transformers Is Space-Time Attention All You Need for Video Understanding?

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T10:37:19.909396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:37:19.909396Z digest=sha256:8643c43174abca698c968336b087b97d4d0f759242881989c72a9ab641f658f6

Observation 54c9698f-6890-46eb-becd-a9439c9e4701 · outbound

This paper cites PolyViT: Co-training Vision Transformers on Images, Videos and Audio.

Sensitive Image Classification by Vision Transformers PolyViT: Co-training Vision Transformers on Images, Videos and Audio

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T10:37:19.914676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:37:19.914676Z digest=sha256:0d13f8c937a83bc509b190cdfa5ef3a24b7499392a5b67392aac5c7b5d741f3e

Observation d1c46bea-02dc-4088-a560-d39b6b25ea66 · outbound

This paper cites Omnimae: Single model masked pretraining on images and videos,.

Sensitive Image Classification by Vision Transformers Omnimae: Single model masked pretraining on images and videos,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.359371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T10:37:19.919303Z digest=sha256:f646c93c67b561221630af8cf7412aa484b9b2e0e911e0aebd9b7790fe061a7c

Observation ca03eb9b-1a15-4b0a-a9dc-4fe64b80c01b · outbound

This paper cites Omnivore: A single model for many visual modalities,.

Sensitive Image Classification by Vision Transformers Omnivore: A single model for many visual modalities,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.342848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T10:37:19.923664Z digest=sha256:f0dabbf2f59b6a9cdac02766045b75a115ee7285c0a9a746755cb61e9099b548

Observation aa71f67d-3c31-43f1-8f6d-432f80907d63 · outbound

This paper cites M&M Mix: A Multimodal Multiview Transformer Ensemble.

Sensitive Image Classification by Vision Transformers M&M Mix: A Multimodal Multiview Transformer Ensemble

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T10:37:19.928139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:37:19.928139Z digest=sha256:eee0bff7f608e98504c2ab8fc8caedf25f10de5b2e1fe54885178b8d2196d776

Observation 5aaa08d9-12da-44e2-9225-2e1a0f4b434e · outbound

This paper cites MultiMAE: Multi-modal multi-task masked autoencoders,.

Sensitive Image Classification by Vision Transformers MultiMAE: Multi-modal multi-task masked autoencoders,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.325566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T10:37:19.933264Z digest=sha256:21bd5af8a3db8d0680fb35cde4352f9ae00121174ec25ef2c867ad5118301307

Observation 790da5fd-1fa7-4c0d-be08-ddd584149612 · outbound

This paper cites Video swin transformer,.

Sensitive Image Classification by Vision Transformers Video swin transformer,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T10:37:19.937968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:37:19.937968Z digest=sha256:bf3cc3f4ede00d5835271ec81ca0181eb132ba33c3edb3f45a8c8b9730061664

Observation dbf9293b-db3f-483b-98cd-4e59398db252 · outbound

This paper cites BEVT: BERT pretraining of video transform- ers,.

Sensitive Image Classification by Vision Transformers BEVT: BERT pretraining of video transform- ers,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.298900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T10:37:19.942883Z digest=sha256:8e1dcc371c8b7de0b40015dca1f9db3e0b829cec35c398b237474c5d6887d7e9

Observation 41d996c0-5fb5-4e87-a7a6-47f6a52e68c2 · outbound

This paper cites Attention is all you need,.

Sensitive Image Classification by Vision Transformers Attention is all you need,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.281089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T10:37:19.947832Z digest=sha256:39d9897b7dd546673d0c15048d5be7382737e51547f5fe88974b6f1046b8675a

Observation bfac9bde-7f4b-49ab-866a-e3c508af8b25 · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale,.

Sensitive Image Classification by Vision Transformers An image is worth 16x16 words: Transformers for image recognition at scale,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T10:37:19.953563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:37:19.953563Z digest=sha256:dd09dc518478179262557b042202151b4da0e436ede2301d128a742c2ddced42

Observation 5c48f9d8-efd4-4a6b-948e-a2e6adefd456 · outbound

This paper cites Swin transformer: Hierarchical vision transformer using shifted windows,.

Sensitive Image Classification by Vision Transformers Swin transformer: Hierarchical vision transformer using shifted windows,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T10:37:19.958117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:37:19.958117Z digest=sha256:d210a023858fdc68e83ee4f6f6f144ea82a7de10f1d4ab806b0223c445e3a9d2

Observation 7be411ab-43b3-414e-aa24-e33a47363ac2 · outbound

This paper cites Fast vision transformers with HiLo at- tention,.

Sensitive Image Classification by Vision Transformers Fast vision transformers with HiLo at- tention,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.243234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T10:37:19.962272Z digest=sha256:7b09a7af1101cf3a290b3b7e0f13345be842369968d8ee83ab05023b52ec7584

Observation 352f6087-b9a7-46c6-ba11-4eefa851e2c7 · outbound

This paper cites State-of-the-art in nudity classification: A comparative analysis,.

Sensitive Image Classification by Vision Transformers State-of-the-art in nudity classification: A comparative analysis,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.227175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T10:37:19.968555Z digest=sha256:ad25639157afd0b8f3dbbdcc560dc8405da2590e3733667bb9c30deb6df5638c

Observation 4279de66-d7ea-4b40-a2c3-8e65d9416397 · outbound

This paper cites The Bumble’s private detector model.

Sensitive Image Classification by Vision Transformers The Bumble’s private detector model

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.210155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T10:37:19.972883Z digest=sha256:2d11bdef3f786768873100b976d8dc45603c7a01bb088e31480faacf204b398d

Observation b35b030c-7f1d-4ef9-83c8-b7a96553033e · outbound

This paper cites EfficientNet: Rethinking model scaling for con- volutional neural networks,.

Sensitive Image Classification by Vision Transformers EfficientNet: Rethinking model scaling for con- volutional neural networks,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.194560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T10:37:19.977230Z digest=sha256:297dac19cb0e80f31404a21320594bdb27d54d914d196e9ad3617c5546343c16

Pith citing papers

No inbound Pith citation observations are available.