Pith. sign in

Paper Citation Record · LEDGER

Sensitive Image Classification by Vision Transformers

As of 15 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 0 inbound Pith citation observations for arXiv:2412.16446.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.16446 v1

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T10:37:19.977230Z

measured 35 of 35 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

35 of 35 outbound references displayed

  • verified exact2
  • verified fuzzy23
  • unresolved10
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a3247c8f-959b-49a7-8e3b-b584f6c791e6 · outbound

This paper cites Findings from WeProtect global alliance/ technology coalition survey of technology companies.

Sensitive Image Classification by Vision Transformers Findings from WeProtect global alliance/ technology coalition survey of technology companies

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.590035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.824734Z digest=sha256:9c1e96f892ed4e4ad7303227531db73fd2b24ea6b48dee9a554e4a9238084fc6

Observation 93612713-5a75-499e-9b9c-f6c04ed05ba9 · outbound

This paper cites The tale of Telegram governance: When the rule of thumb fails,.

Sensitive Image Classification by Vision Transformers The tale of Telegram governance: When the rule of thumb fails,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.576665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.829751Z digest=sha256:4d6115f47ba6b07ad9836919242934eba88b11e484300def65769bb0afa6c150

Observation 3d139b80-6b47-4932-a132-105c0127f3f6 · outbound

This paper cites Detecting child sexual abuse material: A comprehensive survey,.

Sensitive Image Classification by Vision Transformers Detecting child sexual abuse material: A comprehensive survey,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.562766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.834465Z digest=sha256:9a2afb210766e0dd64cc9b40d63e7186db89ce374e90d520d1b729f7fa78f9fa

Observation dedefd7a-479c-4c57-abae-50b524c24806 · outbound

This paper cites Smart content recognition from images using a mixture of convolutional neural networks,.

Sensitive Image Classification by Vision Transformers Smart content recognition from images using a mixture of convolutional neural networks,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.548476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.839109Z digest=sha256:b1fdabca90bf8cb178bf2bc2879815cbcb65dabc55d78270125f78eb740c50a5

Observation 07ebffce-6460-4b8a-a787-f4e18e7ebd20 · outbound

This paper cites 20k nudity dataset,.

Sensitive Image Classification by Vision Transformers 20k nudity dataset,

Reference 5

Resolution
verified exact
raw_fallback, observed 2026-08-11T10:37:20.180497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.843632Z digest=sha256:7bc8b1ae97144762c01b936dadc265201d0682b91d761cf0238b86a8b987d415

Observation 1e6a80df-73c7-4ad2-85c9-5c17aec191f8 · outbound

This paper cites AttM- CNN: Attention and metric learning based CNN for pornography, age and child sexual abuse (CSA) detection in images,.

Sensitive Image Classification by Vision Transformers AttM- CNN: Attention and metric learning based CNN for pornography, age and child sexual abuse (CSA) detection in images,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.533728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.848013Z digest=sha256:7c103f794562fb8836b1270441463d2b435048bfcdc287922793833493baef5b

Observation e823b342-7dda-4768-a42f-4d5fb91e7bcd · outbound

This paper cites Description of the neural network based on AB/DL pictures. Possible implications for forensic sexology,.

Sensitive Image Classification by Vision Transformers Description of the neural network based on AB/DL pictures. Possible implications for forensic sexology,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.519350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.852580Z digest=sha256:3a0a9fec396cefe16a7a0142e59f91fa35cb446f283aec44f225194da409102f

Observation 725679e3-c2f1-4203-ac36-cf55d46cd546 · outbound

This paper cites LSPD: A large-scale pornographic dataset for detection and classification,.

Sensitive Image Classification by Vision Transformers LSPD: A large-scale pornographic dataset for detection and classification,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.503920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.857381Z digest=sha256:f1cbddfe3984c8327e2b7fe2515e851b724c806a6341924e736f31c3f0756289

Observation b0e23209-dd61-431b-81cf-380d73dd6919 · outbound

This paper cites an unresolved cited work.

Sensitive Image Classification by Vision Transformers Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-11T10:37:20.487588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.861365Z digest=sha256:d203cca7f04d1523f0a2def4f361fc9af1d77d064b422904ccc62cd9931b08c1

Observation 275d7f96-4e65-4209-adeb-05c41a44f927 · outbound

This paper cites Detecting pornographic images by localizing skin rois,.

Sensitive Image Classification by Vision Transformers Detecting pornographic images by localizing skin rois,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.472012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.865681Z digest=sha256:efc638133f4d530b1a795b9e7536aa7627326763c8b5db95bf2be70c2179cdd7

Observation fee19391-3a80-4351-8615-e122190b463f · outbound

This paper cites NuDetective: A Forensic Tool to Help Combat Child Pornography through Automatic Nudity Detection,.

Sensitive Image Classification by Vision Transformers NuDetective: A Forensic Tool to Help Combat Child Pornography through Automatic Nudity Detection,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.456656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.869443Z digest=sha256:ab4be2df50ef149b8d997ae55f1e02a241d0da54dad3c44cd0d121986b88edd3

Observation 061454a2-7421-44bd-a022-7fcacb5e3c8e · outbound

This paper cites Open nsfw model,.

Sensitive Image Classification by Vision Transformers Open nsfw model,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.443407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.873444Z digest=sha256:ecc026c6050811f119cb837e7c4b60406fb34a4f8bf8089a31b3875a37f95cb3

Observation d5f4db3f-303e-4547-b344-a0b95be9bf8c · outbound

This paper cites Laying foundations for effective machine learning in law enforce- ment. Majura – A labelling schema for child exploitation materials,.

Sensitive Image Classification by Vision Transformers Laying foundations for effective machine learning in law enforce- ment. Majura – A labelling schema for child exploitation materials,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.429531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.877368Z digest=sha256:ff2cfcae4b1b5e750b2b756446909c44e81daae6a1a91c1d01b0a0ffe79eed2c

Observation 2738a141-98ee-4173-aad6-281e53a34d06 · outbound

This paper cites Neural Machine Translation by Jointly Learning to Align and Translate.

Sensitive Image Classification by Vision Transformers Neural Machine Translation by Jointly Learning to Align and Translate

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T10:37:19.881081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:37:19.881081Z digest=sha256:c9f1b9d70cbd7b273f51d0fba5a04d3017ca837d18328540c1e0dde245aab144

Observation ca7ea071-2fa1-4600-ac11-9a01e9b04636 · outbound

This paper cites Survey on the attention based RNN model and its applications in computer vision.

Sensitive Image Classification by Vision Transformers Survey on the attention based RNN model and its applications in computer vision

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-11T10:37:20.088055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.885023Z digest=sha256:5b60275b20f69a414e6f9c311f4d493b2ba917624abf47b2d4e91003bf960f15

Observation 5db4cb8a-494b-46fc-ad12-514e57ff39f4 · outbound

This paper cites The Kinetics Human Action Video Dataset.

Sensitive Image Classification by Vision Transformers The Kinetics Human Action Video Dataset

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T10:37:19.888909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:37:19.888909Z digest=sha256:b30849f800b5c0b3b4b4380705b8ca12d6cf5fe2b9198c0f2f0b8af72d535ccb

Observation 755b89b8-1194-42f0-b51e-774c19727d60 · outbound

This paper cites The “something something.

Sensitive Image Classification by Vision Transformers The “something something

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.413802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.893297Z digest=sha256:1d59e27b1c6e9f98b3b3b1b8d7fe84cdecb0c48fc2e123dd6735ab2f41db6766

Observation 6477ee2c-bc56-4763-92a7-6c1aa4fca996 · outbound

This paper cites Multimodal learning with trans- formers: A survey,.

Sensitive Image Classification by Vision Transformers Multimodal learning with trans- formers: A survey,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T10:37:19.897311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:37:19.897311Z digest=sha256:1cc74dd89036547cfc56deb037a4a215cfd2d2e8186932062ff5c06806ce8c2e

Observation b75aecde-cf7e-4111-be61-bfa67dfaf42d · outbound

This paper cites Frozen CLIP models are efficient video learners,.

Sensitive Image Classification by Vision Transformers Frozen CLIP models are efficient video learners,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.388758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.901198Z digest=sha256:cbc3fa5907369afab423e93856ab08581bf891ed1eaf8bbe17401047660c5d03

Observation bf12f670-a356-47c9-bd56-039b168bb1f6 · outbound

This paper cites VideoMAE: Masked autoencoders are data-efficient learners for self-supervised video pre- training,.

Sensitive Image Classification by Vision Transformers VideoMAE: Masked autoencoders are data-efficient learners for self-supervised video pre- training,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.374744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.905164Z digest=sha256:4301b36dd773d79bac7a30ef10033c9144b8a297c8c471b96fb008eab44795f5

Observation d766ce6d-4c5f-40be-8d5d-1431e86b02eb · outbound

This paper cites Is Space-Time Attention All You Need for Video Understanding?.

Sensitive Image Classification by Vision Transformers Is Space-Time Attention All You Need for Video Understanding?

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T10:37:19.909396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:37:19.909396Z digest=sha256:eb9977ad21d3e54f084149d089131012d690728b10d7bfe1bd489b30adb91f65

Observation 54c9698f-6890-46eb-becd-a9439c9e4701 · outbound

This paper cites PolyViT: Co-training Vision Transformers on Images, Videos and Audio.

Sensitive Image Classification by Vision Transformers PolyViT: Co-training Vision Transformers on Images, Videos and Audio

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T10:37:19.914676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:37:19.914676Z digest=sha256:bbc498a492230a0f0eeb1e55bb96e0b222cfb1fe74f740312752c4b335394fd7

Observation d1c46bea-02dc-4088-a560-d39b6b25ea66 · outbound

This paper cites Omnimae: Single model masked pretraining on images and videos,.

Sensitive Image Classification by Vision Transformers Omnimae: Single model masked pretraining on images and videos,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.359371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.919303Z digest=sha256:6b4d2efad347207aa7e5ba1f800cd710800df7265fecc75e3fe7e2998d2624de

Observation ca03eb9b-1a15-4b0a-a9dc-4fe64b80c01b · outbound

This paper cites Omnivore: A single model for many visual modalities,.

Sensitive Image Classification by Vision Transformers Omnivore: A single model for many visual modalities,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.342848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.923664Z digest=sha256:c3d5a742567a61a0269509f4fff145a8ae69b86eebd442e64e167cc7536cce47

Observation aa71f67d-3c31-43f1-8f6d-432f80907d63 · outbound

This paper cites M&M Mix: A Multimodal Multiview Transformer Ensemble.

Sensitive Image Classification by Vision Transformers M&M Mix: A Multimodal Multiview Transformer Ensemble

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T10:37:19.928139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:37:19.928139Z digest=sha256:4bfb9afc678a5bfe28647d0e2c3feec606256e62d90dbb0f5a2e9389d23b9fe2

Observation 5aaa08d9-12da-44e2-9225-2e1a0f4b434e · outbound

This paper cites MultiMAE: Multi-modal multi-task masked autoencoders,.

Sensitive Image Classification by Vision Transformers MultiMAE: Multi-modal multi-task masked autoencoders,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.325566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.933264Z digest=sha256:05aa24477360ee131403aa2598e3a7e031fee525fd456fcbbab097a2a9370b57

Observation 790da5fd-1fa7-4c0d-be08-ddd584149612 · outbound

This paper cites Video swin transformer,.

Sensitive Image Classification by Vision Transformers Video swin transformer,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T10:37:19.937968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:37:19.937968Z digest=sha256:03f24bbcf7c89900dfe956f3da3b6c8acd478eb4f23e81e5a0773394c298bccf

Observation dbf9293b-db3f-483b-98cd-4e59398db252 · outbound

This paper cites BEVT: BERT pretraining of video transform- ers,.

Sensitive Image Classification by Vision Transformers BEVT: BERT pretraining of video transform- ers,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.298900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.942883Z digest=sha256:55e662faf558c902854be3e5440907e1de33cc0c5c23be6afbd906788cc5245c

Observation 41d996c0-5fb5-4e87-a7a6-47f6a52e68c2 · outbound

This paper cites Attention is all you need,.

Sensitive Image Classification by Vision Transformers Attention is all you need,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.281089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.947832Z digest=sha256:bac6798b9dc3ff9b1fd8fb68a3312ca3a7592670a01e28a977f06dc5a52add9a

Observation bfac9bde-7f4b-49ab-866a-e3c508af8b25 · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale,.

Sensitive Image Classification by Vision Transformers An image is worth 16x16 words: Transformers for image recognition at scale,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T10:37:19.953563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:37:19.953563Z digest=sha256:8b185a11f729c1c8398f245e41343450019e9a6d7d66ba611f989d6b07897d59

Observation 5c48f9d8-efd4-4a6b-948e-a2e6adefd456 · outbound

This paper cites Swin transformer: Hierarchical vision transformer using shifted windows,.

Sensitive Image Classification by Vision Transformers Swin transformer: Hierarchical vision transformer using shifted windows,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T10:37:19.958117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:37:19.958117Z digest=sha256:dbffa701fa336340a32f8e40ed65967249b6b8c1dc3fb3e3d7bd81f648ec5a47

Observation 7be411ab-43b3-414e-aa24-e33a47363ac2 · outbound

This paper cites Fast vision transformers with HiLo at- tention,.

Sensitive Image Classification by Vision Transformers Fast vision transformers with HiLo at- tention,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.243234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.962272Z digest=sha256:2e50e37070418631de7e89457f1f79f04ab09bb31fed2bd200af12171fd6fdb0

Observation 352f6087-b9a7-46c6-ba11-4eefa851e2c7 · outbound

This paper cites State-of-the-art in nudity classification: A comparative analysis,.

Sensitive Image Classification by Vision Transformers State-of-the-art in nudity classification: A comparative analysis,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.227175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.968555Z digest=sha256:5ccc8383851a91e515003b590a6f6e60f26930cad548b5734a155245753a0d25

Observation 4279de66-d7ea-4b40-a2c3-8e65d9416397 · outbound

This paper cites The Bumble’s private detector model.

Sensitive Image Classification by Vision Transformers The Bumble’s private detector model

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.210155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.972883Z digest=sha256:0861c8f92850e98c44adc6155df04bd4e8b339124b3d1402ed2ec01444d70bce

Observation b35b030c-7f1d-4ef9-83c8-b7a96553033e · outbound

This paper cites EfficientNet: Rethinking model scaling for con- volutional neural networks,.

Sensitive Image Classification by Vision Transformers EfficientNet: Rethinking model scaling for con- volutional neural networks,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.194560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.977230Z digest=sha256:c4867c5aaad13e48cfa5a279e1f728de723372ad9330f57b8d7df4a3b3258f8f

Pith citing papers

No inbound Pith citation observations are available.