Pith. sign in

Paper Citation Record · LEDGER

Sensitive Image Classification by Vision Transformers

As of 15 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 0 inbound Pith citation observations for arXiv:2412.16446.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.16446 v1

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T10:37:19.977230Z

measured 35 of 35 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

35 of 35 outbound references displayed

  • verified exact2
  • verified fuzzy23
  • unresolved10
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a3247c8f-959b-49a7-8e3b-b584f6c791e6 · outbound

This paper cites Findings from WeProtect global alliance/ technology coalition survey of technology companies.

Sensitive Image Classification by Vision Transformers Findings from WeProtect global alliance/ technology coalition survey of technology companies

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.590035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.824734Z digest=sha256:354e42f908a0b8b338f822077ff4e7202d9e32da3a79d187e7f9966a87c1231e

Observation 93612713-5a75-499e-9b9c-f6c04ed05ba9 · outbound

This paper cites The tale of Telegram governance: When the rule of thumb fails,.

Sensitive Image Classification by Vision Transformers The tale of Telegram governance: When the rule of thumb fails,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.576665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.829751Z digest=sha256:52458a323c9d73ce1af41b1fc5a8d8443d8024caee043824473556f5bf2d7f00

Observation 3d139b80-6b47-4932-a132-105c0127f3f6 · outbound

This paper cites Detecting child sexual abuse material: A comprehensive survey,.

Sensitive Image Classification by Vision Transformers Detecting child sexual abuse material: A comprehensive survey,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.562766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.834465Z digest=sha256:92cff0df4e7e08dcd4526810ec93864d7e7885956189ae6127b93f7afc4adfe8

Observation dedefd7a-479c-4c57-abae-50b524c24806 · outbound

This paper cites Smart content recognition from images using a mixture of convolutional neural networks,.

Sensitive Image Classification by Vision Transformers Smart content recognition from images using a mixture of convolutional neural networks,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.548476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.839109Z digest=sha256:aaea72876a869e548dd217f51701d0286cdb1b83f70c3154ce392dd6a69e5a92

Observation 07ebffce-6460-4b8a-a787-f4e18e7ebd20 · outbound

This paper cites 20k nudity dataset,.

Sensitive Image Classification by Vision Transformers 20k nudity dataset,

Reference 5

Resolution
verified exact
raw_fallback, observed 2026-08-11T10:37:20.180497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.843632Z digest=sha256:890bc81c8b77b53a27686ffa1c83f06c285de633f7e61b30b38284aa598389a2

Observation 1e6a80df-73c7-4ad2-85c9-5c17aec191f8 · outbound

This paper cites AttM- CNN: Attention and metric learning based CNN for pornography, age and child sexual abuse (CSA) detection in images,.

Sensitive Image Classification by Vision Transformers AttM- CNN: Attention and metric learning based CNN for pornography, age and child sexual abuse (CSA) detection in images,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.533728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.848013Z digest=sha256:12bc3ec0478de19e5bb2a858dc09c40c7c321cc3d8a06fe4cbb3c37c52b3d058

Observation e823b342-7dda-4768-a42f-4d5fb91e7bcd · outbound

This paper cites Description of the neural network based on AB/DL pictures. Possible implications for forensic sexology,.

Sensitive Image Classification by Vision Transformers Description of the neural network based on AB/DL pictures. Possible implications for forensic sexology,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.519350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.852580Z digest=sha256:94374bfbf5afc81336a321c213c727c6b88960e3ac1f1b034715c7d316004db6

Observation 725679e3-c2f1-4203-ac36-cf55d46cd546 · outbound

This paper cites LSPD: A large-scale pornographic dataset for detection and classification,.

Sensitive Image Classification by Vision Transformers LSPD: A large-scale pornographic dataset for detection and classification,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.503920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.857381Z digest=sha256:ba652960a7249f2d1a251b31a3740fc02ac07b872cd2ba674f1051a71e53e5dd

Observation b0e23209-dd61-431b-81cf-380d73dd6919 · outbound

This paper cites an unresolved cited work.

Sensitive Image Classification by Vision Transformers Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-11T10:37:20.487588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.861365Z digest=sha256:7d00d718832d5ba3ce54eea2447bbf8f9fb8c80ec877881f16065bf21fe8460e

Observation 275d7f96-4e65-4209-adeb-05c41a44f927 · outbound

This paper cites Detecting pornographic images by localizing skin rois,.

Sensitive Image Classification by Vision Transformers Detecting pornographic images by localizing skin rois,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.472012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.865681Z digest=sha256:9e45a250ec7b22d242163af8ac7bf7d298654470ae1afcba6c907958a69a4fc4

Observation fee19391-3a80-4351-8615-e122190b463f · outbound

This paper cites NuDetective: A Forensic Tool to Help Combat Child Pornography through Automatic Nudity Detection,.

Sensitive Image Classification by Vision Transformers NuDetective: A Forensic Tool to Help Combat Child Pornography through Automatic Nudity Detection,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.456656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.869443Z digest=sha256:8c4f1f014f7adcf4f912f6ceb7696571c379fee55dbd06ace3bda866c106d25c

Observation 061454a2-7421-44bd-a022-7fcacb5e3c8e · outbound

This paper cites Open nsfw model,.

Sensitive Image Classification by Vision Transformers Open nsfw model,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.443407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.873444Z digest=sha256:dfd29f59582b35d82eb2982b22a2f03f4fb3914f567e4815587bc618b7a6d9a4

Observation d5f4db3f-303e-4547-b344-a0b95be9bf8c · outbound

This paper cites Laying foundations for effective machine learning in law enforce- ment. Majura – A labelling schema for child exploitation materials,.

Sensitive Image Classification by Vision Transformers Laying foundations for effective machine learning in law enforce- ment. Majura – A labelling schema for child exploitation materials,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.429531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.877368Z digest=sha256:36a6ace7f0b22837a9e12002edffe354d7ba5d58a9241b29d0b7822215bec65f

Observation 2738a141-98ee-4173-aad6-281e53a34d06 · outbound

This paper cites Neural Machine Translation by Jointly Learning to Align and Translate.

Sensitive Image Classification by Vision Transformers Neural Machine Translation by Jointly Learning to Align and Translate

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T10:37:19.881081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:37:19.881081Z digest=sha256:c885155ebefbba1326ee6993379239acf21d157929a36def52e9bd50f9989fa3

Observation ca7ea071-2fa1-4600-ac11-9a01e9b04636 · outbound

This paper cites Survey on the attention based RNN model and its applications in computer vision.

Sensitive Image Classification by Vision Transformers Survey on the attention based RNN model and its applications in computer vision

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-11T10:37:20.088055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.885023Z digest=sha256:60e2e69feed8444b5ec487dbe6a33bd7e213db52fdd4416326595db1caaccee8

Observation 5db4cb8a-494b-46fc-ad12-514e57ff39f4 · outbound

This paper cites The Kinetics Human Action Video Dataset.

Sensitive Image Classification by Vision Transformers The Kinetics Human Action Video Dataset

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T10:37:19.888909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:37:19.888909Z digest=sha256:2e5ae1b02b63a6b7a8fc571114a5ac35a5db47b7a010b8a114e80be580a91d25

Observation 755b89b8-1194-42f0-b51e-774c19727d60 · outbound

This paper cites The “something something.

Sensitive Image Classification by Vision Transformers The “something something

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.413802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.893297Z digest=sha256:332edfb800bde194a7f8165202e12c05aab4cd130f2030f77ffb45824c3058ba

Observation 6477ee2c-bc56-4763-92a7-6c1aa4fca996 · outbound

This paper cites Multimodal learning with trans- formers: A survey,.

Sensitive Image Classification by Vision Transformers Multimodal learning with trans- formers: A survey,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T10:37:19.897311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:37:19.897311Z digest=sha256:40769ec059b1ea09a457ecf7e1d99a8ffa969e6249ba9acc04ab4014ca8ca98d

Observation b75aecde-cf7e-4111-be61-bfa67dfaf42d · outbound

This paper cites Frozen CLIP models are efficient video learners,.

Sensitive Image Classification by Vision Transformers Frozen CLIP models are efficient video learners,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.388758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.901198Z digest=sha256:aa83670fd0087a4207ef2452438d975d50e90d86dc3497873d4974332eb33405

Observation bf12f670-a356-47c9-bd56-039b168bb1f6 · outbound

This paper cites VideoMAE: Masked autoencoders are data-efficient learners for self-supervised video pre- training,.

Sensitive Image Classification by Vision Transformers VideoMAE: Masked autoencoders are data-efficient learners for self-supervised video pre- training,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.374744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.905164Z digest=sha256:1a396840029c5d64433affe948a4fd6bf74010255c33e1e27f4f95e71c2d99f2

Observation d766ce6d-4c5f-40be-8d5d-1431e86b02eb · outbound

This paper cites Is Space-Time Attention All You Need for Video Understanding?.

Sensitive Image Classification by Vision Transformers Is Space-Time Attention All You Need for Video Understanding?

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T10:37:19.909396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:37:19.909396Z digest=sha256:9db74451b8dcbc4ae0ff7d462a01c451943f3d8dc3a7b7cec42426f2f30d5eff

Observation 54c9698f-6890-46eb-becd-a9439c9e4701 · outbound

This paper cites PolyViT: Co-training Vision Transformers on Images, Videos and Audio.

Sensitive Image Classification by Vision Transformers PolyViT: Co-training Vision Transformers on Images, Videos and Audio

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T10:37:19.914676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:37:19.914676Z digest=sha256:5b6ac611d905a46a6fbcca10d58288a77a1f80d5da7bf560fa81eb2b6bd83a89

Observation d1c46bea-02dc-4088-a560-d39b6b25ea66 · outbound

This paper cites Omnimae: Single model masked pretraining on images and videos,.

Sensitive Image Classification by Vision Transformers Omnimae: Single model masked pretraining on images and videos,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.359371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.919303Z digest=sha256:79f85fbe02ad7facaa6f114ad59c89a6c0e055666e82116e2eac475f1e14ddf0

Observation ca03eb9b-1a15-4b0a-a9dc-4fe64b80c01b · outbound

This paper cites Omnivore: A single model for many visual modalities,.

Sensitive Image Classification by Vision Transformers Omnivore: A single model for many visual modalities,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.342848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.923664Z digest=sha256:9470e8ec8c9e4fb75bf1431a2265455f4091cd5891c37b335c4287ef9b509c28

Observation aa71f67d-3c31-43f1-8f6d-432f80907d63 · outbound

This paper cites M&M Mix: A Multimodal Multiview Transformer Ensemble.

Sensitive Image Classification by Vision Transformers M&M Mix: A Multimodal Multiview Transformer Ensemble

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T10:37:19.928139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:37:19.928139Z digest=sha256:7dc44c3ba4d84e79e71e653f942e008ae54901a68e8ca18dc5868004271588da

Observation 5aaa08d9-12da-44e2-9225-2e1a0f4b434e · outbound

This paper cites MultiMAE: Multi-modal multi-task masked autoencoders,.

Sensitive Image Classification by Vision Transformers MultiMAE: Multi-modal multi-task masked autoencoders,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.325566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.933264Z digest=sha256:554487c5d96d95be81daf71bb1b3edeaef5ec4e99d5274a0633d518e191c168c

Observation 790da5fd-1fa7-4c0d-be08-ddd584149612 · outbound

This paper cites Video swin transformer,.

Sensitive Image Classification by Vision Transformers Video swin transformer,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T10:37:19.937968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:37:19.937968Z digest=sha256:96ba247a69da7eb75af67c70dc25537ef594471737e90a8cd9cbe238661a245c

Observation dbf9293b-db3f-483b-98cd-4e59398db252 · outbound

This paper cites BEVT: BERT pretraining of video transform- ers,.

Sensitive Image Classification by Vision Transformers BEVT: BERT pretraining of video transform- ers,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.298900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.942883Z digest=sha256:c9ac457e666c9dc0f1ac3d3f0b7bd5263777eaf0b83ef29b2c7d9f589d38d8f3

Observation 41d996c0-5fb5-4e87-a7a6-47f6a52e68c2 · outbound

This paper cites Attention is all you need,.

Sensitive Image Classification by Vision Transformers Attention is all you need,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.281089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.947832Z digest=sha256:3877d8a8db2624996774406d6d866cfb911c026eed74a5aab5e6d1fdf1d70e12

Observation bfac9bde-7f4b-49ab-866a-e3c508af8b25 · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale,.

Sensitive Image Classification by Vision Transformers An image is worth 16x16 words: Transformers for image recognition at scale,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T10:37:19.953563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:37:19.953563Z digest=sha256:c69f27e4f2ba539866339d619c47db7dfb0a4409c87ec313f0a0407b9787683c

Observation 5c48f9d8-efd4-4a6b-948e-a2e6adefd456 · outbound

This paper cites Swin transformer: Hierarchical vision transformer using shifted windows,.

Sensitive Image Classification by Vision Transformers Swin transformer: Hierarchical vision transformer using shifted windows,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T10:37:19.958117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:37:19.958117Z digest=sha256:708b443bf96d126702ebcdf5c1ed746a35a860e1dd9f610c4bb4fd193da50513

Observation 7be411ab-43b3-414e-aa24-e33a47363ac2 · outbound

This paper cites Fast vision transformers with HiLo at- tention,.

Sensitive Image Classification by Vision Transformers Fast vision transformers with HiLo at- tention,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.243234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.962272Z digest=sha256:40cdf02cee98e99a0e780f85d26ff48d8d6736bb9898d858bdec1bd57a782915

Observation 352f6087-b9a7-46c6-ba11-4eefa851e2c7 · outbound

This paper cites State-of-the-art in nudity classification: A comparative analysis,.

Sensitive Image Classification by Vision Transformers State-of-the-art in nudity classification: A comparative analysis,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.227175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.968555Z digest=sha256:2dfd429791a92b0e6fc05a246f8b219d357c8e6aa6d50a4b90dd00f6aef387a1

Observation 4279de66-d7ea-4b40-a2c3-8e65d9416397 · outbound

This paper cites The Bumble’s private detector model.

Sensitive Image Classification by Vision Transformers The Bumble’s private detector model

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.210155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.972883Z digest=sha256:221564b72e0e9c6e6b40d652194ee714c038a162313ddd1fbfdc38c2ca061002

Observation b35b030c-7f1d-4ef9-83c8-b7a96553033e · outbound

This paper cites EfficientNet: Rethinking model scaling for con- volutional neural networks,.

Sensitive Image Classification by Vision Transformers EfficientNet: Rethinking model scaling for con- volutional neural networks,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:37:20.194560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T10:37:19.977230Z digest=sha256:ca12a5e1f57abd0a0712c82db2fd8028ba9b619aa9ba90b4d6f083270394efcb

Pith citing papers

No inbound Pith citation observations are available.