Pith. sign in

Paper Citation Record · LEDGER

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion

As of 18 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 0 inbound Pith citation observations for arXiv:2505.12051.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.12051 v1

Coverage vector

measured 31 of 31 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:44:16.660639Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

31 of 31 outbound references displayed

  • verified exact1
  • verified fuzzy17
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 066c9484-8849-4d45-b0cc-6496fc88dfb6 · outbound

This paper cites Hate begets hate: A temporal study of hate speech,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Hate begets hate: A temporal study of hate speech,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:44:17.097670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:44:16.534713Z digest=sha256:5de3ba663072366abde33560c69d07641dd158f94d013a3c07dec6e59726049a

Observation 1fcbd0d9-c8f2-4601-a083-5d5bf5cd0866 · outbound

This paper cites You can’t stay here: The efficacy of reddit’s 2015 ban examined through hate speech,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion You can’t stay here: The efficacy of reddit’s 2015 ban examined through hate speech,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:44:17.083321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:44:16.539999Z digest=sha256:d34c2b6fd7cd9310cef143b778d8957f3096bf56fea7d8f8580f2101144caab2

Observation ad8e5623-4b33-46c6-a5c3-2d6931bdddcf · outbound

This paper cites Early prediction of hate speech propagation,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Early prediction of hate speech propagation,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:44:17.062907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:44:16.543985Z digest=sha256:17a3148403c4feb83b8cc6b0a57fc6a6feb306bccfe8657183d43336db584fb2

Observation eb3d5fd8-9aac-468f-937e-071b434fe0aa · outbound

This paper cites Hatemm: A multi-modal dataset for hate video classification,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Hatemm: A multi-modal dataset for hate video classification,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:44:17.020196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:44:16.547969Z digest=sha256:741be89d0477ffad9393e01a566325a7459c06905a888fd2be20b986edddf887

Observation 5e90f612-6436-43ad-8513-1169a762c191 · outbound

This paper cites Hate speech detection: Challenges and solutions,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Hate speech detection: Challenges and solutions,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T20:44:16.552237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:44:16.552237Z digest=sha256:9f8dfdd6ae64b6dd588d3f73f9ee7457d8615c4deae5fdb36f91d16a0a30b40b

Observation b16f78d3-4e71-4e1c-b37e-89fe90490159 · outbound

This paper cites Deep learning for hate speech detection in tweets,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Deep learning for hate speech detection in tweets,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T20:44:16.556058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:44:16.556058Z digest=sha256:87b6b45c06f860e8a7f85a71381ebd47fdfd286c02b6bf84c4b46ce481d8807c

Observation 6a1ce0e7-1bea-4ffa-b096-4767135708c3 · outbound

This paper cites Hate me, hate me not: Hate speech detection on face- book,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Hate me, hate me not: Hate speech detection on face- book,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:44:16.984437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:44:16.560524Z digest=sha256:c699a252b0dd8750555d7fbd17dfd19e40483a45871f7148d376b47ec7c858d5

Observation 03a52ca4-073f-4d7b-b1fd-799c2041bc54 · outbound

This paper cites Understanding and detecting hateful content using contrastive learning,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Understanding and detecting hateful content using contrastive learning,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:44:16.971927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:44:16.564983Z digest=sha256:5f40b22604c6072c6c3cb322646ed2f8c6ddd5c7a316ff77344147bd6b88fe0f

Observation a08a5f70-7d99-4386-bfb2-5d9982a559ef · outbound

This paper cites The hateful memes challenge: Detecting hate speech in multimodal memes,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion The hateful memes challenge: Detecting hate speech in multimodal memes,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T20:44:16.569367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:44:16.569367Z digest=sha256:33fbf60a8e620045b6cd5ec531b4cc837eb99238128c9564a691cf6135cb010f

Observation f9a0cd61-e0f5-4b2e-9f41-9d1e15e4634b · outbound

This paper cites Detection of hate speech texts using machine learning algorithm,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Detection of hate speech texts using machine learning algorithm,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:44:16.951464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:44:16.573572Z digest=sha256:69e92eefa0a5b8c9f5fd58b7dd25cbe859588cc0c34eecbc6b3380c0feb761df

Observation 8a48b8b9-7cce-4b7e-93ae-44f92530f371 · outbound

This paper cites Towards generalisable hate speech detection: a review on obstacles and solutions,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Towards generalisable hate speech detection: a review on obstacles and solutions,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:44:16.938528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:44:16.577320Z digest=sha256:6a6ea592df3391ef8694382276f642ed00177d38e6d7380420495152bb663141

Observation 0bc0c4af-7a94-45e9-b4b9-b70830830b96 · outbound

This paper cites HateCheck: Functional Tests for Hate Speech Detection Models.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion HateCheck: Functional Tests for Hate Speech Detection Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T20:44:16.581302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:44:16.581302Z digest=sha256:3c8bfa8c6ba95e3141b87761d48c4e9bf2e2bf4a1dcc4c07ff906a5cdfaa5e19

Observation 19de1005-a3f0-4007-ba9d-2c27ef48a9aa · outbound

This paper cites All you need is.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion All you need is

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:44:16.925654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:44:16.585770Z digest=sha256:ab86a4a1f9372ed7231c517b32806947e0445ce3419b6e882959873f4c82ce51

Observation d159d891-9ae5-4ccc-9c40-9957f18454be · outbound

This paper cites Exploring deep multimodal fusion of text and photo for hate speech classification,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Exploring deep multimodal fusion of text and photo for hate speech classification,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:44:16.912654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:44:16.589482Z digest=sha256:1630ae9f6a78eb2986f266f50c962783d4d2142a3771d7d2404d893d5eaa2762

Observation b0ac9bc6-8a17-4d60-9cbe-f0a32f610fb7 · outbound

This paper cites Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T20:44:16.593704Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:44:16.593704Z digest=sha256:099da9cb40a8042ef0e8ae7da4b329ad3247536400d82e738afb643dd25675c2

Observation 16f86c24-a771-4274-9e8b-1eb6c329b09d · outbound

This paper cites Exploring hate speech detection in multimodal publications,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Exploring hate speech detection in multimodal publications,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:44:16.891206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:44:16.598394Z digest=sha256:4318458c1fc3b688a88b83e416d79d5cd197e2de8b1945e92fb1d8adeb786360

Observation 3f8a2c54-8b74-4ae4-b01c-4d71e004b20b · outbound

This paper cites Multimodal hate speech detection via multi-scale visual kernels and knowledge distillation architecture,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Multimodal hate speech detection via multi-scale visual kernels and knowledge distillation architecture,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:44:16.877369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:44:16.602265Z digest=sha256:2afde1bdfe8c1943c88cbbaddceff48acc846545950074c5d4e74a5eaa03e107

Observation 5a32029d-3127-41a0-8951-553eee384732 · outbound

This paper cites A social emotion classification approach using multi-model fusion,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion A social emotion classification approach using multi-model fusion,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:44:16.865168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:44:16.606550Z digest=sha256:d4769d91e4365e306cef21ffbabb3c9ccc867cd39ad97a973813f33847bb0f46

Observation f65183d1-800c-4f56-a9a1-640e95f7749e · outbound

This paper cites Modulated Fusion using Transformer for Linguistic-Acoustic Emotion Recognition.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Modulated Fusion using Transformer for Linguistic-Acoustic Emotion Recognition

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-08-15T20:44:16.747449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:44:16.610597Z digest=sha256:12f199b93b08b2eae86087fcdafa349f71ad92dd41aec41d3efda1eba6cbee71

Observation df945bed-361d-40ad-81fc-324a13cebf2e · outbound

This paper cites Detecting fake news on chinese social media based on hybrid feature fusion method,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Detecting fake news on chinese social media based on hybrid feature fusion method,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:44:16.851570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:44:16.614754Z digest=sha256:52417257d4855e126043bc3e3479a69073691216d816649ac1d91d99db62f7cc

Observation 85c4056e-4278-429e-82ff-4b92748ad9e4 · outbound

This paper cites Multi-feature fusion via hierarchical regression for multimedia anal- ysis,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Multi-feature fusion via hierarchical regression for multimedia anal- ysis,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:44:16.838476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:44:16.618736Z digest=sha256:b4fb9bee9a21db75a4da915d86ffed102d84647ea5c7c8582261d29d2d560981

Observation ba4d12f9-b980-45bb-a229-30ceee9ed06a · outbound

This paper cites Visual and textual deep feature fusion for document image classification,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Visual and textual deep feature fusion for document image classification,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:44:16.826064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:44:16.622341Z digest=sha256:a15222836e16b5934db981a9c83b5159ea43a6a2fd1ba7b2906e48a691122df7

Observation 886d757e-31a7-4af3-af09-21a4d79fdc2f · outbound

This paper cites A practical tutorial on autoencoders for nonlinear feature fusion: Taxon- omy, models, software and guidelines,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion A practical tutorial on autoencoders for nonlinear feature fusion: Taxon- omy, models, software and guidelines,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:44:16.813017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:44:16.626360Z digest=sha256:ad0e8a5e25594d8bcb6a88c4a4b38a5ca15de1571ba4d013b991bd16021e45c2

Observation 14757439-fed6-413a-8188-f6b229d8b450 · outbound

This paper cites an unresolved cited work.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:44:16.798170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:44:16.630628Z digest=sha256:692420686593a8890e41bf35dd965d2b5c8cd811c7620210125889599f32043e

Observation e07fa401-2fe5-4d0f-922d-3b606b783893 · outbound

This paper cites Robust speech recognition via large-scale weak supervi- sion,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Robust speech recognition via large-scale weak supervi- sion,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T20:44:16.635855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:44:16.635855Z digest=sha256:442032d6f15ec116c80fa05c823044d6106fdf7bc464bbc060ab97d1f29a89fa

Observation dfee4084-871c-4843-b48f-626801321707 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T20:44:16.639583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:44:16.639583Z digest=sha256:20331bde86e343f21e71bfdd9816e5ef739aa5ecb7005f88dc3744fffea63cdc

Observation 44e71b19-f685-43e8-88bc-96f0e43a0430 · outbound

This paper cites Voice Recognition Algorithms using Mel Frequency Cepstral Coefficient (MFCC) and Dynamic Time Warping (DTW) Techniques.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Voice Recognition Algorithms using Mel Frequency Cepstral Coefficient (MFCC) and Dynamic Time Warping (DTW) Techniques

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T20:44:16.644034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:44:16.644034Z digest=sha256:d7b5de5631f57cc521e2463e15e0b5f61d29b74508aa146b4641561cc3940a5b

Observation 346bfa8b-c691-497b-8160-3f78e4baa911 · outbound

This paper cites Squeeze-and-excitation networks,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Squeeze-and-excitation networks,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T20:44:16.648422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:44:16.648422Z digest=sha256:432642f43aeece6bde62897c751e455e790b5b3ae38eb6334372836d36a791a2

Observation 04fb1ae2-9ccb-4fa4-bd26-c56b5308990d · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T20:44:16.652686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:44:16.652686Z digest=sha256:9ddd8ddc973cc6f00ff844d69e062261fe375b9a2f409f6243fa1f1b3ea80c3d

Observation 022408d4-5b7f-41ac-b107-8d0ec46ba297 · outbound

This paper cites Language mod- els are few-shot learners,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Language mod- els are few-shot learners,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T20:44:16.656815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:44:16.656815Z digest=sha256:59d55270d55c6a9f0ea3c9dc13349df81e91c7f54ec9efbddbb17e03161e5263

Observation 40fb30f4-0f6e-424d-841c-dec50513c52e · outbound

This paper cites UMAP: Uniform Manifold Approximation and Projection for Dimension Reduction.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion UMAP: Uniform Manifold Approximation and Projection for Dimension Reduction

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T20:44:16.660639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:44:16.660639Z digest=sha256:91964820986f7791ce0368a33f461cb1ddec962ad27f00202ffac44c99ebaada

Pith citing papers

No inbound Pith citation observations are available.