Pith. sign in

Paper Citation Record · LEDGER

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion

As of 20 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 0 inbound Pith citation observations for arXiv:2505.12051.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.12051 v1

Coverage vector

measured 31 of 31 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:44:16.660639Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

31 of 31 outbound references displayed

  • verified exact1
  • verified fuzzy17
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 066c9484-8849-4d45-b0cc-6496fc88dfb6 · outbound

This paper cites Hate begets hate: A temporal study of hate speech,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Hate begets hate: A temporal study of hate speech,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:44:17.097670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T20:44:16.534713Z digest=sha256:ce9b7f5d5fb0f7e1a2ffca8d805faf0e86cc378e8536bb97d84a245d860418c8

Observation 1fcbd0d9-c8f2-4601-a083-5d5bf5cd0866 · outbound

This paper cites You can’t stay here: The efficacy of reddit’s 2015 ban examined through hate speech,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion You can’t stay here: The efficacy of reddit’s 2015 ban examined through hate speech,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:44:17.083321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T20:44:16.539999Z digest=sha256:1240836d726188d66e3a4dfcdf265aea902f696cb4ec9ed049ad686718f48474

Observation ad8e5623-4b33-46c6-a5c3-2d6931bdddcf · outbound

This paper cites Early prediction of hate speech propagation,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Early prediction of hate speech propagation,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:44:17.062907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T20:44:16.543985Z digest=sha256:ef01a049788626baca8b9fd0ce70fe5efb01d43aefd9acaedf12add0286dcab1

Observation eb3d5fd8-9aac-468f-937e-071b434fe0aa · outbound

This paper cites Hatemm: A multi-modal dataset for hate video classification,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Hatemm: A multi-modal dataset for hate video classification,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:44:17.020196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T20:44:16.547969Z digest=sha256:b13ebdac1f548ed607bce6bcdff4cd2ea15b6fcc24c37d47c17b67bf7b36c1a1

Observation 5e90f612-6436-43ad-8513-1169a762c191 · outbound

This paper cites Hate speech detection: Challenges and solutions,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Hate speech detection: Challenges and solutions,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T20:44:16.552237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:44:16.552237Z digest=sha256:c14bfd116ae5a12f4d2ca770e2f0c9c94b29a623d9870587025ca1715568c1b5

Observation b16f78d3-4e71-4e1c-b37e-89fe90490159 · outbound

This paper cites Deep learning for hate speech detection in tweets,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Deep learning for hate speech detection in tweets,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T20:44:16.556058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:44:16.556058Z digest=sha256:15458684e01c1c3f6e0d595f3cdbdab0ab7eafbbb52cf60a544bddf16ffaaae0

Observation 6a1ce0e7-1bea-4ffa-b096-4767135708c3 · outbound

This paper cites Hate me, hate me not: Hate speech detection on face- book,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Hate me, hate me not: Hate speech detection on face- book,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:44:16.984437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T20:44:16.560524Z digest=sha256:f457cab4a2c64603389444b3c580b7e779966d895313f7f756afc46d72c9c331

Observation 03a52ca4-073f-4d7b-b1fd-799c2041bc54 · outbound

This paper cites Understanding and detecting hateful content using contrastive learning,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Understanding and detecting hateful content using contrastive learning,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:44:16.971927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T20:44:16.564983Z digest=sha256:2413347c2069fe7e3fe11e9501aadd5885d21662b72762a5f51fd4d145e2d78d

Observation a08a5f70-7d99-4386-bfb2-5d9982a559ef · outbound

This paper cites The hateful memes challenge: Detecting hate speech in multimodal memes,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion The hateful memes challenge: Detecting hate speech in multimodal memes,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T20:44:16.569367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:44:16.569367Z digest=sha256:1a90b521f9e46469800c62a2ff0a4c7c964f486a028e362c7beda8df991be65e

Observation f9a0cd61-e0f5-4b2e-9f41-9d1e15e4634b · outbound

This paper cites Detection of hate speech texts using machine learning algorithm,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Detection of hate speech texts using machine learning algorithm,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:44:16.951464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T20:44:16.573572Z digest=sha256:b53101d119e475a2a6a07f2a0a3c1d9cee3d28a5c91682a2318b2db44385890c

Observation 8a48b8b9-7cce-4b7e-93ae-44f92530f371 · outbound

This paper cites Towards generalisable hate speech detection: a review on obstacles and solutions,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Towards generalisable hate speech detection: a review on obstacles and solutions,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:44:16.938528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T20:44:16.577320Z digest=sha256:4d40fa7f96a80b645104da04446babd5ac8705b0dad109a48870ab53b468dd06

Observation 0bc0c4af-7a94-45e9-b4b9-b70830830b96 · outbound

This paper cites HateCheck: Functional Tests for Hate Speech Detection Models.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion HateCheck: Functional Tests for Hate Speech Detection Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T20:44:16.581302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:44:16.581302Z digest=sha256:b47598c2e0961fec45be86750cd3c2c8b48892f85911e84157078c6df626e1d1

Observation 19de1005-a3f0-4007-ba9d-2c27ef48a9aa · outbound

This paper cites All you need is.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion All you need is

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:44:16.925654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T20:44:16.585770Z digest=sha256:2bd8f424cd74a6c93b1f2082e635f145f3eeaaf4b16ce877ac0874f03328bc6a

Observation d159d891-9ae5-4ccc-9c40-9957f18454be · outbound

This paper cites Exploring deep multimodal fusion of text and photo for hate speech classification,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Exploring deep multimodal fusion of text and photo for hate speech classification,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:44:16.912654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T20:44:16.589482Z digest=sha256:4043b4ed8e4ddc2cb3466cb0282b73c2021c8e647a1357ac6920db1a5f1e460a

Observation b0ac9bc6-8a17-4d60-9cbe-f0a32f610fb7 · outbound

This paper cites Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T20:44:16.593704Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:44:16.593704Z digest=sha256:b516c7bd3abcc013eb489d231f7893ea3b4c9b58273b1fbac41b8058734406b0

Observation 16f86c24-a771-4274-9e8b-1eb6c329b09d · outbound

This paper cites Exploring hate speech detection in multimodal publications,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Exploring hate speech detection in multimodal publications,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:44:16.891206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T20:44:16.598394Z digest=sha256:4b69921dbda765c3fc698f88a7d359d2b10df023c1a741fb6c0e6ea4f1be0044

Observation 3f8a2c54-8b74-4ae4-b01c-4d71e004b20b · outbound

This paper cites Multimodal hate speech detection via multi-scale visual kernels and knowledge distillation architecture,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Multimodal hate speech detection via multi-scale visual kernels and knowledge distillation architecture,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:44:16.877369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T20:44:16.602265Z digest=sha256:dc5fabd265006d062282a4bbf7055dc90d122667b65e8bb62560e146d4b33479

Observation 5a32029d-3127-41a0-8951-553eee384732 · outbound

This paper cites A social emotion classification approach using multi-model fusion,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion A social emotion classification approach using multi-model fusion,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:44:16.865168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T20:44:16.606550Z digest=sha256:4e00d9e50d28bc858a7a2ad7e351d4b34b3441fe0e36f4803d67c041a3e5eb0e

Observation f65183d1-800c-4f56-a9a1-640e95f7749e · outbound

This paper cites Modulated Fusion using Transformer for Linguistic-Acoustic Emotion Recognition.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Modulated Fusion using Transformer for Linguistic-Acoustic Emotion Recognition

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-08-15T20:44:16.747449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T20:44:16.610597Z digest=sha256:f7f6d2be2204660813462eddb4d4b7bd7545f39a6120cf84b177806b85501c17

Observation df945bed-361d-40ad-81fc-324a13cebf2e · outbound

This paper cites Detecting fake news on chinese social media based on hybrid feature fusion method,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Detecting fake news on chinese social media based on hybrid feature fusion method,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:44:16.851570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T20:44:16.614754Z digest=sha256:5aa02cad921b23c40281f5d683570242673d56b9f07fda78840d8a0951e45d09

Observation 85c4056e-4278-429e-82ff-4b92748ad9e4 · outbound

This paper cites Multi-feature fusion via hierarchical regression for multimedia anal- ysis,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Multi-feature fusion via hierarchical regression for multimedia anal- ysis,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:44:16.838476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T20:44:16.618736Z digest=sha256:787859be20a386b6e85a973e823fbc40c2465f971370cb77a746ba4d91aca797

Observation ba4d12f9-b980-45bb-a229-30ceee9ed06a · outbound

This paper cites Visual and textual deep feature fusion for document image classification,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Visual and textual deep feature fusion for document image classification,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:44:16.826064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T20:44:16.622341Z digest=sha256:05cf164c2730463dc4d1f910d052752c8a60feea8ae5e95d1c01993ad41d9579

Observation 886d757e-31a7-4af3-af09-21a4d79fdc2f · outbound

This paper cites A practical tutorial on autoencoders for nonlinear feature fusion: Taxon- omy, models, software and guidelines,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion A practical tutorial on autoencoders for nonlinear feature fusion: Taxon- omy, models, software and guidelines,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:44:16.813017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T20:44:16.626360Z digest=sha256:4f25438613f16c4f1387bc83bfc19a84ea0a524d6cdd1a3b1b15c13b5f90e578

Observation 14757439-fed6-413a-8188-f6b229d8b450 · outbound

This paper cites an unresolved cited work.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:44:16.798170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T20:44:16.630628Z digest=sha256:5c6efca2d0b5c0420190291bbb8173be1584460f99329e21118264f740bd7622

Observation e07fa401-2fe5-4d0f-922d-3b606b783893 · outbound

This paper cites Robust speech recognition via large-scale weak supervi- sion,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Robust speech recognition via large-scale weak supervi- sion,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T20:44:16.635855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:44:16.635855Z digest=sha256:02dd6ac68feda90bd6bf5b0f47b3d4f0659ad62d1b358b6da8dfc0e7bc104f6e

Observation dfee4084-871c-4843-b48f-626801321707 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T20:44:16.639583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:44:16.639583Z digest=sha256:bcdf5631c42b677310ad74b516828940a5fdfc9214d41dadab17d37be9eaab70

Observation 44e71b19-f685-43e8-88bc-96f0e43a0430 · outbound

This paper cites Voice Recognition Algorithms using Mel Frequency Cepstral Coefficient (MFCC) and Dynamic Time Warping (DTW) Techniques.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Voice Recognition Algorithms using Mel Frequency Cepstral Coefficient (MFCC) and Dynamic Time Warping (DTW) Techniques

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T20:44:16.644034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:44:16.644034Z digest=sha256:ca754ac09f5e8d5e7037649033bcd96fcfdddc733ab73535c82840ee5bc47142

Observation 346bfa8b-c691-497b-8160-3f78e4baa911 · outbound

This paper cites Squeeze-and-excitation networks,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Squeeze-and-excitation networks,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T20:44:16.648422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:44:16.648422Z digest=sha256:e19bc12643d5a7277f7d1ce37abcdf2e207ab02758258a6ab8db4c26334586e8

Observation 04fb1ae2-9ccb-4fa4-bd26-c56b5308990d · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T20:44:16.652686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:44:16.652686Z digest=sha256:e5fd8cd70fb01b8444225ea71327ef256e8c9b0aa8230142bef618163e3a807b

Observation 022408d4-5b7f-41ac-b107-8d0ec46ba297 · outbound

This paper cites Language mod- els are few-shot learners,.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion Language mod- els are few-shot learners,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T20:44:16.656815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:44:16.656815Z digest=sha256:489434eccf1cfb0da825ba846177d65ef725b00ff2944fa1dbf9b55112cff568

Observation 40fb30f4-0f6e-424d-841c-dec50513c52e · outbound

This paper cites UMAP: Uniform Manifold Approximation and Projection for Dimension Reduction.

Enhanced Multimodal Hate Video Detection via Channel-wise and Modality-wise Fusion UMAP: Uniform Manifold Approximation and Projection for Dimension Reduction

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T20:44:16.660639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:44:16.660639Z digest=sha256:1cddfcb02332793f34c112d8efbf8edf5c19f70e20b5de935531085c2258a6c2

Pith citing papers

No inbound Pith citation observations are available.