Pith. sign in

Paper Citation Record · LEDGER

Multi-modal Attention for Speech Emotion Recognition

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2009.04107.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2009.04107 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:34:08.695045Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-16T16:07:53.138866Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ddc34ba7-9a96-4450-a29a-63af3615c266 · inbound

In-Context Edit: Enabling Instructional Image Editing with In-Context Generation in Large Scale Diffusion Transformer cites this paper.

In-Context Edit: Enabling Instructional Image Editing with In-Context Generation in Large Scale Diffusion Transformer Multi-modal Attention for Speech Emotion Recognition

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:07:53.140578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T16:07:53.054355Z digest=sha256:4cf7281ac331aaf542058660508d23dcc573002d04069bc12de10ad7fe8080f7

Observation de3e0fe3-0290-4195-b903-e405c77dad86 · inbound

OmniConsistency: Learning Style-Agnostic Consistency from Paired Stylization Data cites this paper.

OmniConsistency: Learning Style-Agnostic Consistency from Paired Stylization Data Multi-modal Attention for Speech Emotion Recognition

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:34:08.695045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:34:08.695045Z digest=sha256:961dc1972143cfe0245a748dbce64c6c59fe2f1171a58a13707c4e4575335e53

Observation d99b507b-246d-4ca3-ba17-1fc6b982fe04 · inbound

Learning Annotation Consensus for Continuous Emotion Recognition cites this paper.

Learning Annotation Consensus for Continuous Emotion Recognition Multi-modal Attention for Speech Emotion Recognition

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T13:43:00.880411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:43:00.880411Z digest=sha256:a70ba917cc1f2d43bcbb9258c574533bdf283f7bab808f51073e8be7d6d604c4

Observation 70218a74-fba3-4563-817a-34557ad3eb31 · inbound

RelationAdapter: Learning and Transferring Visual Relation with Diffusion Transformers cites this paper.

RelationAdapter: Learning and Transferring Visual Relation with Diffusion Transformers Multi-modal Attention for Speech Emotion Recognition

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T11:26:28.052890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:26:28.052890Z digest=sha256:11c30f9b170e38d8e04fa72bac9101156c337ea52b9b0e0cff39a700b503ee88

Observation 9ced870f-7f9a-4704-bda6-22122cf3619b · inbound

Image Editing As Programs with Diffusion Models cites this paper.

Image Editing As Programs with Diffusion Models Multi-modal Attention for Speech Emotion Recognition

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T10:51:38.472327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:51:38.472327Z digest=sha256:05e8edd2b7eef9ae8400f7cbaf97d7533eb34efe60a13d17ac8a2583a4ea06b8

Observation e91dbf3d-efc3-490d-bece-3e8589f93713 · inbound

WordCraft: Interactive Artistic Typography with Attention Awareness and Noise Blending cites this paper.

WordCraft: Interactive Artistic Typography with Attention Awareness and Noise Blending Multi-modal Attention for Speech Emotion Recognition

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T17:57:32.728477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:57:32.728477Z digest=sha256:cd6cb7742c01dd9280ad992ad62a19d4a291e1a760e45eb22b5bdcd4b0fe271b