Pith. sign in

Paper Citation Record · LEDGER

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition

As of 18 August 2026, this Paper Citation Record lists 68 of 68 outbound references and 0 inbound Pith citation observations for arXiv:2501.10408.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.10408 v1

Coverage vector

measured 68 of 68 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T22:03:46.423506Z

measured 68 of 68 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

68 of 68 outbound references displayed

  • verified exact1
  • verified fuzzy52
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5d9f458a-2b9d-42e0-97c9-c9e1efbf47d0 · outbound

This paper cites Survey on speech emotion recognition: Features, classification schemes, and databases,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Survey on speech emotion recognition: Features, classification schemes, and databases,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:46.158505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:46.158505Z digest=sha256:c52c626dc4cb577d00eaec18ec997e7f7afae995c610cd92d38906a309817f45

Observation 7fcfe993-29b8-4819-9a9d-b516735e2180 · outbound

This paper cites Automated screening for distress: A perspective for the future,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Automated screening for distress: A perspective for the future,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:47.235024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.163148Z digest=sha256:ff6a190e70fe4ffb3b69269847bb2029fb0275802678cb525e1ffa58afd40be9

Observation 2136bb00-d04d-4c7c-8408-55edfd3d2f94 · outbound

This paper cites A comprehensive review of speech emotion recognition systems,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition A comprehensive review of speech emotion recognition systems,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:47.222037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.167285Z digest=sha256:529b5bd0b1d8bf9ceabd2355f40f62f4c8fdefbe4ce9c32cd9a2e5aefe779ad0

Observation 6bfd99bb-f33e-4e79-b448-8ddcb0ce83e9 · outbound

This paper cites A systematic review on affective computing: Emotion models, databases, and recent advances,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition A systematic review on affective computing: Emotion models, databases, and recent advances,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:47.208815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.171506Z digest=sha256:1198e4ad6e3494a75ba34aa594ed66c52c2c1bb192861508549622acbfd6e141

Observation f4bdc4c7-2d10-403c-aced-091b912dd74e · outbound

This paper cites Speech emotion recognition based on hmm and svm,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Speech emotion recognition based on hmm and svm,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:47.196125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.175572Z digest=sha256:3613bf774271c911e2dc79e38b6501ef2a157a2682b79d7b41d927b40c6812a7

Observation 8e9a929a-100d-4543-a310-aea698f690b2 · outbound

This paper cites Speech emotion recognition using fourier parameters,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Speech emotion recognition using fourier parameters,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:47.183388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.179384Z digest=sha256:16ce01ff2f37c9313de795a4aa8125b861ccb89e13440fa2a65af7a3d1458d54

Observation cafb2ea4-c30f-4b5d-804b-71580730622a · outbound

This paper cites Implementation and comparison of speech emotion recognition system using gaussian mix- ture model (gmm) and k-nearest neighbor K-NN techniques,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Implementation and comparison of speech emotion recognition system using gaussian mix- ture model (gmm) and k-nearest neighbor K-NN techniques,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:47.171983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.183781Z digest=sha256:8910a47221e698d16faa25fe7e835e6a07eb23e608302c214607fab5ff5053f1

Observation 53bd97bb-f986-429d-bb1e-b3e633253f65 · outbound

This paper cites Speech emotion recognition using deep learning techniques: A review,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Speech emotion recognition using deep learning techniques: A review,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:46.187538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:46.187538Z digest=sha256:e8fc4839117d3251207eb416974fa41f7fe968fc10463ba467862a87593ba6c1

Observation eba3a012-aedf-4141-aac3-2c045f8600fa · outbound

This paper cites Deep learning approaches for speech emotion recognition: State of the art and research challenges,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Deep learning approaches for speech emotion recognition: State of the art and research challenges,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:47.151911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.191149Z digest=sha256:be9ef80d107b960c9ea55eb9c86c82624653e4c9abc017e20cd1b99e83b280d8

Observation 10111b15-b7e0-4516-a1bb-3847a9d1bf30 · outbound

This paper cites Speech emotion recognition: A review,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Speech emotion recognition: A review,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:47.140142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.194664Z digest=sha256:1cf463325d82bf50abe991c9ebc2564f6e34c2b96aebb05e46bbbc51dd3f5ec5

Observation 7ca597b3-62f9-495e-a587-5efba9288d31 · outbound

This paper cites Self-supervised speech representation learning: A review,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Self-supervised speech representation learning: A review,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:47.129132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.198403Z digest=sha256:dcf9a7a3a0e85236157ffeab33b909f6fe7383c0ef40133cb89b8f6d4df8c6a8

Observation 54ea93b7-b461-4e08-ac93-dcf46326a465 · outbound

This paper cites Cross-corpus speech emotion recognition using semi-supervised transfer non-negative matrix factorization with adapta- tion regularization.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Cross-corpus speech emotion recognition using semi-supervised transfer non-negative matrix factorization with adapta- tion regularization

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:47.117859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.202037Z digest=sha256:f947475ce45dfae07593019807f8bc591892812688a61d501f2f420ce8c2b30b

Observation 5aafa406-d37a-4a45-b8ff-02121f1d7764 · outbound

This paper cites Multisource i-vectors domain adaptation using maximum mean discrepancy based autoencoders,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Multisource i-vectors domain adaptation using maximum mean discrepancy based autoencoders,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:47.107450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.205681Z digest=sha256:0aefed555a9d44eb9b0c7248cb25bea7b026b966a317bfdf46b696e69f7f80e0

Observation 6004547b-3651-4129-b6c9-89687eefa887 · outbound

This paper cites Self-supervised learning for multimedia recommendation,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Self-supervised learning for multimedia recommendation,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:46.209137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:46.209137Z digest=sha256:460f5c135fd340f2f0662cf4fda8a555a9059bceb99f09d88db489dbbfe73030

Observation f35da031-9485-42b1-8e35-3962b573b406 · outbound

This paper cites V oicepm: A robust privacy measurement on voice anonymity,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition V oicepm: A robust privacy measurement on voice anonymity,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:47.089422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.212888Z digest=sha256:fc96b11d40538541e0e8d7164ca8de37cc9458c3b37d486ac6475b6a835ca910

Observation 9b40a672-56f6-44db-bbc9-79e3aa93cda2 · outbound

This paper cites EmoBox: Multilingual Multi-corpus Speech Emotion Recognition Toolkit and Benchmark.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition EmoBox: Multilingual Multi-corpus Speech Emotion Recognition Toolkit and Benchmark

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:46.216697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:46.216697Z digest=sha256:e9a01b1ec6255f175b5d1a80068e6f74ef28473300f030246a689187c625fb0d

Observation 93248707-0b57-4517-b528-d7068f787745 · outbound

This paper cites Distilhubert: Speech rep- resentation learning by layer-wise distillation of hidden-unit bert,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Distilhubert: Speech rep- resentation learning by layer-wise distillation of hidden-unit bert,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:47.078403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.221057Z digest=sha256:549a3c8c72fd525a885b16ffcfd51cdc8702b9b301cad4a2317e580a21fbab56

Observation 400e1269-1d19-4870-beeb-31daa841da73 · outbound

This paper cites Cross- corpus speech emotion recognition with hubert self-supervised represen- tation,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Cross- corpus speech emotion recognition with hubert self-supervised represen- tation,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:47.054835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.228550Z digest=sha256:40ff068ec84628049b846918c8c42098119f4fcd37e56f5a6f49a3b91c6498cc

Observation 9ca3722d-d12f-4a91-a515-6e85776180bb · outbound

This paper cites Representation learning through cross-modal conditional teacher-student training for speech emotion recognition,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Representation learning through cross-modal conditional teacher-student training for speech emotion recognition,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:47.032530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.236144Z digest=sha256:35b4a79887082ff8e4f126eebb1742065bc51370eeaa95eea76d23a6c291d0ca

Observation e9528ff4-ea07-44b5-8e2c-e783a2cdc060 · outbound

This paper cites Multi-lingual multi-task speech emotion recognition using wav2vec 2.0,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Multi-lingual multi-task speech emotion recognition using wav2vec 2.0,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:47.043852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.239811Z digest=sha256:d7f6075e022167c3da7b5d9756d90b747ad76f40d7583977222591dad8fe8034

Observation 97025186-4438-4199-ab80-91f3e70e8bf6 · outbound

This paper cites A systematic literature review of speech emotion recognition approaches,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition A systematic literature review of speech emotion recognition approaches,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:47.020648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.243277Z digest=sha256:6a9dd227266ce28a91305144e897c87a2d91ec0e5e5a1403e81b5b660b0db000

Observation 43f0ceff-e65e-4675-806d-ec65f6e7d3bf · outbound

This paper cites Speech emotion recognition using sequential capsule net- works,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Speech emotion recognition using sequential capsule net- works,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:47.008668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.246846Z digest=sha256:c0108bec08bc75cc5a1ab42dab8a26458698e650cf3b3d98d0ff5e409a5e1e6b

Observation f0550e3e-0284-4f9e-958b-b921ddc39129 · outbound

This paper cites Transformer based unsupervised pre-training for acoustic representation learning,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Transformer based unsupervised pre-training for acoustic representation learning,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:46.996977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.250373Z digest=sha256:7fb8e4dcd6b30c12d82644dd3e58e4e3dbe96f9848d640898e02173a96642406

Observation 4b49e13e-484d-49cb-9b74-d3734ff4ceff · outbound

This paper cites Contrastive unsupervised learning for speech emotion recognition,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Contrastive unsupervised learning for speech emotion recognition,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:46.985001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.253761Z digest=sha256:bbf782353f8ee76765168398c8cf249798d2e1af5a618e391622e54803d3b680

Observation 343226f7-59e1-4756-ace0-0a571e0216ac · outbound

This paper cites Cross-corpus classification of realistic emotions–some pilot experiments,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Cross-corpus classification of realistic emotions–some pilot experiments,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:46.972663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.256684Z digest=sha256:2b29204fd435e3143219dfc19460e27a60ef27c72ae489989408a250510855b1

Observation 90c716c7-c6bf-4905-94a4-07fe633321bb · outbound

This paper cites Using multiple databases for training in emotion recognition: To unite or to vote?.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Using multiple databases for training in emotion recognition: To unite or to vote?

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:46.962020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.259863Z digest=sha256:4eaa399b17701de2a91b6b16aba9d061f1a5088c16cb6f4b3dc474ec1242fbc6

Observation e89833ab-4ca0-44a7-984c-3b6c2cf0ecf0 · outbound

This paper cites Cross lingual speech emotion recognition: Urdu vs. western languages,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Cross lingual speech emotion recognition: Urdu vs. western languages,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:46.937417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.265679Z digest=sha256:f916e9c3a1f699a912e54046fc46d7e6f4965815ec0a834a693836f6a1cfe22d

Observation 0b3c17dd-5112-46e6-a076-36c1360e710b · outbound

This paper cites A study on cross-corpus speech emotion recognition and data augmenta- tion,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition A study on cross-corpus speech emotion recognition and data augmenta- tion,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:46.923837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.268532Z digest=sha256:bdedd9af0fdb91c9cd4a7844eb7fba18f21f6fb2a34fcc8631a424460b112926

Observation f14ac591-040d-4b74-b210-26fa306e9e72 · outbound

This paper cites Wavlm: Large-scale self-supervised pre- training for full stack speech processing,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Wavlm: Large-scale self-supervised pre- training for full stack speech processing,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:46.271286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:46.271286Z digest=sha256:8d6aa6345911ad8956ed9629a737dc97bb8f12babcc8ec1f728e5f18cad90389

Observation fa02346d-452d-47ac-91a8-54bdf61fe6c4 · outbound

This paper cites Emotion Recognition from Speech Using Wav2vec 2.0 Embeddings.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Emotion Recognition from Speech Using Wav2vec 2.0 Embeddings

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:46.274112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:46.274112Z digest=sha256:476b6031a8575d4520ca3a42da44ce42d9cdf7bd21d3b4f3ecbd35bdb1aacb0b

Observation 1a3dadf8-57c4-421f-9ee0-4d15698724f0 · outbound

This paper cites Hubert: Self-supervised speech representation learning by masked prediction of hidden units,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Hubert: Self-supervised speech representation learning by masked prediction of hidden units,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:46.277544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:46.277544Z digest=sha256:5dd5cdf53f54907152a017e83e941cf7473fb2038bf710211bbe2fb536045cb9

Observation c68edca5-7a1a-426c-a7ff-02e20c3b098d · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:46.280610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:46.280610Z digest=sha256:6c1196d77ed4b5f62028d76b4cd7c7b0b2717d71c60393adbb942d7bc2e6422f

Observation 037f3f20-ed5d-4395-84af-fe1224603bdc · outbound

This paper cites Unveiling em- bedded features in wav2vec2 and hubert msodels for speech emotion recognition,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Unveiling em- bedded features in wav2vec2 and hubert msodels for speech emotion recognition,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:46.894905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.284910Z digest=sha256:6b7a1a8d5f1549706ab1bb37dd2859e1fd1b17adbc05b1cf1a251c76958ceb48

Observation 338da2a7-2f0a-4d57-b961-6368ccf41018 · outbound

This paper cites Layer-wise analysis of a self- supervised speech representation model,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Layer-wise analysis of a self- supervised speech representation model,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:46.879795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.288780Z digest=sha256:bf8af7df2e2f537436be8fa62ddc938b1fd51631e534c7dd26e552552c0e4b3f

Observation 1fa10ff4-f3f1-48d7-ba42-7b07a6951420 · outbound

This paper cites Multiple acoustic features speech emotion recognition using cross-attention transformer,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Multiple acoustic features speech emotion recognition using cross-attention transformer,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:46.867830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.292238Z digest=sha256:fbbd6244a64269d1f522e51d22a963c2fbfa92269e6acda325a335fd96eb7851

Observation f685dcdd-657d-4155-84ac-8a6c61275ade · outbound

This paper cites Speech emotion recognition using local and global features,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Speech emotion recognition using local and global features,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:46.854580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.295572Z digest=sha256:eb17d7098e0278068faea42e4a85521edaed3badc3469933d7fc1fc9df39a09b

Observation cb2f5bd9-4d17-4cab-8f74-c2793d99fdaf · outbound

This paper cites Modeling prosodic features with joint factor analysis for speaker verification,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Modeling prosodic features with joint factor analysis for speaker verification,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:46.299358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:46.299358Z digest=sha256:a948f9824e13e2deb095cea5accb49260b7e7a8a4f9ba7bd041a12317f904ea3

Observation d292971c-9eb5-4d50-bf3c-69d234cad1e6 · outbound

This paper cites A novel feature selection method for speech emotion recognition,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition A novel feature selection method for speech emotion recognition,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:46.833622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.302957Z digest=sha256:aa4f6d558494cb3a4fb82bd5bd17bf297cddfe3bf2183e7c35cb68737cf78cc0

Observation dfc004f2-7f58-48e6-89a4-2625872800b5 · outbound

This paper cites Analysis of linguistic and prosodic features of bilingual arabic–english speakers for speech emotion recognition,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Analysis of linguistic and prosodic features of bilingual arabic–english speakers for speech emotion recognition,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:46.819871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.306418Z digest=sha256:48b56676ade3a959907dec194e42c0110d425f559038bc024e0ffa85b5474486

Observation 8af5c4ff-fc55-4de5-b71d-676536114824 · outbound

This paper cites Towards an automatic evaluation of the dysarthria level of patients with parkinson’s disease,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Towards an automatic evaluation of the dysarthria level of patients with parkinson’s disease,

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:46.310892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:46.310892Z digest=sha256:ddf754ffb82cab53dae85da3d8b781647fad662953b06d14f4fb82e9ff0c1b08

Observation 5337d1f6-a749-4037-a1d9-b67e06b47cbd · outbound

This paper cites Speech emotion recognition based on multiple acoustic features and deep convolutional neural network,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Speech emotion recognition based on multiple acoustic features and deep convolutional neural network,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:46.799357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.314890Z digest=sha256:7766a82de79b62c7c2fdf155ec3ef787a1255fbbf1b851c266bbd6e1a4c0657c

Observation c56a6026-5de0-47d0-b64f-14c5bcbfcb22 · outbound

This paper cites Speech emotion recognition using mel frequency log spectrogram and deep convolutional neural network,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Speech emotion recognition using mel frequency log spectrogram and deep convolutional neural network,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:46.788227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.318695Z digest=sha256:de8cce6780e158e5a8bf7cf75729ad6a5444669d08367df31e6898918faafe09

Observation 2e3501cb-d60e-466a-8c92-4c2ec6c94b6a · outbound

This paper cites Learning deep features to recognise speech emotion using merged deep CNN,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Learning deep features to recognise speech emotion using merged deep CNN,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:46.778509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.323360Z digest=sha256:e4f2db2bed118d94275f2d7a13e136fec24ad80ff049aec23156fccc8d9453df

Observation aaa42b2d-2d11-4800-a419-be10441767e2 · outbound

This paper cites The SpeakIn Speaker Verification System for Far-Field Speaker Verification Challenge 2022.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition The SpeakIn Speaker Verification System for Far-Field Speaker Verification Challenge 2022

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-08-10T22:03:46.491560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.327790Z digest=sha256:76c9a46a8f02c458f2546eb1685012a89eea478000e14e3a4a71ac00a9979d5d

Observation 1adcb3af-0096-48bd-982d-03183cda01fb · outbound

This paper cites IEMOCAP: Interactive emotional dyadic motion capture database,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition IEMOCAP: Interactive emotional dyadic motion capture database,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:46.768265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.332005Z digest=sha256:f941705f8ddcf540345cf6bc6dfc9017a49fb048a3e06575269e9aaec20efd79

Observation 10d4eb0b-110f-4a64-8e1b-463a04f36d52 · outbound

This paper cites The ryerson audio-visual database of emotional speech and song RA VDESS: A dynamic, multimodal set of facial and vocal expressions in north american english,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition The ryerson audio-visual database of emotional speech and song RA VDESS: A dynamic, multimodal set of facial and vocal expressions in north american english,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:46.755318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.335878Z digest=sha256:4b01e97244745c98d8b567f35749e28862dda13e3e588766f380cbf88469fd2d

Observation bc1a721e-f373-4eb8-9588-e8a26d4e8a97 · outbound

This paper cites Real- time end-to-end speech emotion recognition with cross-domain adapta- tion,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Real- time end-to-end speech emotion recognition with cross-domain adapta- tion,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:46.742387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.339578Z digest=sha256:9bedd1cc123372d1950dd85d2d87e691691496a4d0fcfc33153962ceee3503b6

Observation caa75239-bb26-49a2-a7f8-401f5e1cd636 · outbound

This paper cites A database of german emotional speech.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition A database of german emotional speech

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:46.729141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.343453Z digest=sha256:eab7ad4b8c1bc9fa5ecfe6acc8aebb40f5de694793aa8ef7aac16e46095b311c

Observation b185e9a2-5251-4935-9818-eec808a3723e · outbound

This paper cites Emovo corpus: an italian emotional speech database,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Emovo corpus: an italian emotional speech database,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:46.716831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.347276Z digest=sha256:5399fc9d2bf63d3516b3b9b163c8285f5b2d8622f044015c6b7efa38864f2504

Observation 7db1f795-113e-4958-87b5-8f0d9812ddd4 · outbound

This paper cites The mexican emotional speech database (mesd): elaboration and assessment based on machine learning,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition The mexican emotional speech database (mesd): elaboration and assessment based on machine learning,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:46.703398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.351021Z digest=sha256:925e5d0bc0097d0dc77f0c36e3590d86d988ff7feaf583a225425da51c5e1a94

Observation 2b29674c-226a-49af-92e2-a3054f1ea25a · outbound

This paper cites Emotional voice conversion: Theory, databases and esd,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Emotional voice conversion: Theory, databases and esd,

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:46.354658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:46.354658Z digest=sha256:c5d2741617555e416434fce92a25fbf17136a97f7e9c100642e3809eaa0bc922

Observation b28167ab-7754-40e5-bda5-81796de291b3 · outbound

This paper cites Towards discriminative representations and unbiased predictions: Class-specific angular softmax for speech emotion recognition.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Towards discriminative representations and unbiased predictions: Class-specific angular softmax for speech emotion recognition

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:46.681547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.358869Z digest=sha256:ff0ca637649aea6c2a79df1d838d33c2a7dd7c220509dbbb25afb2e3b259b69c

Observation f877a1a3-1829-47ba-af7a-88ba5f563faf · outbound

This paper cites Improving speech emotion recognition using graph attentive bi-directional gated recurrent unit network,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Improving speech emotion recognition using graph attentive bi-directional gated recurrent unit network,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:46.669951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.362435Z digest=sha256:4bb89e1f1118288bdee27bd7de39e985fd77a3624919ab36720d63e93eef6288

Observation 910aa210-3996-44b8-a6af-1394c5706a06 · outbound

This paper cites Hgfm: A hierarchical grained and feature model for acoustic emotion recognition,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Hgfm: A hierarchical grained and feature model for acoustic emotion recognition,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:46.657924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.366274Z digest=sha256:4570b19d2d6e22c60d46e6fd16c4ce0f3cc222b07f4d7444ae0755988a1e1f28

Observation a9dcfab3-f59b-4bc0-a18a-3285263b20b0 · outbound

This paper cites Cross- corpus speech emotion recognition with hubert self-supervised represen- tation,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Cross- corpus speech emotion recognition with hubert self-supervised represen- tation,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:46.646773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.370150Z digest=sha256:c5179a90b7ddfda3075ad962130858f6ef583daf891b61e35ffe8e92ceebacef

Observation b319171d-682a-4f14-bbf7-d215ffdd5c9b · outbound

This paper cites Temporal modeling matters: A novel temporal emotional modeling approach for speech emotion recognition,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Temporal modeling matters: A novel temporal emotional modeling approach for speech emotion recognition,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:46.635793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.374518Z digest=sha256:38ae3df3616a6a0497e00b433e98e9835016c4dd995b11abb1c112e0cd1e57a5

Observation 12670868-5e41-4cfc-ba70-eaa9de08a6bb · outbound

This paper cites Learning multi-scale features for speech emotion recognition with connection attention mechanism,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Learning multi-scale features for speech emotion recognition with connection attention mechanism,

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:46.622964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.378567Z digest=sha256:eb104bac2d34be04a84cbbb4a9666244032f3e18636ee5841cbc6d7801ee127d

Observation 9b089bd3-6681-4a5b-bdf6-203a974d190d · outbound

This paper cites Exploring wav2vec 2.0 fine tuning for improved speech emotion recognition,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Exploring wav2vec 2.0 fine tuning for improved speech emotion recognition,

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:47.065910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.382309Z digest=sha256:113000d6a7f4f58ebe0518c25bf847422e858786a4ec7f93356004cda031b77b

Observation ef116f2c-77f8-4266-ba98-430ea4343ea2 · outbound

This paper cites Unsupervised adversarial domain adaptation for cross-lingual speech emotion recognition,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Unsupervised adversarial domain adaptation for cross-lingual speech emotion recognition,

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:46.610408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.385929Z digest=sha256:173aa4796c1ed8f022b6ec0cf6c7931c5555fff443bd0e629ab2b3074f18ee1f

Observation 66fd3c04-cf38-4ab0-897b-42ef636f72b9 · outbound

This paper cites Speech emotion recognition from 3D log-mel spectrograms with deep learning network,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Speech emotion recognition from 3D log-mel spectrograms with deep learning network,

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:46.597142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.390039Z digest=sha256:39bef47b1d187ce69faab9a77cece521048493d169c469650fd33274107a717c

Observation 1a243e78-a50c-4836-b71b-d274313efb63 · outbound

This paper cites Fusing visual attention CNN and bag of visual words for cross-corpus speech emotion recognition,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Fusing visual attention CNN and bag of visual words for cross-corpus speech emotion recognition,

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:46.583333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.393861Z digest=sha256:93300ffb541ee10cd42987d7dbfedaf8a8a3ece6f27850be04257bfae1de8ddb

Observation 0dd1f7e1-74d4-475d-8937-b7ce804bb2bf · outbound

This paper cites Cross corpus multi-lingual speech emotion recognition using ensemble learning,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Cross corpus multi-lingual speech emotion recognition using ensemble learning,

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:46.949879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.397391Z digest=sha256:e44f7aff479e9bc1bb1cb05cf7111f8da2e7587fcda3250c1e93c675d11c285f

Observation 7fe177a5-5e7c-4965-8f01-0b397131db6e · outbound

This paper cites Cross-corpus speech emotion recognition based on few-shot learning and domain adaptation,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Cross-corpus speech emotion recognition based on few-shot learning and domain adaptation,

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:46.572085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.400427Z digest=sha256:a46fa5b29d48f829b91ceb23a75ce6125f4d0a14fbc00061116ac6db3935f12c

Observation 01a88162-f488-419d-9c08-11aa484af918 · outbound

This paper cites Crema-d: Crowd-sourced emotional multimodal actors dataset,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Crema-d: Crowd-sourced emotional multimodal actors dataset,

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:46.404732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:46.404732Z digest=sha256:38ec28b66a3ae17dc3dd6fbf777258d2ea4f2611c0540dbd223fcea6c68297a2

Observation e771e516-c8f7-4423-9b4d-68f39fb96eb9 · outbound

This paper cites Enhancing cross-language multimodal emotion recognition with dual attention transformers,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Enhancing cross-language multimodal emotion recognition with dual attention transformers,

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:03:46.550903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-10T22:03:46.409923Z digest=sha256:ffde788d673ab07ecd32064ecaf6957f89f1f0ce88288b562747d143ad991342

Observation 76638cb0-65de-4f8a-8bdc-db5c787a6d3e · outbound

This paper cites RoBERTa: A Robustly Optimized BERT Pretraining Approach.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition RoBERTa: A Robustly Optimized BERT Pretraining Approach

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:46.414254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:46.414254Z digest=sha256:17b05c7a30712399b24900fe035382c371656f659e55325cf75ab72abb62b3a6

Observation 438ddb48-b1ce-4bc2-871a-c440909606d6 · outbound

This paper cites XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:46.418497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:46.418497Z digest=sha256:2df289c1b4f94b61f3869abd6c8cdb755e9120a1cb866288598ac2fb7b09fa4b

Observation f224eeb2-635a-4306-8855-a694e17bbd5d · outbound

This paper cites Robust speech recognition via large-scale weak supervi- sion,.

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition Robust speech recognition via large-scale weak supervi- sion,

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-10T22:03:46.423506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:03:46.423506Z digest=sha256:2eee6b5718e717751ec6473ca65f0c9f732fbd5e5a8731f1acf57a5f816171ad

Pith citing papers

No inbound Pith citation observations are available.