Pith. sign in

Paper Citation Record · LEDGER

The MSP-Podcast Corpus

As of 17 August 2026, this Paper Citation Record lists 100 of 111 outbound references and 14 inbound Pith citation observations for arXiv:2509.09791.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.09791 v1

Coverage vector

measured 100 of 111 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T16:09:02.342775Z

measured 114 of 114 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T11:50:20.922023Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-11T01:57:52.098079Z

Reference resolution

100 of 111 outbound references displayed

  • verified exact2
  • verified fuzzy59
  • unresolved39
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8f9a7e37-0baa-4759-9b61-e8d8cbe5743c · outbound

This paper cites Toward effective automatic recognition systems of emotion in speech,.

The MSP-Podcast Corpus Toward effective automatic recognition systems of emotion in speech,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T16:08:59.863001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:08:59.863001Z digest=sha256:038f33974e41c3ee0bc295425b4265e5c5bab6446ce80c72ae4fa605cf361c1a

Observation fd61180b-aa67-4271-b21f-3effcd823910 · outbound

This paper cites CREMA-D: Crowd-sourced emotional multimodal actors dataset,.

The MSP-Podcast Corpus CREMA-D: Crowd-sourced emotional multimodal actors dataset,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:01.796760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:01.796760Z digest=sha256:9bb39b4458e28614eb5fe1aa4715768c1d7c772f96b56cd07fad05555779806e

Observation 7fbf5c49-2f80-4879-8195-411f4471d488 · outbound

This paper cites Emotionalprosodyspeechandtranscripts,.

The MSP-Podcast Corpus Emotionalprosodyspeechandtranscripts,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:01.826349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:01.826349Z digest=sha256:dcc148d0ac35f9ab3412acabd732a7e20a973eeff758db7056ebad03603b4bad

Observation 35a67b79-5e4e-41d5-998b-ff02a5af72ff · outbound

This paper cites AdatabaseofGermanemotionalspeech,.

The MSP-Podcast Corpus AdatabaseofGermanemotionalspeech,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:01.840882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:01.840882Z digest=sha256:bbe7b2443447932782cd2546d4650ed3382987b05bc346ffb0e604c449b733be

Observation 47edf37a-4e27-4ddf-ae9e-f3bdd4bf2b19 · outbound

This paper cites The Ryerson audio-visual database of emotional speech and song (RAVDESS): A dy- namic, multimodal set of facial and vocal expressions in North American English,.

The MSP-Podcast Corpus The Ryerson audio-visual database of emotional speech and song (RAVDESS): A dy- namic, multimodal set of facial and vocal expressions in North American English,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:01.855884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:01.855884Z digest=sha256:3b639d170e665b257de17423db36e61934b875e453ee1ef7f7414b699e405bb7

Observation 674c11e0-3576-45e6-a936-d5356d948436 · outbound

This paper cites IEMOCAP: Interactive emotional dyadic motion capture database,.

The MSP-Podcast Corpus IEMOCAP: Interactive emotional dyadic motion capture database,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:01.907003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:01.907003Z digest=sha256:811f2af96f895d8472748a1dd28627106b5d19aba1185acc6dd7aa66e9b81387

Observation 2486cae9-cfc4-467a-8798-e6c2a5a697ec · outbound

This paper cites Anewemotion database: considerations, sources and scope,.

The MSP-Podcast Corpus Anewemotion database: considerations, sources and scope,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:01.925881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:01.925881Z digest=sha256:ecb5ca5e6a399068474d29fe05380dd72718802f92560f4405563cae5a01780e

Observation 106b6ab1-fa53-49cb-aba6-0d188d0f9923 · outbound

This paper cites MSP-IMPROV: An acted corpus of dyadic interactions to study emotion percep- tion,.

The MSP-Podcast Corpus MSP-IMPROV: An acted corpus of dyadic interactions to study emotion percep- tion,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:01.946215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:01.946215Z digest=sha256:b0afff9cb13443aa622391fccd49dfd4517ab88e1404622a8aa72a7d992b5161

Observation adbe7e65-b2f4-4497-ae22-85b981040ca4 · outbound

This paper cites In- troducing the RECOLA multimodal corpus of remote col- laborative and affective interactions,.

The MSP-Podcast Corpus In- troducing the RECOLA multimodal corpus of remote col- laborative and affective interactions,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:01.958210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:01.958210Z digest=sha256:514cb1a3307adbf3a8d91d3957eabdcf6401af6c4fe6e7a183cb6acba810d90d

Observation 0ed8b957-8944-4674-bd1c-e471b269b718 · outbound

This paper cites NNIME: The NTHU-NTUA Chinese interactive multimodal emotion corpus,.

The MSP-Podcast Corpus NNIME: The NTHU-NTUA Chinese interactive multimodal emotion corpus,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:01.967723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:01.967723Z digest=sha256:50543f8e9335e12e39132bbf33c12437ef4f0a1b6b5ba93f6bee6e4c1610d9e5

Observation 978463f2-199a-438b-accc-e287e0647145 · outbound

This paper cites Memor: A dataset for multimodal emotion reasoning in videos,.

The MSP-Podcast Corpus Memor: A dataset for multimodal emotion reasoning in videos,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:01.971721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:01.971721Z digest=sha256:e24a6935de7c22ccdd1ad4c0504ff88e5dd0dd25277ec94b6608b18686541887

Observation 1847b916-f835-4814-b594-e2eab35556fe · outbound

This paper cites The Vera am Mit- tag German audio-visual emotional speech database,.

The MSP-Podcast Corpus The Vera am Mit- tag German audio-visual emotional speech database,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:01.975340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:01.975340Z digest=sha256:c2219ef20d57f0c8ac80395e8f792898d8c373f1f27511eabd3b72d10e7ef84d

Observation 31fe17f7-a58a-4bd6-9335-d2a67bb31cd5 · outbound

This paper cites MELD: A multimodal multi-party dataset for emotion recognition in conversations,.

The MSP-Podcast Corpus MELD: A multimodal multi-party dataset for emotion recognition in conversations,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:01.985323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:01.985323Z digest=sha256:b810a4ae17b3bf5bbbc8456ac0cb5c7e534a79a4da25902cd3546b32437d7a39

Observation fcce79cf-49d4-421e-ab31-2fed1a9ec2df · outbound

This paper cites MSP-face corpus: A natural audiovisual emotional database,.

The MSP-Podcast Corpus MSP-face corpus: A natural audiovisual emotional database,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:01.989608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:01.989608Z digest=sha256:0f659c1e7a818115e1a7fcb0974eebf5f7e8bc2a896821b04f5ce306b25d8a67

Observation 35f6950b-60b7-44ec-bb07-3aec2241814a · outbound

This paper cites An intelligent infrastructure toward large scale naturalistic affective speech corpora collection,.

The MSP-Podcast Corpus An intelligent infrastructure toward large scale naturalistic affective speech corpora collection,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:01.992747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:01.992747Z digest=sha256:e723b6b09116750ac666b57da88b84d1a95087e9e9c26d6b89c67760f4b8cb82

Observation 9af801ac-6950-4860-a099-ef64c1975112 · outbound

This paper cites Deep affect prediction in-the-wild: Aff-wild database and challenge, deep architectures, and beyond,.

The MSP-Podcast Corpus Deep affect prediction in-the-wild: Aff-wild database and challenge, deep architectures, and beyond,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:01.997141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:01.997141Z digest=sha256:0438b461e978cc51772981d3896a99896d312253277e51f545d5c3c53dd9dcef

Observation 9e3e9d27-0db1-4082-a04b-e45b0e5c2cb6 · outbound

This paper cites Building naturalistic emotionally balanced speech corpus by retrieving emotional speech from existing podcast recordings,.

The MSP-Podcast Corpus Building naturalistic emotionally balanced speech corpus by retrieving emotional speech from existing podcast recordings,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:02.005328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:02.005328Z digest=sha256:27a14c806abe01fb601511311ba4c4c1dfb465f77670c458ab38787c97b43525

Observation b6dc6f42-bb49-4561-989e-67d0152688d3 · outbound

This paper cites The MSP- conversation corpus,.

The MSP-Podcast Corpus The MSP- conversation corpus,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:02.009053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:02.009053Z digest=sha256:429a2f0d3ffa99dedc5c828d40e79a0701230f18b748ac3533dd9dc5f7d9f1ba

Observation 016abcd5-8231-454c-ae05-bd9fd5c2f9bd · outbound

This paper cites Buildinganaturalis- tic emotional speech corpus by retrieving expressive behaviors from existing speech corpora,.

The MSP-Podcast Corpus Buildinganaturalis- tic emotional speech corpus by retrieving expressive behaviors from existing speech corpora,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:02.011781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:02.011781Z digest=sha256:52f0642ccf0706270195358bae321c3237f905432c2d0c81fe5c7bc0f6623e8d

Observation 8316d031-91fc-496b-94a3-25abb03d908c · outbound

This paper cites Hybrid Dataset for Speech Emotion Recognition in Russian Language,.

The MSP-Podcast Corpus Hybrid Dataset for Speech Emotion Recognition in Russian Language,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:02.015052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:02.015052Z digest=sha256:63bc5f8b2b3c9c40b58b2e33dead41c0bda8966f8431c4a14de8a411ce7408c4

Observation 4e714a53-0317-48e1-80ab-d0fc9f266a9f · outbound

This paper cites Crowdsourcing emotional speech,.

The MSP-Podcast Corpus Crowdsourcing emotional speech,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:02.020211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:02.020211Z digest=sha256:2cf9a6f1b6b180ebc20dbf8350ffb45883218a80d3addb5c90cc8f22af239234

Observation 9178e9e7-759a-497d-b732-5868b6b7d46c · outbound

This paper cites MIKU-PAL: An Automated and Standardized Multi-Modal Method for Speech Paralinguistic and Affect Labeling,.

The MSP-Podcast Corpus MIKU-PAL: An Automated and Standardized Multi-Modal Method for Speech Paralinguistic and Affect Labeling,

Reference 22

Resolution
verified exact
raw_fallback, observed 2026-08-15T16:09:02.599555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.024229Z digest=sha256:d779618e2a34e8e14eb0bcfbf0156d13501e88f6ee07343b3a8037d76f8e5254

Observation 0a09cced-6bc7-4208-9d1d-dd297c7f938a · outbound

This paper cites CMU-MOSEAS: A multimodal language dataset for Spanish, Portuguese, German and French,.

The MSP-Podcast Corpus CMU-MOSEAS: A multimodal language dataset for Spanish, Portuguese, German and French,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:02.027199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:02.027199Z digest=sha256:edc02a3574e71b9568db04a69954384a1846ee147233e32c2b7b3bfa3b6c3bed

Observation 346372b3-eaa2-4502-9d31-20df9e0ef12f · outbound

This paper cites Multimodal language analysis in the wild: CMU-MOSEI dataset and interpretable dynamicfusiongraph,.

The MSP-Podcast Corpus Multimodal language analysis in the wild: CMU-MOSEI dataset and interpretable dynamicfusiongraph,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:02.029284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:02.029284Z digest=sha256:df004999bc633f2fbe99840de98f7bcee05203e5710820157ef961bfcc09de4f

Observation 775b6cf7-566b-4180-ad5c-c70085b77f18 · outbound

This paper cites THAI Speech Emotion Recognition (THAI-SER) corpus.

The MSP-Podcast Corpus THAI Speech Emotion Recognition (THAI-SER) corpus

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-15T16:09:02.549725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.031873Z digest=sha256:65ce5909b71cb92c77a2a9d38b42ae51d5733c48c990dfd2ab72894c0f59b041

Observation fbc19aea-56ae-463c-a118-ebeaa1b5cf45 · outbound

This paper cites Real-lifeemotionsdetectionwith lexical and paralinguistic cues on human-human call center dialogs,.

The MSP-Podcast Corpus Real-lifeemotionsdetectionwith lexical and paralinguistic cues on human-human call center dialogs,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:02.035810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:02.035810Z digest=sha256:bf838f40796d90ffaf55e85ffab3cb2c4fbb3a2a7a33e2ab94b8a3cdc6dc73a4

Observation c4d2d229-b0da-4a85-a421-3e58b864c489 · outbound

This paper cites Audiovisual recognition of spontaneous in- terest within conversations,.

The MSP-Podcast Corpus Audiovisual recognition of spontaneous in- terest within conversations,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:02.038094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:02.038094Z digest=sha256:a7ea06f0c77edcfb2ebe22d303cad1cefe5749bc43f3f6cc7c2686c509a62f94

Observation c5e8051c-cf17-4869-912c-94ab1b0b3759 · outbound

This paper cites Releasing a thoroughly annotated and processed spontaneous emotional database: the FAU Aibo emotion corpus,.

The MSP-Podcast Corpus Releasing a thoroughly annotated and processed spontaneous emotional database: the FAU Aibo emotion corpus,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:02.040402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:02.040402Z digest=sha256:ddeedb636423f496178681a06c473483c00b278a1d95b1703ea735c0d3a8add1

Observation 8357e5d9-ee15-483c-8d4c-b16828f05018 · outbound

This paper cites MEC 2017: Multimodal Emotion Recognition Challenge,.

The MSP-Podcast Corpus MEC 2017: Multimodal Emotion Recognition Challenge,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:02.042972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:02.042972Z digest=sha256:72acb4a5adeb68f1513f65c30c6a661ae674a0e6e168bc56fd6e9623f92ab4d9

Observation c7ac165f-34fb-49d4-a8e3-8de22533c26f · outbound

This paper cites Demos: An italian emotional speech cor- pus: Elicitation methods, machine learning, and perception,.

The MSP-Podcast Corpus Demos: An italian emotional speech cor- pus: Elicitation methods, machine learning, and perception,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:02.045342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:02.045342Z digest=sha256:7d031c84ed1e8519b97be009db4ced81b1e599b906c1c845ba8118ee34e6eed4

Observation da077649-2c25-43ce-9c2b-70873ac983f5 · outbound

This paper cites Emozionalmente: A Crowdsourced Corpus of Simulated Emotional Speech in Italian,.

The MSP-Podcast Corpus Emozionalmente: A Crowdsourced Corpus of Simulated Emotional Speech in Italian,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:02.048460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:02.048460Z digest=sha256:b58b4db40186d9e0914986aa67368b44eb5325cf365090e76a8bd51ef48e1122

Observation 166951b6-f0e5-40f0-99a5-37794bd4299d · outbound

This paper cites WHiSER: White House Tapes speech emotion recognition corpus,.

The MSP-Podcast Corpus WHiSER: White House Tapes speech emotion recognition corpus,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:02.052114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:02.052114Z digest=sha256:ea59b498a269c08a27d1e31ef947a5289b11c9dfd314db3d287f9600b7d72806

Observation 1a66bc5f-325b-46f2-8d62-c2cfe2d97a6d · outbound

This paper cites The SEMAINE database: Annotated multi- modal records of emotionally colored conversations between a person and a limited agent,.

The MSP-Podcast Corpus The SEMAINE database: Annotated multi- modal records of emotionally colored conversations between a person and a limited agent,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:02.062600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:02.062600Z digest=sha256:5892a885aada73b687fb3884eae1bf9b851c19430af934c60722538606f79c8b

Observation f1520d92-2b65-4bd3-bd38-dd140495e40b · outbound

This paper cites Joint processing of audio-visual information for the recognition of emotional expressions in human-computer inter- action,.

The MSP-Podcast Corpus Joint processing of audio-visual information for the recognition of emotional expressions in human-computer inter- action,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:02.066544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:02.066544Z digest=sha256:5594be7613f7445061190a97d09b3fcc05db297d775f57b6c081cbd4fc293da8

Observation 7c4039b2-349a-44e6-a217-ecb39fa9f100 · outbound

This paper cites UrduSER: A comprehensive dataset for speech emotion recognition in Urdu language,.

The MSP-Podcast Corpus UrduSER: A comprehensive dataset for speech emotion recognition in Urdu language,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:02.071096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:02.071096Z digest=sha256:ae1dd58bbdcc05475ba273ed22c2ade7334d511ab624dc701da56f42c3430f00

Observation 01b7f7fd-4650-4224-af70-bb5d201abcfb · outbound

This paper cites MOSI: Multimodal Corpus of Sentiment Intensity and Subjectivity Analysis in Online Opinion Videos.

The MSP-Podcast Corpus MOSI: Multimodal Corpus of Sentiment Intensity and Subjectivity Analysis in Online Opinion Videos

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:02.074779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:02.074779Z digest=sha256:33e31df89b287fd401b205b451201e0b16ca3b2040c7a3de5cf6ec05ef1f8e0a

Observation 05c97f4c-9621-48e5-9c4f-274f928a8609 · outbound

This paper cites Toronto emotional speech set (tess),.

The MSP-Podcast Corpus Toronto emotional speech set (tess),

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.634577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.092874Z digest=sha256:8c5242e45dac9b424ff52fcebf89c53aaef605c4eae26ccbbbcd150c7ad0034d

Observation 1f7fed9c-a965-4bf7-b28a-b5790297cb7e · outbound

This paper cites Challenges in real- lifeemotionannotationandmachinelearningbaseddetection,.

The MSP-Podcast Corpus Challenges in real- lifeemotionannotationandmachinelearningbaseddetection,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.625529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.103096Z digest=sha256:c5abb01a13963c213932349bd667387028b00af684a855f21e528b2ef234850e

Observation 94cf7f72-be03-4873-8841-70fcf1824531 · outbound

This paper cites Desperately seeking emotions or: actors, wizards and human beings,.

The MSP-Podcast Corpus Desperately seeking emotions or: actors, wizards and human beings,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.617873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.110773Z digest=sha256:74824425f5227a8326196403ebbe4ec5ca388e85641666fb738c5346ffb74d99

Observation 974b5d2c-1e93-4262-a2d1-c93482a11c43 · outbound

This paper cites Recordingaudio-visualemotional databases from actors: a closer look,.

The MSP-Podcast Corpus Recordingaudio-visualemotional databases from actors: a closer look,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.610024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.117050Z digest=sha256:4791e29e508d5688f5aa3e52ba0d93569cad11ebd3c0d180e5338c0308b0d862

Observation c41e13bf-3abe-49fe-8bf6-b79ba6656b42 · outbound

This paper cites CHEAVD: a Chinese natural emotional audio–visual database,.

The MSP-Podcast Corpus CHEAVD: a Chinese natural emotional audio–visual database,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.601211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.120581Z digest=sha256:9ee524865f1e3306f74bc1be7b7115732dfc9c7d62c71ffe88818642c105de57

Observation f2000a9a-5909-419a-8720-4c99eee06326 · outbound

This paper cites Domain-specific adaptation in speech emotion recognition using emotional distribution alignment,.

The MSP-Podcast Corpus Domain-specific adaptation in speech emotion recognition using emotional distribution alignment,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.591633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.123587Z digest=sha256:7bcdbec904c55452ad52451d63bcd6da25794bce8d0e17b71c6cbd8ff572e02d

Observation c86c6350-2c34-4332-9522-ae1a46c7b0c7 · outbound

This paper cites Increasing the reliability of crowdsourcing evaluations using online quality as- sessment,.

The MSP-Podcast Corpus Increasing the reliability of crowdsourcing evaluations using online quality as- sessment,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.580911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.127155Z digest=sha256:d009b7c7864e3bd185bf69827932f9d710b3b2647e17899cc3d09ebe3b0d2d68

Observation e8db9cfc-5043-4653-91a8-c966b32f8bfb · outbound

This paper cites Sample-level deep convolutional neural networks for music auto-tagging using raw waveforms,.

The MSP-Podcast Corpus Sample-level deep convolutional neural networks for music auto-tagging using raw waveforms,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.570704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.136337Z digest=sha256:ee04b5c4f22cc3b21b1378c4fc9cbab7387bf15e05bfc3a1bb9487102b1336e3

Observation dcfa844a-0786-4860-8857-c51b713cb3b2 · outbound

This paper cites Deep learning for minimum mean-square error approaches to speech enhancement,.

The MSP-Podcast Corpus Deep learning for minimum mean-square error approaches to speech enhancement,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.561611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.139953Z digest=sha256:2cc17ce30f738d3d31b6c549710c28353647fec6dc9b717602225c2f006df232

Observation e186b383-360b-414b-bb4e-2479e5c121eb · outbound

This paper cites librosa: Audio and music signal analysis in python,.

The MSP-Podcast Corpus librosa: Audio and music signal analysis in python,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.551665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.143279Z digest=sha256:920e2d9237b427632761a90a6f213e664d685a135d831f6475bdcda766a2bb54

Observation 20c317fc-52d9-4a17-965a-553114828fbf · outbound

This paper cites Robust speech recognition via large-scale weak supervision,.

The MSP-Podcast Corpus Robust speech recognition via large-scale weak supervision,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.542564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.146445Z digest=sha256:c00c288c7a0a7ca9c983b8384327bd93dbbc31fdfa632b22f2de81c6767c054e

Observation 94691933-31d6-47af-aa52-2bff37e5b7c4 · outbound

This paper cites HuggingFace's Transformers: State-of-the-art Natural Language Processing.

The MSP-Podcast Corpus HuggingFace's Transformers: State-of-the-art Natural Language Processing

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:02.151167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:02.151167Z digest=sha256:9a8e823df258821e75f7d12c348306431bdd3157b2472d5270a37d2a2516567f

Observation 23a18f5a-fd25-4007-9439-6074cece97e0 · outbound

This paper cites Python classes for Praat TextGrid and TextTier files (and HTK .mlf files),.

The MSP-Podcast Corpus Python classes for Praat TextGrid and TextTier files (and HTK .mlf files),

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.533031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.155545Z digest=sha256:5c62647ec2e16e0c791d2605e131984ad73fd84ca30bc9b92c477b6525ba207f

Observation 7cd1e95d-f8af-4de4-997c-4b8adb17e0fb · outbound

This paper cites Praat, a system for doing phonetics by com- puter,.

The MSP-Podcast Corpus Praat, a system for doing phonetics by com- puter,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.522633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.159167Z digest=sha256:130698062c8decd83414d3f81214a2fa4b7d27047468902a349744c4bf8f2091

Observation f84f7acd-ec3f-405c-a186-9db90c15514d · outbound

This paper cites Robust signal-to-noise ratio estima- tion based on waveform amplitude distribution analysis,.

The MSP-Podcast Corpus Robust signal-to-noise ratio estima- tion based on waveform amplitude distribution analysis,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.513154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.162773Z digest=sha256:0b8da10d7425bffa3aed06b3307a635742175ae60c3898f03bd1643ca2fd297f

Observation 1831527a-28ec-4f4e-88f0-82bfff0a0a92 · outbound

This paper cites Powerset multi-class cross entropy loss for neural speaker diarization,.

The MSP-Podcast Corpus Powerset multi-class cross entropy loss for neural speaker diarization,

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:02.165233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:02.165233Z digest=sha256:dcb52892a512a58d62d5ebbe058561b1972074838777166deeb38a6404d19ed6

Observation d1e556a7-7852-4d4d-90bb-9c196c4fd8a5 · outbound

This paper cites pyannote.audio 2.1 speaker diarization pipeline: principle, benchmark, and recipe,.

The MSP-Podcast Corpus pyannote.audio 2.1 speaker diarization pipeline: principle, benchmark, and recipe,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.499415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.168649Z digest=sha256:3e67ad6f7a7415e5dd9eaeb20bc93f8b2f44ad3682e851310d829f323c2e2848

Observation 2aaa7dfd-4712-4f31-ba65-7f6bbe17e0ad · outbound

This paper cites An effective gender recognition approach using voice data via deeper lstm networks,.

The MSP-Podcast Corpus An effective gender recognition approach using voice data via deeper lstm networks,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.491634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.172157Z digest=sha256:4401ff85587d279a6e02476a9132824989a49db6f6eb245b778fbe677a54754d

Observation 58611463-401b-40a8-8104-96a733418bf7 · outbound

This paper cites Improving speech emotion recognition using self-supervised learning with domain-specific audiovisual tasks,.

The MSP-Podcast Corpus Improving speech emotion recognition using self-supervised learning with domain-specific audiovisual tasks,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.483470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.175471Z digest=sha256:4ecb2ab34ef7992b57950558073e72a35fbd3cf756297f73e58bdef1a0463e84

Observation 0c8f1065-7cba-4046-8e88-fb133b6299fb · outbound

This paper cites Semi-supervised speech emo- tion recognition with ladder networks,.

The MSP-Podcast Corpus Semi-supervised speech emo- tion recognition with ladder networks,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.475401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.177945Z digest=sha256:341180a1e69370a3b142a8aaa39d3ab5068209600caca64e677e7976f9ea2d9d

Observation 96ba0c39-4954-410e-887c-363d52db9ad8 · outbound

This paper cites Dawn of the transformer era in speech emo- tion recognition: Closing the valence gap,.

The MSP-Podcast Corpus Dawn of the transformer era in speech emo- tion recognition: Closing the valence gap,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.467977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.180207Z digest=sha256:f4fe1a8b2e71b52f2226399e3309af1b5f0759283418ceb22246b3321fc046cb

Observation cfb5278b-68da-4bd7-89ae-b8f177067ce1 · outbound

This paper cites Unsupervised do- main adaptation for preference learning based speech emotion recognition,.

The MSP-Podcast Corpus Unsupervised do- main adaptation for preference learning based speech emotion recognition,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.460027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.184508Z digest=sha256:5279c3c274486dd2b194f6e4ba68c4ed56c646a601342497327bdbb7277f5e6d

Observation 16368ba3-1662-4f32-9e47-3ee98c27b1a1 · outbound

This paper cites TweetEval: Unified benchmark and comparative evaluation for tweet classification,.

The MSP-Podcast Corpus TweetEval: Unified benchmark and comparative evaluation for tweet classification,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.449617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.188964Z digest=sha256:a053cc7c2efcaa50a8c454c6022a66f60b75cdaa41033e316a28c543af939def

Observation f2600e10-f24e-4575-b913-765cfee2d639 · outbound

This paper cites Odyssey 2024 - speech emotion recognition challenge: Dataset, baseline framework, and results,.

The MSP-Podcast Corpus Odyssey 2024 - speech emotion recognition challenge: Dataset, baseline framework, and results,

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.442185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.192742Z digest=sha256:753adb903e6a08e5d6399f1ea6043dc2c056134c32d17e880f59bfbceb82aa9a

Observation 2189fa86-4f07-452d-9937-b286a813ffa1 · outbound

This paper cites Separation of emotional and reconstruction embeddings on ladder network to improve speech emotion recognition robust- ness in noisy conditions,.

The MSP-Podcast Corpus Separation of emotional and reconstruction embeddings on ladder network to improve speech emotion recognition robust- ness in noisy conditions,

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.433457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.195881Z digest=sha256:dac7fdaf26877c22f8451ffc43e32a072ea04e5fa3e6756d050c350240c67848

Observation 9f1465db-a6fa-4fc5-9031-922a37bb34eb · outbound

This paper cites Ladder networks for emotion recognition: Using unsupervised auxiliary tasks to improve predictions of emotional attributes,.

The MSP-Podcast Corpus Ladder networks for emotion recognition: Using unsupervised auxiliary tasks to improve predictions of emotional attributes,

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.409549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.199220Z digest=sha256:514b0dd4dd9b317885330448bfa0e13acbd27ed5ce21ed3540e81760bbc49316

Observation ff698bed-5818-4955-a943-f6650f12c903 · outbound

This paper cites Scripted dialogs versus improvi- sation: Lessons learned about emotional elicitation techniques from the IEMOCAP database,.

The MSP-Podcast Corpus Scripted dialogs versus improvi- sation: Lessons learned about emotional elicitation techniques from the IEMOCAP database,

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.394753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.202210Z digest=sha256:9f00e3f638fac1791c898c426569733e5166bf0f7e85a78db08c23ddfad31d27

Observation 4de9781b-2ce5-4597-a60d-cc972aaee201 · outbound

This paper cites TSATC: Twitter Sentiment Analysis Training Cor- pus,.

The MSP-Podcast Corpus TSATC: Twitter Sentiment Analysis Training Cor- pus,

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.386512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.206226Z digest=sha256:16e47d04d6c798847c23ecf89b18a3655799e8bc0d8afb3d12abdf0554f0a5ca

Observation 808b6ddd-bf90-4a4a-b28e-b62e39efa3b0 · outbound

This paper cites Curriculum learning for speech emotion recognition from crowdsourced labels,.

The MSP-Podcast Corpus Curriculum learning for speech emotion recognition from crowdsourced labels,

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.378825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.210314Z digest=sha256:6846ed4a68325de38c8b52c9630b4d993dd4d01e3fa6307ee5b6f874be24e281

Observation 5f2372ff-e33c-4ba0-b5c6-dab4a5b4a60d · outbound

This paper cites Exploiting co- occurrence frequency of emotions in perceptual evaluations to train a speech emotion classifier,.

The MSP-Podcast Corpus Exploiting co- occurrence frequency of emotions in perceptual evaluations to train a speech emotion classifier,

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.369519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.213519Z digest=sha256:0bb2ca04776d351e94457526c597e7d50d37848e8472386cd0f9732c3ad7788a

Observation 22fea14f-7ac6-4a6e-a51b-6cc19aa02183 · outbound

This paper cites Modeling subjec- tiveness in emotion recognition with deep neural networks: Ensembles vs soft labels,.

The MSP-Podcast Corpus Modeling subjec- tiveness in emotion recognition with deep neural networks: Ensembles vs soft labels,

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.359822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.216080Z digest=sha256:910e635eb15f6287104ce2bb3ec2a41e484528fb5170996b42ead4afd3c54247

Observation 541b0fdb-cfd7-450a-95b7-c3adf4495968 · outbound

This paper cites Formulating emotion perception as a probabilistic model with application to categorical emotion classification,.

The MSP-Podcast Corpus Formulating emotion perception as a probabilistic model with application to categorical emotion classification,

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.351273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.219470Z digest=sha256:7438e15ca1532b86c39ddd87f7de7ee0c9f57af3cfe8533898c7a7ab81c11f52

Observation fa4f41f6-2367-48e8-9fb3-42d3d23c5924 · outbound

This paper cites Generative approach using soft-labels to learn uncertainty in predicting emotional attributes,.

The MSP-Podcast Corpus Generative approach using soft-labels to learn uncertainty in predicting emotional attributes,

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.342225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.223728Z digest=sha256:3b28bacb1b411f3050ecddb9e7ae51199dce3c5b826a0058625abee05e5d6e0c

Observation b9776d2a-8a41-4f14-9ffe-e888245f6cab · outbound

This paper cites Embracing ambiguity and subjectivity using the all-inclusive aggregation rule for evaluating multi-label speech emotion recognition systems,.

The MSP-Podcast Corpus Embracing ambiguity and subjectivity using the all-inclusive aggregation rule for evaluating multi-label speech emotion recognition systems,

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.332507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.226453Z digest=sha256:9804cdeb457c339bfc6435d5e94e407297c21b701180bb0cdb6fe5d935b0cf17

Observation fffb0549-8148-4c42-a17d-e36a3a3a8aa6 · outbound

This paper cites Minority views matter: Evaluating speech emotion classifiers with human subjective annotations by an all-inclusive aggregation rule,.

The MSP-Podcast Corpus Minority views matter: Evaluating speech emotion classifiers with human subjective annotations by an all-inclusive aggregation rule,

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.323197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.229402Z digest=sha256:92deb7a32c97a9b7a7d4d64cbde9f270bdda903f81a1a5901135d34e65871acf

Observation 03b990ab-9499-4408-ae14-e8cf52e5e9f3 · outbound

This paper cites Over-sampling emotional speech data based on subjective evaluations provided by multiple in- dividuals,.

The MSP-Podcast Corpus Over-sampling emotional speech data based on subjective evaluations provided by multiple in- dividuals,

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.312647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.232122Z digest=sha256:2f3880731ccdde9143ec5df988303533ef711037c7bcf75243f644f5b1b93ce6

Observation 50c95e64-16b2-4a74-9eda-16c0227de0dd · outbound

This paper cites Predictingemotionallysalient regions using qualitative agreement of deep neural network re- gressors,.

The MSP-Podcast Corpus Predictingemotionallysalient regions using qualitative agreement of deep neural network re- gressors,

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.302299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.234604Z digest=sha256:dfdd069e08bf17b3ab6ae4a7d4aeab058d496faf5126f7acbfcfe1b0af6f8bb7

Observation 9d2a53bc-f679-4a21-a5c9-78e9d54ef0b4 · outbound

This paper cites Preference-learning with qualitative agreement for sen- tence level emotional annotations,.

The MSP-Podcast Corpus Preference-learning with qualitative agreement for sen- tence level emotional annotations,

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.293711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.237652Z digest=sha256:21e55461cb9874e07b6919eddd9fc30c0ded45c6e25aac75a15aecf1995e00fb

Observation e0abebcd-6785-4fe6-928f-cc18bb680fb7 · outbound

This paper cites Forced-choice response format in the study of facial expression,.

The MSP-Podcast Corpus Forced-choice response format in the study of facial expression,

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.283996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.240694Z digest=sha256:aa53d451546bcce79d621736eda70431a65523aa021e8e70f72e98f8fc5e383f

Observation 8bbb6be9-0da7-4a7b-ac4d-af4a45f18465 · outbound

This paper cites Measuring emotion: the self- assessment manikin and the semantic differential,.

The MSP-Podcast Corpus Measuring emotion: the self- assessment manikin and the semantic differential,

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.274749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.245166Z digest=sha256:ace3c6587c4e82f2cf2ac778a22f3fedcb59974415b454787f171e564bcbd315

Observation 0c29f89a-86f0-4308-8313-ee940ffab17d · outbound

This paper cites Predicting speaker recogni- tion reliability by considering emotional content,.

The MSP-Podcast Corpus Predicting speaker recogni- tion reliability by considering emotional content,

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.265362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.247821Z digest=sha256:9c183caf3c1bb85c2bdb50edad24ba98abf4423abc7b32e2a260b1b2123a69df

Observation 93be1a60-3eb6-419f-ba0a-a67f7ad0ec03 · outbound

This paper cites X-Vectors meet emotions: A study on dependencies between emotion and speaker recognition,.

The MSP-Podcast Corpus X-Vectors meet emotions: A study on dependencies between emotion and speaker recognition,

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.255350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.250661Z digest=sha256:2cda6ca0d06ef01cdae1f9ce331d6a14eab9035c56c447ef545e5f59c3d02ff6

Observation 66a6f84f-7f73-47ea-9be2-697ff96dac02 · outbound

This paper cites A study of speaker verification performance with expressive speech,.

The MSP-Podcast Corpus A study of speaker verification performance with expressive speech,

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.243726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.253031Z digest=sha256:8b6b7cf58970a2b51c17ae0c29d4262c410d56e87f8a10490c8ab9618cfd84cd

Observation 0ce12f66-6a6f-44d3-9113-f70e4a9f7ad8 · outbound

This paper cites Exploring the intersection between speaker verification and emotion recognition,.

The MSP-Podcast Corpus Exploring the intersection between speaker verification and emotion recognition,

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.235513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.257497Z digest=sha256:0813737c9c5d4aa43a11e3b75da96409e5397440ea9a151cc83365897730dba6

Observation c6c19f00-538f-4be3-a046-318222dab254 · outbound

This paper cites Revealingemotional clusters in speaker embeddings: A contrastive learning strategy for speech emotion recognition,.

The MSP-Podcast Corpus Revealingemotional clusters in speaker embeddings: A contrastive learning strategy for speech emotion recognition,

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.227863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.260572Z digest=sha256:32a9629b1acbc3e7cce5faf5d76f928191a4b892c5ac678bbbec569d4d3a1c4b

Observation 03aab752-b553-40f5-aeef-16dfacd1c498 · outbound

This paper cites Can emotion fool anti-spoofing?.

The MSP-Podcast Corpus Can emotion fool anti-spoofing?

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.219448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.264007Z digest=sha256:c8aa453e80837c3f83ab7517332bf7281619d346be1a41c56148f20f45fe97ac

Observation 14c139b6-370d-44fd-adbb-fd4abe0c5a6e · outbound

This paper cites We need variationsinspeechsynthesis:Sub-centermodellingforspeaker embeddings,.

The MSP-Podcast Corpus We need variationsinspeechsynthesis:Sub-centermodellingforspeaker embeddings,

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:02.267371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:02.267371Z digest=sha256:bddd500b5937639cb411b98509823d5791049101c0b0a02adfdf74243f8e4f50

Observation d43932dc-74f3-436e-b5ce-df1f4e64c52d · outbound

This paper cites The Interspeech 2025 challenge on speech emotion recognition in naturalistic conditions,.

The MSP-Podcast Corpus The Interspeech 2025 challenge on speech emotion recognition in naturalistic conditions,

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.210308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.271044Z digest=sha256:3c3751b9f100ba8c8ab3eb04a98698e343641f41e2de32e9eb8d4c9b243d24c0

Observation a3c5bc37-64ee-463b-985f-efa5f25ac9d8 · outbound

This paper cites New standard for speech recognition and translation from the nvidia nemo canary model,.

The MSP-Podcast Corpus New standard for speech recognition and translation from the nvidia nemo canary model,

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.201492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.273595Z digest=sha256:56d339750a5595865f667770708ff43be4e2668977da133750b448dd3c2982d6

Observation 21865c44-c969-4b42-8450-eecc12be02a2 · outbound

This paper cites Open automatic speech recognition leaderboard,.

The MSP-Podcast Corpus Open automatic speech recognition leaderboard,

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.192826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.277518Z digest=sha256:1012338d1cd94022b9ed1559c6a9b58d86a747833684ea9f54e46eeb131b7114

Observation 5cba6771-a77d-48a6-b288-b8b4484243a4 · outbound

This paper cites Phonetically-anchored domain adaptation for cross- lingual speech emotion recognition,.

The MSP-Podcast Corpus Phonetically-anchored domain adaptation for cross- lingual speech emotion recognition,

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.183309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.280672Z digest=sha256:12047d135fa516d85eada3f7391a4df3d4a6ff776ff99e0b1e02986a1beb3672

Observation 05f1d8d5-06f7-4559-af0c-19eb0a9706cf · outbound

This paper cites Phonetic anchor-based transfer learning to facilitate unsupervised cross-lingual speech emotion recogni- tion,.

The MSP-Podcast Corpus Phonetic anchor-based transfer learning to facilitate unsupervised cross-lingual speech emotion recogni- tion,

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.173887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.294801Z digest=sha256:8b627402556b3bd3e2c0a1c3e7208ba0b7fc6edf7239a69573bf593d73ba9124

Observation f41369d6-f77a-4ce2-9e7a-d01076aae6ab · outbound

This paper cites Analysis of phonetic level similarities across lan- guages in emotional speech,.

The MSP-Podcast Corpus Analysis of phonetic level similarities across lan- guages in emotional speech,

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.163688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.297664Z digest=sha256:fc124934ee96d0bac7161b91872721bee77e38d9437194e3524cfbb5b1bf8439

Observation a4f23c6c-c460-4006-a19e-14f72ef4f740 · outbound

This paper cites Montreal forced aligner: Trainable text-speech alignment using kaldi.

The MSP-Podcast Corpus Montreal forced aligner: Trainable text-speech alignment using kaldi

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.153110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.304595Z digest=sha256:d91ec384402b7d44ddd83bce79a00d61fb8fe79aca263c2be21c84c16aa7720c

Observation 9c607813-77fe-428a-b123-eb71a9b0f94f · outbound

This paper cites Modeling uncertainty in predicting emotional attributes from spontaneous speech,.

The MSP-Podcast Corpus Modeling uncertainty in predicting emotional attributes from spontaneous speech,

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.144171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.308204Z digest=sha256:15e166569c654962253aa6f40adb6aacaa424d3b0d54d511428cc4f095c37902

Observation de6a1134-3f8a-4c80-b88a-fa74ca3cb80b · outbound

This paper cites WavLM: Large-scale self-supervised pre- training for full stack speech processing,.

The MSP-Podcast Corpus WavLM: Large-scale self-supervised pre- training for full stack speech processing,

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.133403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.311572Z digest=sha256:e44ab323e585ddeb67295b536510be8dea2988b481315bb980a4803279bb7029

Observation 2fc8d4f3-4412-4b86-8d50-86df83a11f51 · outbound

This paper cites Robust wav2vec 2.0: Analyzing Domain Shift in Self-Supervised Pre-Training.

The MSP-Podcast Corpus Robust wav2vec 2.0: Analyzing Domain Shift in Self-Supervised Pre-Training

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:02.314694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:02.314694Z digest=sha256:ed21a093f258d247ee7805a525e428e3eb4cc591d722fdee51ac5ba7ff4d658b

Observation 8086c82a-0245-43b1-81e7-df9d1c950788 · outbound

This paper cites HuBERT: Self-supervised speech repre- sentation learning by masked prediction of hidden units,.

The MSP-Podcast Corpus HuBERT: Self-supervised speech repre- sentation learning by masked prediction of hidden units,

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.124112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.318825Z digest=sha256:258a7ba64323c538dd391507107430a4c7a34ef72477fb14ea07bbdaabaa0ddc

Observation 2ff73fd4-2dff-4e77-b423-5fdda3947d30 · outbound

This paper cites EMO-SUPERB: An In-depth Look at Speech Emotion Recognition.

The MSP-Podcast Corpus EMO-SUPERB: An In-depth Look at Speech Emotion Recognition

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:02.323632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:02.323632Z digest=sha256:b330ea91d4aa5a19f8115e71f17421f8da1fcb971a9b675ec5553d8df33507b9

Observation e5b91602-35de-4225-8745-c5ff0559b8b2 · outbound

This paper cites Generalization of self-supervised learning-based representations for cross-domain speech emotion recognition,.

The MSP-Podcast Corpus Generalization of self-supervised learning-based representations for cross-domain speech emotion recognition,

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.113947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.327721Z digest=sha256:3d1aca823ea9da0787888578a6b3361c4849740a54ae8cc67cf8467489c49d6f

Observation 8eeab1e9-2410-4fa3-8654-10054766b936 · outbound

This paper cites Analyzing the effect of affective priming on emotional annotations,.

The MSP-Podcast Corpus Analyzing the effect of affective priming on emotional annotations,

Reference 97

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.104773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.331255Z digest=sha256:4914d63127300ec785067a64a92803e121bea3e7ca3e80c0a60314f2b4d2a55e

Observation e6632b4e-8889-4b63-b083-e919786e2038 · outbound

This paper cites Affective priming in emotional annotations and its effect on speech emotion recognition,.

The MSP-Podcast Corpus Affective priming in emotional annotations and its effect on speech emotion recognition,

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.097223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.334964Z digest=sha256:7cc8472506bdc23ccb23ad59fbb1cb7a5beaca31391ac7e8bb0ed39aee8d4765

Observation c7b43be0-e7ac-409d-8424-9afd6cdb1d94 · outbound

This paper cites Preferencelearning labelsbyanchoringonconsecutiveannotations,.

The MSP-Podcast Corpus Preferencelearning labelsbyanchoringonconsecutiveannotations,

Reference 99

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.087605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.338550Z digest=sha256:a470eebf21ab375ce8dff19a33212981cea1605de7c48c8021f421b33de6426e

Observation d074dbfe-d3fa-4316-86dd-3e1e9369828f · outbound

This paper cites Analyzing continuous-time and sentence-level annotations for speech emotion recognition,.

The MSP-Podcast Corpus Analyzing continuous-time and sentence-level annotations for speech emotion recognition,

Reference 100

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:09:03.080883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T16:09:02.342775Z digest=sha256:b6f1966dfae0b4232bd59a405af43fdd4b7fb3cb795ca564e26e36e647417e95

Pith citing papers

Observation bf78bfb0-a6fe-411d-bc00-bd4db8e168b3 · inbound

Meta-cavity Quantum Electrodynamics cites this paper.

Meta-cavity Quantum Electrodynamics The MSP-Podcast Corpus

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-15T12:12:51.663341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T12:12:51.663341Z digest=sha256:124e8c4c28145f5644798cc3968bcb16b6ad7b05a949337ea0ace3a36ac56643

Observation 847cb35a-d870-4e9a-bf06-32a45fc49155 · inbound

Toward using Speech to Sense Student Emotion in Remote Learning Environments cites this paper.

Toward using Speech to Sense Student Emotion in Remote Learning Environments The MSP-Podcast Corpus

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:31:04.451365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-10T15:58:53.767679Z digest=sha256:b8c9a15ff492b619ce6b89323e9baf3f574efa52ff04828f8000b71eb680ca5e

Observation 8640e920-5c6f-4556-8d9c-f9c3f20d2f37 · inbound

The False Resonance: A Critical Examination of Emotion Embedding Similarity for Speech Generation Evaluation cites this paper.

The False Resonance: A Critical Examination of Emotion Embedding Similarity for Speech Generation Evaluation The MSP-Podcast Corpus

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:11:27.095516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-07T12:34:40.888089Z digest=sha256:f7d00e9fb87776d71f93767fc83789de8428cfe2e347de5b0160bbf1e61a3ee8

Observation f2b54d2e-2817-4ff3-9195-a2ce806c93ea · inbound

The False Resonance: A Critical Examination of Emotion Embedding Similarity for Speech Generation Evaluation cites this paper.

The False Resonance: A Critical Examination of Emotion Embedding Similarity for Speech Generation Evaluation The MSP-Podcast Corpus

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-02T15:21:28.905076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T15:21:28.905076Z digest=sha256:4bc6abb54cee5585928d613a8bb47eaf48c178f5b0a9f08969d00377e09e0659

Observation 1a0222cd-fd0b-4f1d-aa67-142518ea7067 · inbound

AffectCodec: Emotion-Preserving Neural Speech Codec for Expressive Speech Modeling cites this paper.

AffectCodec: Emotion-Preserving Neural Speech Codec for Expressive Speech Modeling The MSP-Podcast Corpus

Reference 37

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T01:07:00.295887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-05-13T01:04:54.506749Z digest=sha256:b663d9920bfb75c0cba310a78782f08051e16c31e3ead0968e6bfb577ac83d95

Observation b49097d6-0ceb-49f1-ba0c-94dd058e53dd · inbound

Multimodal Hidden Markov Models for Persistent Emotional State Tracking cites this paper.

Multimodal Hidden Markov Models for Persistent Emotional State Tracking The MSP-Podcast Corpus

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T20:39:28.085502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-05-14T20:33:24.062878Z digest=sha256:f7d17a7386d42f03e1e568bcfd592be1c6b9e1769fd372b95e20b2d1f0063e7f

Observation 3966a3c6-210a-4b7b-97b3-57dad552fbd8 · inbound

AgentSteerTTS: A Multi-Agent Closed-Loop Framework for Composite-Instruction Text-to-Speech cites this paper.

AgentSteerTTS: A Multi-Agent Closed-Loop Framework for Composite-Instruction Text-to-Speech The MSP-Podcast Corpus

Reference 85

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T21:19:03.223680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-05-20T21:14:58.814362Z digest=sha256:fba250cda7ae62cbda895d80ca5c450d089f8c436e0b36b680f6da4af7e5f3d9

Observation fcdba4ee-6390-453d-983f-e1d0f2f722e1 · inbound

AffectCodec: Emotion-Preserving Neural Speech Codec with Block-Diagonal Residual FSQ cites this paper.

AffectCodec: Emotion-Preserving Neural Speech Codec with Block-Diagonal Residual FSQ The MSP-Podcast Corpus

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-25T03:05:16.755032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-25T03:04:01.544231Z digest=sha256:fe89183644f8c4d6747eedea3aa56e5ab378291b5c05e225bcbb6c3357e033a4

Observation ccbdf5ae-8b9e-4cc0-9745-cb1a1ea6065d · inbound

Sympatheia: Emotionally Adaptive Voice Assistant with Continuous Affect Conditioning cites this paper.

Sympatheia: Emotionally Adaptive Voice Assistant with Continuous Affect Conditioning The MSP-Podcast Corpus

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T18:02:27.104597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-28T17:56:47.512549Z digest=sha256:0641a50d682af5b00d5934898a05285e71e2623371025da797d3c1919db2fc66

Observation aea90970-332a-44c9-99d3-fbb7809638cf · inbound

SHALA-LLM: Smartly Handling Ambiguous Labels in Aligning LLMs cites this paper.

SHALA-LLM: Smartly Handling Ambiguous Labels in Aligning LLMs The MSP-Podcast Corpus

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T07:06:44.215670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-28T07:11:05.234129Z digest=sha256:c3929fdcae186bf88ce332f5661be14a7862cace3e570703c18d15fd8908ddc6

Observation f12bc82b-a385-48a9-9610-1cf85b0c66a4 · inbound

Comparative Reasoning: Making an Audio Language Model Better at Comparing Emotions cites this paper.

Comparative Reasoning: Making an Audio Language Model Better at Comparing Emotions The MSP-Podcast Corpus

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-07-04T18:00:00.272369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-25T23:18:55.383851Z digest=sha256:83f965960e1de0a9bf6b19f7c3d8c97d495348f5326b8e36f8dede51fd65f267

Observation 15189c9e-3930-468d-a65d-e8e1a58f8f57 · inbound

Learning from Annotation Uncertainty: Entropy-Aware Curriculum for Speech Emotion Recognition cites this paper.

Learning from Annotation Uncertainty: Entropy-Aware Curriculum for Speech Emotion Recognition The MSP-Podcast Corpus

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:06:03.337238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-29T00:53:21.116337Z digest=sha256:17807ab70601408699eb05cd251c04285bd52749caabc01ef4855d0c51d31da3

Observation 6bf36bf0-23ef-4bd0-a977-ea0adbef5d99 · inbound

Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts cites this paper.

Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts The MSP-Podcast Corpus

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-07-11T01:57:52.125343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-11T01:53:00.127636Z digest=sha256:2a172e30571b5342210aab53f0f6c022631cf0f609a0af2ebedc0ececde73a16

Observation 7f716022-9104-4dd8-ac8f-6df4a970c680 · inbound

AffectDF: The Most Comprehensive Benchmark for Speech Deepfake Detection against Emotionally Expressive Attacks cites this paper.

AffectDF: The Most Comprehensive Benchmark for Speech Deepfake Detection against Emotionally Expressive Attacks The MSP-Podcast Corpus

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-08T11:50:20.922023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T11:50:20.922023Z digest=sha256:b0ea9989ddf1a50ca27ec6cc8e27ce5c7bdf2c2008f5d535541c80e3c6c9b3a3