Pith. sign in

Paper Citation Record · LEDGER

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models

As of 14 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 0 inbound Pith citation observations for arXiv:2508.07273.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.07273 v1

Coverage vector

measured 37 of 37 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T22:19:03.131803Z

measured 37 of 37 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

37 of 37 outbound references displayed

  • verified exact0
  • verified fuzzy27
  • unresolved10
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c4826f91-e930-4571-af64-868e66878355 · outbound

This paper cites Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T22:18:57.881185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:18:57.881185Z digest=sha256:f62c1b924e3e2362e4014b4170c4054347bde29a298af3a5e4f2b1e0843d7cf8

Observation 1cb81b11-8db1-4cfb-8362-02103d244f99 · outbound

This paper cites Qwen2-Audio Technical Report.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models Qwen2-Audio Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T22:18:57.990172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:18:57.990172Z digest=sha256:b27d2faf25583f8e3ede22ca34d3200aa74a361cc7bad5c6028fc5472213eab6

Observation 7fcdb65a-1231-4ffe-8460-ed12d2a6d3e6 · outbound

This paper cites GPT-4 Technical Report.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models GPT-4 Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T22:18:58.130239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:18:58.130239Z digest=sha256:b2090fcd4c9c8cc4c34ba625391b8d0b30506d80a480379f183037bf4d49b71d

Observation 1436705d-17b9-439d-b5a3-4e5eb7a9d8b1 · outbound

This paper cites SALMONN: Towards Generic Hearing Abilities for Large Language Models.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models SALMONN: Towards Generic Hearing Abilities for Large Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T22:18:58.276120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:18:58.276120Z digest=sha256:f6d354c155bec1fc83570c9d06da914c7c540de002bce36775081d2004977196

Observation c9d18b3b-045c-439f-89c4-e4529e4a8c42 · outbound

This paper cites Advancing Large Language Models to Capture Varied Speaking Styles and Respond Properly in Spoken Conversations.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models Advancing Large Language Models to Capture Varied Speaking Styles and Respond Properly in Spoken Conversations

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T22:18:58.424699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:18:58.424699Z digest=sha256:87e04905c78396f23618193bf57374fef79b3be84c547600f71ea291304ab7a9

Observation e3f07d77-2dea-496e-8e8e-f7d5bfd38919 · outbound

This paper cites Paralinguistics- aware speech-empowered large language models for natural conversa- tion,.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models Paralinguistics- aware speech-empowered large language models for natural conversa- tion,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:19:11.147456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T22:18:58.577376Z digest=sha256:1ce49e2ece9e581afaf8a2f9a3fee7dce66437927aaecb2e5eadaebe23ff7ee7

Observation e89fc06b-8bf9-41ef-91c5-30cafc19f03d · outbound

This paper cites Frozen Large Language Models Can Perceive Paralinguistic Aspects of Speech.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models Frozen Large Language Models Can Perceive Paralinguistic Aspects of Speech

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T22:18:58.748810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:18:58.748810Z digest=sha256:29bea1e87c5bb2e6a327a872883fb36ef8256b92a1c81d548682c216320e33ef

Observation f74ae3df-7fb6-47ad-a15e-ca1f4bd8c37e · outbound

This paper cites BLSP-Emo: Towards empathetic large speech-language models,.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models BLSP-Emo: Towards empathetic large speech-language models,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:19:10.990667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T22:18:58.914092Z digest=sha256:71663a691426cdb8e465e9efae29048fc765d9f4089e47d59152377d031f391c

Observation ff0d1a75-3f6f-4567-879c-7ec9dd830f1c · outbound

This paper cites DeSTA: Enhancing speech language models through descriptive speech-text alignment,.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models DeSTA: Enhancing speech language models through descriptive speech-text alignment,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:19:10.821861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T22:18:59.038997Z digest=sha256:fe6f7991a26d0db9d0e1df9f7b7ee12a7457dafc4e8477f3ed08fa3c3d82c74e

Observation f7cde303-a5b5-4319-838c-36ce2085b52c · outbound

This paper cites DeSTA2: Developing instruction-following speech language model without speech instruction-tuning data,.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models DeSTA2: Developing instruction-following speech language model without speech instruction-tuning data,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:19:10.568200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T22:18:59.198133Z digest=sha256:3459471677614861cad9433f864b800cb378e4a2fae5449d20d1b86ad31c859a

Observation 396c684e-e5fa-4f94-9cb2-6c0ca3d20be2 · outbound

This paper cites Beyond silent letters: Amplifying LLMs in emotion recognition with vocal nuances,.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models Beyond silent letters: Amplifying LLMs in emotion recognition with vocal nuances,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:19:10.233939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T22:18:59.330714Z digest=sha256:f6ff9513c8daa7e4880206200fcea76541dbd6e0ceac1087c857741a4a927a8b

Observation 13f75bb6-1233-458d-9f6c-93ae9b31edc1 · outbound

This paper cites CLAP4Emo: ChatGPT-Assisted Speech Emotion Retrieval with Natural Language Supervision,.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models CLAP4Emo: ChatGPT-Assisted Speech Emotion Retrieval with Natural Language Supervision,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:19:09.973950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T22:18:59.466028Z digest=sha256:a4af0c1f80acf581a5fb64fef2196111b5d057ef589150b3e0c978d357387488

Observation 7080ceca-6ea4-400f-9df0-0ebcb419c949 · outbound

This paper cites SECap: Speech emotion captioning with large language model,.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models SECap: Speech emotion captioning with large language model,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:19:09.690497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T22:18:59.633364Z digest=sha256:3526965fcb92a3353120aa965578eefdcfedd847db12119bd2c04b3c4341cf7c

Observation d5f71264-0a61-41af-9a2f-98dad20db104 · outbound

This paper cites Empower typed descriptions by large language models for speech emotion recognition,.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models Empower typed descriptions by large language models for speech emotion recognition,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:19:09.371189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T22:18:59.813350Z digest=sha256:cb801e589497a4e6ae7f906074624325ed52f28cb55c7c56d3c221c5eb266646

Observation 10f686a7-c02d-4952-944b-47344877811a · outbound

This paper cites V oxDialogue: Can spoken dialogue systems understand information be- yond words?,.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models V oxDialogue: Can spoken dialogue systems understand information be- yond words?,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:19:09.084928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T22:18:59.969824Z digest=sha256:523793b62ed77172af7ca790c6cb2e1961b1a975af7e9dfec2cd17905bd6a904

Observation 1b006634-4d6e-4d48-9493-50b63c761dd7 · outbound

This paper cites SIFT-50M: A Large-Scale Multilingual Dataset for Speech Instruction Fine-Tuning.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models SIFT-50M: A Large-Scale Multilingual Dataset for Speech Instruction Fine-Tuning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T22:19:00.083040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:19:00.083040Z digest=sha256:5e51ee00854891f0a707614f56bb8e3f2dc443cd75000065287fb1f37ce2bc05

Observation cdf56ab7-5479-448e-8915-d9f2f0a21bb0 · outbound

This paper cites Joint audio and speech understanding,.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models Joint audio and speech understanding,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:19:08.715779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T22:19:00.249738Z digest=sha256:8acff9fc407ea259b6f44d5e55df4e9ef324fb65176efc0516ff48c7b22ff9e0

Observation 51b62e32-e945-4c14-b2b6-cb1845bf5548 · outbound

This paper cites Audiobench: A universal benchmark for audio large language models,.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models Audiobench: A universal benchmark for audio large language models,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:19:08.375935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T22:19:00.426892Z digest=sha256:6ba21d337cced57e8a67017cd799d3a7504135b36997125a84f95b94a3b0e250

Observation 284eb8ae-8932-40b7-91df-fa2f241005f9 · outbound

This paper cites Dynamic-superb: Towards a dy- namic, collaborative, and comprehensive instruction-tuning benchmark for speech,.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models Dynamic-superb: Towards a dy- namic, collaborative, and comprehensive instruction-tuning benchmark for speech,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:19:08.072863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T22:19:00.562787Z digest=sha256:c64f95ec319d0844fd84ab1eedad18b1175f7bed4a514ddbf90f01a52ba3cd94

Observation d038238a-12da-4994-a9e6-9e592a756ec2 · outbound

This paper cites AIR-Bench: Benchmarking large audio-language models via generative comprehension,.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models AIR-Bench: Benchmarking large audio-language models via generative comprehension,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:19:07.775254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T22:19:00.758702Z digest=sha256:9e67f67e5ba469370d09df74501579460f0ec3024229941be2d1ffd950ece7e5

Observation f66111f0-5d6c-4ac9-ba14-ec29dfebf371 · outbound

This paper cites Listen, think, and understand,.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models Listen, think, and understand,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:19:07.482732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T22:19:00.929096Z digest=sha256:4c0c59bb3b6ec377b5a9ae8099c598666bcafbe3ab0ce339b8b08c83087fe131

Observation 9f16ec29-7210-4b16-b1a6-a47cc5d313d9 · outbound

This paper cites MMAU: A massive multi-task audio understanding and reasoning benchmark,.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models MMAU: A massive multi-task audio understanding and reasoning benchmark,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:19:07.156600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T22:19:01.094132Z digest=sha256:cc4244a31054aecd84a6776ccb3b3881ce8dcd60e5f426c7dde8be2728d38ec8

Observation 33fb1d05-d676-4deb-aa92-c26ab73481ce · outbound

This paper cites Contextual paralinguistic data creation for multi-modal Speech-LLM: Data condensation and spoken QA generation,.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models Contextual paralinguistic data creation for multi-modal Speech-LLM: Data condensation and spoken QA generation,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:19:06.815227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T22:19:01.230959Z digest=sha256:6bee7a7602ce390807dfdd636072bd10d299c8febf9c971c419b370fc046c552

Observation 1dafea79-6319-40c9-8757-d75adf8e29ef · outbound

This paper cites WavLM: Large-scale self-supervised pre-training for full stack speech processing,.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models WavLM: Large-scale self-supervised pre-training for full stack speech processing,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:19:06.538955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T22:19:01.371170Z digest=sha256:c5e1770b144b260a1e548bc79f1e1aaa62bae62a600e4c733b1a3dadbb2ae937

Observation add265d2-c0c9-41b1-84ba-4cc16fbbbcf8 · outbound

This paper cites emotion2vec: Self-supervised pre-training for speech emotion representation,.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models emotion2vec: Self-supervised pre-training for speech emotion representation,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:19:06.201026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T22:19:01.507690Z digest=sha256:1d7d9b02150b7974b68eaf9a6506952755515e91ed1ee86f532032e1e5a17954

Observation 1e1bcdfa-e379-4a72-b7ea-215da3a5e14d · outbound

This paper cites Dawn of the transformer era in speech emotion recognition: Closing the valence gap,.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models Dawn of the transformer era in speech emotion recognition: Closing the valence gap,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:19:05.918481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T22:19:01.664852Z digest=sha256:a04af3e7dc0f81b384d0094a5707dc258eca053d93d5d25a0b9c1d0a809c561f

Observation 4afd48f4-8dce-4d01-844d-81078dea12cf · outbound

This paper cites Whis- perX: Time-accurate speech transcription of long-form audio,.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models Whis- perX: Time-accurate speech transcription of long-form audio,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:19:05.565564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T22:19:01.765730Z digest=sha256:ea55e6aa824dd37bd320f063ac69c8a485bcbce4d15b8f316e73fdbf559e1d14

Observation ee166452-0ddc-4232-907a-c0e509467641 · outbound

This paper cites Sentence-bert: Sentence embeddings using siamese bert-networks,.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models Sentence-bert: Sentence embeddings using siamese bert-networks,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:19:05.239830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T22:19:01.901008Z digest=sha256:f1748f71207900aff71c15d3baf06af1e1068aab2a937ea1ded7e0952b9151cf

Observation b3964042-d69d-4891-8c9f-0de663cf8d1a · outbound

This paper cites MERaLiON-AudioLLM: Bridging Audio and Language with Large Language Models.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models MERaLiON-AudioLLM: Bridging Audio and Language with Large Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T22:19:02.006029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:19:02.006029Z digest=sha256:e0aee218db0ce0c8559da32bab11795e3200822ad54c5425382526a903d97a49

Observation 86bc0bfb-bce2-4a74-a663-beb72b7a2b4a · outbound

This paper cites Robust speech recognition via large- scale weak supervision,.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models Robust speech recognition via large- scale weak supervision,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:19:04.928009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T22:19:02.191735Z digest=sha256:e7e259d2060d1e06042a56a9171c20936553558fa686ebbe8984d2a1c2d73e20

Observation 7d080e31-18cc-48e8-8da7-b94a6389b350 · outbound

This paper cites Gemma: Open Models Based on Gemini Research and Technology.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models Gemma: Open Models Based on Gemini Research and Technology

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T22:19:02.331972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:19:02.331972Z digest=sha256:ee5e6bcdf75b01f152103f567913c45b0cee5d855f7feb6bd184d52776b7a3f8

Observation d0ce1fd1-838f-4d9e-8d6e-d592d06956e6 · outbound

This paper cites Advancing Singlish Understanding: Bridging the Gap with Datasets and Multimodal Models.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models Advancing Singlish Understanding: Bridging the Gap with Datasets and Multimodal Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T22:19:02.450512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:19:02.450512Z digest=sha256:ff8d812f5945d815f1ca36fe4044a2b8f150e8ef246c6ef1dd5cb5039d04c1d3

Observation 3c5d796a-1369-4324-b30d-6519a858bcf1 · outbound

This paper cites Building the Singapore English national speech corpus,.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models Building the Singapore English national speech corpus,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:19:04.706734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T22:19:02.570145Z digest=sha256:ff9394be143215c8bb0ad44d1f629e3cfce080f0eacf54da77c789abfb682510

Observation 91c7c5e1-003d-4014-8862-e8d6e77e61ec · outbound

This paper cites Librispeech: An ASR corpus based on public domain audio books,.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models Librispeech: An ASR corpus based on public domain audio books,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:19:04.426180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T22:19:02.725176Z digest=sha256:a37ecb81dd7352f3e5649142916778625225182445a30ba1ffaaae9d592170ad

Observation 3dfe46cb-0dd2-4c44-8e9f-04d47fd8a869 · outbound

This paper cites IEMOCAP: interactive emotional dyadic motion capture database,.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models IEMOCAP: interactive emotional dyadic motion capture database,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:19:04.166254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T22:19:02.841800Z digest=sha256:a730146f99b69282de5cddcedb87fba0df0884cc09d6f25623e1e242480bbb5c

Observation a005d119-3757-45ff-81ab-c69bfe6c25ea · outbound

This paper cites V oxceleb: A large-scale speaker identification dataset,.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models V oxceleb: A large-scale speaker identification dataset,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:19:03.907253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T22:19:03.024584Z digest=sha256:a48ad4bb9d07f6dbb1d63ab216e5a4672672804424317606160c6a70c761119d

Observation ce965b70-3779-478a-972c-2ff491567a98 · outbound

This paper cites The MSP-Podcast corpus for speech emotion recognition,.

Incorporating Contextual Paralinguistic Understanding in Large Speech-Language Models The MSP-Podcast corpus for speech emotion recognition,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:19:03.588814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T22:19:03.131803Z digest=sha256:fa9bce61c76e670fb84f2367ec8b48c7b033e5d1aecfea43608e8973e89ec1b0

Pith citing papers

No inbound Pith citation observations are available.