Pith. sign in

Paper Citation Record · LEDGER

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning

As of 10 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 2 inbound Pith citation observations for arXiv:2507.04048.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.04048 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:01:29.438826Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:01:29.331354Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-13T01:07:00.272497Z

Reference resolution

42 of 42 outbound references displayed

  • verified exact2
  • verified fuzzy29
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 92b8aeb1-c476-4ab4-a49c-c1763b421afb · outbound

This paper cites this is a sound of.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning this is a sound of

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.777066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:01:29.327598Z digest=sha256:3102bd8ab18711a3b6712ac71808d2f169141f6222117322a900fb84fd8bde70

Observation ad2f2715-09be-4f1c-89e4-ceed8bf8eea9 · outbound

This paper cites CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T20:01:29.331354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:01:29.331354Z digest=sha256:e531fab3c6cbc6d58629b1295f025b38639a45a9106d0d76ebb02da7ce5473e4

Observation 79d52062-635a-4fc2-8219-aea16cf4813c · outbound

This paper cites Happy” or “Sad.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Happy” or “Sad

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.768125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:01:29.334984Z digest=sha256:c5c1fbb731cb2958927f84a434628658c953a2d89cdbb98a95f2f515a5839f25

Observation 670fde26-2dd7-438f-99dc-9ffda49bd000 · outbound

This paper cites excited” and “happy.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning excited” and “happy

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.759674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:01:29.338492Z digest=sha256:394a008fe0f92301e76d9a44229e659e437f26c7be1706bb520d425cf6b07ca2

Observation 9b8f5ce6-b8e1-4eca-8a2c-603583d0ac75 · outbound

This paper cites Integrating Acoustic Context Prompt Tuning (ACPT), CLEP-DG improves general- ization across diverse acoustic environments without additional labeled speech data.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Integrating Acoustic Context Prompt Tuning (ACPT), CLEP-DG improves general- ization across diverse acoustic environments without additional labeled speech data

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.751978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:01:29.341488Z digest=sha256:976fc62e9bd61b4fb07073775eb1286282a519293c375df1304c2c412c59032d

Observation 8c4ede18-eee3-4a24-a63f-310a23eb363b · outbound

This paper cites Affective computing: A review,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Affective computing: A review,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.744069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:01:29.344546Z digest=sha256:2ce48bba132c7bb6b0c32674c1142dbc04055771c511268a154eefdca8c65b1d

Observation 99d1de42-a775-4ee7-ab23-3589916be9f4 · outbound

This paper cites Preece, Y.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Preece, Y

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.736494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:01:29.347248Z digest=sha256:375e80990eae52d532826f1e5ee6b02b2fc1e286ce9bdd2d0db1165af9cc14f1

Observation f8212573-553b-4109-88da-8f8b7a38ef16 · outbound

This paper cites Clap learning audio concepts from natural language supervision,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Clap learning audio concepts from natural language supervision,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T20:01:29.349695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:01:29.349695Z digest=sha256:733d77e181aae85cd7af98ff3f357f8e203c5ac56a8f07f3a41c818b4e67b0a6

Observation 370f554c-f3d9-4e8c-95ce-f3c59a0f963c · outbound

This paper cites Large-scale contrastive language-audio pretraining with feature fusion and keyword-to-caption augmentation,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Large-scale contrastive language-audio pretraining with feature fusion and keyword-to-caption augmentation,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T20:01:29.352747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:01:29.352747Z digest=sha256:f0946b508c8a6246a3c831e6f5e764f7fada863ec47b9159adeb409bfc4159ca

Observation 94dcbb3d-0e1a-4fbb-ab3a-f09af57abd36 · outbound

This paper cites ParaCLAP -- Towards a general language-audio model for computational paralinguistic tasks.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning ParaCLAP -- Towards a general language-audio model for computational paralinguistic tasks

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:01:29.518461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:01:29.355551Z digest=sha256:39965261f375cb3d9d41d07cff7c6a5565f1956a773d0d636a7b62167dde797a

Observation 4cba598e-f394-4aaa-b568-172460786805 · outbound

This paper cites GEmo-CLAP: Gender-Attribute-Enhanced Contrastive Language-Audio Pretraining for Accurate Speech Emotion Recognition.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning GEmo-CLAP: Gender-Attribute-Enhanced Contrastive Language-Audio Pretraining for Accurate Speech Emotion Recognition

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:01:29.507173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:01:29.358691Z digest=sha256:19e0021774f17bb93c94fc213cea7e8f4afd755b15795003bf794f9a73f8da7b

Observation dfa20ce6-0dce-4dfa-9608-140020328e92 · outbound

This paper cites Cross-modal features interaction-and-aggregation network with self-consistency train- ing for speech emotion recognition,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Cross-modal features interaction-and-aggregation network with self-consistency train- ing for speech emotion recognition,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.719068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:01:29.361954Z digest=sha256:b3375062af3da99733aec4bd7da727d628c710c5f2ddc68736573605e81a00fe

Observation 7797ec94-1d13-421d-a9c2-632b61bfbf16 · outbound

This paper cites Conditional prompt learning for vision-language models,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Conditional prompt learning for vision-language models,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.711360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:01:29.364310Z digest=sha256:992ade0852651b5e03d5f9b03b0bce4402a28078718fd96b3fcff5498dc35dd2

Observation 2ce0f5f2-f8ce-4c16-95f4-35219796bc82 · outbound

This paper cites Learning to prompt for vision-language models,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Learning to prompt for vision-language models,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.703426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:01:29.367220Z digest=sha256:4b5f573c5387d45d7fde3a9d996acb202a07e6375fc33f623f6644c78bab6850

Observation b888d2f4-3420-418e-990a-d53cfaca40a6 · outbound

This paper cites Diagnosing and Rectifying Vision Models using Language.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Diagnosing and Rectifying Vision Models using Language

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T20:01:29.369990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:01:29.369990Z digest=sha256:ef06a152d0733a0f68ac2ea1613369c4067ded9b22456b33c2a20eddb5f3d77d

Observation c28a94be-cb91-447c-9971-1e4bccd0005e · outbound

This paper cites Using language to ex- tend to unseen domains.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Using language to ex- tend to unseen domains

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.695918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:01:29.372739Z digest=sha256:8593662dd4dcd4b882c8bf9c99a500964f40fa2a57e0a881a06cc20ec0054f3c

Observation 64e1a4bd-0bca-4c0e-b7af-3d4419df860d · outbound

This paper cites Promptstyler: Prompt-driven style generation for source-free domain generaliza- tion,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Promptstyler: Prompt-driven style generation for source-free domain generaliza- tion,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.688417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:01:29.375131Z digest=sha256:621a82ea4d3cc067e091845bc3da0c0730dce891ab7e4a8c9926cc7cb921f33d

Observation dff51774-32af-4eae-bbd5-54d657626344 · outbound

This paper cites Dpstyler: dynamic prompt- styler for source-free domain generalization,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Dpstyler: dynamic prompt- styler for source-free domain generalization,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.680801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:01:29.377373Z digest=sha256:b44516575d134b64e4a8dbd7ccc68fc308fed1f71e6e3678d8506a2529e7c3b8

Observation 8816ffcc-4c9a-4caa-8900-cebb04616ce6 · outbound

This paper cites Whisper: Robust speech recognition via large-scale weak supervision,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Whisper: Robust speech recognition via large-scale weak supervision,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.672462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:01:29.380093Z digest=sha256:e6219a9558bc775f03b6b2b2277e4c2ca610b87a446ce0f9f83a76f31a155299

Observation 8fb4169a-d501-497a-b5f9-9597b10410fd · outbound

This paper cites Wavlm: Large-scale self- supervised pre-training for full stack speech processing,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Wavlm: Large-scale self- supervised pre-training for full stack speech processing,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T20:01:29.382456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:01:29.382456Z digest=sha256:e1809146564aaf85c8eda360c29de0b993bfb236dfefaf51bd5aee0ad7128261

Observation 70ad7e6a-9751-4580-a8ab-441ee5be647c · outbound

This paper cites Hubert: Self-supervised speech represen- tation learning by masked prediction of hidden units,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Hubert: Self-supervised speech represen- tation learning by masked prediction of hidden units,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T20:01:29.384916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:01:29.384916Z digest=sha256:3af0baedd9da2162aa62bca3c6aa626ec5c57430d5aef728317e7364e8145e1a

Observation 01573aa2-c727-433e-883e-63f65e8768f0 · outbound

This paper cites Wav2clip: Learning robust audio representations via contrastive learning,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Wav2clip: Learning robust audio representations via contrastive learning,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.655603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:01:29.387390Z digest=sha256:2e89ecd6e36f7a60d6deabab2e74a8350218185314ce47e4d7c4b4ede6192a0a

Observation 2625ad7b-fcfd-4752-9299-1440e23b9889 · outbound

This paper cites Audioclip: Ex- tending clip to audio for zero-shot learning,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Audioclip: Ex- tending clip to audio for zero-shot learning,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.648806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:01:29.390000Z digest=sha256:da5bb362c552c6cf60e7aa2b7f08300d8a500432f7fb64016d7492d535c5bad7

Observation 91557793-16a5-440d-985b-1a8294ce4506 · outbound

This paper cites Compa-clap: Composi- tional prompts for improving audio-text alignment,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Compa-clap: Composi- tional prompts for improving audio-text alignment,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.641902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:01:29.392561Z digest=sha256:69524f384ae61290b6db3031556616cd9ad4ba87470c2478d23a56978eafc757

Observation cc81ca27-3f19-46a3-b1bd-790c4c42a003 · outbound

This paper cites Deep Convolutional Ranking for Multilabel Image Annotation.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Deep Convolutional Ranking for Multilabel Image Annotation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T20:01:29.395015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:01:29.395015Z digest=sha256:2dacf61f4cdf303f12b54663a4d8e514eb1a6fcd972eed1ae6c5bdaf11358903

Observation aa93ad76-4c87-46f2-9331-c24d3baab56a · outbound

This paper cites Arcface: Additive angular margin loss for deep face recognition,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Arcface: Additive angular margin loss for deep face recognition,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.634308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:01:29.397824Z digest=sha256:a3190f7cb422f042d8ff45e6958930628b8bf85651178cac5a7621b20254e0f1

Observation d08ccacf-c164-47bd-98c7-9b4f6ed5406b · outbound

This paper cites IEMOCAP: In- teractive emotional dyadic motion capture database,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning IEMOCAP: In- teractive emotional dyadic motion capture database,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.626550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:01:29.400110Z digest=sha256:48cbae99a35d3ea938bcdf86a80182d3a65f1057dcabde84f47b59a117e473e5

Observation 055f513e-b28d-43f8-a98a-adbf5f065974 · outbound

This paper cites MELD: A Multimodal Multi-Party Dataset for Emotion Recognition in Conversations,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning MELD: A Multimodal Multi-Party Dataset for Emotion Recognition in Conversations,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.618864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:01:29.402579Z digest=sha256:c8d68c4e321c1f8f8891c95ca8a0b0e86cf292ad1d640a913bf1c765a5afbf36

Observation b6872e72-4806-4a77-8d43-e4957be45e4e · outbound

This paper cites MEAD: A Large-Scale Audio-Visual Dataset for Affective Un- derstanding and Emotional Expression Analysis,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning MEAD: A Large-Scale Audio-Visual Dataset for Affective Un- derstanding and Emotional Expression Analysis,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.611138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:01:29.405056Z digest=sha256:68bff6ae2a88186a4245f3cf09445895952131402dece295a655b17d5f5d468c

Observation 3d296c53-f88e-4c98-ac46-035dbca314f6 · outbound

This paper cites CMU-MOSEI: A Dataset for Multimodal Sentiment Analysis and Emotion Recognition,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning CMU-MOSEI: A Dataset for Multimodal Sentiment Analysis and Emotion Recognition,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.603415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:01:29.407680Z digest=sha256:cb4642b48c04c296618011896dc9bf7288702b483d175104baecdefb09e5731a

Observation d0869239-af0e-4643-b230-96a001f5cd9f · outbound

This paper cites Wav2clip: Learning robust audio representations from clip,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Wav2clip: Learning robust audio representations from clip,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.595828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:01:29.410016Z digest=sha256:e0cb8257294643db86834ff695a2095fc1c4915ee82c24fe138d0f20728cc1e5

Observation 3a01696e-7f0d-4ddf-99fb-bd60b513d358 · outbound

This paper cites AudioCLIP: Extending CLIP to Image, Text and Audio.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning AudioCLIP: Extending CLIP to Image, Text and Audio

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T20:01:29.412385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:01:29.412385Z digest=sha256:2f0f132a784acf9ab780976365e121407d8936c9c798bd67db3b3cfd4b40a791

Observation ca24a637-25ba-4b48-aaa4-e835a7fa2ce1 · outbound

This paper cites CompA: Addressing the Gap in Compositional Reasoning in Audio-Language Models.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning CompA: Addressing the Gap in Compositional Reasoning in Audio-Language Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T20:01:29.415153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:01:29.415153Z digest=sha256:27d07e51caf37b0eeb6021dfcb2bad14807cea203ede09b8fc19919e7dfef918

Observation 770b239e-4e4b-4de8-91cf-9a3cc21ec562 · outbound

This paper cites The Ryerson Audio-Visual Database of Emotional Speech and Song (RA VDESS),.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning The Ryerson Audio-Visual Database of Emotional Speech and Song (RA VDESS),

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.587583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:01:29.418479Z digest=sha256:f080ae2040f5ac64447afa0bc28d1f5f4cc088881d7306286230788049e435e3

Observation 55cbc3c2-9c45-482e-8f43-e23149d54465 · outbound

This paper cites Toronto emotional speech set (tess)-younger talker happy,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Toronto emotional speech set (tess)-younger talker happy,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.580223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:01:29.420933Z digest=sha256:9dc224eb63bf9187e5871f5ed5ccdf3dc929d58fe06c476c294f02fc817fe638

Observation b01a99ee-2eac-4821-b6cc-aa66c965a4e4 · outbound

This paper cites SA VEE: Surrey Audio-Visual Expressed Emotion,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning SA VEE: Surrey Audio-Visual Expressed Emotion,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.572715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:01:29.423438Z digest=sha256:9ea8e45d60031465deb7a11e15e902bff37e5bb084800b7209d8edc220e84bb0

Observation 75b1e647-c221-40d3-9a69-e280d3211535 · outbound

This paper cites DST: Deformable Speech Transformer for Emotion Recognition,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning DST: Deformable Speech Transformer for Emotion Recognition,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.564895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:01:29.426195Z digest=sha256:9eec3273349e8c7333e6dc9fb8527c64dc580fbf33d774364868bffbcd3f3559

Observation 5f541385-8e8c-430e-b204-058de5de6653 · outbound

This paper cites Tem- poral modeling matters: A novel temporal emotional modeling approach,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Tem- poral modeling matters: A novel temporal emotional modeling approach,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.556456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:01:29.429206Z digest=sha256:2c86bb043c7c2af7d53167e9686301d1a18f48d6ac28c80d7788ad05c1c2ea09

Observation c969f048-ef1b-4930-b924-17b2e368aec5 · outbound

This paper cites The ryerson audio-visual database of emotional speech and song (ravdess): A dynamic, multimodal set of facial and vocal expressions in north american english,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning The ryerson audio-visual database of emotional speech and song (ravdess): A dynamic, multimodal set of facial and vocal expressions in north american english,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T20:01:29.431567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:01:29.431567Z digest=sha256:5f56b60aa9867fe4da914c2722984cc130a82f4a1a121028d5cd5f1f161e5d55

Observation 129db151-23ea-4960-96d6-925a35c3d28f · outbound

This paper cites Speaker-dependent audio- visual emotion recognition.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Speaker-dependent audio- visual emotion recognition

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.543120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:01:29.434143Z digest=sha256:97b8333e0a6803dc8a278ff053601238e9e6cf2830c4ba60429855c043f1e005

Observation 4dd6b5c7-e19c-4e85-bd4a-1ff104e030e5 · outbound

This paper cites Panns: Large-scale pretrained audio neural networks for audio pattern recognition,.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning Panns: Large-scale pretrained audio neural networks for audio pattern recognition,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:01:29.535020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:01:29.436495Z digest=sha256:b9c085ede22ed46f698dab63e1b40505a7f5ebb965cc6ac15dcb07c1f21d154f

Observation 2d292fc6-b95e-4061-932f-ee7dd5174719 · outbound

This paper cites RoBERTa: A Robustly Optimized BERT Pretraining Approach.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning RoBERTa: A Robustly Optimized BERT Pretraining Approach

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T20:01:29.438826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:01:29.438826Z digest=sha256:47e2ec7de55660b1b1cfdc1bf2bfb59e9b7f4bdde919a261c7d70f60a1aa4bda

Pith citing papers

Observation ad2f2715-09be-4f1c-89e4-ceed8bf8eea9 · inbound

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning cites this paper.

CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T20:01:29.331354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:01:29.331354Z digest=sha256:e531fab3c6cbc6d58629b1295f025b38639a45a9106d0d76ebb02da7ce5473e4

Observation 9226301f-dbc6-4bb4-a138-ba43f41028d8 · inbound

AffectCodec: Emotion-Preserving Neural Speech Codec for Expressive Speech Modeling cites this paper.

AffectCodec: Emotion-Preserving Neural Speech Codec for Expressive Speech Modeling CLEP-DG: Contrastive Learning for Speech Emotion Domain Generalization via Soft Prompt Tuning

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:07:00.273850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-13T01:04:54.506749Z digest=sha256:ad4448164ab874d61d62040d9cf68db5778804fd872d0aa1407107bc93237039