Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:17:56.416700Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 1 inbound Pith citation observation for arXiv:2505.19437.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:17:56.416700Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:17:54.000523Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-07T14:17:56.702242Z
30 of 30 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation a4d2841a-34d1-4abc-8a00-52bae8c84bb5 · outbound
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4fd6735a-f6ad-48e7-9a57-56f91ff8816f · outbound
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval 1, the proposed RA-CLAP model is designed to learn a joint representation of speech and text through a con- trastive learning and self-distillation framework
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 255fd815-9fcf-4735-8fdd-12a953119b6c · outbound
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 22cfcac5-4ecd-4f32-a2f6-cd10482d2d70 · outbound
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dcab4470-541c-4159-ab6e-9a460c34c306 · outbound
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval Multi-level knowledge distillation for speech emotion recogni- tion in noisy conditions,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e2d36bc6-127e-4640-b312-de50206a9f3c · outbound
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval Iterative prototype refinement for ambiguous speech emotion recognition,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8dcd976f-92b6-47d7-ba03-c5630c96c011 · outbound
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval Fine- grained disentangled representation learning for multimodal emo- tion recognition,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7b381f07-e144-40b8-b0d9-714a6aeccfdf · outbound
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval Enhancing multimodal emotion recognition through multi- granularity cross-modal alignment,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 24a5fce8-6dbb-496e-a982-5e77b7170dd9 · outbound
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval Enhancing emotion recognition in incomplete data: A novel cross-modal alignment, reconstruction, and refinement framework,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b5e1bef2-6569-419b-a241-a2292c9f7476 · outbound
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval Emotion- preserving prosody anonymization network for voice privacy pro- tection,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7b76d34f-d7e0-4a80-be68-e6aa57a9c381 · outbound
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval Speaker-Text Retrieval via Contrastive Learning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c13774c-0ad4-4b6f-be57-d16cd86060c4 · outbound
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval ParaCLAP -- Towards a general language-audio model for computational paralinguistic tasks
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 885dc978-f9c3-4645-afed-bb1808695639 · outbound
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval Prompttts: Control- lable text-to-speech with text descriptions,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2dde2d89-3f11-4d8f-8a91-71334313d05a · outbound
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval Prompttts++: Controlling speaker identity in prompt-based text-to-speech using natural language descriptions,
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 217c859d-ace3-4f9b-935d-7b0465d73791 · outbound
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval Stylecap: Automatic speaking-style captioning from speech based on speech and lan- guage self-supervised learning models,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation eecc696b-4dec-42a6-86ce-e67a873841e5 · outbound
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval Learning transferable visual models from natural language supervision,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58ce6195-63f3-4120-a029-d8cfed917ff5 · outbound
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval Large-scale contrastive language-audio pretraining with feature fusion and keyword-to-caption augmentation,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ec33af7-db3f-4148-acc8-169c6d83f315 · outbound
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval Gemo-clap: Gender-attribute-enhanced contrastive language- audio pretraining for accurate speech emotion recognition,
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0bf44369-fa18-4cc9-92a8-1c3955fd0b31 · outbound
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval LibriTTS: A Corpus Derived from LibriSpeech for Text-to-Speech
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b98c4c3d-eb89-4736-9d6d-234d00ccfc05 · outbound
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval Textrolspeech: A text style control speech corpus with codec language text-to-speech models,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03232a1d-a6c3-40eb-88ec-7f590001e8c1 · outbound
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval Cstr vctk corpus: English multi-speaker corpus for cstr voice cloning toolkit (ver- sion 0.92),
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b41c8ce9-08fa-448e-9a36-22230ca671ba · outbound
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval Seen and unseen emo- tional style transfer for voice conversion with a new emotional speech dataset,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bc78f9b0-6e12-40ea-8dd1-b6dfefc3985e · outbound
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval Toronto emotional speech set (tess)-younger talker happy,
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 926bbc5f-94fc-49f7-98c5-5813ee848590 · outbound
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval Mead: A large-scale audio-visual dataset for emotional talking-face generation,
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f03dea00-5445-4c91-a680-281a4ca24cea · outbound
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval Speaker-dependent audio- visual emotion recognition
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4d642703-dd2c-4e55-aa53-48d00d88f84c · outbound
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval Categorical and dimensional ratings of emotional speech: Behavioral findings from the morgan emotional speech set,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3f5edecf-56f6-4bc2-b46c-2e855d6fd05c · outbound
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval Speechcraft: A fine-grained expressive speech dataset with natural language description,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 72b3b378-ef9c-4449-8708-5f8c65c46d46 · outbound
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval AISHELL-3: A Multi-speaker Mandarin TTS Corpus and the Baselines
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6c69078-b779-436a-8e7e-6dc2e647e0ad · outbound
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval GigaSpeech: An Evolving, Multi-domain ASR Corpus with 10,000 Hours of Transcribed Audio
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fea97d2f-3f14-46fd-8cc3-9a25b645b90f · outbound
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval LibriTTS-R: A Restored Multi-Speaker Text-to-Speech Corpus
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4d2841a-34d1-4abc-8a00-52bae8c84bb5 · inbound
RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval RA-CLAP: Relation-Augmented Emotional Speaking Style Contrastive Language-Audio Pretraining For Speech Retrieval
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.