Pith. sign in

Paper Citation Record · LEDGER

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation

As of 16 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 1 inbound Pith citation observation for arXiv:2508.07302.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.07302 v2

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T22:17:38.566540Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-18T16:27:37.596817Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T16:31:37.259420Z

Reference resolution

30 of 30 outbound references displayed

  • verified exact4
  • verified fuzzy18
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 747a176c-8695-4ef4-a090-5eede55c3179 · outbound

This paper cites FastSpeech 2: Fast and High-Quality End-to-End Text to Speech.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation FastSpeech 2: Fast and High-Quality End-to-End Text to Speech

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T22:17:38.247156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:17:38.247156Z digest=sha256:8f2706e0d7b100aab68a208bd248b05503a0a2804335465b3c2709efc7b13c8e

Observation cb5c16d8-bd7b-449b-8519-f84e610c92a5 · outbound

This paper cites AdaSpeech: Adaptive Text to Speech for Custom Voice.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation AdaSpeech: Adaptive Text to Speech for Custom Voice

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T22:17:38.258146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:17:38.258146Z digest=sha256:d4acbb4402ac819d8bf43767aa14cd66d0b2988fb0da58be2a57e86ab2f7298e

Observation d3080bf9-3ac0-430d-91a3-617464b9f6d1 · outbound

This paper cites Conditional variational autoencoder with adversarial learning for end-to-end text-to-speech,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Conditional variational autoencoder with adversarial learning for end-to-end text-to-speech,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.858055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T22:17:38.274410Z digest=sha256:1f49dabbede634db76b61adc90a40cbe73918aad20d5605c2ece2233a247e221

Observation 5ec9cfc4-6a90-424a-9aa4-a7ac5a171124 · outbound

This paper cites Generspeech: Towards style transfer for generalizable out-of-domain text-to-speech,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Generspeech: Towards style transfer for generalizable out-of-domain text-to-speech,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.826565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T22:17:38.291568Z digest=sha256:4aa0070d7a13f1033b6f2fe5b0bff1042a3c56416802953d590ce87a2cd27706

Observation 64a33c59-d966-4a47-b022-18342097d1b9 · outbound

This paper cites Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Speak Foreign Languages with Your Own Voice: Cross-Lingual Neural Codec Language Modeling

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T22:17:38.301601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:17:38.301601Z digest=sha256:409b23b80915032973e47fdcf41d523669349e5e8eee4a56510ee305ae214586

Observation e7f00805-43b9-4391-b1f9-c4d742d2f610 · outbound

This paper cites XTTS: a Massively Multilingual Zero-Shot Text-to-Speech Model.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation XTTS: a Massively Multilingual Zero-Shot Text-to-Speech Model

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T22:17:38.311772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:17:38.311772Z digest=sha256:19c3681001fc526eda6afe631994fc8e9241730c03054b521a8036cb3d1a6c3c

Observation 59f5b7dc-2217-4e9f-bb9a-e6bfae52afd9 · outbound

This paper cites Zero-shot Cross-lingual Voice Transfer for TTS.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Zero-shot Cross-lingual Voice Transfer for TTS

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-08-05T22:17:39.030809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T22:17:38.318698Z digest=sha256:1aac051ef8bd24311276857ad92bdefe163f808b36b974f3ee51a46e289aef39

Observation 605ed599-7d83-4687-81d3-bacee00df06c · outbound

This paper cites Multilingual video dubbing—a tech- nology review and current challenges,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Multilingual video dubbing—a tech- nology review and current challenges,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.793134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T22:17:38.325689Z digest=sha256:14e22fe209f63b6cd99894fe219a7fafabd8654f5b829ae82158cd82ab1023d1

Observation 98aae8f0-105e-4709-aba3-f3d9113366d6 · outbound

This paper cites DSE-TTS: Dual Speaker Embedding for Cross-Lingual Text-to-Speech.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation DSE-TTS: Dual Speaker Embedding for Cross-Lingual Text-to-Speech

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-05T22:17:38.977395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T22:17:38.331461Z digest=sha256:d3f9771a435925f930c16e7528d3327803b5d9ef968f1f24533fb5c3cd3db300

Observation a34f24ce-f9b9-4e55-9784-b403575276b2 · outbound

This paper cites Diclet-tts: Diffusion model based cross- lingual emotion transfer for text-to-speech – a study between english and mandarin,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Diclet-tts: Diffusion model based cross- lingual emotion transfer for text-to-speech – a study between english and mandarin,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.755773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T22:17:38.341071Z digest=sha256:0892a217327cd91f51f0d7461617c4876e93ef97f0f18dd6f8e9869071a451f2

Observation dfd6b0ba-a7d5-4eaf-8b30-2dfe786e5db2 · outbound

This paper cites Llasa: Scaling Train-Time and Inference-Time Compute for Llama-based Speech Synthesis.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Llasa: Scaling Train-Time and Inference-Time Compute for Llama-based Speech Synthesis

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T22:17:38.347773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:17:38.347773Z digest=sha256:646707f643854fbee80617a11e8abd4d5bb79a2e3f764a65f7f28d526033036a

Observation 75e50fdd-80fb-4768-9237-e5027783d363 · outbound

This paper cites DelightfulTTS: The Microsoft Speech Synthesis System for Blizzard Challenge 2021.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation DelightfulTTS: The Microsoft Speech Synthesis System for Blizzard Challenge 2021

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-08-05T22:17:39.128510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T22:17:38.360877Z digest=sha256:141e1901dc2eb573f2fd79723eb007d3717f958b3fed3fa4758eef6ffbd2c816

Observation c3cd9acd-ddc0-4b8c-aa05-0f5377f9ae59 · outbound

This paper cites iemotts: Toward robust cross- speaker emotion transfer and control for speech synthesis based on disentanglement between prosody and timbre,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation iemotts: Toward robust cross- speaker emotion transfer and control for speech synthesis based on disentanglement between prosody and timbre,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.723619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T22:17:38.374104Z digest=sha256:f9c6bedb2a7417fbe63b7d2a368da54c16af26e29ff90bafabba3bf280f275c4

Observation ebd62e70-9ff8-444c-b5ad-d094344832c6 · outbound

This paper cites Zero-shot emotion transfer for cross-lingual speech synthesis,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Zero-shot emotion transfer for cross-lingual speech synthesis,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.694834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T22:17:38.382804Z digest=sha256:4a42d45dadc98240b127dc1cb5a1e4e7f6921c7f8d858b43317d367dc3a736bf

Observation ffc56a9f-a356-4bce-977d-4e0b7e8ef712 · outbound

This paper cites Zet- speech: Zero-shot adaptive emotion-controllable text-to-speech synthesis with diffusion and style-based models,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Zet- speech: Zero-shot adaptive emotion-controllable text-to-speech synthesis with diffusion and style-based models,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.668409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T22:17:38.396075Z digest=sha256:bcaada9ffee4f5e3694afa0f42cc99652fd6245b1ced86b331a928cbf8923254

Observation b5f15223-498d-4312-b7d4-6d7df3408c0e · outbound

This paper cites Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Learning to Speak Fluently in a Foreign Language: Multilingual Speech Synthesis and Cross-Language Voice Cloning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T22:17:38.408435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:17:38.408435Z digest=sha256:9abf7f8a2741a2a04cdd36272ed99e20656f72cff3a2f1810050480f42722587

Observation b563c52b-a9a8-4dd0-b931-3a11f9ffc017 · outbound

This paper cites Metts: Multilingual emotional text-to-speech by cross- speaker and cross-lingual emotion transfer,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Metts: Multilingual emotional text-to-speech by cross- speaker and cross-lingual emotion transfer,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.590546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T22:17:38.432824Z digest=sha256:a1417e49d3037d93f5432ed1e042d5c4226c8dd0b42ad9a476d6e9f6e9970cf0

Observation c289e77e-be12-49de-854d-42d9081d5e61 · outbound

This paper cites Cosyvoice 3: Towards in-the-wild speech generation via scaling-up and post-training,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Cosyvoice 3: Towards in-the-wild speech generation via scaling-up and post-training,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.550878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T22:17:38.448447Z digest=sha256:fa278e8e5b816d645d8118e40279a5c1c46f12e2590215952e92c476fda57ae6

Observation f8971949-c88b-47e5-9cba-070cd6ad9ad7 · outbound

This paper cites Maskgct: Zero-shot text-to-speech with masked generative codec transformer,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Maskgct: Zero-shot text-to-speech with masked generative codec transformer,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.518076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T22:17:38.455700Z digest=sha256:d341b4dd56bbf5af2411f1da8294939d405639ddb59322efe2fcc346b20d064c

Observation 6d350465-2bb9-4d84-9e98-ca93d880dd75 · outbound

This paper cites Glm-4-voice: Towards intelligent and human-like end-to-end spoken chatbot,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Glm-4-voice: Towards intelligent and human-like end-to-end spoken chatbot,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.465254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T22:17:38.466878Z digest=sha256:58a0fdf9b7581a6da2f752264a7f9864cdc61a0d95ef8a10ede0ef72d1adbe1b

Observation fccb8667-09ba-4ccc-80fd-6853f0519b74 · outbound

This paper cites Speak foreign languages with your own voice: Cross-lingual neural codec language modeling,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Speak foreign languages with your own voice: Cross-lingual neural codec language modeling,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.626151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T22:17:38.476870Z digest=sha256:3b35dec12b8d982fb4c6906aa4d9031ae8d1d0815f7995d64b08fc463ba27cae

Observation 5007cd8a-5f5a-4b75-a446-eaf0afcbbef1 · outbound

This paper cites Retrieval- augmented generation for knowledge-intensive nlp tasks,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Retrieval- augmented generation for knowledge-intensive nlp tasks,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.405977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T22:17:38.486028Z digest=sha256:bb1514e7bae078c7fa7020be4f03ab1c2a319ff9e0618231b829277cc1d60a5d

Observation 4c753727-01f5-46c9-b865-9e959a85ceaa · outbound

This paper cites AutoStyle-TTS: Retrieval-Augmented Generation based Automatic Style Matching Text-to-Speech Synthesis.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation AutoStyle-TTS: Retrieval-Augmented Generation based Automatic Style Matching Text-to-Speech Synthesis

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-05T22:17:38.830913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T22:17:38.493215Z digest=sha256:86983623bbc52319a98b3e15ebcdc4275157751bdcd37b1a66956c2ff4628189

Observation 1e42d927-5c44-4fff-a769-f3f74672a38c · outbound

This paper cites Wavrag: Audio-integrated retrieval augmented generation for spoken dialogue models,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Wavrag: Audio-integrated retrieval augmented generation for spoken dialogue models,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.343676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T22:17:38.503618Z digest=sha256:a9380516a257c4bfa93f85665631fc6ca1650a0cd62d63148686ddefe5d8e0ad

Observation cccfc76b-36da-420c-b2e4-c4131e596751 · outbound

This paper cites Codec does matter: Exploring the semantic shortcoming of codec for audio language model,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Codec does matter: Exploring the semantic shortcoming of codec for audio language model,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.304422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T22:17:38.511641Z digest=sha256:3c5183628d062589595a555acfb5ac726b884d0dc9c38bbbfe246a3e45312303

Observation ccf1bd42-12ea-41ad-adf4-7b8b96c96d3a · outbound

This paper cites Emo2vec: Learning emotional embeddings via multi-emotion cate- gory,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Emo2vec: Learning emotional embeddings via multi-emotion cate- gory,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.279034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T22:17:38.527270Z digest=sha256:2eca95ae3e38fc1044251d9c70c36ec3aac74905da94ca13b78b4e45ea5936c3

Observation 41f923b0-f431-41f5-9199-848c59f8635e · outbound

This paper cites Dspgan: A gan-based universal vocoder for high-fidelity tts by time-frequency domain supervision from dsp,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Dspgan: A gan-based universal vocoder for high-fidelity tts by time-frequency domain supervision from dsp,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.259195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T22:17:38.538673Z digest=sha256:ab089c29423e336412bc002a01c103e992ebc6cb10bf13647edfdae51fefb313

Observation 673bf918-60b9-4a34-b0c1-4645917f6e50 · outbound

This paper cites Typhoon 2: A Family of Open Text and Multimodal Thai Large Language Models.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Typhoon 2: A Family of Open Text and Multimodal Thai Large Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-05T22:17:38.548736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:17:38.548736Z digest=sha256:0a6b1904a5751b41e31851f633d6e8e8494ff0a259bcc30845da533ddeebf761

Observation d50a04fe-1cb7-4954-97cf-7339f98a8a42 · outbound

This paper cites Wespeaker: A research and production oriented speaker embedding learning toolkit,.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Wespeaker: A research and production oriented speaker embedding learning toolkit,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T22:17:39.234880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T22:17:38.555521Z digest=sha256:dbf8e9aada86d914e7123fdd8f87c246eaef7dbd3403b702e56ad111db286158

Observation 36270e26-3b47-4623-b5ba-0d932e9939a4 · outbound

This paper cites Zipformer: A faster and better encoder for automatic speech recognition.

XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation Zipformer: A faster and better encoder for automatic speech recognition

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T22:17:38.566540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:17:38.566540Z digest=sha256:62a153439a87672cabb521565e85fc29228014bfa82e5db16473e0444477d652

Pith citing papers

Observation a8438032-1c43-4fd7-a270-741ef50b734a · inbound

Towards Building Speech Large Language Models for Multitask Understanding in Low-Resource Languages cites this paper.

Towards Building Speech Large Language Models for Multitask Understanding in Low-Resource Languages XEmoRAG: Cross-Lingual Emotion Transfer with Controllable Intensity Using Retrieval-Augmented Generation

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-18T16:31:37.262268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-18T16:27:37.596817Z digest=sha256:62e5017c508e3f4ff150116e96f5d9371ab61de6e9ca74e6b930424e8c099313