Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T20:49:51.683856Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 81 of 81 outbound references and 3 inbound Pith citation observations for arXiv:2502.04976.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T20:49:51.683856Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T18:28:56.798473Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-15T17:18:53.193305Z
81 of 81 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation fb87eba8-4fc5-4cc3-a2a9-697bfd994a21 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Qwen Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31147a08-07ff-4534-9c32-ccbbf8074a64 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 056093af-70b6-445d-bb1a-871fb52b9cc5 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Gonzalez, Ion Stoica, and Eric P
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66ebfe9e-e93b-4e7d-9fb7-86552c74a084 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29821088-ca91-40cf-a029-9405535b50de · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ed0215c-ad68-4e10-b656-50786cd0b997 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark DreamLLM: Synergistic Multimodal Comprehension and Creation
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d775e18f-52dd-4ee1-ae6b-23aa0f2522ac · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0a18e02f-bb40-4054-a0cf-d9235503413d · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bed9aafa-425c-4deb-a3d3-5c9862d2d37c · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f6d23236-04d6-4b56-a5d5-57ccb2b86412 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation da6d4218-8a5a-4b4e-bf2f-c80d7351c8f3 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark EmpathyEar: An Open-source Avatar Multimodal Empathetic Chatbot
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 503c6581-b745-444c-ac62-337a5dc32756 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark In Proceedings of the 2024 Joint International Conference on Computational Linguistics, Language Resources and Evaluation (LREC-COLING 2024): Tutorial Summaries
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 26abf529-89f8-465c-82c9-a342d2febada · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 026f6c2b-820e-432e-97a8-e7b6408c99f8 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e3d29065-5f7a-4c8f-8de0-e02348cc64a3 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark LoRA: Low-Rank Adaptation of Large Language Models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4189bc9b-c508-4943-b6ba-5366ab3247f7 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18f12589-cb3a-4c75-ad01-c5cd9541715d · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark DiaASQ : A Benchmark of Conversational Aspect-based Sentiment Quadruple Analysis
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 639262ca-b2a4-499c-90d3-901024897bd1 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Auto-Encoding Variational Bayes
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00378d85-45e0-4e90-a7a5-2c533c9f3cee · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6357f5ef-71d5-49d2-b059-d866fe875d4c · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark A Diversity-Promoting Objective Function for Neural Conversation Models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f81845f-abf1-49a0-bc00-421084672c58 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation dc6c9300-5af4-4b18-aca5-a9ab9bb1ea3d · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark A Survey on Benchmarks of Multimodal Large Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c16bd71b-5db5-4cce-acc5-dd42d2412db8 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Video-LLaVA: Learning United Visual Representation by Alignment Before Projection
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7cc22715-9a48-4d91-9b05-62f5586f76c5 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d7a011df-0e4a-4441-ad5c-8c795905297a · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 774632b9-a894-4948-8628-4ed7e5677e85 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark MoEL: Mixture of Empathetic Listeners
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8c2ad0f5-e375-4a3a-b729-bea345d8a0a8 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1e125c31-4ad3-4b76-9d4d-69cb9c4b4d23 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark The Voice Conversion Challenge 2018: Promoting Development of Parallel and Nonparallel Methods
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 519ba272-27db-40c5-ae17-409449357493 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3377d4c9-0abc-42e4-8b07-7bbe50a1a9ac · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark PanoSent: A Panoptic Sextuple Extraction Benchmark for Multimodal Conversational Aspect-based Sentiment Analysis
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation adda083f-3989-4b76-9f15-1470a0218d7d · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark MIME: MIMicking Emotions for Empathetic Response Generation
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad011661-eb30-4f38-b868-0fa2e5831a66 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark DreamTalk: When Emotional Talking Head Generation Meets Diffusion Probabilistic Models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b57493c-9938-4813-b92a-dd57ea777fc7 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Harnessing the Power of Large Language Models for Empathetic Response Generation: Empirical Investigations and Improvements
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87ec28d0-353a-4027-adbc-14929db4ddfe · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ebd0be06-2ead-48bf-b3f9-bc832c413447 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3f87b12-e718-4dba-b923-859d5d12b45d · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 01c515dd-d82c-4aaf-a231-3b81af018c48 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7e4d92b7-ad2c-4f5c-b3b6-f4b5fa189708 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Towards Empathetic Open-domain Conversation Models: a New Benchmark and Dataset
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a940ed49-f0b6-4a23-9e7b-7e18e28f4774 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80047f13-f03c-4e15-a98f-c317e2a2cb7f · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1f9e4a3e-c6ca-467a-ae58-0b6a4160d5bb · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 840e6f52-38dc-4251-88b6-e2d43a08c4e3 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc5ae234-023d-442f-9b5c-527e3301931e · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1343a509-e0af-4f92-aea5-ceffad9ab69f · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e03bcb67-6ab2-4b67-97b7-44db29d4d48d · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Towards Semantic Equivalence of Tokenization in Multimodal LLM
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ba78d90-f4a9-4817-b575-1e48d260b4e2 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c1ec96a0-5b61-47cc-b32d-ca99d38961b3 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32f64eee-8cdd-4936-a868-d7c5cc0d2cbb · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6fa01968-e790-43c7-840f-f341027f7931 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8ea6d1d5-c5a6-4848-bca1-e89bf56e8cd2 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Enhancing Empathetic Response Generation by Augmenting LLMs with Small-scale Empathetic Models
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 139bc4c6-ce29-4d47-ac3c-e0c0ccb5b7eb · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Faithful Logical Reasoning via Symbolic Chain-of-Thought
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e637f55-148f-4a4e-bfad-048a9e2e7e83 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark A Survey of Large Language Models
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7bece7b9-8aef-43ac-b64e-ace0af46964f · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Exploiting Emotion-Semantic Correlations for Empathetic Response Generation
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2533db02-107c-4585-a428-c44e8307f18e · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark CASE: Aligning Coarse-to-Fine Cognition and Affection for Empathetic Response Generation
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 972c834a-7b94-4739-87de-b0325cc573ab · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 96151797-fc49-416e-9120-772d8c712263 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark ECQED: Emotion-Cause Quadruple Extraction in Dialogs
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41e073f6-2b6d-4735-9a4a-4e2ab616a733 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7931415-92b0-4347-b766-87de8e798a2f · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark dia_id":
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b1153dac-d97b-4c20-b916-4eba867c6306 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 38d1a22a-a33d-44db-ba4c-2661f8c4c309 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b5e822e1-d806-478b-8b4b-35fbae1cd6b4 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark It’s like they have no empathy or think about what if it was them
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a4fdbe06-2545-46be-ba4a-060b3e34e4e4 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Achievements and Self-Realization
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 40a52432-062e-4a44-a55c-1e18d064b5e9 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 17733aba-a035-4fb3-bf77-9ad0bfc0205f · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2be31d4b-7a47-4c47-a6bc-762fee5f8602 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 074d344d-e9e9-4799-8589-d15636c8c968 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Avg. Score
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 649fcae0-1cb1-4ad0-ae2c-4dfa1520ef2e · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark As shown, the model’s performance peaks when the number of tokens reaches 16
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b7132b85-21e9-4e4d-bc31-2fe924d2828f · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1daad0ad-245f-4243-a96b-e0ffdb572373 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5b7a6db1-39ae-4259-b7e8-59f202305a64 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b7d9000f-65a5-4168-9fe4-edcb5fa5a0c0 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f8f960a5-79f9-4072-b7d1-17889a6b1ee1 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 76328e9c-ba14-44f1-b375-d964621910fd · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark 4.Goal to Response: Validate the speaker’s sense of relief and preparedness, acknowledging the stressful situation they avoided
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2dc9975c-9905-41f8-84f1-950c27fe3330 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6b757b37-4780-49a9-8028-102a74959d60 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2540466f-e338-4f96-8906-664f4489e4e6 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark 4.Goal to Response: To provide empathy and acknowledge the speaker’s excitement and enjoyment of the trip
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a130b068-e74f-44ae-9ffe-0ac0e290fbd6 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e92b5957-dfe8-4510-b368-284088896f31 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Unresolved cited work
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 83ef674b-50de-4c29-aee8-73302ef10a5e · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark 4.Goal to Response: Offer empathy and validation for the speaker’s feelings, acknowledging the challenge of parenting and the importance of communication
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b6542455-5c43-442a-88bb-038ad27c90d1 · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark PandaGPT: One Model To Instruction-Follow Them All
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84b2134c-7597-4a83-a21e-9fefb77419fd · outbound
Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark Proceedings of the Advances in neural information processing systems
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a5c5a808-e9d9-47ff-b2fc-35e12ed5f1a6 · inbound
Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark
Reference 128
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d9dd7114-7e1d-4658-bdd7-c897424e51b0 · inbound
A Multi-Agent Framework with Structured Reasoning and Reflective Refinement for Multimodal Empathetic Response Generation Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e42e6a58-8452-43d4-b7fd-a389285c8dce · inbound
EmpaAva: An Open-source Agentic 3D-Avatar Empathetic Live Chatbot Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.