Pith. sign in

Paper Citation Record · LEDGER

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer?

As of 11 August 2026, this Paper Citation Record lists 20 of 20 outbound references and 0 inbound Pith citation observations for arXiv:2501.04138.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.04138 v1

Coverage vector

measured 20 of 20 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T21:43:25.919378Z

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

20 of 20 outbound references displayed

  • verified exact0
  • verified fuzzy7
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fbe93ba6-5467-474e-b398-ce14c6bd1447 · outbound

This paper cites In Findings of the Association for Computational Linguistics: ACL 2023, pages 8256–.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? In Findings of the Association for Computational Linguistics: ACL 2023, pages 8256–

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:43:26.557221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:43:25.456995Z digest=sha256:88aa929a7e5c0cb8d1f13271b9a8c73b5fccd8c9e84e0b333d37f6149f33ef0f

Observation 9a741e32-a20c-4f9e-b086-c03cf1679886 · outbound

This paper cites Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T21:43:25.469157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:43:25.469157Z digest=sha256:103a45a4d4f47ff764ebf598c21112ef27dc94605c32421ef6594fbd12ba6a13

Observation eff4710e-d4fb-46b2-9c86-fd29ee8bec99 · outbound

This paper cites SOUL: Towards Sentiment and Opinion Understanding of Language.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? SOUL: Towards Sentiment and Opinion Understanding of Language

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-08-10T21:43:26.179670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:43:25.515987Z digest=sha256:7a1065a3ffcc400ad700e79dd5469deb7cd6cc66eb8b26336a78ddd4958604c8

Observation 0b73e684-7f8d-436d-b5c3-da0daa631ca5 · outbound

This paper cites HCAM -- Hierarchical Cross Attention Model for Multi-modal Emotion Recognition.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? HCAM -- Hierarchical Cross Attention Model for Multi-modal Emotion Recognition

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T21:43:25.555945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:43:25.555945Z digest=sha256:7fd7a1e9c5afd2ed69934e0cac04b115fd97b59b2915015f3dbc7e184edbfc27

Observation 3bbf07de-fedf-4d76-9bbd-affcaac80e81 · outbound

This paper cites Listen, Think, and Understand.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? Listen, Think, and Understand

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T21:43:25.603086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:43:25.603086Z digest=sha256:e2fb36a616c22c1d221c319e003f1fd236c8532aa5dfee4f1853032a132a5f88

Observation 6b02cde4-29bc-4008-8bf1-e461ac3a7622 · outbound

This paper cites In Findings of the Association for Computational Linguistics: EMNLP 2022, pages 289–299.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? In Findings of the Association for Computational Linguistics: EMNLP 2022, pages 289–299

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:43:26.496325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:43:25.720846Z digest=sha256:a41745c7f7419b5287dff3247f24a65831e0b5f3a0ad7b6a0a878c94a10cd4af

Observation 4b541765-0c3c-4d42-a69a-154dc9cf24e5 · outbound

This paper cites ScaleVLAD: Improving Multimodal Sentiment Analysis via Multi-Scale Fusion of Locally Descriptors.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? ScaleVLAD: Improving Multimodal Sentiment Analysis via Multi-Scale Fusion of Locally Descriptors

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T21:43:25.752770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:43:25.752770Z digest=sha256:04231197cf30a406f5b9252853acf04e1c947f55a88f53de1bb74cb5f206c93a

Observation fab8e1db-5c1d-46bc-8eb5-111a0a484aa2 · outbound

This paper cites In 2021 IEEE Automatic Speech Recognition and Understanding Workshop (ASRU), pages 350–357.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? In 2021 IEEE Automatic Speech Recognition and Understanding Workshop (ASRU), pages 350–357

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:43:26.418865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:43:25.757190Z digest=sha256:9bc7ec2ebd7495a45d30bd62072c6bc9ad8ed05e1d719ea1acfe7643b53ad376

Observation 833f3756-dcad-4d1f-bb05-ed4b27f1e90c · outbound

This paper cites M-SENA: An Integrated Platform for Multimodal Sentiment Analysis.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? M-SENA: An Integrated Platform for Multimodal Sentiment Analysis

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T21:43:25.768375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:43:25.768375Z digest=sha256:f08b0dd0440d69925cd65f2588e80e53a1c8df618174861881f9751eb2a7565a

Observation 14e7e7f4-71bc-43d2-b84b-6a49421b81c4 · outbound

This paper cites Emotion Recognition from Speech Using Wav2vec 2.0 Embeddings.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? Emotion Recognition from Speech Using Wav2vec 2.0 Embeddings

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T21:43:25.777825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:43:25.777825Z digest=sha256:09a5ca0b0f1ea58d5d3cddc333cfc0e26c1458b6d082b1a7c5378c6320046d0d

Observation 7da15da4-8ca4-4a91-9553-8bb79d6e8d48 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T21:43:25.789415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:43:25.789415Z digest=sha256:fedc49454e1b0c0aac818e53c29f2e487af87f70afe17091cce036f6403af64e

Observation 4a195208-744b-4aa3-8dba-b2d9d432ceb3 · outbound

This paper cites Is ChatGPT a Good Sentiment Analyzer? A Preliminary Study.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? Is ChatGPT a Good Sentiment Analyzer? A Preliminary Study

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T21:43:25.796331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:43:25.796331Z digest=sha256:6133aebcfb44e2164b6b94a6bb2a3160917f9ccffa70a33ec94ee9dd877e9a4d

Observation 701acd15-b9f8-43d1-988f-e0bbbf381a75 · outbound

This paper cites SpeechGPT: Empowering Large Language Models with Intrinsic Cross-Modal Conversational Abilities.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? SpeechGPT: Empowering Large Language Models with Intrinsic Cross-Modal Conversational Abilities

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T21:43:25.919378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:43:25.919378Z digest=sha256:d9f12add346fe172f7b1a4c0f89f00ab6aa2ad9fa94e62cedea48363cca68225

Observation 1794769f-95d4-4e71-a7a3-3382b4541830 · outbound

This paper cites In 2017 IEEE International Conference on Multimedia and Expo (ICME), pages 949–954.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? In 2017 IEEE International Conference on Multimedia and Expo (ICME), pages 949–954

Reference 2017

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:43:26.324957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:43:25.793068Z digest=sha256:1aef5d371487014f9c0043503373979fff97e1476fc96ee07354f88f898f49f9

Observation ae462fc6-3f3d-46a8-8220-381f37b7d077 · outbound

This paper cites In Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? In Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing

Reference 2019

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:43:26.236757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:43:25.848604Z digest=sha256:71a5fd34d4ef5783e16da496dbdabf23cb44fe43127be074938d7dd3100bc131

Observation 9fe3bd8c-7d49-46c8-aae1-33d8ad25675e · outbound

This paper cites BERT-Like.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? BERT-Like

Reference 2020

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:43:26.396574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:43:25.786413Z digest=sha256:3f5616b17a009d9c6e5a932e80e148db69c2946929aece3af7891de1c715dd65

Observation ed06ba36-3c5f-4a30-a177-9fd74326803c · outbound

This paper cites Improving Multimodal Fusion with Hierarchical Mutual Information Maximization for Multimodal Sentiment Analysis.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? Improving Multimodal Fusion with Hierarchical Mutual Information Maximization for Multimodal Sentiment Analysis

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-10T21:43:25.649485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:43:25.649485Z digest=sha256:d06b72e864429b1e19d3d6f1f0fc58763db9740ea30ef2dd4b45c81802f7fe35

Observation 854d5b97-c836-4351-8eeb-af833e40746b · outbound

This paper cites In Findings of the Association for Computational Linguistics: EMNLP 2022, pages 5105–5114.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? In Findings of the Association for Computational Linguistics: EMNLP 2022, pages 5105–5114

Reference 2022

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:43:26.529333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:43:25.465275Z digest=sha256:e739627947fe8b68af9ca8730b2247f3f396169d2dab3de6155eacc8e6816d24

Observation 3edb7c4c-1dc7-4ede-ba0b-c6e8dbaf03a9 · outbound

This paper cites GPT-4 Technical Report.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? GPT-4 Technical Report

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-10T21:43:25.417581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:43:25.417581Z digest=sha256:c2fd760e2ff6fe0c1588b892dca136fa6cc4bed7a73b596d443009a01b592ac5

Observation 45535fc0-2535-4a72-b57c-11c0815b15bc · outbound

This paper cites Are Human Conversations Special? A Large Language Model Perspective.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? Are Human Conversations Special? A Large Language Model Perspective

Reference 2024

Resolution
metadata mismatch
local_arxiv, observed 2026-08-10T21:43:26.089463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:43:25.689842Z digest=sha256:798f353cfe3f2165f35b456089be46bc5765c51db9a1d0124970df8d4ab5dbe3

Pith citing papers

No inbound Pith citation observations are available.