Pith. sign in

Paper Citation Record · LEDGER

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer?

As of 11 August 2026, this Paper Citation Record lists 20 of 20 outbound references and 0 inbound Pith citation observations for arXiv:2501.04138.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.04138 v1

Coverage vector

measured 20 of 20 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T21:43:25.919378Z

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

20 of 20 outbound references displayed

  • verified exact0
  • verified fuzzy7
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fbe93ba6-5467-474e-b398-ce14c6bd1447 · outbound

This paper cites In Findings of the Association for Computational Linguistics: ACL 2023, pages 8256–.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? In Findings of the Association for Computational Linguistics: ACL 2023, pages 8256–

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:43:26.557221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:43:25.456995Z digest=sha256:bd912df55cce1e37aa5accf20cdb533acb4bcb722b41624e22d660987944a6d6

Observation 9a741e32-a20c-4f9e-b086-c03cf1679886 · outbound

This paper cites Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T21:43:25.469157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:43:25.469157Z digest=sha256:b9ca73f38d7c459f25ce23479851f2232fe4debdaa07a38ae1675be718c0312b

Observation eff4710e-d4fb-46b2-9c86-fd29ee8bec99 · outbound

This paper cites SOUL: Towards Sentiment and Opinion Understanding of Language.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? SOUL: Towards Sentiment and Opinion Understanding of Language

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-08-10T21:43:26.179670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:43:25.515987Z digest=sha256:9d62a507c2577aa0e59d6a0d0685e635d0f28c960d6a9d122181e1ca426039a2

Observation 0b73e684-7f8d-436d-b5c3-da0daa631ca5 · outbound

This paper cites HCAM -- Hierarchical Cross Attention Model for Multi-modal Emotion Recognition.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? HCAM -- Hierarchical Cross Attention Model for Multi-modal Emotion Recognition

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T21:43:25.555945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:43:25.555945Z digest=sha256:7c996bbe4c6adcd67ec321ca18655c4d8568c34dfdb897400ce60616c37dbaca

Observation 3bbf07de-fedf-4d76-9bbd-affcaac80e81 · outbound

This paper cites Listen, Think, and Understand.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? Listen, Think, and Understand

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T21:43:25.603086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:43:25.603086Z digest=sha256:c09ada37e6aace67456affd6859d38bcc68e0d891cf75b133ce12c4574b2978e

Observation 6b02cde4-29bc-4008-8bf1-e461ac3a7622 · outbound

This paper cites In Findings of the Association for Computational Linguistics: EMNLP 2022, pages 289–299.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? In Findings of the Association for Computational Linguistics: EMNLP 2022, pages 289–299

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:43:26.496325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:43:25.720846Z digest=sha256:5d584ef0c9a448a8a0b7034c05b2170b2f424d4753cc1d20afa1f89deaf45960

Observation 4b541765-0c3c-4d42-a69a-154dc9cf24e5 · outbound

This paper cites ScaleVLAD: Improving Multimodal Sentiment Analysis via Multi-Scale Fusion of Locally Descriptors.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? ScaleVLAD: Improving Multimodal Sentiment Analysis via Multi-Scale Fusion of Locally Descriptors

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T21:43:25.752770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:43:25.752770Z digest=sha256:829c0cb9e7aa79dc9839e12b91ea05fdd0096339471c2171d0b5d4006ab29116

Observation fab8e1db-5c1d-46bc-8eb5-111a0a484aa2 · outbound

This paper cites In 2021 IEEE Automatic Speech Recognition and Understanding Workshop (ASRU), pages 350–357.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? In 2021 IEEE Automatic Speech Recognition and Understanding Workshop (ASRU), pages 350–357

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:43:26.418865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:43:25.757190Z digest=sha256:0d327dbc6eb67767a92b930e566fbc254b10ae573be86f202b49223832e9521c

Observation 833f3756-dcad-4d1f-bb05-ed4b27f1e90c · outbound

This paper cites M-SENA: An Integrated Platform for Multimodal Sentiment Analysis.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? M-SENA: An Integrated Platform for Multimodal Sentiment Analysis

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T21:43:25.768375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:43:25.768375Z digest=sha256:1eda03ff6250f04dfb2589767919531420ce6faffdd8ca9a9bc2a922f2e8e28e

Observation 14e7e7f4-71bc-43d2-b84b-6a49421b81c4 · outbound

This paper cites Emotion Recognition from Speech Using Wav2vec 2.0 Embeddings.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? Emotion Recognition from Speech Using Wav2vec 2.0 Embeddings

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T21:43:25.777825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:43:25.777825Z digest=sha256:355a41536add30b7c7dabd70248429c69f1932ded9197c49fe67e36d35bb5aa0

Observation 7da15da4-8ca4-4a91-9553-8bb79d6e8d48 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T21:43:25.789415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:43:25.789415Z digest=sha256:d87341aed54e2bf19e35971e2e98193dbf4671ed3607a55559efbbdff2b2ae5f

Observation 4a195208-744b-4aa3-8dba-b2d9d432ceb3 · outbound

This paper cites Is ChatGPT a Good Sentiment Analyzer? A Preliminary Study.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? Is ChatGPT a Good Sentiment Analyzer? A Preliminary Study

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T21:43:25.796331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:43:25.796331Z digest=sha256:69fba943fba61e9d97311f2fb24b5f0dcd4be7b86f4ee52d46e67daca377b952

Observation 701acd15-b9f8-43d1-988f-e0bbbf381a75 · outbound

This paper cites SpeechGPT: Empowering Large Language Models with Intrinsic Cross-Modal Conversational Abilities.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? SpeechGPT: Empowering Large Language Models with Intrinsic Cross-Modal Conversational Abilities

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T21:43:25.919378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:43:25.919378Z digest=sha256:83bbdc95a1485a6899fd05b7551da128548d8050faf2944f02dac9042057e1f1

Observation 1794769f-95d4-4e71-a7a3-3382b4541830 · outbound

This paper cites In 2017 IEEE International Conference on Multimedia and Expo (ICME), pages 949–954.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? In 2017 IEEE International Conference on Multimedia and Expo (ICME), pages 949–954

Reference 2017

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:43:26.324957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:43:25.793068Z digest=sha256:07398ffa8faacf6041d88e7eff133d434ea599c454746390361958ddeabc1206

Observation ae462fc6-3f3d-46a8-8220-381f37b7d077 · outbound

This paper cites In Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? In Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing

Reference 2019

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:43:26.236757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:43:25.848604Z digest=sha256:6ee7de9d96af574ca056566eadead7e1c7635dc97240c5e899d9759f43be6e7f

Observation 9fe3bd8c-7d49-46c8-aae1-33d8ad25675e · outbound

This paper cites BERT-Like.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? BERT-Like

Reference 2020

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:43:26.396574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:43:25.786413Z digest=sha256:a90c2b9c0352d02f2fdfb36c65f2d3a029569855153f92901de90a557b53ec86

Observation ed06ba36-3c5f-4a30-a177-9fd74326803c · outbound

This paper cites Improving Multimodal Fusion with Hierarchical Mutual Information Maximization for Multimodal Sentiment Analysis.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? Improving Multimodal Fusion with Hierarchical Mutual Information Maximization for Multimodal Sentiment Analysis

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-10T21:43:25.649485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:43:25.649485Z digest=sha256:4f6bc39637ea6d90a018bb75c795a04b1ce89c69ae9ef650ed006c9309f9da2b

Observation 854d5b97-c836-4351-8eeb-af833e40746b · outbound

This paper cites In Findings of the Association for Computational Linguistics: EMNLP 2022, pages 5105–5114.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? In Findings of the Association for Computational Linguistics: EMNLP 2022, pages 5105–5114

Reference 2022

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:43:26.529333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:43:25.465275Z digest=sha256:9f727f7b5c3de147bdd927cc043119392130047989338d7ce90480797cccea48

Observation 3edb7c4c-1dc7-4ede-ba0b-c6e8dbaf03a9 · outbound

This paper cites GPT-4 Technical Report.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? GPT-4 Technical Report

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-10T21:43:25.417581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:43:25.417581Z digest=sha256:9f7581c32c3014999422a5e064a4794aee08f22cfe380ab9b64fc8d06dbe3d58

Observation 45535fc0-2535-4a72-b57c-11c0815b15bc · outbound

This paper cites Are Human Conversations Special? A Large Language Model Perspective.

"Yeah Right!" -- Do LLMs Exhibit Multimodal Feature Transfer? Are Human Conversations Special? A Large Language Model Perspective

Reference 2024

Resolution
metadata mismatch
local_arxiv, observed 2026-08-10T21:43:26.089463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T21:43:25.689842Z digest=sha256:b5100e1ff28d17cd2e1dbac1d6bebd4319309b38c8aa0adb877a41ae4a601e5e

Pith citing papers

No inbound Pith citation observations are available.