Pith. sign in

Paper Citation Record · LEDGER

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations

As of 19 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 0 inbound Pith citation observations for arXiv:2508.09789.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.09789 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T20:52:15.689430Z

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

42 of 42 outbound references displayed

  • verified exact3
  • verified fuzzy2
  • unresolved37
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d055120f-4662-453a-a739-5b0970657c43 · outbound

This paper cites YouTube-8M: A Large-Scale Video Classification Benchmark.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations YouTube-8M: A Large-Scale Video Classification Benchmark

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:11.898760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:11.898760Z digest=sha256:ad0c720c01220597ed9dbbf1e1f7efe68baa1a3fb59dface1686b17061b7f022

Observation 53afb53d-4905-43c0-bc65-1518f0ee4249 · outbound

This paper cites MiniGPT4-Video: Advancing Multimodal LLMs for Video Understanding with Interleaved Visual-Textual Tokens.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations MiniGPT4-Video: Advancing Multimodal LLMs for Video Understanding with Interleaved Visual-Textual Tokens

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:11.947587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:11.947587Z digest=sha256:63b6d616e410e6b666919a7d9aa8ee5ac3e3eb0840e30e152715c55530332a80

Observation 3d50f3c6-2f7b-43f7-824d-93dd7f258105 · outbound

This paper cites Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:12.075305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:12.075305Z digest=sha256:3b565f0096cadf9bdf5eb565dc579679f5374ac0776940ea14ed3b312da7f9d9

Observation 4aea27eb-a000-40ae-97d5-04cafad7297d · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:12.159645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:12.159645Z digest=sha256:7653bd619b6a4a75427bf38ec13438261fb49bcf40acef3d5c757ca92c3d1850

Observation b29ad16e-a2c7-41a8-9d88-dc69fd749e74 · outbound

This paper cites VITA-1.5: Towards GPT-4o Level Real-Time Vision and Speech Interaction.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations VITA-1.5: Towards GPT-4o Level Real-Time Vision and Speech Interaction

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:12.243640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:12.243640Z digest=sha256:77d87b0b305d9325c7ca89bccf11041accf4309a213562fd251c5167e9791951

Observation 6575eda3-5343-4888-aed5-b0c5c45f7e7b · outbound

This paper cites Efficient and Effective Adaptation of Multimodal Foundation Models in Sequential Recommendation.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Efficient and Effective Adaptation of Multimodal Foundation Models in Sequential Recommendation

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:52:16.342181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T20:52:12.334243Z digest=sha256:d781c00b9b22a1dae60f8bf0bf91d886f723acfe9798d64ac43599ff71d1328d

Observation 79b52895-6a8c-4903-a674-86fdbc6965f4 · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:12.424673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:12.424673Z digest=sha256:9b6f71724f08e9eac7bb3d05485d44098c114893ea4c4ecff1403a75dda45a8d

Observation 49c69ce7-a34f-4d91-bc08-3387ece6a749 · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:52:18.133375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T20:52:12.502619Z digest=sha256:44b974b621e9338ba543ea579522a01a3c381681fd900a4449c38b45ae45f25a

Observation 4aa7b084-5a4e-4ce3-95fa-6f81d215c617 · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:12.593453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:12.593453Z digest=sha256:07515dd7eadaa5bcb5ff9969aebb865afca787f892b940f389ff8fd3a1558b6a

Observation d83c8282-6656-421f-8d58-e1728c634d8e · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:12.694792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:12.694792Z digest=sha256:bd2207f6d9611e560389e1e6fe3830cb2680126eede72d68792df94fbd6ed2ef

Observation daceec4b-7953-42e4-aafc-dbeb36d89a7a · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:52:17.968385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T20:52:12.749307Z digest=sha256:4cd834fb6238cd2ccc3c61be34e96c45be3d1339bdfa25df6b64575bf2086256

Observation 06700e41-578c-4819-8852-7093726e7337 · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:52:17.807895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T20:52:12.799997Z digest=sha256:5c1ada693bc23a5f51ab498caff97b0c140d760bb41d19401ee28dede241b97c

Observation 76eb33c5-c78b-439b-82b7-0bd79bf9d56e · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:52:17.678604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T20:52:12.872360Z digest=sha256:76da6a52d91b0f1fef1a32392432931b51249fcf6cc7d33c64f94335bcfe9aed

Observation 48250b28-9d41-47ee-8236-f27f6671d5d9 · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:12.924826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:12.924826Z digest=sha256:c26bc691ec6d7783745645249a7cdb76744f4672c51a08c0ccb7e43f3c7fe5cd

Observation 0722b8c9-6e91-4875-9b84-370a8afe9d33 · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:13.023547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:13.023547Z digest=sha256:672f66fe40d6058a3774a7dca06b0b816fbe7f9d579655f21cdf2b5075bce7de

Observation a24d9033-1080-4dbb-b9ae-991b08888aea · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:13.052971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:13.052971Z digest=sha256:abba2a965d2616e0cbb155edbc524004ce5cb5e1686d4f1cbabd0d58d8e02eec

Observation bcbbc6c5-14c3-40dc-a0a5-f4e39207f35e · outbound

This paper cites A Content-Driven Micro-Video Recommendation Dataset at Scale.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations A Content-Driven Micro-Video Recommendation Dataset at Scale

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:13.166472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:13.166472Z digest=sha256:7eb11300b16a20ef56c5d9267e4ae89dd708c52df4fd75800dd4fcaaae2648a1

Observation 6e5ab421-3e45-40ec-be4f-91e26bdca6c6 · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:13.243785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:13.243785Z digest=sha256:ede4ac741cc576736bfda28bdf9dfb9870a1ea9cd1eaa417d5ad438b0dae7fb5

Observation a42f05f0-5664-461f-a87a-80e70907432c · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:52:17.489037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T20:52:13.318919Z digest=sha256:1ce385684ae6c659f440ce69fb1166b5bccb771e634197b772dc6e0a7249f2b4

Observation aac5cf76-5701-4a05-a699-4286a25a570f · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:13.429880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:13.429880Z digest=sha256:a92de36c6077bc6b791beeda1acdecf5f1b46c20003104e1fc42bd9c05a9bc5f

Observation 76c6d758-ba70-4ca7-aad7-346e57c4aae1 · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:13.545944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:13.545944Z digest=sha256:62eb056b0b39fa22e60d78ae712d4f4e452751d0bb5eb738c4244c00d1bb5bb1

Observation 127ea34d-e05c-456e-a5a0-9b5b23904366 · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:13.621368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:13.621368Z digest=sha256:fddd3fd484cd733c02c2d5bc1853b96c70097da70b31a880e2ceb03d9b959d8d

Observation c1959898-0515-44cc-a1f7-c8d94f9be27c · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:52:17.311605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T20:52:13.698637Z digest=sha256:59465c681123ede4581a7435a97632200e74e56785bbad2b476e708b23144fea

Observation 2594f7c6-066a-4613-b4fe-76768d84fb52 · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:52:17.121366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T20:52:13.771710Z digest=sha256:7d8b942793c10533566c1c8f7d267710f119d74ad497fe30583ec3093f9d3126

Observation 97a12b12-cc48-407f-ac8f-54f3433ea9cc · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:13.893223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:13.893223Z digest=sha256:bb5a117ab8f2ed60b95412813ffcbf5944b2d0d37758a8729b23fa536d3d43f5

Observation ab16a572-6f9e-49a6-b90c-cb56da4db9b2 · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:13.977897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:13.977897Z digest=sha256:a6ba29fdadbfb019ebb8cc311c5a288832ce488696323b097daf30f22bbe7bea

Observation 1f32c961-9e88-48bf-8b6d-ee745e3bce15 · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:52:16.960276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T20:52:14.071955Z digest=sha256:11dfbb28f80d5d3bd1050260907cad358f743359ab491b4cb1f772adb43825d9

Observation 88e5fd52-a81c-4a8f-87ed-baeef4e07f6c · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:14.202485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:14.202485Z digest=sha256:ae34f26786ef05007126f485b52c77ff449d6e7d5267ed4a29e6991249692856

Observation 7ff7e7a7-7591-4063-abbe-a5a33a89ffb3 · outbound

This paper cites C-Pack: Packed Resources For General Chinese Embeddings.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations C-Pack: Packed Resources For General Chinese Embeddings

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:14.307646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:14.307646Z digest=sha256:70751eb342ae80b0732442bb6fdfd988a10213c5c201cce057bb986ffccd3ed6

Observation 8a67f889-752a-44de-a832-b79c6c5c6f1b · outbound

This paper cites Qwen2.5-Omni Technical Report.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Qwen2.5-Omni Technical Report

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:14.402114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:14.402114Z digest=sha256:2fea302696d0f039e7bd3e5a5d26704f06257afd44ca1fc680bf91fe9c0fb3be

Observation 5fb8cb3b-4fec-4145-b173-f02a7cad2cb9 · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:14.485175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:14.485175Z digest=sha256:2cf5327b875c709d9bc9357b08081a3442c81168be1bfc40b3724a6d800db404

Observation 915ea390-591b-407b-adfc-d31a928fd135 · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:14.732843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:14.732843Z digest=sha256:5d34df5bd4be9a0cd27dd079eadcb8cfd7c1240d5659c56e9bd76db1299a982e

Observation 36bc5dd1-7a63-4e1a-b7d2-e823b8e7094a · outbound

This paper cites Tenrec: A Large-scale Multipurpose Benchmark Dataset for Recommender Systems.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Tenrec: A Large-scale Multipurpose Benchmark Dataset for Recommender Systems

Reference 34

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:52:16.114245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T20:52:14.872860Z digest=sha256:3d2575cfaacd364303c60fedb4228b98c1506d7fa34f602011454f370c8e5aa5

Observation 9ad5f067-a670-4a12-bca3-54a75516314e · outbound

This paper cites Qwen3 Embedding: Advancing Text Embedding and Reranking Through Foundation Models.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Qwen3 Embedding: Advancing Text Embedding and Reranking Through Foundation Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:14.945706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:14.945706Z digest=sha256:58ec669c8dd6669cab0808c9e0a6e10fa21835141d22435c26255f70a4b02058

Observation c3119bdb-4b74-4f4c-850b-44f8f265f529 · outbound

This paper cites Gundavarapu, Liangzhe Yuan, Hao Zhou, Shen Yan, Jen- nifer J.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Gundavarapu, Liangzhe Yuan, Hao Zhou, Shen Yan, Jen- nifer J

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:52:16.667824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T20:52:15.070733Z digest=sha256:fe60e884fc3acd6d0cb8c2b0b437665d7271e0fedecc8a24f6676cf64aabacf8

Observation 28361b96-a139-4637-afff-522bf6200c46 · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:15.235622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:15.235622Z digest=sha256:48a7dfbcbe575e031f4b9554880d813612dddd3a8a594625bfa5dea29164acc4

Observation bcebc30e-fec5-4ecb-b202-52b03f8cc715 · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:52:16.538053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T20:52:15.360641Z digest=sha256:610cc08d919c83f9cc5e2e9b67c05ab4b7e5baeecb0638238c427e9b10c7f8d9

Observation b261718e-d345-4574-8354-0c6507ea43ee · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:15.523539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:15.523539Z digest=sha256:936d1be34d64b33ae9eacc425e76bce4a83be0090b7372c749cb7ab45de27884

Observation 90c7369b-9612-4c3c-97dc-37408295e3d3 · outbound

This paper cites Is Extending Modality The Right Path Towards Omni-Modality?.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Is Extending Modality The Right Path Towards Omni-Modality?

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:52:15.874065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T20:52:15.689430Z digest=sha256:3082d1a07f5a70fc8643f91bb6559aaad1b44421143d2f0820a309e1be3515d1

Observation d0348ab4-8f7e-4f04-9d18-a5074b52bb5e · outbound

This paper cites InProceedings of the 28th ACM International conference on Multimedia.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations InProceedings of the 28th ACM International conference on Multimedia

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:12.646185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:12.646185Z digest=sha256:45f5867e12c4550fee2523b6ed9c4f4f9e1c4f822480293972f8f176d04c0535

Observation 2470fc22-7c73-4485-93db-eaba583f1b46 · outbound

This paper cites In Joint European Conference on Machine Learning and Knowledge Discovery in Databases.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations In Joint European Conference on Machine Learning and Knowledge Discovery in Databases

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:14.686234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:14.686234Z digest=sha256:cb70f0f8db462615b2145ed4c8adf8c2caef1faaf9b8dbcf36acca598db60b19

Observation e6791eb2-d322-4c60-8638-505775694151 · outbound

This paper cites In Machine Learning and Knowledge Discovery in Databases: European Conference, ECML PKDD 2022, Grenoble, France, September 19–23, 2022, Proceedings, Part I.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations In Machine Learning and Knowledge Discovery in Databases: European Conference, ECML PKDD 2022, Grenoble, France, September 19–23, 2022, Proceedings, Part I

Reference 2023

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:52:16.791348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T20:52:14.845228Z digest=sha256:0d16b317cafc3fafd2dcdb8c916d5e7b93b9704bd9b820f06c360918323a3316

Pith citing papers

No inbound Pith citation observations are available.