Pith. sign in

Paper Citation Record · LEDGER

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations

As of 8 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 0 inbound Pith citation observations for arXiv:2508.09789.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.09789 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T20:52:15.689430Z

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

42 of 42 outbound references displayed

  • verified exact3
  • verified fuzzy2
  • unresolved37
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d055120f-4662-453a-a739-5b0970657c43 · outbound

This paper cites YouTube-8M: A Large-Scale Video Classification Benchmark.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations YouTube-8M: A Large-Scale Video Classification Benchmark

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:11.898760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:11.898760Z digest=sha256:db8a1b3f11324f0f377e6c600e321b60c5410ce95f4bd49c2fa366bbab2231ab

Observation 53afb53d-4905-43c0-bc65-1518f0ee4249 · outbound

This paper cites MiniGPT4-Video: Advancing Multimodal LLMs for Video Understanding with Interleaved Visual-Textual Tokens.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations MiniGPT4-Video: Advancing Multimodal LLMs for Video Understanding with Interleaved Visual-Textual Tokens

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:11.947587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:11.947587Z digest=sha256:b73d32858fe239d75e58262f4a5f3b43cf0de3bf4fa53711e2b0335a3c2f239c

Observation 3d50f3c6-2f7b-43f7-824d-93dd7f258105 · outbound

This paper cites Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:12.075305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:12.075305Z digest=sha256:95baef1df9439fe222f11e5449db474dc9884f30eaecd4d9dc8042b734c9a76b

Observation 4aea27eb-a000-40ae-97d5-04cafad7297d · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:12.159645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:12.159645Z digest=sha256:b741d9942d57167c526bfe510f9ac6f6a5066e0bab3751a47ef3a71e1dc457cd

Observation b29ad16e-a2c7-41a8-9d88-dc69fd749e74 · outbound

This paper cites VITA-1.5: Towards GPT-4o Level Real-Time Vision and Speech Interaction.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations VITA-1.5: Towards GPT-4o Level Real-Time Vision and Speech Interaction

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:12.243640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:12.243640Z digest=sha256:e4419bbe97fca71ab2db84ec1c4692ad345cb5feb5e1694e757ebd9e548e91ac

Observation 6575eda3-5343-4888-aed5-b0c5c45f7e7b · outbound

This paper cites Efficient and Effective Adaptation of Multimodal Foundation Models in Sequential Recommendation.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Efficient and Effective Adaptation of Multimodal Foundation Models in Sequential Recommendation

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:52:16.342181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:52:12.334243Z digest=sha256:69e174020d63c9a60dcd2b5c08000b6890930c166a70905469fc4473a0db2dbc

Observation 79b52895-6a8c-4903-a674-86fdbc6965f4 · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:12.424673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:12.424673Z digest=sha256:d89c8155d66bd7b357a7c8dbce9c7d4e27937b55891ec3487fd5500bc2943557

Observation 49c69ce7-a34f-4d91-bc08-3387ece6a749 · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:52:18.133375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:52:12.502619Z digest=sha256:73e504ebef5fc56915954a737cf30b947ad6b7ef59a36201622bf63311a317b4

Observation 4aa7b084-5a4e-4ce3-95fa-6f81d215c617 · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:12.593453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:12.593453Z digest=sha256:36a4f893c852f33bf2c4d94f83e35162f1611dce0396dd667bad4f0e6a0289ee

Observation d83c8282-6656-421f-8d58-e1728c634d8e · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:12.694792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:12.694792Z digest=sha256:fa37385509cbd4a924879d4b3cb967f87637d40eb90fbdfb2bcc693c8c002255

Observation daceec4b-7953-42e4-aafc-dbeb36d89a7a · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:52:17.968385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:52:12.749307Z digest=sha256:de6c0175c576a74b97fb51fbb1ebcd9d78f23d8c3d8401e3dd89f8f8629c1dfb

Observation 06700e41-578c-4819-8852-7093726e7337 · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:52:17.807895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:52:12.799997Z digest=sha256:6351e5392ccc5b9d22424358ed87e88d547b1aca07e3c5ff41195e6539fbb266

Observation 76eb33c5-c78b-439b-82b7-0bd79bf9d56e · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:52:17.678604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:52:12.872360Z digest=sha256:51e5a9c0d7a4d18fb8de2babdac80114ff4a93a8edc95f51d15785c75c05cf2c

Observation 48250b28-9d41-47ee-8236-f27f6671d5d9 · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:12.924826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:12.924826Z digest=sha256:0f844b9996e77270261208b961945e596b05c91c29277d2342c9b60190a90891

Observation 0722b8c9-6e91-4875-9b84-370a8afe9d33 · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:13.023547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:13.023547Z digest=sha256:0cc15d895b8b10ca425937b07ea94de76b17457602692f7e484e51b326308dfc

Observation a24d9033-1080-4dbb-b9ae-991b08888aea · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:13.052971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:13.052971Z digest=sha256:14f693b71d313e95b28c077bc63f2739daf8128e69a4bac189dc05e14ebe168c

Observation bcbbc6c5-14c3-40dc-a0a5-f4e39207f35e · outbound

This paper cites A Content-Driven Micro-Video Recommendation Dataset at Scale.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations A Content-Driven Micro-Video Recommendation Dataset at Scale

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:13.166472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:13.166472Z digest=sha256:bf9dae6c53fa7874b5a9a95ab7a0681e3689afc3dc5b8d8130eef32aaeeb6410

Observation 6e5ab421-3e45-40ec-be4f-91e26bdca6c6 · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:13.243785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:13.243785Z digest=sha256:280867a65c663e78664dd010cc2f093911b1753b1e7cf74f7d1aad8fc0fa963a

Observation a42f05f0-5664-461f-a87a-80e70907432c · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:52:17.489037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:52:13.318919Z digest=sha256:079109b76d324d934aca44fa643a41e898e25d20a783e2af4b24c571cc135dcc

Observation aac5cf76-5701-4a05-a699-4286a25a570f · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:13.429880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:13.429880Z digest=sha256:40ee7538dbb6170b07d42bb113ae2447d6e123d141ef0bf40a1b4f27816ab393

Observation 76c6d758-ba70-4ca7-aad7-346e57c4aae1 · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:13.545944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:13.545944Z digest=sha256:9cad09c0339b5839aa22e427a2061a35fd26b8bc12b5894e4c20f9c0d7b6720c

Observation 127ea34d-e05c-456e-a5a0-9b5b23904366 · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:13.621368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:13.621368Z digest=sha256:ff3841ab6f2d4805dedfac44833bb403f77a852a2387631d19a32114e540b9d3

Observation c1959898-0515-44cc-a1f7-c8d94f9be27c · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:52:17.311605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:52:13.698637Z digest=sha256:5af55d7bb5f345340f2e4e5b8c1452ae16f9451ef1f438f5c6d5f0fad34c2572

Observation 2594f7c6-066a-4613-b4fe-76768d84fb52 · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:52:17.121366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:52:13.771710Z digest=sha256:874035e384dce5c5589ec55072324704c921b215671215da2a1074c1e68b40c7

Observation 97a12b12-cc48-407f-ac8f-54f3433ea9cc · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:13.893223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:13.893223Z digest=sha256:bfaf9c39b03e2e83f24cfcb100ea21935fa1b13496ac7adb4f122308f1cf8b38

Observation ab16a572-6f9e-49a6-b90c-cb56da4db9b2 · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:13.977897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:13.977897Z digest=sha256:6a90676dd34bffb215c37fc0656d0a2bb003b2b60bf3dae1110bdce30374ea71

Observation 1f32c961-9e88-48bf-8b6d-ee745e3bce15 · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:52:16.960276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:52:14.071955Z digest=sha256:03271adcf3718e33f6405192058897dd402c40e5f05cd1ee568eb4496e2aa6f0

Observation 88e5fd52-a81c-4a8f-87ed-baeef4e07f6c · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:14.202485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:14.202485Z digest=sha256:b7987e64600bd732cbad71a6e8d8244f1e984122c9bb86a49b2bac3fd3f0162b

Observation 7ff7e7a7-7591-4063-abbe-a5a33a89ffb3 · outbound

This paper cites C-Pack: Packed Resources For General Chinese Embeddings.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations C-Pack: Packed Resources For General Chinese Embeddings

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:14.307646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:14.307646Z digest=sha256:0e6fb7ff3c2ed12f950396ae9cca082ac958c09870c828a4ebafe194162d62fc

Observation 8a67f889-752a-44de-a832-b79c6c5c6f1b · outbound

This paper cites Qwen2.5-Omni Technical Report.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Qwen2.5-Omni Technical Report

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:14.402114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:14.402114Z digest=sha256:e466b98260d7e30009323792aea6c2e8742b658abd23d08323eaf26cf6b3bf28

Observation 5fb8cb3b-4fec-4145-b173-f02a7cad2cb9 · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:14.485175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:14.485175Z digest=sha256:0e66269a99cfaac188e91906a2db0187046bea1cd49d9245f3a512d3b59b2a03

Observation 915ea390-591b-407b-adfc-d31a928fd135 · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:14.732843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:14.732843Z digest=sha256:d5f373ed6c94dc41f50505e7402ca1355e075b09f00b7c842876b6b3e010b733

Observation 36bc5dd1-7a63-4e1a-b7d2-e823b8e7094a · outbound

This paper cites Tenrec: A Large-scale Multipurpose Benchmark Dataset for Recommender Systems.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Tenrec: A Large-scale Multipurpose Benchmark Dataset for Recommender Systems

Reference 34

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:52:16.114245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:52:14.872860Z digest=sha256:1a39318f6cd2bc2cd8e75eee243452479b2f87c3c7b51b5fab3f10025e9b57fa

Observation 9ad5f067-a670-4a12-bca3-54a75516314e · outbound

This paper cites Qwen3 Embedding: Advancing Text Embedding and Reranking Through Foundation Models.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Qwen3 Embedding: Advancing Text Embedding and Reranking Through Foundation Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:14.945706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:14.945706Z digest=sha256:6a8ad43660b2934b4e6af0b76837494e8e2ac78188fabc06fef64e3190c34142

Observation c3119bdb-4b74-4f4c-850b-44f8f265f529 · outbound

This paper cites Gundavarapu, Liangzhe Yuan, Hao Zhou, Shen Yan, Jen- nifer J.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Gundavarapu, Liangzhe Yuan, Hao Zhou, Shen Yan, Jen- nifer J

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:52:16.667824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:52:15.070733Z digest=sha256:3fd4fef36f7ef74cf2ecca99cda6fbde13783a925a5883b48d6e125272704ea4

Observation 28361b96-a139-4637-afff-522bf6200c46 · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:15.235622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:15.235622Z digest=sha256:73989435ad211907b9d6caa8d873e1d197878e4494262b3e1fa4267f9d50a482

Observation bcebc30e-fec5-4ecb-b202-52b03f8cc715 · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-05T20:52:16.538053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:52:15.360641Z digest=sha256:dd44b29645e74910c55718814d42b5373d50e54c266853fb94b67e39458e553d

Observation b261718e-d345-4574-8354-0c6507ea43ee · outbound

This paper cites an unresolved cited work.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:15.523539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:15.523539Z digest=sha256:d526d1ce6a03a32b8f701f58473f5b540661c1c72a4e98a5c6a5ee45a76046a9

Observation 90c7369b-9612-4c3c-97dc-37408295e3d3 · outbound

This paper cites Is Extending Modality The Right Path Towards Omni-Modality?.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations Is Extending Modality The Right Path Towards Omni-Modality?

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:52:15.874065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:52:15.689430Z digest=sha256:ff13be9ee9373d7a8bf3aba8d73b7e65ee92b9eb2a44cf616794d6c4cb4a87c5

Observation d0348ab4-8f7e-4f04-9d18-a5074b52bb5e · outbound

This paper cites InProceedings of the 28th ACM International conference on Multimedia.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations InProceedings of the 28th ACM International conference on Multimedia

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:12.646185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:12.646185Z digest=sha256:4e37e0817c161e379ba3f171632800aa4bf9da0d88c30a77cb5b7240eb8c1d38

Observation 2470fc22-7c73-4485-93db-eaba583f1b46 · outbound

This paper cites In Joint European Conference on Machine Learning and Knowledge Discovery in Databases.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations In Joint European Conference on Machine Learning and Knowledge Discovery in Databases

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-05T20:52:14.686234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:52:14.686234Z digest=sha256:5f816f166413943a92935ff816df73c760262cff9fce6bd75ae1ebac61b82c26

Observation e6791eb2-d322-4c60-8638-505775694151 · outbound

This paper cites In Machine Learning and Knowledge Discovery in Databases: European Conference, ECML PKDD 2022, Grenoble, France, September 19–23, 2022, Proceedings, Part I.

Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations In Machine Learning and Knowledge Discovery in Databases: European Conference, ECML PKDD 2022, Grenoble, France, September 19–23, 2022, Proceedings, Part I

Reference 2023

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:52:16.791348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:52:14.845228Z digest=sha256:6879b1443cb74498da20d5bcbd76fc184eefec25cfaff87e28038b5ae72ee938

Pith citing papers

No inbound Pith citation observations are available.