Pith. sign in

Paper Citation Record · LEDGER

ViLLA-MMBench: A Unified Benchmark Suite for LLM-Augmented Multimodal Movie Recommendation

As of 21 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 1 inbound Pith citation observation for arXiv:2508.04206.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.04206 v1

Coverage vector

measured 22 of 22 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T00:52:21.762363Z

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-27T14:42:10.975540Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T03:47:35.693365Z

Reference resolution

22 of 22 outbound references displayed

  • verified exact0
  • verified fuzzy9
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9dd8ac1d-68c7-4b39-a7a1-e8bdaa057672 · outbound

This paper cites Deldjoo, M.

ViLLA-MMBench: A Unified Benchmark Suite for LLM-Augmented Multimodal Movie Recommendation Deldjoo, M

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:52:24.986481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-06T00:52:19.889594Z digest=sha256:ac5dc3e4e1a240b9ee06793c760eee3372495f3c715f900f1370dd950b95074b

Observation 1e813b31-b548-4de8-b8aa-6736edb77484 · outbound

This paper cites an unresolved cited work.

ViLLA-MMBench: A Unified Benchmark Suite for LLM-Augmented Multimodal Movie Recommendation Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:52:24.847467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-06T00:52:19.948165Z digest=sha256:aed3fd1a59143aa87fa00fc22d79ad4dae21cd4f101fe8b293720f6a4228c3ad

Observation 9f96843b-62c2-4de9-8064-393663ddc718 · outbound

This paper cites an unresolved cited work.

ViLLA-MMBench: A Unified Benchmark Suite for LLM-Augmented Multimodal Movie Recommendation Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:52:24.668593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-06T00:52:20.048543Z digest=sha256:0bc013dc246bf6846a2978336940e680233a21a064678b55ded3c180f69bc3d3

Observation adbabba6-fbd4-4405-91ef-5fd09d143c26 · outbound

This paper cites A Content-Driven Micro-Video Recommendation Dataset at Scale.

ViLLA-MMBench: A Unified Benchmark Suite for LLM-Augmented Multimodal Movie Recommendation A Content-Driven Micro-Video Recommendation Dataset at Scale

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T00:52:20.162364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:52:20.162364Z digest=sha256:4c49d8f2dc971947abc1ab32cf7fd3bc842fbc76b49cc9c4936a5df1f889cb98

Observation 0e9a49e9-3391-45db-85a9-f3436f039c0b · outbound

This paper cites YouTube-8M: A Large-Scale Video Classification Benchmark.

ViLLA-MMBench: A Unified Benchmark Suite for LLM-Augmented Multimodal Movie Recommendation YouTube-8M: A Large-Scale Video Classification Benchmark

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T00:52:20.245801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:52:20.245801Z digest=sha256:7a8eed4501dfa9b4c540bc99e405fd214f453ad130cfa9ea0d76edeea26ce601

Observation 8aaae8e2-9307-4cc3-b0db-3db535bcccb5 · outbound

This paper cites Zhou, ``Mmrec: Simplifying multimodal recommendation,'' in Proceedings of the 5th ACM International Conference on Multimedia in Asia Workshops, 2023, pp.

ViLLA-MMBench: A Unified Benchmark Suite for LLM-Augmented Multimodal Movie Recommendation Zhou, ``Mmrec: Simplifying multimodal recommendation,'' in Proceedings of the 5th ACM International Conference on Multimedia in Asia Workshops, 2023, pp

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:52:24.450553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-06T00:52:20.344170Z digest=sha256:bc6ecf2061c8b37eacc0700153bac4a2f23e2af348bb45ef5ad50025fff9b6d7

Observation a4461401-a356-4acb-b327-ce6a8b4a7874 · outbound

This paper cites an unresolved cited work.

ViLLA-MMBench: A Unified Benchmark Suite for LLM-Augmented Multimodal Movie Recommendation Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:52:24.336502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-06T00:52:20.428385Z digest=sha256:e48aa078134e4f3da6f5ddb558f9febde3e10e43e58a9d91ddbfe2b954c7e787

Observation 7b08006f-5c4c-48bc-a756-05218d322e98 · outbound

This paper cites an unresolved cited work.

ViLLA-MMBench: A Unified Benchmark Suite for LLM-Augmented Multimodal Movie Recommendation Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:52:24.204048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-06T00:52:20.538380Z digest=sha256:024a16ebf473c15e700bb9339214edce375a565ba24e62fbca9d059ea78d5d68

Observation e6fe0da9-da07-479a-a052-772459f8317b · outbound

This paper cites Attimonelli, D.

ViLLA-MMBench: A Unified Benchmark Suite for LLM-Augmented Multimodal Movie Recommendation Attimonelli, D

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T00:52:20.649091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:52:20.649091Z digest=sha256:19ed35456ae1e1d291841053621388f60b92808c2e4eba7356603cc4e6d8f65a

Observation 538cbbb1-9f7b-481a-bdca-0132e2c4d31c · outbound

This paper cites an unresolved cited work.

ViLLA-MMBench: A Unified Benchmark Suite for LLM-Augmented Multimodal Movie Recommendation Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:52:24.031599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-06T00:52:20.770326Z digest=sha256:f013c4a535b01716aadea55cd541cd9fb6db6b21f8ca25318199af20a3bb530e

Observation 1218bdfd-c71c-4097-97b6-65c54ac74393 · outbound

This paper cites Attimonelli, D.

ViLLA-MMBench: A Unified Benchmark Suite for LLM-Augmented Multimodal Movie Recommendation Attimonelli, D

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:52:23.857260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-06T00:52:20.865714Z digest=sha256:e86c867769a473314fa919b0a60b41c5a0473a24e2d7ae9f33db41552cd54255

Observation bf739d6f-05fd-40e5-bf9b-2567a8531986 · outbound

This paper cites Rec-GPT4V: Multimodal Recommendation with Large Vision-Language Models.

ViLLA-MMBench: A Unified Benchmark Suite for LLM-Augmented Multimodal Movie Recommendation Rec-GPT4V: Multimodal Recommendation with Large Vision-Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T00:52:20.947043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:52:20.947043Z digest=sha256:528ba828de2bf72aeac9298c6e52671e495d284572f6d779e3466de8d92e2296

Observation d03b59d0-c496-4b31-9c06-627dd1f195a6 · outbound

This paper cites an unresolved cited work.

ViLLA-MMBench: A Unified Benchmark Suite for LLM-Augmented Multimodal Movie Recommendation Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:52:23.649803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-06T00:52:21.014836Z digest=sha256:656e873d2dbf9d58c27b2f222125611b7f9b16d8e3609c92435fec98a23a68e1

Observation 1aaf3904-e487-4c21-ba94-888843fd62fd · outbound

This paper cites Koren, R.

ViLLA-MMBench: A Unified Benchmark Suite for LLM-Augmented Multimodal Movie Recommendation Koren, R

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:52:23.530219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-06T00:52:21.101533Z digest=sha256:82606cc8c332622027d02b5878956d9e41842f1f353c4d8c21261c7bdd0e3f97

Observation 21b13d8a-e9a6-4836-8322-2e34bfb70cda · outbound

This paper cites Liang, R.

ViLLA-MMBench: A Unified Benchmark Suite for LLM-Augmented Multimodal Movie Recommendation Liang, R

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:52:23.382740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-06T00:52:21.189017Z digest=sha256:37c64db26d47fbd6791191aec2cd6a3aa7831d4c6b23f1d749e93f7f5213ae0e

Observation 0f8ef9bf-fccc-46bb-8a91-cd02ad2377a6 · outbound

This paper cites McAuley and J.

ViLLA-MMBench: A Unified Benchmark Suite for LLM-Augmented Multimodal Movie Recommendation McAuley and J

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:52:23.262514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-06T00:52:21.253714Z digest=sha256:8d0bd33fb653247a61f944b89e154fb69a319a0e10407ad05d24ec202953dc8b

Observation 1aa69449-8cb3-4908-aa91-3a0a81f495c0 · outbound

This paper cites an unresolved cited work.

ViLLA-MMBench: A Unified Benchmark Suite for LLM-Augmented Multimodal Movie Recommendation Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:52:23.046561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-06T00:52:21.322666Z digest=sha256:a3809cf7ad128292f7a2344175b910e3aa1325cb655f4b78be6184769e672826

Observation 2dc740be-4445-4193-98f9-a0d944016b86 · outbound

This paper cites He and J.

ViLLA-MMBench: A Unified Benchmark Suite for LLM-Augmented Multimodal Movie Recommendation He and J

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:52:22.888091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-06T00:52:21.405771Z digest=sha256:ae14bc5323f8b32e3db037776ee1aae3c28e2679b1ff78bdddc37276311c186d

Observation 8435100c-16c9-488b-a6de-bc7bf56e024a · outbound

This paper cites also-viewed.

ViLLA-MMBench: A Unified Benchmark Suite for LLM-Augmented Multimodal Movie Recommendation also-viewed

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:52:22.679691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-06T00:52:21.487572Z digest=sha256:e63ac6817d16d0a1c92b3c6159cc5617bcfd6aff314cce6128cab7838ebf722d

Observation e48c22ea-44a0-4f56-ba6a-cb375d2083a2 · outbound

This paper cites an unresolved cited work.

ViLLA-MMBench: A Unified Benchmark Suite for LLM-Augmented Multimodal Movie Recommendation Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:52:22.491129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-06T00:52:21.592810Z digest=sha256:6fd2d446bba56c620105044c799cd735b495206b37f1fdb3d62347933251f2d9

Observation 00837fba-31a8-48ef-b280-481cc59c26af · outbound

This paper cites an unresolved cited work.

ViLLA-MMBench: A Unified Benchmark Suite for LLM-Augmented Multimodal Movie Recommendation Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:52:22.275490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-06T00:52:21.667327Z digest=sha256:d8f77d8b678b0b890c6f9b19c0b7f76851b3ef61453f942112d5f10f565d651d

Observation 889947f3-08f7-4825-b433-ffccdace1c39 · outbound

This paper cites Salah, Q.-T.

ViLLA-MMBench: A Unified Benchmark Suite for LLM-Augmented Multimodal Movie Recommendation Salah, Q.-T

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:52:22.077445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-06T00:52:21.762363Z digest=sha256:4e5c256ff483c5b5a3243fdf2495f28bcf3c4e97514896bf79fd8d45cf061356

Pith citing papers

Observation 673ba97e-e191-4ed3-80c7-841203bd7350 · inbound

Popcorn: A Configurable Benchmark for Visual Evidence in Multimodal Movie Recommendation cites this paper.

Popcorn: A Configurable Benchmark for Visual Evidence in Multimodal Movie Recommendation ViLLA-MMBench: A Unified Benchmark Suite for LLM-Augmented Multimodal Movie Recommendation

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-03T03:47:35.694978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-27T14:42:10.975540Z digest=sha256:a2fe15f70e539aa623dd6d91eecfd11faea149b5402d86b3c5178baa411ecedd