Pith. sign in

Paper Citation Record · LEDGER

Learning Item Representations Directly from Multimodal Features for Effective Recommendation

As of 18 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 2 inbound Pith citation observations for arXiv:2505.04960.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.04960 v1

Coverage vector

measured 53 of 53 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T23:22:53.947087Z

measured 55 of 55 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T17:33:30.789820Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T00:42:54.760557Z

Reference resolution

53 of 53 outbound references displayed

  • verified exact0
  • verified fuzzy31
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 93c91814-c9c1-4087-857b-bd56368d6bc0 · outbound

This paper cites Multimodal machine learning: A survey and taxonomy,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Multimodal machine learning: A survey and taxonomy,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T23:22:53.740003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:22:53.740003Z digest=sha256:a6f880fdbb52861db915e693efea86bb99aa57c750c5729b3f3af52be6ebf294

Observation 4aef370d-1342-4e88-8b0e-637ad9357248 · outbound

This paper cites Recommender sys- tems leveraging multimedia content,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Recommender sys- tems leveraging multimedia content,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:22:54.526127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:22:53.744568Z digest=sha256:961ad960854ef6ce77db83ef0fdb99c303726cd98b8e97e75528ff02f69d7ec8

Observation f87bd943-93f8-405f-a808-9281c165b9ee · outbound

This paper cites Cornac: A comparative framework for multimodal recommender systems,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Cornac: A comparative framework for multimodal recommender systems,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:22:54.514643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:22:53.748499Z digest=sha256:9ab00f8e62043460c0ff06bd2b0cdb0376cfb6c11f84e384863982b6002df2a7

Observation 3cdb4ad6-5fa6-452d-9db9-80d6665d5c9e · outbound

This paper cites A Comprehensive Survey on Multimodal Recommender Systems: Taxonomy, Evaluation, and Future Directions.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation A Comprehensive Survey on Multimodal Recommender Systems: Taxonomy, Evaluation, and Future Directions

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T23:22:53.752365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:22:53.752365Z digest=sha256:e5111567fa36e8128b8d714da9de1a56cd76f8f3bffb63256e39e28ced4eb9e1

Observation f2669a38-9ce6-4b09-be62-341e26386226 · outbound

This paper cites Multimodal recommender systems: A survey,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Multimodal recommender systems: A survey,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:22:54.503202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:22:53.757492Z digest=sha256:a42c0909d8a714e2c5d5e26a974bea452a4fcfa0f20efff4faf3d9d3dde9d148

Observation ba104de6-cfb1-465c-9d7e-ea8125279e75 · outbound

This paper cites Multimodal pretraining, adaptation, and generation for recommendation: A survey,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Multimodal pretraining, adaptation, and generation for recommendation: A survey,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T23:22:53.762187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:22:53.762187Z digest=sha256:2e36bbe82f01871a0edabdcc52832728903ab23659c390221c3f6cff5e096c6e

Observation 4216c2e7-8f5c-4202-8087-bee2623016c2 · outbound

This paper cites Mining latent structures for multimedia recommendation,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Mining latent structures for multimedia recommendation,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:22:54.483154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:22:53.766884Z digest=sha256:2b180a83c17739912c17f874d716317b4c61fb6aee40721fa8a99c0848ba4b40

Observation 265134ae-4f16-4c99-a53e-d8ee7fde100e · outbound

This paper cites A tale of two graphs: Freezing and denoising graph structures for multimodal recommendation,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation A tale of two graphs: Freezing and denoising graph structures for multimodal recommendation,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T23:22:53.772611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:22:53.772611Z digest=sha256:1296df74db3d6c1fed01a7641516c045aa789b38981fd183da548e828f6a5f4d

Observation 3337275d-fd81-454a-b4e1-953be9142a27 · outbound

This paper cites Enhancing Dyadic Relations with Homogeneous Graphs for Multimodal Recommendation.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Enhancing Dyadic Relations with Homogeneous Graphs for Multimodal Recommendation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T23:22:53.776557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:22:53.776557Z digest=sha256:3b366f783264e3445d875229a849d12d15abb3f2eb54b081c45e10f6942b40b8

Observation 63792155-1818-4088-b9d1-063aaf28c146 · outbound

This paper cites Vbpr: visual bayesian personalized ranking from implicit feedback,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Vbpr: visual bayesian personalized ranking from implicit feedback,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T23:22:53.780808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:22:53.780808Z digest=sha256:2bb3b7a35e0273aca3a56bc6f03d14b4b6614409e517bde65ce2420ec0850584

Observation f4b32952-3191-4f92-84ed-6af21cf9e070 · outbound

This paper cites Multi-view graph convolutional network for multimedia recommendation,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Multi-view graph convolutional network for multimedia recommendation,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T23:22:53.785801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:22:53.785801Z digest=sha256:cb266a1a570b7dd2b6a3be78291474f14254144a1c8a371c54f080293066f458

Observation 00186cad-7d66-4788-86c8-05384321eb8c · outbound

This paper cites Graph-refined convolutional network for multimedia recommendation with implicit feedback,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Graph-refined convolutional network for multimedia recommendation with implicit feedback,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:22:54.447892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:22:53.790830Z digest=sha256:293e17046dced00cdd60a3096a8e8d061720e34d95b50a869f9bc66410b90a0b

Observation 1f222f44-5174-4f18-bc42-a4c09231da22 · outbound

This paper cites Attention-guided multi- step fusion: a hierarchical fusion network for multimodal recommenda- tion,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Attention-guided multi- step fusion: a hierarchical fusion network for multimodal recommenda- tion,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:22:54.436336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:22:53.794645Z digest=sha256:8d9a499c0fe18dc6bf1c70737bca26bf0422ec6b0d9b86fc475785d28675ddf1

Observation 523dc7d5-47de-447d-a981-e531c5b0fcbd · outbound

This paper cites Lgmrec: Local and global graph learning for multimodal recommendation,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Lgmrec: Local and global graph learning for multimodal recommendation,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T23:22:53.798385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:22:53.798385Z digest=sha256:775982a8886fe929e8ca314d7adeb805c89e10f47a958271cdf5f607ba666b4d

Observation d4958c89-8798-40b7-9231-841f996db93b · outbound

This paper cites Bpr: Bayesian personalized ranking from implicit feedback,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Bpr: Bayesian personalized ranking from implicit feedback,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:22:54.417587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:22:53.803347Z digest=sha256:ab489451e3041a6dbbdd2cf5845eda3595a191a778d3c5f983154accd91a08dd

Observation db1de974-6a8b-4063-881c-800400d1f4f1 · outbound

This paper cites Mmgcn: Multi-modal graph convolution network for personalized recommenda- tion of micro-video,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Mmgcn: Multi-modal graph convolution network for personalized recommenda- tion of micro-video,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:22:54.405955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:22:53.807140Z digest=sha256:d16aa8d08d1cf2130fc4b9e27b46a6b1eed9a36fbccc610e5379e19707269405

Observation ac9e5b68-741b-4a20-bf87-8279ded99d5a · outbound

This paper cites Towards universal sequence representation learning for recommender systems,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Towards universal sequence representation learning for recommender systems,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T23:22:53.811185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:22:53.811185Z digest=sha256:3fbe45ccc0492c97252c6316e4822617dcccc8a42083d2e67045851c0772e0fd

Observation c70135c5-c278-48d2-af3c-f3a0964a7769 · outbound

This paper cites Learning vector- quantized item representation for transferable sequential recommenders,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Learning vector- quantized item representation for transferable sequential recommenders,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:22:54.384283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:22:53.814692Z digest=sha256:3a1488cc2e2a0ee32771416781e4d4965ffd9f9a7b6f7f9fabf1823b7ad37fc5

Observation 802b95ea-2932-44b5-8198-8563053f6603 · outbound

This paper cites Are id embeddings nec- essary? whitening pre-trained text embeddings for effective sequential recommendation,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Are id embeddings nec- essary? whitening pre-trained text embeddings for effective sequential recommendation,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:22:54.370383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:22:53.818776Z digest=sha256:e977612389b8a86a250ece673227d1c8874f6532ebf0674c6aba7891e02ab1d5

Observation 3d070473-bde6-4d85-83de-a5a99b2aa18f · outbound

This paper cites The elephant in the room: rethinking the usage of pre-trained language model in sequential recommendation,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation The elephant in the room: rethinking the usage of pre-trained language model in sequential recommendation,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:22:54.358228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:22:53.822709Z digest=sha256:d3fba567a55bcdabc07bc1bd54ca7123554fdf13fefc7e6bebe52a70a84b826f

Observation 87ab3267-9e89-4fc4-b37a-1f150e4e9c11 · outbound

This paper cites Dual-view whitening on pre- trained text embeddings for sequential recommendation,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Dual-view whitening on pre- trained text embeddings for sequential recommendation,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:22:54.346845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:22:53.826716Z digest=sha256:80cc8596ae5638d524fe53e250bd41ef8a400a605fcaa3c8c5ae94d05325af8f

Observation 1597e60f-00f3-41aa-9c20-30e173748911 · outbound

This paper cites Where to go next for recommender systems? id-vs. modality-based recommender models revisited,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Where to go next for recommender systems? id-vs. modality-based recommender models revisited,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:22:54.334636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:22:53.830673Z digest=sha256:1c7e3cb254d2330aa9e1d6053b0881987146ed9037e462031c0e2c76c0da0eae

Observation f2924912-368f-45c1-8f3e-d8988d1838d1 · outbound

This paper cites Semantic- guided feature distillation for multimodal recommendation,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Semantic- guided feature distillation for multimodal recommendation,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:22:54.322802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:22:53.834485Z digest=sha256:0d473ac13a29ef06d5da2dc02a7e995881b1e4bfd23ef9b1414981d95eac70f3

Observation b0361750-02b3-45a4-8abc-bca156cacc06 · outbound

This paper cites Rectifier nonlinearities improve neural network acoustic models,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Rectifier nonlinearities improve neural network acoustic models,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T23:22:53.838418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:22:53.838418Z digest=sha256:fffa5aab73cb976c41c3bfa7ac48cf13e6f57a5161e0b3da9353d7112dbd29d4

Observation 7b61a210-56ad-4766-981d-588aff012c1f · outbound

This paper cites Discrete cosine transform,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Discrete cosine transform,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T23:22:53.841874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:22:53.841874Z digest=sha256:7c85a832b16972feeca1d95721da0705a6daac8fe38135627223177173222697

Observation 744e1cb2-da55-4abc-b51c-37ff748f833b · outbound

This paper cites Faster neural networks straight from jpeg,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Faster neural networks straight from jpeg,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:22:54.295350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:22:53.845716Z digest=sha256:9b73cf0883ad8e2e51b3f113909e0f7b57f86880ffd47e1919e3d892e020025d

Observation e39e0307-9acb-4bb3-9dbe-7890fdc3896b · outbound

This paper cites Discrete cosin trans- former: Image modeling from frequency domain,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Discrete cosin trans- former: Image modeling from frequency domain,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:22:54.281657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:22:53.849599Z digest=sha256:3e32b253a5ceee7f667ee45c07486aa0f8c26a0a0fadbfb02828f43fcd2276e9

Observation 9129b117-bb34-4a95-ab5a-182f30107ba2 · outbound

This paper cites Dctvit: Discrete cosine transform meet vision transformers,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Dctvit: Discrete cosine transform meet vision transformers,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:22:54.269049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:22:53.853369Z digest=sha256:158cd4c7122741eb3694848dfef1dd77bf7591a399523215754f896243040112

Observation 821be921-733d-44ae-9f95-281aa597d207 · outbound

This paper cites Dct-former: Ef- ficient self-attention with discrete cosine transform,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Dct-former: Ef- ficient self-attention with discrete cosine transform,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:22:54.256262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:22:53.857555Z digest=sha256:d27a531791679821633146da628aa5df284a512fa4e3330bfcf6918190e3f0ed

Observation 7f80f9cf-3f39-432e-acb6-ddcf1c446a77 · outbound

This paper cites Latent structure mining with contrastive modality fusion for multimedia recom- mendation,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Latent structure mining with contrastive modality fusion for multimedia recom- mendation,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:22:54.244661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:22:53.861312Z digest=sha256:f1ec8814d91b6e9563a99cb384b82d240660f2f9582a8d00dce8eb3140b5bdb3

Observation 4ff291cc-c5db-429a-a79c-c96b54becb2d · outbound

This paper cites Disentangled graph variational auto-encoder for multimodal recommendation with interpretability,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Disentangled graph variational auto-encoder for multimodal recommendation with interpretability,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T23:22:53.864703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:22:53.864703Z digest=sha256:85db637eeec240af3bffcc16a0cc210da450cb175b4e029832d9b1db9acac9d7

Observation e5cb1942-f5cf-448c-adc5-9831712b6375 · outbound

This paper cites Lightgcn: Simplifying and powering graph convolution network for recommenda- tion,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Lightgcn: Simplifying and powering graph convolution network for recommenda- tion,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:22:54.227001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:22:53.868139Z digest=sha256:891a52ecc0e20b38dde81f5a728baf26d9dc14e2c5cc25024f8c8b439900bcd8

Observation 3708abfb-ccf8-4ff2-abe8-1f6fc3cf2b0d · outbound

This paper cites Latent structure mining with contrastive modality fusion for multimedia recom- mendation,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Latent structure mining with contrastive modality fusion for multimedia recom- mendation,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:22:54.214997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:22:53.871720Z digest=sha256:d1f2c403cad9da2f6b7544a06a16c26888c0fd877ba7cfb16bde8498750215c0

Observation 10276aaf-a585-42ac-a38a-c6d3e81b02fb · outbound

This paper cites Attention is all you need,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Attention is all you need,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:22:54.203562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:22:53.875243Z digest=sha256:0bd715a546969fef703b2d58c0060e65c38fffd74a9406ddcf9ea8cff517e37c

Observation 575a773b-cc1d-4a48-8a57-88c800101715 · outbound

This paper cites Ups and downs: Modeling the visual evolution of fashion trends with one-class collaborative filtering,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Ups and downs: Modeling the visual evolution of fashion trends with one-class collaborative filtering,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T23:22:53.878850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:22:53.878850Z digest=sha256:bc1ebf4ee409a294137a3db88d9c0d60eb91e525cd48656a38369bf0a0182c90

Observation c63c7ae4-e336-4ab2-8c07-2755ac148912 · outbound

This paper cites A Content-Driven Micro-Video Recommendation Dataset at Scale.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation A Content-Driven Micro-Video Recommendation Dataset at Scale

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T23:22:53.882620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:22:53.882620Z digest=sha256:11b91067c3cd5c388f8f110f434fa06a09607097ba921edb662621edb4e9ce98

Observation d41f47e6-2a30-4c58-912a-1a96fe8d5d58 · outbound

This paper cites Mmrec: Simplifying multimodal recommendation,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Mmrec: Simplifying multimodal recommendation,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:22:54.183506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:22:53.886992Z digest=sha256:f7930a76cf802a7c6b010c6c0de1fe6ef21619174310a81a53d935e3ace94cf3

Observation c6030955-3239-427d-b9c5-cb59b21787b4 · outbound

This paper cites Improving Text Embeddings with Large Language Models.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Improving Text Embeddings with Large Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T23:22:53.890612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:22:53.890612Z digest=sha256:8ab3de7cb03eafe2b63a47af1539e4d06d7f731e5b7b09a55bff718ce73cb369

Observation a6d2c961-28ab-4704-8e28-1f86be8f674e · outbound

This paper cites LLM2Vec: Large Language Models Are Secretly Powerful Text Encoders.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation LLM2Vec: Large Language Models Are Secretly Powerful Text Encoders

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T23:22:53.894638Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:22:53.894638Z digest=sha256:bdc704fd87e2efea5a248474cca12831a799b2056759c1904af77341c2025856

Observation b2f1843f-bde6-49b3-a1f7-8e80acef1aef · outbound

This paper cites The Llama 3 Herd of Models.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation The Llama 3 Herd of Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T23:22:53.898604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:22:53.898604Z digest=sha256:0a4f4b8dd655495b021a09385766ca50fef66281b0ad9a0960e17bcfd6051dd5

Observation b2da99d7-4d10-4006-9fa9-374f9f5c1bd9 · outbound

This paper cites Self-supervised learning for multimedia recommendation,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Self-supervised learning for multimedia recommendation,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T23:22:53.902421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:22:53.902421Z digest=sha256:f7c901aa1d9e813e223acccf8c4be220099bb7f6ff062505e4e182dd323b6dc2

Observation d369e12c-c926-4f97-8e2d-580bfd975918 · outbound

This paper cites Bootstrap latent representations for multi-modal recommen- dation,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Bootstrap latent representations for multi-modal recommen- dation,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T23:22:53.906207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:22:53.906207Z digest=sha256:b6c94c3c9851e2514a51c62302a494ec318d015d66e564d30502fae04fab9fd9

Observation 0967f53e-9333-4a44-b758-dca09505899d · outbound

This paper cites Understanding the difficulty of training deep feedforward neural networks,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Understanding the difficulty of training deep feedforward neural networks,

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T23:22:53.910295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:22:53.910295Z digest=sha256:83e7c8c946c510eeac34305cb33c5e39d436b9cb43e10b5cf7ddc213f8f9fec0

Observation 496fda9c-a4e2-4470-81fa-3b9c514c928c · outbound

This paper cites Selfcf: A simple framework for self-supervised collaborative filtering,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Selfcf: A simple framework for self-supervised collaborative filtering,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:22:54.146655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:22:53.913804Z digest=sha256:c0c99f75c06f7e86bc2c2fb6a3ed1cf9b07a5f013d4b0d8b0466f4a35e3e69c2

Observation 655bb0aa-7e62-4f8c-989b-e6a842ffd0c5 · outbound

This paper cites Layer-refined graph convolu- tional networks for recommendation,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Layer-refined graph convolu- tional networks for recommendation,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:22:54.134142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:22:53.917893Z digest=sha256:ba9c3a587aeabf9d53141b2d295d5c85ad42082e9df4e5b2b56338c54e2a3912

Observation 17f52306-5b28-409c-a3f2-04482c8bd8e7 · outbound

This paper cites Multi- modal food recommendation using clustering and self-supervised learn- ing,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Multi- modal food recommendation using clustering and self-supervised learn- ing,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:22:54.120151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:22:53.921622Z digest=sha256:57caaa149eff54f6b6ad2db8bc3ba21b06b17033484c590acf662a05a345769a

Observation 34bb7191-7071-4658-ae74-1ffb74b464d2 · outbound

This paper cites Multi-modal food recommendation with health-aware knowledge dis- tillation,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Multi-modal food recommendation with health-aware knowledge dis- tillation,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:22:54.107586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:22:53.925079Z digest=sha256:b27640785c276abc24ae9570e3a2bf7c27bc09ce62ba9ff0d8a7ca5c8e10e8cf

Observation e695a58a-0889-4dc8-8f30-3f8048a07366 · outbound

This paper cites Multi-modal discrete collaborative filtering for efficient cold-start recommendation,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Multi-modal discrete collaborative filtering for efficient cold-start recommendation,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:22:54.094898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:22:53.928482Z digest=sha256:48290800ac857a23208711212d7031af4b2b04bddc78a08546b66854d07a3b8a

Observation 45f3e09b-c930-4f92-a8b3-60d0aad26d7a · outbound

This paper cites Multimodal pre-training for sequential recommendation via contrastive learning,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Multimodal pre-training for sequential recommendation via contrastive learning,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:22:54.082312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:22:53.931977Z digest=sha256:25489601881393a4b6d95f9254422040c446d0e34bc171adbd18b861c15e2e39

Observation 8bc74160-0df0-4f96-b7e4-4223d51b9941 · outbound

This paper cites Deepstyle: Learning user preferences for visual recommendation,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Deepstyle: Learning user preferences for visual recommendation,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:22:54.069909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:22:53.935853Z digest=sha256:56f7520acb49dc8698a8cf695244263b7bfedd5eada100985163a350dda0f74a

Observation f76fc906-bea8-4cac-912d-b7868ccf89d7 · outbound

This paper cites Atten- tive collaborative filtering: Multimedia recommendation with item-and component-level attention,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Atten- tive collaborative filtering: Multimedia recommendation with item-and component-level attention,

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-15T23:22:53.939199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:22:53.939199Z digest=sha256:a11f086ed8df1804e84e2da5747c0a14e6c8760ea9b91bde5f380f0b8b12fad7

Observation cda897e1-b82f-45aa-a094-e641ba6a1b95 · outbound

This paper cites Personalized fashion recommendation with visual explanations based on multimodal attention network: Towards visually explainable rec- ommendation,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Personalized fashion recommendation with visual explanations based on multimodal attention network: Towards visually explainable rec- ommendation,

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-15T23:22:53.942784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:22:53.942784Z digest=sha256:26577863f444d36e230fa11303946415f24bb793e29527c03c8e9cbee408eaa7

Observation f1a5e8bc-c7d5-4bbe-b731-cd0eb25ada4b · outbound

This paper cites Mm-frec: Multi-modal enhanced fashion item recommendation,.

Learning Item Representations Directly from Multimodal Features for Effective Recommendation Mm-frec: Multi-modal enhanced fashion item recommendation,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:22:54.044371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:22:53.947087Z digest=sha256:281290a0b5983a35e0bdee0d9ad4fbfa678ead41ea181636a384168a737d7c03

Pith citing papers

Observation b73b543e-44ac-4702-b9dd-d6fdbfb0e548 · inbound

EGRA:Toward Enhanced Behavior Graphs and Representation Alignment for Multimodal Recommendation cites this paper.

EGRA:Toward Enhanced Behavior Graphs and Representation Alignment for Multimodal Recommendation Learning Item Representations Directly from Multimodal Features for Effective Recommendation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T17:33:30.789820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:33:30.789820Z digest=sha256:807fd1e27d09f14163c0307e7fb4530ae37c9ecbe0183ce132840e2312c00ae9

Observation 41c1802d-0035-41f6-9d20-e5395dc4e6d5 · inbound

Modality-Aware Identity Construction and Counterfactual Structure Learning for ID-Free Multimodal Recommendation cites this paper.

Modality-Aware Identity Construction and Counterfactual Structure Learning for ID-Free Multimodal Recommendation Learning Item Representations Directly from Multimodal Features for Effective Recommendation

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-20T00:42:54.763990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-20T00:38:34.705348Z digest=sha256:7a957b2f1266c7344b4deda128336cd2cff5b156c0f0f7abe44922231248df06