Pith. sign in

Paper Citation Record · LEDGER

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model

As of 21 August 2026, this Paper Citation Record lists 64 of 64 outbound references and 1 inbound Pith citation observation for arXiv:2504.17826.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.17826 v1

Coverage vector

measured 64 of 64 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:49:42.241870Z

measured 65 of 65 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T08:24:15.644051Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

64 of 64 outbound references displayed

  • verified exact2
  • verified fuzzy46
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d6ebd7ef-2cab-4eda-bdae-30207a101239 · outbound

This paper cites A comprehensive review of circular economy research in the textile and clothing industry,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model A comprehensive review of circular economy research in the textile and clothing industry,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.971514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:41.983098Z digest=sha256:4b0ad96b531df42115ea872f5a4c0450b1c53c22db3a19336f92d60148ae9341

Observation fcf3f090-d082-49b4-9eef-9bbd07afcae3 · outbound

This paper cites Hybrid recommender systems: Survey and experiments,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Hybrid recommender systems: Survey and experiments,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.959647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:41.988436Z digest=sha256:7ac432b098f374a7711cdb8038b3223512c440a280d56de7890505b9c29e14c9

Observation d218ef83-bad9-47f6-8aab-61276da57a93 · outbound

This paper cites The evolution and future of retailing and retailing education,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model The evolution and future of retailing and retailing education,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.947233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:41.992451Z digest=sha256:302b714489f41f869d5a814f101ca5eea242e10f1f73512979b890b1a57c0aeb

Observation 9b3de4e0-b6be-4c35-82ad-b2b98c55be52 · outbound

This paper cites Knowledge management and fashion retail performance: the moderating role of product complexity,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Knowledge management and fashion retail performance: the moderating role of product complexity,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.929603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:41.996628Z digest=sha256:9f750c36ecb47149ef2537b69abd8102812513228581486ab36123a1b4df2403

Observation 23df6d15-1792-4a65-935f-0afdd43daaa5 · outbound

This paper cites Assembled or unassembled? different types of outfit coordination presentations in online fashion retailing,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Assembled or unassembled? different types of outfit coordination presentations in online fashion retailing,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.919311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.001584Z digest=sha256:33734dfe2d025b68eea2e1e319a8d961b05caf0be47238b45d9498f6d55e4d16

Observation 5ca171eb-9269-4c97-a562-889fcffe719a · outbound

This paper cites What is the future of fashion retailing with generative ai? understanding consumer response through twitter data,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model What is the future of fashion retailing with generative ai? understanding consumer response through twitter data,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.908261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.005660Z digest=sha256:26a4b88daa132ea907e0cc625cacf94b9498d29d4b173bcb3bfaab17849ae460

Observation 4ad9fe94-f38e-4dd2-a3c9-4d6bcad1399a · outbound

This paper cites Modeling fashion compat- ibility with explanation by using bidirectional lstm,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Modeling fashion compat- ibility with explanation by using bidirectional lstm,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.897964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.009696Z digest=sha256:a1c3c5ac994a95049878b817573bdb0ed2c4e5fe952d803508a04c56f83e1eb5

Observation d601fc5b-cceb-45db-aaa9-3992d260555b · outbound

This paper cites Fashion forward with ai creations using gan,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Fashion forward with ai creations using gan,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.886767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.013260Z digest=sha256:6be81e657f098b1ff61b664345b9a3a9ab0ed3096d0e65f590b759882dab4dfb

Observation 9bdba344-d270-45c8-831f-9acbb7fc60a0 · outbound

This paper cites Per- sonalized clothing recommendation fusing the 4-season color system and users’ biological characteristics,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Per- sonalized clothing recommendation fusing the 4-season color system and users’ biological characteristics,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.874852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.016917Z digest=sha256:e3dc01c89979aac7fbfa959c70be67e0525fc4515a2789eb502b6f47052825e2

Observation 59e376e9-40cc-4bf9-b4ac-aff76ccb13df · outbound

This paper cites The impact of servitization on perceived quality, purchase intentions and recommendation intentions in the ready- to-wear sector,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model The impact of servitization on perceived quality, purchase intentions and recommendation intentions in the ready- to-wear sector,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.864199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.020743Z digest=sha256:3d452b393fe20485b303b95e1687349f83ec3845922db06a80b9b7fabedd5a39

Observation 38e67e14-3b44-4658-b088-1724cfc4be93 · outbound

This paper cites An intelligent recommendation system in e- commerce using ensemble learning,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model An intelligent recommendation system in e- commerce using ensemble learning,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T10:49:42.024872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:49:42.024872Z digest=sha256:0036444409ec0421fac6fe9f8a8b6632a246e315c66ee9e966a79f6946fb6f81

Observation 49df58c6-66c3-4346-b453-a6c04b04f727 · outbound

This paper cites Learning binary code for personalized fashion recommendation,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Learning binary code for personalized fashion recommendation,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.845495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.029645Z digest=sha256:7daab63b76400911fdf5f301e20b606b3c8f22c255da9949739b4c3f8c732b88

Observation 4223bc97-087a-481f-a5ba-41adcf0c4843 · outbound

This paper cites Learning similarity conditions without explicit supervision,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Learning similarity conditions without explicit supervision,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.834856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.033751Z digest=sha256:a7102c06dfde3eaa52bb724a5fef63293b644bc7562ea6fc963b79551ca0f047

Observation 03888240-d1d9-4143-8b86-8b4e385d7eb9 · outbound

This paper cites Fashion outfit complementary item retrieval,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Fashion outfit complementary item retrieval,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.823103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.037131Z digest=sha256:985ff8dbb9f50260872e034f7f73647d1ab999eece58799a3e037fdf94724561

Observation 92a3a12c-c34c-420d-af64-dcc1710a4e3e · outbound

This paper cites Learning type-aware embeddings for fashion compatibil- ity,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Learning type-aware embeddings for fashion compatibil- ity,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.810800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.041190Z digest=sha256:485780d2ef2bbba288b0a945c86df87520986e9a87add50375850f4e7136fc89

Observation 04f470de-a5e3-48ae-b05e-2d1148969b2a · outbound

This paper cites Language Models are Few-Shot Learners.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Language Models are Few-Shot Learners

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T10:49:42.044957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:49:42.044957Z digest=sha256:d5824fdaae6d62ecdb846fb6aed4b6a831dd95c41cfadbf8b20a531a0c803776

Observation 34b984d3-b27b-4c79-a915-5db6fe7bb0d9 · outbound

This paper cites Diffusion models beat gans on image synthesis,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Diffusion models beat gans on image synthesis,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T10:49:42.049567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:49:42.049567Z digest=sha256:da28ce7aafabed8378a00b542f81d691038b05fdc8d32fd91bccff6787fe1707

Observation 916039de-4ea0-4f87-8ab9-33d0d13958cc · outbound

This paper cites Outfitgan: Learning compatible items for generative fashion outfits,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Outfitgan: Learning compatible items for generative fashion outfits,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.790837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.053741Z digest=sha256:cd501ed70180d9f9da699f46aae0028bccbfd8c169324377718dc4719dbaf410

Observation 35468094-ca76-446a-8c53-f5ee30af0b12 · outbound

This paper cites Diffusion models for generative outfit recommendation,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Diffusion models for generative outfit recommendation,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.779867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.057222Z digest=sha256:ba8c6b086bb18edb0bbf544ae2d4a1ba662e7119c679e741443f6ec8c3243d12

Observation b4065064-c0a2-496e-b449-ba67dcf1f27b · outbound

This paper cites Generative Recommendation: Towards Next-generation Recommender Paradigm.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Generative Recommendation: Towards Next-generation Recommender Paradigm

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-16T10:49:42.060893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:49:42.060893Z digest=sha256:8140720a2e4e83c1c52da9c096fab5fe4d748b5aa8b6de3db909da23017460d0

Observation 7cc38f3c-14b4-4d98-bd98-e1e07b7ba541 · outbound

This paper cites Integrating Domain Knowledge into Large Language Models for Enhanced Fashion Recommendations.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Integrating Domain Knowledge into Large Language Models for Enhanced Fashion Recommendations

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-08-16T10:49:42.393472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.065843Z digest=sha256:cc91f8f9d2e986e783a14efc0b6cd894e4bdd6b366e84a1b9818f27077f34482

Observation 14a1073a-e31e-421d-891f-8695acba9703 · outbound

This paper cites Fashion recommendation systems, models and methods: A review,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Fashion recommendation systems, models and methods: A review,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.765987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.069789Z digest=sha256:93d00ceb625d1dc56bdc348aaf33c841e4424d638b32e3f28a38e74fedc4ae1b

Observation 113184b5-5842-4344-a21c-39fc651835f4 · outbound

This paper cites A review of modern fashion recommender systems,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model A review of modern fashion recommender systems,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-16T10:49:42.073654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:49:42.073654Z digest=sha256:5ad32c1e3d2c3af3089a9f8025204735f03b3c5341e42cdef3c9955d1bec2308

Observation 6921501c-af88-4c7b-9c59-448a31e6dc66 · outbound

This paper cites Generative ai-based style recommendation using fashion item detection and classification,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Generative ai-based style recommendation using fashion item detection and classification,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.753901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.078462Z digest=sha256:2252427e93643bca846baee66d5755944ce79b4de3ea5168e9b83b922a133e45

Observation a8c9c96e-8d5b-48ba-a249-86473d7b7951 · outbound

This paper cites Fashioning consumer choices: recommendation, motivation, and purchase intention toward instagram commerce. a mediation analysis,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Fashioning consumer choices: recommendation, motivation, and purchase intention toward instagram commerce. a mediation analysis,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.742186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.082462Z digest=sha256:6da4e7326fc07e3ba631d74298e719b3a02bff18d440e63e8a33749887a0e92f

Observation f1f6d656-2781-4b52-8244-2621037c24ec · outbound

This paper cites Image-based recommendations on styles and substitutes,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Image-based recommendations on styles and substitutes,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-16T10:49:42.087018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:49:42.087018Z digest=sha256:679b1a2ea3a18c708380500c6130a0a6d90e27cba04e1e1eda1b45f2b64351e8

Observation 0bca17aa-a28f-42c9-b5b6-5b88cb44b3c1 · outbound

This paper cites Category-aware mul- timodal attention network for fashion compatibility modeling,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Category-aware mul- timodal attention network for fashion compatibility modeling,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.721727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.090758Z digest=sha256:2885ec0b44b2fc2098b636caf2bdec8daa0a3b80441945a449c7f82de51f8237

Observation 042a6278-ebe6-4b38-978c-22a76d2f86fa · outbound

This paper cites Toward explainable fashion recommenda- tion,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Toward explainable fashion recommenda- tion,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.708638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.094854Z digest=sha256:66284d436b96feb90d8544c604bb448675b992775be860f5f150b82ebe20192b

Observation a3927eb3-d3e8-4692-8aef-b1b0668c5d92 · outbound

This paper cites Collaborative fashion recommendation: A functional tensor factorization approach,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Collaborative fashion recommendation: A functional tensor factorization approach,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.698118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.098706Z digest=sha256:93ca70e0c20ef2963dc2b3afd6cb36ac371f1e2dce1f025d0379732932d64212

Observation 1316061f-c01c-4588-9f1f-a089a484a63e · outbound

This paper cites Personalized outfit recom- mendation with learnable anchors,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Personalized outfit recom- mendation with learnable anchors,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.686209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.102562Z digest=sha256:4063fced1ac5e634e265c1c41ef9f6a26c4d7e3523aa2568bd5e60b7bef81c74

Observation 689557c6-caa8-4125-a7ac-66f34705b0de · outbound

This paper cites Pog: personalized outfit generation for fashion recommendation at alibaba ifashion,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Pog: personalized outfit generation for fashion recommendation at alibaba ifashion,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.675323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.106272Z digest=sha256:05d62ac6924fd461989bfdd12220e04d46af0b127e11f3de8eabe8c1545cd91c

Observation 1ad6f48e-fef5-4b09-bca2-e00221bffbd0 · outbound

This paper cites Hierarchical fashion graph network for personalized outfit recommendation,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Hierarchical fashion graph network for personalized outfit recommendation,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.665113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.109441Z digest=sha256:cc09cb877b43e39298263e77b19a91f00505802b8edf689c90c8b9eee0d152d1

Observation ef82ee6d-951b-4985-b583-5a47653e21aa · outbound

This paper cites Learning visual body-shape-aware embeddings for fashion compatibility,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Learning visual body-shape-aware embeddings for fashion compatibility,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.654748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.113587Z digest=sha256:a815799108c3827ee2f3ca1efd1c5eb9cbef1b7a54af361ad97100a85f27b390

Observation 5fb99c97-18b9-41f1-b963-e99a32d5fb3a · outbound

This paper cites What dress fits me best? fashion recommendation on the clothing style for personal body shape,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model What dress fits me best? fashion recommendation on the clothing style for personal body shape,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.644253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.117841Z digest=sha256:dbc268fa57eaee9d7938a4c06616d4d22a1adbe804c39a24b1d4eab12ead1715

Observation 29a7a6b5-98b0-4287-bff5-4720dd444ac4 · outbound

This paper cites Hairstyle suggestion using statisti- cal learning,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Hairstyle suggestion using statisti- cal learning,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.634137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.122070Z digest=sha256:bcb7231ab36338b9d9526e3559ac1a7cf9b7d90d90e792b66157a692f041c971

Observation 9cd4eaef-68b1-4fc3-b6d2-2e5aa45f5ecc · outbound

This paper cites Wow! you are so beautiful today!.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Wow! you are so beautiful today!

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.624114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.126593Z digest=sha256:338fe7ab5c29ceda35ed8365397d50b31810fdcb18126d24cce47c1a37a90a44

Observation 479a73d6-b855-4fb7-a791-04f0a7b11274 · outbound

This paper cites Show me the best outfit for a certain scene: A scene-aware fashion recommender system,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Show me the best outfit for a certain scene: A scene-aware fashion recommender system,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.612449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.130137Z digest=sha256:68bbea63eb13c89ffde851fc78e7723b00baa4e52f7ab307cb7ac001ee3779f9

Observation 3ed2e207-a413-427c-b8c7-7d2406d0f19c · outbound

This paper cites A unified framework for outfit design and advice,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model A unified framework for outfit design and advice,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.601950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.134807Z digest=sha256:5de4af5187736c5469a6fbb9cd405b0ba1d294dba3db32eede8d3d08b185177e

Observation e3bada71-a438-4903-a925-ac57ad0a463a · outbound

This paper cites Explainable outfit recommendation with joint outfit matching and comment genera- tion,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Explainable outfit recommendation with joint outfit matching and comment genera- tion,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.591247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.138383Z digest=sha256:0be2bf52f175127135c2aa0558ecaf5e1262a7f373840a5691e05bfb2368337f

Observation 646017b4-2335-4d6c-972f-5092d7ea5605 · outbound

This paper cites Fashion outfit generation for e-commerce,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Fashion outfit generation for e-commerce,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.579433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.142267Z digest=sha256:78036a7e2d1370a802d6fe026716ad5fe8f75ee87bb48e4dc1426506e8f332ba

Observation 2a76fc3a-6daa-4be9-a6e9-ef6e2d1dbd83 · outbound

This paper cites Fashion compatibility modeling through a multi-modal try-on-guided scheme,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Fashion compatibility modeling through a multi-modal try-on-guided scheme,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.568817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.145936Z digest=sha256:f28a21ca664c88ce865c78817c4d9ff1c976496d47bd2cafd5fd4d0732ef53a7

Observation 5acbc0c6-c13e-49f4-9325-de760b9d5a96 · outbound

This paper cites Knowledge-guided compatibility mod- eling,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Knowledge-guided compatibility mod- eling,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.557797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.149821Z digest=sha256:0b7defd3aeedfa9d0081417e5dac1cd0ce7a5ab500079aa20bda23c5862b1c72

Observation c0981c9f-6b8b-468d-8b05-8fca8f371a20 · outbound

This paper cites Outfittransformer: Outfit representations for fashion rec- ommendation,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Outfittransformer: Outfit representations for fashion rec- ommendation,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.546631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.154104Z digest=sha256:11c3f0b72facd3ac902752c14613042210a5c76447c346e46f2df779a1afce2e

Observation fa579f24-0b81-4238-9cc0-a69009d914fc · outbound

This paper cites Leveraging multimodal features and item-level user feedback for bundle construc- tion,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Leveraging multimodal features and item-level user feedback for bundle construc- tion,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.535235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.158153Z digest=sha256:763e039dc95b9d028b3dc720e72cde66f57394d9f07a6467f1307fe4f6fa0eb2

Observation 376f3bef-d490-4731-9cc0-b7ef31a06180 · outbound

This paper cites AI-Yo: Embedding Psychosocial Aspects In the Fashion Stylist Chatbot Design,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model AI-Yo: Embedding Psychosocial Aspects In the Fashion Stylist Chatbot Design,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.523104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.162563Z digest=sha256:f0c95a758b7f21d6e0fb8d9496fc7c822992a0943258fc087da240abb14db983

Observation d412ecd8-1a2f-4cd2-a6c0-36e51fbaf25f · outbound

This paper cites Multimodal conversational fashion recommendation with positive and negative natural-language feedback,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Multimodal conversational fashion recommendation with positive and negative natural-language feedback,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.510418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.166391Z digest=sha256:bd40929dacffed41a666456066887ceba6d5a71b47ef08dc0b8ee48b5a4c2d53

Observation 7a8317ee-331e-46d2-958b-9c597e62340b · outbound

This paper cites Multi-modal dialog state tracking for interactive fashion recom- mendation,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Multi-modal dialog state tracking for interactive fashion recom- mendation,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.499440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.170570Z digest=sha256:7764c6b68089655bb804bf52014431a9ad73c322c8b5d3ff3ef8c21d335f8ae3

Observation 0a18937b-87f2-4f53-9a82-2c1d43f8067c · outbound

This paper cites Show-o: One Single Transformer to Unify Multimodal Understanding and Generation.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Show-o: One Single Transformer to Unify Multimodal Understanding and Generation

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-16T10:49:42.174850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:49:42.174850Z digest=sha256:93b286447d908c152b089a9a26d6555cdb20c57c0b4bca4f73bfdfd0e9c5c867

Observation a3550611-3e6f-42f4-ad15-0d675fd31e71 · outbound

This paper cites JanusFlow: Harmonizing Autoregression and Rectified Flow for Unified Multimodal Understanding and Generation.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model JanusFlow: Harmonizing Autoregression and Rectified Flow for Unified Multimodal Understanding and Generation

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-16T10:49:42.179374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:49:42.179374Z digest=sha256:c7e58ea40dd8744537082263933aa02764260d88d031e18e3f6c742b222f1946

Observation 2de4414d-15ec-4860-9b09-af3631469b4f · outbound

This paper cites Vitron: A Unified Pixel-level Vision LLM for Understanding, Generating, Segmenting, Editing.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Vitron: A Unified Pixel-level Vision LLM for Understanding, Generating, Segmenting, Editing

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-16T10:49:42.183744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:49:42.183744Z digest=sha256:789c90002533885f7a60d0da9610a09e40c5da9b5ce942c99386a5e34eea37b8

Observation 8023ea8e-1476-45c1-911b-9da72d5165bc · outbound

This paper cites UniFashion: A Unified Vision-Language Model for Multimodal Fashion Retrieval and Generation.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model UniFashion: A Unified Vision-Language Model for Multimodal Fashion Retrieval and Generation

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-16T10:49:42.188238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:49:42.188238Z digest=sha256:7d9719e00a633c868928559c946d9c34f7845df0b5094370ffe088b9d16e1f10

Observation d103c76e-a62f-449e-bcc4-9db9d62144dd · outbound

This paper cites Fashionai: A hierarchical dataset for fashion understanding,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Fashionai: A hierarchical dataset for fashion understanding,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.488566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.192936Z digest=sha256:d4b1b0d6f1426eb817d35488b10f3f747389bf7443802f17fda8fbbadb305662

Observation 32fbaedb-b21c-4dc0-90ea-ee853f37cc24 · outbound

This paper cites Theme-Matters: Fashion Compatibility Learning via Theme Attention.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Theme-Matters: Fashion Compatibility Learning via Theme Attention

Reference 53

Resolution
verified exact
local_arxiv, observed 2026-08-16T10:49:42.335894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.197024Z digest=sha256:61c5c00f2b56ca7a1e95d8773b333a779eedfb48b23f2cdd048e09e55fadcf66

Observation de069976-6d0f-4a81-ad7a-e0880fa44380 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Learning transferable visual models from natural language supervision,

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-16T10:49:42.201520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:49:42.201520Z digest=sha256:082a4824530be311e5573556c67e0ad600dd179e9661879c1559b5ddebc823a0

Observation dd45b2ed-729f-4f7c-b847-8dc746814640 · outbound

This paper cites Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-16T10:49:42.204933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:49:42.204933Z digest=sha256:dc237ea95ace68ccfcadf461481fb832128cf23db5f630c983a5418b197395c3

Observation 32f8db0a-dd76-4981-8c6a-8124608e2e13 · outbound

This paper cites Textbooks Are All You Need.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Textbooks Are All You Need

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-16T10:49:42.209513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:49:42.209513Z digest=sha256:47bfb694f8b83e6813b9137ea2c9148c3284df48e9208eac78fe95ab9a0b1520

Observation ea24a450-8af1-4f57-a406-babb0debca26 · outbound

This paper cites Chainlit,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Chainlit,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.469884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.213663Z digest=sha256:c1603875381dd9fe43c948c1f75c4e01173983654fadc5c1b05d2b88293333e4

Observation 7c19551d-caf2-4862-adef-aef3707635c8 · outbound

This paper cites Llama-3.2-11B-Vision,.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Llama-3.2-11B-Vision,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.458673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.217508Z digest=sha256:e5b5933548fc1778601f74b3aadf004189a43f015377bd3d24c578a8109d5817

Observation 8abe641f-b175-4ce7-8758-0a6338a3f0b2 · outbound

This paper cites GPT-4o System Card.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model GPT-4o System Card

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-16T10:49:42.221737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:49:42.221737Z digest=sha256:9261cfc032a3e9f9fdfa173e6524bee0f5e43df11cf5ca458fb786544d71a2b8

Observation e2f3e9d6-9200-4b0a-a4ed-f25704b398a7 · outbound

This paper cites MEDEC: A Benchmark for Medical Error Detection and Correction in Clinical Notes.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model MEDEC: A Benchmark for Medical Error Detection and Correction in Clinical Notes

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-16T10:49:42.226577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:49:42.226577Z digest=sha256:c525fe8246560c4b464cae3b52990e55a668c5c23f879003a1e61becdf7ed2f2

Observation b4f98a40-f1da-43ec-848b-1815959cf879 · outbound

This paper cites Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-16T10:49:42.230853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:49:42.230853Z digest=sha256:2f2035a2211b4ea9a4ed257366c95060d5c511249d15dda2da8aabef08f803ed

Observation 3bb3c7ef-01f5-4ad5-9a97-bfb8c3976ac6 · outbound

This paper cites messages=[{‘role’: ‘system’, ‘content’: ‘‘‘ As a fashion expert, generate a user−system conversation for training a fashion stylist model.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model messages=[{‘role’: ‘system’, ‘content’: ‘‘‘ As a fashion expert, generate a user−system conversation for training a fashion stylist model

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.447897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.234869Z digest=sha256:bbad9a4753039187dbcf06983376e7c4d2cf4666f8d566154a48841fff8f8ff4

Observation 95936c60-9f34-4c63-b94e-17a5a97b8363 · outbound

This paper cites messages=[{‘role’: ‘system’, ‘content’: ‘‘‘ Create a user−system conversation for training a personalized fashion stylist model.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model messages=[{‘role’: ‘system’, ‘content’: ‘‘‘ Create a user−system conversation for training a personalized fashion stylist model

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.436927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.238322Z digest=sha256:37be97bfe4601b9c3f2e49e09c2b9b9a2519a4a86d79cd6a4aeaa17cf00385fb

Observation f1c15865-64ea-40eb-ab7c-14e9c73bd39e · outbound

This paper cites UN GAYVOE MVEREI LELS.

FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model UN GAYVOE MVEREI LELS

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:49:42.426483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:49:42.241870Z digest=sha256:405ff360154237e57b3541a608f4542f7e989a2e4a8b966afd4cca353611474e

Pith citing papers

Observation 76cc31d9-f9ee-4b35-ab27-39f79bf088c2 · inbound

VOGUE: A Multimodal Dataset for Conversational Recommendation in Fashion cites this paper.

VOGUE: A Multimodal Dataset for Conversational Recommendation in Fashion FashionM3: Multimodal, Multitask, and Multiround Fashion Assistant based on Unified Vision-Language Model

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T08:24:15.644051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:24:15.644051Z digest=sha256:0e4cd95e5c3ee92a6fbcd9d47e48f371b61fac3f753c81edcfe5803c9383b253