Pith. sign in

Paper Citation Record · LEDGER

MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 36 inbound Pith citation observations for arXiv:2411.02571.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.02571 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 36 of 36 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T05:54:15.993546Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T15:49:57.631122Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 591ce93e-7e28-434c-9f98-8987dc53e7b8 · inbound

UniCoRN: Unified Commented Retrieval Network with LMMs cites this paper.

UniCoRN: Unified Commented Retrieval Network with LMMs MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-08T05:54:15.993546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T05:54:15.993546Z digest=sha256:52ff72caa4a4e3c5746cacc58f6f3436d34ec0b6298422d4e24dd2555649d703

Observation a6817cad-c2f1-499f-afc9-97037fc12468 · inbound

Modality Curation: Building Universal Embeddings for Advanced Multimodal Information Retrieval cites this paper.

Modality Curation: Building Universal Embeddings for Advanced Multimodal Information Retrieval MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T14:14:41.299337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:14:41.299337Z digest=sha256:de9196f250b284a297a6bdefbcb9ff7ad98b6cc907e8199c99365a5a9a018308

Observation e0561101-cd0f-4b08-b7a2-6713c5757cd8 · inbound

mRAG: Elucidating the Design Space of Multi-modal Retrieval-Augmented Generation cites this paper.

mRAG: Elucidating the Design Space of Multi-modal Retrieval-Augmented Generation MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T12:42:25.168442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:42:25.168442Z digest=sha256:1d8127ff213acbb254a7c2ecead91b7458ec894abb492d517e8af64cbf875466

Observation b54f5c2c-b25a-4369-8aeb-aafe3724a387 · inbound

MoCa: Modality-aware Continual Pre-training Makes Better Bidirectional Multimodal Embeddings cites this paper.

MoCa: Modality-aware Continual Pre-training Makes Better Bidirectional Multimodal Embeddings MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:24.419284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:24.419284Z digest=sha256:38c3377d23b9ca6ee2633214ec0ce8ceeb9d247b37e5f9d3754c479cd07378ea

Observation 2d954e84-55b7-4d5a-acaa-8464a1b88474 · inbound

VLM2Vec-V2: Advancing Multimodal Embedding for Videos, Images, and Visual Documents cites this paper.

VLM2Vec-V2: Advancing Multimodal Embedding for Videos, Images, and Visual Documents MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T14:10:15.181457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T14:10:14.929207Z digest=sha256:a98bbbef5231c8617fceb9c06781e059c52edc79a11646c4a23babac30ec4673

Observation 8e347a1a-dff9-4f3b-93bf-c328d78ab3fd · inbound

Empowering Multimodal LLMs with External Tools: A Comprehensive Survey cites this paper.

Empowering Multimodal LLMs with External Tools: A Comprehensive Survey MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 116

Resolution
unresolved
no resolver link, observed 2026-08-05T20:28:53.509159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:28:53.509159Z digest=sha256:2671acbd37c2009444872543e5adbe4d995f3329a278446ff01090dd24429934

Observation 8ff187b7-6219-4f13-b0fc-4bc36a32b6d5 · inbound

Multi-Rationale Explainable Object Recognition via Contrastive Conditional Inference cites this paper.

Multi-Rationale Explainable Object Recognition via Contrastive Conditional Inference MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T18:43:33.749092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T18:43:33.749092Z digest=sha256:e82904ca0b60a28c1aa54186754f57ed4e49c098849c433f5bf4e045994c1843

Observation 23f78e3b-2ba3-4a7f-81d8-b13912b962a4 · inbound

FreeRet: MLLMs as Training-Free Retrievers cites this paper.

FreeRet: MLLMs as Training-Free Retrievers MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:01:23.497573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T13:00:31.952588Z digest=sha256:913b83f6a6a41ea959a0ef695b605b9b6a56c2198d8cd0d56da9d33ba3c1cf35

Observation b9d4bee7-5631-4a32-8c17-ee87f78270b7 · inbound

FreeRet: MLLMs as Training-Free Retrievers cites this paper.

FreeRet: MLLMs as Training-Free Retrievers MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T13:51:45.715775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:51:45.715775Z digest=sha256:2f91b58c198a6f3f7d61b85849b50e35130c6ecf1730b9951f81f2d795dd2c73

Observation 066a815b-5d7f-4b47-b465-4e705632a3ae · inbound

Towards Mitigating Hallucinations in Large Vision-Language Models by Refining Textual Embeddings cites this paper.

Towards Mitigating Hallucinations in Large Vision-Language Models by Refining Textual Embeddings MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T23:35:18.907951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:35:18.907951Z digest=sha256:e775f8a9f2397592110460a518953ce59670dd92455af02645f40d091997d38b

Observation 7602a1db-8b2f-45aa-be7e-5ab2a868d988 · inbound

MOON2.0: Dynamic Modality-balanced Multimodal Representation Learning for E-commerce Product Understanding cites this paper.

MOON2.0: Dynamic Modality-balanced Multimodal Representation Learning for E-commerce Product Understanding MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-03T22:05:16.854105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:05:16.854105Z digest=sha256:629f144105245bae59b5c05d6d4b0ba64e2e05736a299b0e7829cde56af2f9a9

Observation 571ba26f-cb8a-498b-948e-ff8195dc5e6c · inbound

Adapting MLLMs for Nuanced Video Retrieval cites this paper.

Adapting MLLMs for Nuanced Video Retrieval MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-16T22:21:18.872651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T22:20:09.051957Z digest=sha256:8d4e36bae063a0b0768050e744a0d789e6d16c3d664a345f00f1032b2a6e4f0c

Observation dd0e4e57-7762-4cbd-815f-a76f08928956 · inbound

Magic-MM-Embedding: Towards Visual-Token-Efficient Universal Multimodal Embedding with MLLMs cites this paper.

Magic-MM-Embedding: Towards Visual-Token-Efficient Universal Multimodal Embedding with MLLMs MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-03T04:20:54.793364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T04:20:54.793364Z digest=sha256:92878231b1baef3dc9d784f47518dde33327c217cebe4e4d0898bb3c36dbfdcc

Observation bdbc48e4-9afb-4b04-b7c2-854017371b57 · inbound

PLUME: Latent Reasoning Based Universal Multimodal Embedding cites this paper.

PLUME: Latent Reasoning Based Universal Multimodal Embedding MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-13T21:53:19.976365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T21:48:40.722921Z digest=sha256:15a2050e962ffbeaf923bd652032cd3eec391f43c96172e062e6142fe3bdc49c

Observation f741e51f-173c-49fe-adad-56c022dd8aa5 · inbound

Bottleneck Tokens for Unified Multimodal Retrieval cites this paper.

Bottleneck Tokens for Unified Multimodal Retrieval MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:01:03.405897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T15:14:00.615638Z digest=sha256:89c3c8b0a8c14ccec8f497ac6116d3be13e2e0404b11f0b0585a85d635a73f44

Observation 7872be79-f826-4695-9471-b56e8ceac2c5 · inbound

ViLL-E: Video LLM Embeddings for Retrieval cites this paper.

ViLL-E: Video LLM Embeddings for Retrieval MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:21:02.017173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T15:00:43.573409Z digest=sha256:f8bd534e497843f6bdd9df2c48dc1e553b05056be8a243b9cb8cf4c69159278f

Observation edd7da31-3353-4b1b-ab14-541ecba97f9e · inbound

SLQ: Bridging Modalities via Shared Latent Queries for Retrieval with Frozen MLLMs cites this paper.

SLQ: Bridging Modalities via Shared Latent Queries for Retrieval with Frozen MLLMs MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:05:29.280756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T14:03:23.807710Z digest=sha256:f686e0a5ff3e57ca99a6c9049b70f603547f78a45e1e46d74bd87bbaaf0e8d06

Observation f8fee9d3-182e-4f19-b4e4-531bd55d0151 · inbound

SLQ: Bridging Modalities via Shared Latent Queries for Retrieval with Frozen MLLMs cites this paper.

SLQ: Bridging Modalities via Shared Latent Queries for Retrieval with Frozen MLLMs MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-12T00:56:14.404261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T00:54:25.022963Z digest=sha256:e175801d1acb57af913a2bcc08e13d1a3e11596c9599e4741e39bae5c4a5594a

Observation 817840ff-853c-410a-8776-972f69a02598 · inbound

Beyond Chain-of-Thought: Rewrite as a Universal Interface for Generative Multimodal Embeddings cites this paper.

Beyond Chain-of-Thought: Rewrite as a Universal Interface for Generative Multimodal Embeddings MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:06:08.280546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T12:46:30.827346Z digest=sha256:1b0b3685e183069496d17ef6fdcc66336aad479f869a8891296a9b07d6a1a935

Observation 3a35c670-b516-4a38-b881-f27ea1a16359 · inbound

Beyond Chain-of-Thought: Rewrite as a Universal Interface for Generative Multimodal Embeddings cites this paper.

Beyond Chain-of-Thought: Rewrite as a Universal Interface for Generative Multimodal Embeddings MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-12T18:31:37.144968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T18:31:37.144968Z digest=sha256:5e44f02a0bc55e167ac3b590a1b3065550aad338dc7fcc43d0e852c48917626c

Observation a727a064-ce40-48d5-a21b-5c6cda706a59 · inbound

Beyond Chain-of-Thought: Rewrite as a Universal Interface for Generative Multimodal Embeddings cites this paper.

Beyond Chain-of-Thought: Rewrite as a Universal Interface for Generative Multimodal Embeddings MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T15:40:29.196574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T15:40:29.196574Z digest=sha256:6ce68fc0f07e3ff513df5cca013e52208c04113154b212d25a1b9c75b02b3c09

Observation be6b8f37-bd82-45f4-b0a1-eb69e0d897c0 · inbound

MMEB-V3: Measuring the Performance Gaps of Omni-Modality Embedding Models cites this paper.

MMEB-V3: Measuring the Performance Gaps of Omni-Modality Embedding Models MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T21:01:11.445346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-08T07:28:09.725591Z digest=sha256:53307b7348489c29f896daec3e2a4677bf3a1b60b0d1e46d2e05b06f9e4c1103

Observation d30ecb05-8785-4015-a30b-1b2b6393d385 · inbound

MMEB-V3: Measuring the Performance Gaps of Omni-Modality Embedding Models cites this paper.

MMEB-V3: Measuring the Performance Gaps of Omni-Modality Embedding Models MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T15:38:00.262447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T15:38:00.262447Z digest=sha256:d24791e069b1dbfbf3984527a4b6fc85b581f9b1490404b0dba22a2c5179afad

Observation ba782c43-126d-40dd-91b5-65a3ff2c48c8 · inbound

Utility-Oriented Visual Evidence Selection for Multimodal Retrieval-Augmented Generation cites this paper.

Utility-Oriented Visual Evidence Selection for Multimodal Retrieval-Augmented Generation MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-14T19:09:22.929992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-14T19:09:18.975682Z digest=sha256:8a56d5c426f37c336fa6f965e3a9e166d388fc1516305cbbbe93a8cdae168270

Observation fefd53b2-54b2-4e8c-88e6-cfe8f9e72098 · inbound

TTE-Flash: Accelerating Reasoning-based Multimodal Representations via Think-Then-Embed Tokens cites this paper.

TTE-Flash: Accelerating Reasoning-based Multimodal Representations via Think-Then-Embed Tokens MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-20T18:03:36.966857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-20T18:00:18.315737Z digest=sha256:b88766e3e463e32b7af0312d5786c5fa06713c2d2098782cb88d299c0b48abad

Observation 36421437-351c-4ded-bc38-5d7fe16c31fe · inbound

TIGER-FG: Text-Guided Implicit Fine-Grained Grounding for E-commerce Retrieval cites this paper.

TIGER-FG: Text-Guided Implicit Fine-Grained Grounding for E-commerce Retrieval MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-19T23:57:53.255359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T23:54:32.646978Z digest=sha256:3371cd04e16b34d9646bfa7403acc52c862d9841880d886f57cf716a8b85db09

Observation 68e3f404-c0df-4047-b61b-98556488ebec · inbound

Unveil: Unified Visual-Textual Integration and Distillation for Multi-modal Document Retrieval cites this paper.

Unveil: Unified Visual-Textual Integration and Distillation for Multi-modal Document Retrieval MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-06-30T13:24:40.415138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T13:17:04.441743Z digest=sha256:19646e54c37cb6b3439ef6655b446cc63535a0ba7ba51e9232db25e688dd9f43

Observation dc643c07-c8c5-47a8-a898-d3d1f5ace4d0 · inbound

Gemini Embedding 2: A Native Multimodal Embedding Model from Gemini cites this paper.

Gemini Embedding 2: A Native Multimodal Embedding Model from Gemini MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T18:23:50.323587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T18:23:30.681253Z digest=sha256:9052a3fdd46ee76be39e8670f747e3df652aa840c25e64e14ba10a8dc60899b4

Observation 05bd93df-79c5-445d-a3e1-8e26b5cc654e · inbound

HiKEY: Hierarchical Multimodal Retrieval for Open-Domain Document Question Answering cites this paper.

HiKEY: Hierarchical Multimodal Retrieval for Open-Domain Document Question Answering MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T07:43:13.821615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T07:39:14.379668Z digest=sha256:2f4147ce5f5098c71f95a336b37e36f77b63ec7414e537dffb7aacfb3c0e373d

Observation 0a1b5cb4-a5ae-4ee1-b253-7cab5ddb9914 · inbound

DocRetriever: A Plug-and-Play Framework for Multimodal Document Retrieval with Comprehensive Benchmark cites this paper.

DocRetriever: A Plug-and-Play Framework for Multimodal Document Retrieval with Comprehensive Benchmark MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:13:15.090280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T08:09:41.068000Z digest=sha256:f4e4ad376d4fbaccaebeaf712cc51fa85f6fbeb06e1b4d4d01c7a3cd84fe3540

Observation 60feab1f-1f0d-4d6c-aa5e-9c5c62c4187f · inbound

MLT-Dedup: Efficient Large-Scale Online Video Deduplication via Multi-Level Representations and Spatial-Temporal Matching cites this paper.

MLT-Dedup: Efficient Large-Scale Online Video Deduplication via Multi-Level Representations and Spatial-Temporal Matching MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:58:03.542138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T09:43:22.054789Z digest=sha256:9386c162031f024c3eca6cf63f031240ffe79868c82458d46808b911f253fc54

Observation c124fc11-5e34-4994-a06c-07c25adecec3 · inbound

Rethinking RAG in Long Videos: What to Retrieve and How to Use It? cites this paper.

Rethinking RAG in Long Videos: What to Retrieve and How to Use It? MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-07-03T15:28:34.021407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T06:30:33.428489Z digest=sha256:ad0382e89208c4231d7cc809c042f3a6e57f672e195f334925dce2a20085f8b7

Observation b99378b5-4ddd-4d88-9031-cb92beb5cea5 · inbound

ELVA: Exploring Ranking-Driven Universal Multimodal Retrieval cites this paper.

ELVA: Exploring Ranking-Driven Universal Multimodal Retrieval MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-07-04T05:49:36.643419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T15:34:55.062016Z digest=sha256:10d34db61b804da5727933e9f195b478f42c2a5bfddbfce2cba59cda505e7f5f

Observation 0e40d168-2c80-4f17-9044-2d424136352b · inbound

Universal Guideline-Driven Image Clustering via a Hybrid LLM Agent cites this paper.

Universal Guideline-Driven Image Clustering via a Hybrid LLM Agent MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-07-04T15:49:57.633180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T01:19:19.253076Z digest=sha256:efa1927238b812cc5959dfff7852d78667b50d19d1d86c689616dccca7ed9e9d

Observation 7212669b-c4ee-4a8d-b674-f06b960e0043 · inbound

$M^3 QuestionIng$: Multi-modal Multi-span Medical Question Answering cites this paper.

$M^3 QuestionIng$: Multi-modal Multi-span Medical Question Answering MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-06-30T18:15:00.027444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T18:09:46.506953Z digest=sha256:9da6bbb441d2c37524b3f6737224e804edea328a986b84fe27d6639e34b4088c

Observation 4d40886e-5dfc-4178-86e3-e3a93d8769e7 · inbound

ReLoop-UME: Recurrent Depth with Learnable Retrieval Registers for Universal Multimodal Embedding cites this paper.

ReLoop-UME: Recurrent Depth with Learnable Retrieval Registers for Universal Multimodal Embedding MM-Embed: Universal Multimodal Retrieval with Multimodal LLMs

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T00:35:21.397349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T00:35:21.397349Z digest=sha256:7cc71dbbc22fb6d8fba953353cbb31d397bcdd6dbdcef665dfbcade629e44068