Pith. sign in

Paper Citation Record · LEDGER

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories

As of 8 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 0 inbound Pith citation observations for arXiv:2602.10809.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2602.10809 v2

Coverage vector

measured 39 of 39 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-03T01:02:35.739213Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

39 of 39 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved39
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e703c280-ad41-4414-b796-56356ad3d700 · outbound

This paper cites Introducing claude opus 4.5, 2025 a.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Introducing claude opus 4.5, 2025 a

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:32.481330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:32.481330Z digest=sha256:a6f5d18497b9fb3e4228d39273a7db192f29979459f0b5043076b7d63ab974f9

Observation 0c3cd124-f102-4b22-b40c-74306fe7aab0 · outbound

This paper cites Introducing claude sonnet 4.5, 2025 b.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Introducing claude sonnet 4.5, 2025 b

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:32.519876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:32.519876Z digest=sha256:c696a509cc56a298cabb44f0ddf6d0cb1c8ecfd90be0965534b594407880bdb4

Observation b349e52a-5994-4ac3-97de-6beed00b7849 · outbound

This paper cites Qwen3-VL Technical Report.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Qwen3-VL Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:32.574737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:32.574737Z digest=sha256:b98b4117b4ae7578bb0504b4b2cd2e068fbfef14922c8c17f878b516906a6738

Observation 6e93fd0d-42b5-4259-bd2b-04f63bc3beaf · outbound

This paper cites Seed1.6-embedding.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Seed1.6-embedding

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:32.635539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:32.635539Z digest=sha256:7ac2916291b5e9036e5bdfd89922b7dfe5a64d084966e631ab643f065599b9cb

Observation 12d175b6-6b6f-42b9-9584-86eab0bb3b58 · outbound

This paper cites MoCa: Modality-aware Continual Pre-training Makes Better Bidirectional Multimodal Embeddings.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories MoCa: Modality-aware Continual Pre-training Makes Better Bidirectional Multimodal Embeddings

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:32.688658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:32.688658Z digest=sha256:b1202e102014e6c8de3c831365bb1e2e30b6eee9fd138877f14f65a96d9c8419

Observation bf2732b4-67c7-4f6a-95c7-0b56b7d21034 · outbound

This paper cites mme5: Improving multimodal multilingual embeddings via high-quality synthetic data.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories mme5: Improving multimodal multilingual embeddings via high-quality synthetic data

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:32.750201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:32.750201Z digest=sha256:2889b9840b6f77c6f919fa8ffd049c20878e69ccdb621352130230bcb9998a96

Observation 3526ed00-8740-452d-a0e9-f03ea2bb870b · outbound

This paper cites Generative thinking, corrective action: User-friendly composed image retrieval via automatic multi-agent collaboration.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Generative thinking, corrective action: User-friendly composed image retrieval via automatic multi-agent collaboration

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:32.802191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:32.802191Z digest=sha256:c86e837ee36001f7647c152262c599958bb8de2aaf45c497e364f91e9bbe10e6

Observation fdce8250-ad50-4153-8923-823f3152cbf8 · outbound

This paper cites EmbodiedEval: Evaluate Multimodal LLMs as Embodied Agents.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories EmbodiedEval: Evaluate Multimodal LLMs as Embodied Agents

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:32.859526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:32.859526Z digest=sha256:25208c99895d2a304bb31078d885bccff326e98675730c50d17263e677cb8d98

Observation 445417bb-7f86-409f-ba75-5d3366f6297d · outbound

This paper cites N., Awasthi, A., Pan, X., Ahuja, C., Mishra, S.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories N., Awasthi, A., Pan, X., Ahuja, C., Mishra, S

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:32.925071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:32.925071Z digest=sha256:eb622653a1cab4196fef7a5f237916f8cc2f6d1ed512500d422d4ca5aca13e22

Observation 83553ad4-7500-4cc2-b70d-d83a80b48ce5 · outbound

This paper cites Mind2web: Towards a generalist agent for the web.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Mind2web: Towards a generalist agent for the web

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:33.014774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:33.014774Z digest=sha256:c3f8a784969e6e3f127ec7798595eaef9e13d9ea0ade3115cf41b09bdbe24dc4

Observation 1a166859-1437-4990-b6e9-a382c3b0fc2b · outbound

This paper cites Colpali: Efficient document retrieval with vision language models.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Colpali: Efficient document retrieval with vision language models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:33.072098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:33.072098Z digest=sha256:b5fe11c23685859284012282e92430290e2a64930ffd7ec28b11942e9121f830

Observation 1fe1adda-0ff8-4ee8-810d-001bc1a4585f · outbound

This paper cites A new era of intelligence with Gemini 3.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories A new era of intelligence with Gemini 3

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:33.158432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:33.158432Z digest=sha256:154a9b7f8e42a22f3b82a65ab818b3d06e32bddbec21e346a780a09979f4896c

Observation 20dd5299-2df9-4aa2-b4a3-347238f418c1 · outbound

This paper cites Mind2Web 2: Evaluating Agentic Search with Agent-as-a-Judge.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Mind2Web 2: Evaluating Agentic Search with Agent-as-a-Judge

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:33.277823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:33.277823Z digest=sha256:33ba37df612dc931532d95e3a71deba7274c64b42ddf8530fe7487b6576dbe83

Observation 02532e51-fb6a-4e92-8739-f284b0ef4f0b · outbound

This paper cites GPT-4o System Card.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories GPT-4o System Card

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:33.394995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:33.394995Z digest=sha256:2e9068aed7ed8453607b2b020631129023a76dee6d3b239da26f578b6d200bca

Observation 9f24c0db-43fc-4f5d-bd7b-a6602f470807 · outbound

This paper cites V., Sung, Y., Li, Z., and Duerig, T.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories V., Sung, Y., Li, Z., and Duerig, T

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:33.487414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:33.487414Z digest=sha256:a76a795d0029c60f0b7da9931b307fdebdd178f5402ab39d0c4d92ef719925e6

Observation 31d3e9fd-d003-4240-b176-73b5528a99b8 · outbound

This paper cites Vlm2vec: Training vision-language models for massive multimodal embedding tasks.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Vlm2vec: Training vision-language models for massive multimodal embedding tasks

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:33.605943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:33.605943Z digest=sha256:78d24a0efe336010d3fd49c9fcdb85297e22c5eee9e56737604e07ed8c53ea8a

Observation cc69d2e7-97ed-4dad-90db-293002dc2100 · outbound

This paper cites Crafting papers on machine learning.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Crafting papers on machine learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:33.698727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:33.698727Z digest=sha256:ee8285a79ef09fc38cecf72055c710061bd47a27360eb1aeca1886d02e1789e1

Observation 15aa4e4e-bea2-4fb9-bdb9-a3d9b3927088 · outbound

This paper cites an unresolved cited work.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:33.818381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:33.818381Z digest=sha256:5a956b87997bfa5986e55fd98f96f686eec72c502007a8a8e43f8ce1e57e4996

Observation 7ded2419-e48c-4485-958d-a9b5d6c385d7 · outbound

This paper cites Qwen3-VL-Embedding and Qwen3-VL-Reranker: A Unified Framework for State-of-the-Art Multimodal Retrieval and Ranking.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Qwen3-VL-Embedding and Qwen3-VL-Reranker: A Unified Framework for State-of-the-Art Multimodal Retrieval and Ranking

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:33.914474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:33.914474Z digest=sha256:2ffc899484531b47c95a43144e9c2af8de15866742f2f66f14423681b272fe91

Observation 32b42022-5d2f-4b8e-9ab1-c7de86b259c1 · outbound

This paper cites MM-BrowseComp: A Comprehensive Benchmark for Multimodal Browsing Agents.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories MM-BrowseComp: A Comprehensive Benchmark for Multimodal Browsing Agents

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:34.091340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:34.091340Z digest=sha256:17afaf3caa13570d62d7c27fa0a65a61caa74f55f7d00362b98b9225d9bbb15c

Observation 38ffbf4b-2e8d-4ba8-ae1a-df91b7c7e20e · outbound

This paper cites Mm-embed: Universal multimodal retrieval with multimodal LLMS.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Mm-embed: Universal multimodal retrieval with multimodal LLMS

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:34.230034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:34.230034Z digest=sha256:d76818aed5f1cef04b6efe27879cf86c6d4fa2b47c40a0064b838f84befdef17

Observation 16166113-7826-420e-87a1-7f1c84adfb6b · outbound

This paper cites VLM2Vec-V2: Advancing Multimodal Embedding for Videos, Images, and Visual Documents.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories VLM2Vec-V2: Advancing Multimodal Embedding for Videos, Images, and Visual Documents

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:34.320848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:34.320848Z digest=sha256:f2d94f15146e9c6139670cb46a1c9d7893101c972af1c57a341847e7519699d2

Observation 1ac97c6c-3015-4164-ae98-6d82a0434c9c · outbound

This paper cites Introducing GPT-5.2 : The most advanced frontier model for professional work and longrunning agents.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Introducing GPT-5.2 : The most advanced frontier model for professional work and longrunning agents

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:34.417790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:34.417790Z digest=sha256:cad53386e5310fe9f548e66151456014148cf5bed96d10d21160508ba151404b

Observation d1f48c7d-b670-41b2-9661-828190fdb8db · outbound

This paper cites W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J., Krueger, G., and Sutskever, I.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J., Krueger, G., and Sutskever, I

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:34.591875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:34.591875Z digest=sha256:1e275eb70f9f0a6f0f6609790770d1e6649917373822bcd8c167088b7a23ec67

Observation 0b387237-72f7-43ce-87b4-bb5e9b8e8374 · outbound

This paper cites Glm-4.5v and glm-4.1v-thinking: Towards versatile multimodal reasoning with scalable reinforcement learning, 2025.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Glm-4.5v and glm-4.1v-thinking: Towards versatile multimodal reasoning with scalable reinforcement learning, 2025

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:34.769100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:34.769100Z digest=sha256:2411099ce4f37b6c02a8dbb612a04ddf4dd7d2beed86fecfbbf2eadfbe7fd171

Observation 36d67bfd-7b05-455f-8669-7ddc259737ba · outbound

This paper cites A., Friedland, G., Elizalde, B., Ni, K., Poland, D., Borth, D., and Li, L.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories A., Friedland, G., Elizalde, B., Ni, K., Poland, D., Borth, D., and Li, L

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:34.951741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:34.951741Z digest=sha256:10e99946142dc41f315d49ce15283cd3776ff8bbc3e836e72f6b2fc589b0bedd

Observation 9b1c5e39-0d5b-436d-98fc-ec0b8f51f774 · outbound

This paper cites Multimodal Reasoning Agent for Zero-Shot Composed Image Retrieval.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Multimodal Reasoning Agent for Zero-Shot Composed Image Retrieval

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:35.070287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:35.070287Z digest=sha256:e1b56e3922c207f177cb772448cd1d60a1bd47dd3c8b14b28921179c8ef17dde

Observation ef29b82e-b6f1-423d-ab3d-17a186663cef · outbound

This paper cites Uniir: Training and benchmarking universal multimodal information retrievers.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Uniir: Training and benchmarking universal multimodal information retrievers

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:35.158045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:35.158045Z digest=sha256:0884cdb766da62efe543f52d2dce930baac760d4b1210a049f1a417b5ff6da2a

Observation 9fe77158-cac5-4238-a4e8-22e0d2f817b1 · outbound

This paper cites J., Cheng, Z., Shin, D., Lei, F., Liu, Y., Xu, Y., Zhou, S., Savarese, S., Xiong, C., Zhong, V., and Yu, T.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories J., Cheng, Z., Shin, D., Lei, F., Liu, Y., Xu, Y., Zhou, S., Savarese, S., Xiong, C., Zhong, V., and Yu, T

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:35.250679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:35.250679Z digest=sha256:133f86ae7868489f25c1d2ac3b59af6e5a4af30835980f217c7a38ed5703f3b8

Observation cc4dfd74-5688-4ac3-9319-be2578710aad · outbound

This paper cites A survey on agentic multimodal large language models.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories A survey on agentic multimodal large language models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:35.310474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:35.310474Z digest=sha256:bcbe0aaa312c101499cd3709aea3c8369e96538e9ab367070269bed25179742c

Observation 980d3209-a6fa-4191-80b3-f613ddd9f0b7 · outbound

This paper cites A Survey on Multimodal Large Language Models.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories A Survey on Multimodal Large Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:35.381687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:35.381687Z digest=sha256:bc229092e539e9e219931992ff5d0417b1ef774f753d0ba691b63d6c2cd90b5f

Observation 49d1665b-b035-4f00-a2a9-87125fc03ff8 · outbound

This paper cites Sigmoid loss for language image pre-training.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Sigmoid loss for language image pre-training

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:35.441433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:35.441433Z digest=sha256:e443ca0173b529049e046a010942beca2f9c541c48faa4ae947787f013ad7acc

Observation 6b10a885-a4ba-4dfb-95e3-06b3ed909129 · outbound

This paper cites Magiclens: Self-supervised image retrieval with open-ended instructions.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Magiclens: Self-supervised image retrieval with open-ended instructions

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:35.493635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:35.493635Z digest=sha256:e6f4b1bd1826084689e07d720b64384de8a308b38c5eb100bba8b380f87cb6e9

Observation e00c3eb1-e0f5-4c09-8fb1-7a4805c5d85b · outbound

This paper cites GME: Improving Universal Multimodal Retrieval by Multimodal LLMs.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories GME: Improving Universal Multimodal Retrieval by Multimodal LLMs

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:35.562855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:35.562855Z digest=sha256:996cb153cf131bcac57b5b63a73e794375dbdb25bf39f8559bc883b34f1945d0

Observation 3b411d52-5745-4bfb-862b-b6cb155c6133 · outbound

This paper cites V-MAGE: A Game Evaluation Framework for Assessing Vision-Centric Capabilities in Multimodal Large Language Models.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories V-MAGE: A Game Evaluation Framework for Assessing Vision-Centric Capabilities in Multimodal Large Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:35.623295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:35.623295Z digest=sha256:2e4e0e107bd6857f352b2bfc4aa39f5d1f9f54bbd5bb45d86e8da247b30cb1ac

Observation 8ef789e4-2270-4b6f-a5bc-0b33acecb204 · outbound

This paper cites J., and Lian, D.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories J., and Lian, D

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:35.692722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:35.692722Z digest=sha256:df6383bd52820fc1ec34913b5c9e03ca373b12cf0e3a35963d5f9d6c27901f6f

Observation 2ac36528-f433-4fc8-b53f-0c0e46b615e9 · outbound

This paper cites F., Zhu, H., Zhou, X., Lo, R., Sridhar, A., Cheng, X., Ou, T., Bisk, Y., Fried, D., Alon, U., and Neubig, G.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories F., Zhu, H., Zhou, X., Lo, R., Sridhar, A., Cheng, X., Ou, T., Bisk, Y., Fried, D., Alon, U., and Neubig, G

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:35.733752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:35.733752Z digest=sha256:2147b9302fc7db67966366652dd20b432e91ab5d343fd6c683cff21e88a8760f

Observation 5b43014d-2eb1-4f58-95f2-60bce1e64fa0 · outbound

This paper cites Large language models for information retrieval: A survey.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories Large language models for information retrieval: A survey

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:35.736324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:35.736324Z digest=sha256:d2096b354f23dc61e93b0526e23ab86bf4ba0a312effeedfd3f546eece8008d3

Observation 933a7374-a2f8-4049-b55b-4cb929bf10af · outbound

This paper cites write newline.

DeepImageSearch: Benchmarking Multimodal Agents for Context-Aware Image Retrieval in Visual Histories write newline

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:35.739213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T01:02:35.739213Z digest=sha256:d8c9ba52ce542e2fa3c43b5e2065a0598a314d00d828888da46aa4e621105421

Pith citing papers

No inbound Pith citation observations are available.