Pith. sign in

Paper Citation Record · LEDGER

Vision-by-Language for Training-Free Compositional Image Retrieval

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2310.09291.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2310.09291 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T05:54:15.956672Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T16:27:09.023654Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation d2cafc08-83e6-44b9-a357-09c7a8aa6376 · inbound

E5-V: Universal Embeddings with Multimodal Large Language Models cites this paper.

E5-V: Universal Embeddings with Multimodal Large Language Models Vision-by-Language for Training-Free Compositional Image Retrieval

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-16T22:52:20.974516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T22:52:20.935555Z digest=sha256:4cdae5170ae29b8cce54abb9ab4ea2d6281f3eeee6c935a7327965e1f2972806

Observation a2540e4b-15b0-4ffc-b70e-47acf8f5fe44 · inbound

UniCoRN: Unified Commented Retrieval Network with LMMs cites this paper.

UniCoRN: Unified Commented Retrieval Network with LMMs Vision-by-Language for Training-Free Compositional Image Retrieval

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-08T05:54:15.956672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T05:54:15.956672Z digest=sha256:83ea409a7ff1004cbeec112624dcc477afd3504f32e8fbd091add902d70960ed

Observation 703b7aba-149f-44df-9c13-06255edc21b9 · inbound

DetailFusion: A Dual-branch Framework with Detail Enhancement for Composed Image Retrieval cites this paper.

DetailFusion: A Dual-branch Framework with Detail Enhancement for Composed Image Retrieval Vision-by-Language for Training-Free Compositional Image Retrieval

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T14:43:59.484503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:43:59.484503Z digest=sha256:b31683f1ea25bb4900be79660149022d1035db8695856c89daaa4ee26c7e6bcc

Observation 085e0c1d-7a55-444c-b67d-40552bca0c92 · inbound

FACap: A Large-scale Fashion Dataset for Fine-grained Composed Image Retrieval cites this paper.

FACap: A Large-scale Fashion Dataset for Fine-grained Composed Image Retrieval Vision-by-Language for Training-Free Compositional Image Retrieval

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T19:10:00.073729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:10:00.073729Z digest=sha256:adc8aa7d44816cf2e1b124b5b2a05452821974a74e398f5d9cadffb926d0f8d1

Observation b8ea230b-1b02-4063-89f3-427c1cb390ea · inbound

MCoT-RE: Multi-Faceted Chain-of-Thought and Re-Ranking for Training-Free Zero-Shot Composed Image Retrieval cites this paper.

MCoT-RE: Multi-Faceted Chain-of-Thought and Re-Ranking for Training-Free Zero-Shot Composed Image Retrieval Vision-by-Language for Training-Free Compositional Image Retrieval

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T16:42:40.156102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:42:40.156102Z digest=sha256:371065dad1c581be16289ce0b5a0074d83c898cf975ee76af28de1bc8a74d259

Observation cda78990-b7ed-4ae7-9ea9-05acbf4101ff · inbound

Beyond Simple Edits: Composed Video Retrieval with Dense Modifications cites this paper.

Beyond Simple Edits: Composed Video Retrieval with Dense Modifications Vision-by-Language for Training-Free Compositional Image Retrieval

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T18:50:58.144723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:50:58.144723Z digest=sha256:7bc7ae74b36986f0e3c99e455b88135ef61d80aab092c19ccb499b01bcf0afda

Observation d93c6bb1-380b-46be-a1ee-b864874ec20d · inbound

A Sanity Check on Composed Image Retrieval cites this paper.

A Sanity Check on Composed Image Retrieval Vision-by-Language for Training-Free Compositional Image Retrieval

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:36:05.594025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T15:24:27.497284Z digest=sha256:bc3cc2230d7223b4c1e584ee46d28edb6bc079e23863ca76bd6f8fb1b97e46cd

Observation 943d49a4-0d22-4e0a-a289-188b5de38372 · inbound

G-MIXER: Geodesic Mixup-based Implicit Semantic Expansion and Explicit Semantic Re-ranking for Zero-Shot Composed Image Retrieval cites this paper.

G-MIXER: Geodesic Mixup-based Implicit Semantic Expansion and Explicit Semantic Re-ranking for Zero-Shot Composed Image Retrieval Vision-by-Language for Training-Free Compositional Image Retrieval

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:50:19.952322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T11:50:00.850857Z digest=sha256:22080e0645fde186a14406a65be9d6d555687b68e8f2db6d65c87bd28a468742

Observation d8923319-ac51-4b3d-b01c-9f7b0179ea20 · inbound

STiTch: Semantic Transition and Transportation in Collaboration for Training-Free Zero-Shot Composed Image Retrieval cites this paper.

STiTch: Semantic Transition and Transportation in Collaboration for Training-Free Zero-Shot Composed Image Retrieval Vision-by-Language for Training-Free Compositional Image Retrieval

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:33:58.857727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T05:30:08.145613Z digest=sha256:f85eaaf6d86b55e0534e57c8239e0b7bdf3f10e4a869aa4a63a4d2166dd0d895

Observation cac05558-6818-457b-b095-d2c2dfe52d60 · inbound

Never Seen Before: Benchmarking Genuine Zero-Shot Composed Image Retrieval with Consistent Video-Sourced Datasets cites this paper.

Never Seen Before: Benchmarking Genuine Zero-Shot Composed Image Retrieval with Consistent Video-Sourced Datasets Vision-by-Language for Training-Free Compositional Image Retrieval

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-02T16:27:09.025375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T22:44:36.059582Z digest=sha256:208570d2114c3d9fbc160dab25a002b3cd5e3f5dc06ba00c63055b210e6e536a

Observation bbc903ec-637b-47e2-b7f3-8f0cebfa60bc · inbound

CR-Refiner: An Object-Centric Optimal Transport Reranker for Edit-Conditioned 3D Scene Retrieval cites this paper.

CR-Refiner: An Object-Centric Optimal Transport Reranker for Edit-Conditioned 3D Scene Retrieval Vision-by-Language for Training-Free Compositional Image Retrieval

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T13:28:40.030166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:28:40.030166Z digest=sha256:8f3d0e7f4e97b58694a7a859334d46b2b7b0cd703482a27b24195e11d013607e