Pith. sign in

Paper Citation Record · LEDGER

Rethinking Text-Based Image Retrieval in Specific Domain

As of 22 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 0 inbound Pith citation observations for arXiv:2608.10524.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.10524 v1

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T14:23:02.416523Z

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

17 of 17 outbound references displayed

  • verified exact0
  • verified fuzzy3
  • unresolved14
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3e2f9ccd-7aa1-41ba-bb9c-0b1b6bbf2840 · outbound

This paper cites jina-embeddings-v4: Universal Embeddings for Multimodal Multilingual Retrieval.

Rethinking Text-Based Image Retrieval in Specific Domain jina-embeddings-v4: Universal Embeddings for Multimodal Multilingual Retrieval

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.359256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.359256Z digest=sha256:eefc5be9667ee2312ea7e169736bb5052b2fb4208db7b451d0476c3a83d0c6cf

Observation 3d1a5540-a3b5-4556-a3ee-d2cd1a8f0b59 · outbound

This paper cites arXiv:2510.12798.

Rethinking Text-Based Image Retrieval in Specific Domain arXiv:2510.12798

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.378328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.378328Z digest=sha256:62e94241391465966488971fd92fc3d4a010ef9f027375a2cb4f3227a9d64171

Observation 2a8bfb3a-4832-45a1-bde0-5d95e0a3e28f · outbound

This paper cites Qwen3-VL-Embedding and Qwen3-VL-Reranker: A Unified Framework for State-of-the-Art Multimodal Retrieval and Ranking.

Rethinking Text-Based Image Retrieval in Specific Domain Qwen3-VL-Embedding and Qwen3-VL-Reranker: A Unified Framework for State-of-the-Art Multimodal Retrieval and Ranking

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.382460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.382460Z digest=sha256:73929d44b3b2dfc51f02292eafa00897068aa5b0ebfb0e65a280ffa083e53121

Observation 4fdafe87-c3a2-4063-aece-58d4500ae319 · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

Rethinking Text-Based Image Retrieval in Specific Domain DINOv2: Learning Robust Visual Features without Supervision

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.394978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.394978Z digest=sha256:3fc93ab93b0bdd97768b3af7e6afa2730a0d98e29538be46c9cbe49d67f4cc28

Observation 2b4c60c8-a704-4b3c-a2a7-f02aa086b66f · outbound

This paper cites DINOv3.

Rethinking Text-Based Image Retrieval in Specific Domain DINOv3

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.399110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.399110Z digest=sha256:f4058996ca6277aee9ab2bd1af9dbfe987ade0ac43c26bd6d46cb2ccb7bbb986

Observation 0b3638de-b2f7-44ec-88c5-5deb2570c30f · outbound

This paper cites SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features.

Rethinking Text-Based Image Retrieval in Specific Domain SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.403717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.403717Z digest=sha256:6db5438fd758a3540a6653c5caebea42924996ceeaaa158f5d310e4adcecacc7

Observation f7e661eb-b70a-42f9-85e1-b675b0a4ffa6 · outbound

This paper cites When and why vision-language models behave like bags-of-words, and what to do about it?.

Rethinking Text-Based Image Retrieval in Specific Domain When and why vision-language models behave like bags-of-words, and what to do about it?

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.412149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.412149Z digest=sha256:e63073ff6427b83a1ed96f1cc7574001d6d6f1e6229f1f8f8cf684f12ec9f711

Observation ec61eab2-8c51-4266-afc2-ad636b4317c9 · outbound

This paper cites PLIP: Language-Image Pre-training for Person Representation Learning.

Rethinking Text-Based Image Retrieval in Specific Domain PLIP: Language-Image Pre-training for Person Representation Learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.416523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.416523Z digest=sha256:20de712af5a83ab55bc868f8a37d3eabb114408c1c2b3120e364c0c267347f7b

Observation 3887df14-79ad-4484-9829-6ef6c9ccd390 · outbound

This paper cites Loshchilov,I.;andHutter,F.2019.

Rethinking Text-Based Image Retrieval in Specific Domain Loshchilov,I.;andHutter,F.2019

Reference 755

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:23:02.843138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T14:23:02.391050Z digest=sha256:25c2949632e832613f8fa56e7561bf5d7b4dec70adc9efbaf0b8fc7a5ad10cd1

Observation 912a4f61-9f0f-450d-b93c-320f2191e70d · outbound

This paper cites InComputer Vision– ECCV 2014: 13th European Conference, Zurich, Switzer- land, September 6-12, 2014, Proceedings, Part V 13, 740–.

Rethinking Text-Based Image Retrieval in Specific Domain InComputer Vision– ECCV 2014: 13th European Conference, Zurich, Switzer- land, September 6-12, 2014, Proceedings, Part V 13, 740–

Reference 2014

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:23:02.856676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T14:23:02.386965Z digest=sha256:89f311f077b67d65b91226065264eb771bf6c0a81fc6ebdec4fec44c8e5e2d76

Observation 0661cb2a-c744-424b-9b69-11ac60e9a177 · outbound

This paper cites Automatic Spatially-aware Fashion Concept Discovery.

Rethinking Text-Based Image Retrieval in Specific Domain Automatic Spatially-aware Fashion Concept Discovery

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.363970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.363970Z digest=sha256:4438824778761bf7f52e66f225d28f590f4af3936ac30efbe24e125755728caa

Observation 8fe333e9-05dd-4a34-9b72-628ffeaa3e19 · outbound

This paper cites InProceedings of the International Conference on Machine Learning (ICML), 4904–4916.

Rethinking Text-Based Image Retrieval in Specific Domain InProceedings of the International Conference on Machine Learning (ICML), 4904–4916

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.374137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.374137Z digest=sha256:5edaa94b99d89a73f9e3b3a278c4b620df9820ec8ab6d264bcac772912acfd6a

Observation 5d455686-0178-496b-a77c-152bfa02d758 · outbound

This paper cites InProceedings of the 2022 Conference on Empirical Methods in Natural Language Processing, 3876–3887.

Rethinking Text-Based Image Retrieval in Specific Domain InProceedings of the 2022 Conference on Empirical Methods in Natural Language Processing, 3876–3887

Reference 2022

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:23:02.829637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T14:23:02.408018Z digest=sha256:b073b979bc03541d39a14ad8630d70d63ddf1db1b15d3cdd5478ed4696924e8a

Observation 9f2b1b85-6a13-40a9-9293-72d1a9a4f6d1 · outbound

This paper cites Semantically Self-Aligned Network for Text-to-Image Part-aware Person Re-identification.

Rethinking Text-Based Image Retrieval in Specific Domain Semantically Self-Aligned Network for Text-to-Image Part-aware Person Re-identification

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.350551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.350551Z digest=sha256:2017756da3665ae83e027f1e65cd07751a6f9b46d560e24d85b36b38060aea16

Observation b0064054-8f47-4ae7-8996-b2b2f8ce3188 · outbound

This paper cites InProceedings of the AAAI ConferenceonArtificialIntelligence,volume38,1860–1868.

Rethinking Text-Based Image Retrieval in Specific Domain InProceedings of the AAAI ConferenceonArtificialIntelligence,volume38,1860–1868

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.355262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.355262Z digest=sha256:898e5760aa2c4e191fbe4fe0f5289b42b4cef029533d8ee597d5f43860178845

Observation d3e4e7d7-997b-4ada-b890-2a30680b3923 · outbound

This paper cites Qwen3-VL Technical Report.

Rethinking Text-Based Image Retrieval in Specific Domain Qwen3-VL Technical Report

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.345821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.345821Z digest=sha256:95dc5aca4906eb00bcb0da9325ade089bc0705167cdc8ab6158a8269f0f02d99

Observation 67d2c312-7f7a-459f-bd96-0aae8c262a64 · outbound

This paper cites Jia, C.;Yang, Y.; Xia,Y.; Chen, Y.-T.;Parekh, Z.; Pham,H.; Le, Q.; Sung, Y.-H.; Li, Z.; and Duerig, T.

Rethinking Text-Based Image Retrieval in Specific Domain Jia, C.;Yang, Y.; Xia,Y.; Chen, Y.-T.;Parekh, Z.; Pham,H.; Le, Q.; Sung, Y.-H.; Li, Z.; and Duerig, T

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.369622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.369622Z digest=sha256:d7a433ed5803d46e39452128a39a9dae1208d801427eda74068ac7b560d85952

Pith citing papers

No inbound Pith citation observations are available.