Pith. sign in

Paper Citation Record · LEDGER

Rethinking Text-Based Image Retrieval in Specific Domain

As of 22 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 0 inbound Pith citation observations for arXiv:2608.10524.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.10524 v1

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T14:23:02.416523Z

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

17 of 17 outbound references displayed

  • verified exact0
  • verified fuzzy3
  • unresolved14
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3e2f9ccd-7aa1-41ba-bb9c-0b1b6bbf2840 · outbound

This paper cites jina-embeddings-v4: Universal Embeddings for Multimodal Multilingual Retrieval.

Rethinking Text-Based Image Retrieval in Specific Domain jina-embeddings-v4: Universal Embeddings for Multimodal Multilingual Retrieval

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.359256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.359256Z digest=sha256:69614fb4e38b1ee0a42c57c6c70dcdba18609e2f06e5b1a18f642bc6cf6c4f94

Observation 3d1a5540-a3b5-4556-a3ee-d2cd1a8f0b59 · outbound

This paper cites arXiv:2510.12798.

Rethinking Text-Based Image Retrieval in Specific Domain arXiv:2510.12798

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.378328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.378328Z digest=sha256:f8f43aed2b93142409b913d4f7325634a69c0b55225023cbd7415c54a5c909d9

Observation 2a8bfb3a-4832-45a1-bde0-5d95e0a3e28f · outbound

This paper cites Qwen3-VL-Embedding and Qwen3-VL-Reranker: A Unified Framework for State-of-the-Art Multimodal Retrieval and Ranking.

Rethinking Text-Based Image Retrieval in Specific Domain Qwen3-VL-Embedding and Qwen3-VL-Reranker: A Unified Framework for State-of-the-Art Multimodal Retrieval and Ranking

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.382460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.382460Z digest=sha256:a314e9f53198d3232672e9c4b2b8569f6fa1231426605b7c2a499d38bb09de54

Observation 4fdafe87-c3a2-4063-aece-58d4500ae319 · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

Rethinking Text-Based Image Retrieval in Specific Domain DINOv2: Learning Robust Visual Features without Supervision

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.394978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.394978Z digest=sha256:2d087d41905f8445d63b2dd44031080696a5838cecdf08ad9dfe8df540eb6371

Observation 2b4c60c8-a704-4b3c-a2a7-f02aa086b66f · outbound

This paper cites DINOv3.

Rethinking Text-Based Image Retrieval in Specific Domain DINOv3

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.399110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.399110Z digest=sha256:3127644443250119bbee4c84d7b553f471506ff182af0f831563661e34b9392a

Observation 0b3638de-b2f7-44ec-88c5-5deb2570c30f · outbound

This paper cites SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features.

Rethinking Text-Based Image Retrieval in Specific Domain SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.403717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.403717Z digest=sha256:a6e1fe71778d44e6656449028113dc8d5567060435171d72c92dcd75eebe8bdd

Observation f7e661eb-b70a-42f9-85e1-b675b0a4ffa6 · outbound

This paper cites When and why vision-language models behave like bags-of-words, and what to do about it?.

Rethinking Text-Based Image Retrieval in Specific Domain When and why vision-language models behave like bags-of-words, and what to do about it?

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.412149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.412149Z digest=sha256:cb9a21b3cd58a7fd68c291b3c4ecc6fcaace38378cf5d1b6f0a9427acad94a39

Observation ec61eab2-8c51-4266-afc2-ad636b4317c9 · outbound

This paper cites PLIP: Language-Image Pre-training for Person Representation Learning.

Rethinking Text-Based Image Retrieval in Specific Domain PLIP: Language-Image Pre-training for Person Representation Learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.416523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.416523Z digest=sha256:7783e599e13dd502776837a08c3f49d3d34b24ba85f3a3b661e38a97c4734271

Observation 3887df14-79ad-4484-9829-6ef6c9ccd390 · outbound

This paper cites Loshchilov,I.;andHutter,F.2019.

Rethinking Text-Based Image Retrieval in Specific Domain Loshchilov,I.;andHutter,F.2019

Reference 755

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:23:02.843138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T14:23:02.391050Z digest=sha256:6dc837b89de528d2dcf309a36863e7d34093b8b2ef70847509dc27a75e3f7bbe

Observation 912a4f61-9f0f-450d-b93c-320f2191e70d · outbound

This paper cites InComputer Vision– ECCV 2014: 13th European Conference, Zurich, Switzer- land, September 6-12, 2014, Proceedings, Part V 13, 740–.

Rethinking Text-Based Image Retrieval in Specific Domain InComputer Vision– ECCV 2014: 13th European Conference, Zurich, Switzer- land, September 6-12, 2014, Proceedings, Part V 13, 740–

Reference 2014

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:23:02.856676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T14:23:02.386965Z digest=sha256:38efa1715d993316f085d59e0b870574a130142012f4819f2c2fc0cd114f9c53

Observation 0661cb2a-c744-424b-9b69-11ac60e9a177 · outbound

This paper cites Automatic Spatially-aware Fashion Concept Discovery.

Rethinking Text-Based Image Retrieval in Specific Domain Automatic Spatially-aware Fashion Concept Discovery

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.363970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.363970Z digest=sha256:d42072ad02266152e72269f92bedfbf3b206e9f715b308448a3c6bafc1f1cba4

Observation 8fe333e9-05dd-4a34-9b72-628ffeaa3e19 · outbound

This paper cites InProceedings of the International Conference on Machine Learning (ICML), 4904–4916.

Rethinking Text-Based Image Retrieval in Specific Domain InProceedings of the International Conference on Machine Learning (ICML), 4904–4916

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.374137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.374137Z digest=sha256:a72838932dc3780cbca9ffb5dbb01d30472df388a6a0a3e6ee6257a4070cad82

Observation 5d455686-0178-496b-a77c-152bfa02d758 · outbound

This paper cites InProceedings of the 2022 Conference on Empirical Methods in Natural Language Processing, 3876–3887.

Rethinking Text-Based Image Retrieval in Specific Domain InProceedings of the 2022 Conference on Empirical Methods in Natural Language Processing, 3876–3887

Reference 2022

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T14:23:02.829637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T14:23:02.408018Z digest=sha256:a83dbfb8e6d32095e1b99b898dfa07233ab5a799de82a4fcd748a41c0dd6dd37

Observation 9f2b1b85-6a13-40a9-9293-72d1a9a4f6d1 · outbound

This paper cites Semantically Self-Aligned Network for Text-to-Image Part-aware Person Re-identification.

Rethinking Text-Based Image Retrieval in Specific Domain Semantically Self-Aligned Network for Text-to-Image Part-aware Person Re-identification

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.350551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.350551Z digest=sha256:c24f792e52933244c5336d4cbe8f8eadf9da13f4fd67b317640226d0fc8d51bb

Observation b0064054-8f47-4ae7-8996-b2b2f8ce3188 · outbound

This paper cites InProceedings of the AAAI ConferenceonArtificialIntelligence,volume38,1860–1868.

Rethinking Text-Based Image Retrieval in Specific Domain InProceedings of the AAAI ConferenceonArtificialIntelligence,volume38,1860–1868

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.355262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.355262Z digest=sha256:c978c267fd9950fcc7768190fbb5b0fb1a0dd947290674a66a8b4eb029b71535

Observation d3e4e7d7-997b-4ada-b890-2a30680b3923 · outbound

This paper cites Qwen3-VL Technical Report.

Rethinking Text-Based Image Retrieval in Specific Domain Qwen3-VL Technical Report

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.345821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.345821Z digest=sha256:44cb857ccb4592abc729f3e270ae1eeae79d1c3cf5c73c2c4911e1470b1bb217

Observation 67d2c312-7f7a-459f-bd96-0aae8c262a64 · outbound

This paper cites Jia, C.;Yang, Y.; Xia,Y.; Chen, Y.-T.;Parekh, Z.; Pham,H.; Le, Q.; Sung, Y.-H.; Li, Z.; and Duerig, T.

Rethinking Text-Based Image Retrieval in Specific Domain Jia, C.;Yang, Y.; Xia,Y.; Chen, Y.-T.;Parekh, Z.; Pham,H.; Le, Q.; Sung, Y.-H.; Li, Z.; and Duerig, T

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-15T14:23:02.369622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:23:02.369622Z digest=sha256:32545b3920a943e7bf05258fce4004f68e1efef7c9a40fee2b907302fa03984a

Pith citing papers

No inbound Pith citation observations are available.