Pith. sign in

Paper Citation Record · LEDGER

F-VLM: Open-Vocabulary Object Detection upon Frozen Vision and Language Models

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 15 inbound Pith citation observations for arXiv:2209.15639.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2209.15639 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 15 of 15 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T20:48:50.796941Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T00:57:30.385197Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 8d60dfa5-91b3-41c9-b0c4-b2c62c7767e0 · inbound

ForgeryGPT: A Multimodal LLM for Interpretable Image Forgery Detection and Localization cites this paper.

ForgeryGPT: A Multimodal LLM for Interpretable Image Forgery Detection and Localization F-VLM: Open-Vocabulary Object Detection upon Frozen Vision and Language Models

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-23T19:08:21.153909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-23T19:06:10.198724Z digest=sha256:afa926efd28b3291937187e91fee53f9da44fffdf1af29709744cbaddfc266f3

Observation 1b2a88f8-3331-4f65-b549-d1edd7db2922 · inbound

AutoOcc: Automatic Open-Ended Semantic Occupancy Annotation via Vision-Language Guided Gaussian Splatting cites this paper.

AutoOcc: Automatic Open-Ended Semantic Occupancy Annotation via Vision-Language Guided Gaussian Splatting F-VLM: Open-Vocabulary Object Detection upon Frozen Vision and Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-08T20:48:50.796941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T20:48:50.796941Z digest=sha256:47db5dd8b4138e572174aacdca532cdedfbc1e3bc594958d5b3877b13416ad97

Observation 0ca67499-01c0-4d60-bf2d-62371bd9eb94 · inbound

Event-Priori-Based Vision-Language Model for Efficient Visual Understanding cites this paper.

Event-Priori-Based Vision-Language Model for Efficient Visual Understanding F-VLM: Open-Vocabulary Object Detection upon Frozen Vision and Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T05:35:01.342025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:35:01.342025Z digest=sha256:3b2e222aa7c4f44da07fe275bfda8c24308637f6904c3b27d1ac2140328bb3dd

Observation 11c6d7db-7c34-4bcf-985a-a185bbf26bb1 · inbound

MoMa: Modulating Mamba for Adapting Image Foundation Models to Video Recognition cites this paper.

MoMa: Modulating Mamba for Adapting Image Foundation Models to Video Recognition F-VLM: Open-Vocabulary Object Detection upon Frozen Vision and Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T21:51:53.789728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:51:53.789728Z digest=sha256:a4f4ad40eeefd60afdfa6312279339cd06c6df4ba13ecde5e6252047f29e177b

Observation d737176b-afa7-4a34-9af7-dd51e16057d7 · inbound

Visual Textualization for Image Prompted Object Detection cites this paper.

Visual Textualization for Image Prompted Object Detection F-VLM: Open-Vocabulary Object Detection upon Frozen Vision and Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T21:37:05.649074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:37:05.649074Z digest=sha256:6cca76b70984ad668947562e2e4ebc34115aef451e2ef38abaea6beab65bed8a

Observation 8502811c-3294-447e-bbca-22dea30b1f8f · inbound

Free on the Fly: Enhancing Flexibility in Test-Time Adaptation with Online EM cites this paper.

Free on the Fly: Enhancing Flexibility in Test-Time Adaptation with Online EM F-VLM: Open-Vocabulary Object Detection upon Frozen Vision and Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T18:56:34.997383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:56:34.997383Z digest=sha256:f88af9c47f16580e3179d51c641b06bfe574cbf049383e9b0f50d2c23913ab03

Observation e53dd32a-761f-4aab-8c2e-f630c24a0e52 · inbound

RoboChemist: Long-Horizon and Safety-Compliant Robotic Chemical Experimentation cites this paper.

RoboChemist: Long-Horizon and Safety-Compliant Robotic Chemical Experimentation F-VLM: Open-Vocabulary Object Detection upon Frozen Vision and Language Models

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-04T20:10:13.751010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:10:13.751010Z digest=sha256:839db08a3f55217cd89185a8e6ff6868b63cbf37eac51f078b10ddb019b0d261

Observation 7e2b9157-c2d9-4404-8f4c-878f438e33d2 · inbound

Vision Transformers Need More Than Registers cites this paper.

Vision Transformers Need More Than Registers F-VLM: Open-Vocabulary Object Detection upon Frozen Vision and Language Models

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-15T19:10:15.693447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T19:08:50.579190Z digest=sha256:5a871a9238095f2f7f116fbb3aba0f8416dc2e6d9485488e15277a39e64ebbc1

Observation 8a814cf7-50f5-4028-9b0e-aff1988a1a85 · inbound

DetPO: In-Context Learning with Multi-Modal LLMs for Few-Shot Object Detection cites this paper.

DetPO: In-Context Learning with Multi-Modal LLMs for Few-Shot Object Detection F-VLM: Open-Vocabulary Object Detection upon Frozen Vision and Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-13T19:38:38.452093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T19:38:38.452093Z digest=sha256:75fae6d300959d03b07107f95af4de564104ee69ee7854142898bbc31ebe33ff

Observation 5250254f-74f3-4963-822f-782bcacb2188 · inbound

CLIP-RD: Relative Distillation for Efficient CLIP Knowledge Distillation cites this paper.

CLIP-RD: Relative Distillation for Efficient CLIP Knowledge Distillation F-VLM: Open-Vocabulary Object Detection upon Frozen Vision and Language Models

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T00:23:22.867231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T00:20:36.975106Z digest=sha256:9d65beda5ee0c1de06a4ea9da4857e27c3adbf96e52c9a5ab60a13ea00ce4ffc

Observation 82cf2875-9aa2-4a1b-ab96-f9de1362501c · inbound

COVD: Continual Open-Vocabulary Object Detection with Novel Concept Injection cites this paper.

COVD: Continual Open-Vocabulary Object Detection with Novel Concept Injection F-VLM: Open-Vocabulary Object Detection upon Frozen Vision and Language Models

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:13:48.896059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T18:07:07.563662Z digest=sha256:435b2e37c25c4ee9bd96ccd15b8fb9d82eddc49f1f049cedd2889922acd2043b

Observation b7aeed7d-d168-43ee-adea-6bbe2b876d95 · inbound

LV-OSD: Language-Vision-Complementary Open-Set Object Detection cites this paper.

LV-OSD: Language-Vision-Complementary Open-Set Object Detection F-VLM: Open-Vocabulary Object Detection upon Frozen Vision and Language Models

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:53:28.953364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T13:44:59.042817Z digest=sha256:4550dd57110311859b0a929650493fb75953d4ae02505d9b281ee5c5c18af819

Observation 9024ba1f-423c-43e9-b086-b71fa430e75e · inbound

Unveiling the Unknown: Open Vocabulary Object Detection with Scene Graphs cites this paper.

Unveiling the Unknown: Open Vocabulary Object Detection with Scene Graphs F-VLM: Open-Vocabulary Object Detection upon Frozen Vision and Language Models

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:46:56.506522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T01:55:35.172193Z digest=sha256:fda55eeb4fd79ec2366cbb766022590367462a3f5301aa138ac9cd74f990f147

Observation 6104abcd-75d2-436b-9cad-c2f0d142fb0c · inbound

CL-CLIP: CLIP-Based Continual Learning Framework with Cost-Volume Category Decoupling for Object Detection cites this paper.

CL-CLIP: CLIP-Based Continual Learning Framework with Cost-Volume Category Decoupling for Object Detection F-VLM: Open-Vocabulary Object Detection upon Frozen Vision and Language Models

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:07:12.573132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T22:13:34.998844Z digest=sha256:ec98722433a9f621babd6d6108e4bdd18046b954f36a95f0f6a2cd0518a67942

Observation 0497d85d-65e7-40e6-9c95-cb26c5bbe59e · inbound

ExDet: Open-Domain Open-Vocabulary Detection with Cross-modal Extrapolation and Rectification cites this paper.

ExDet: Open-Domain Open-Vocabulary Detection with Cross-modal Extrapolation and Rectification F-VLM: Open-Vocabulary Object Detection upon Frozen Vision and Language Models

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:57:30.386919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T16:53:37.521755Z digest=sha256:237fb72b1c765af5914dc873432ebffb7d04563df58f8ac9f22da8e35e4e1bf4