Pith. sign in

Paper Citation Record · LEDGER

GLIPv2: Unifying Localization and Vision-Language Understanding

As of 13 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2206.05836.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2206.05836 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T14:03:28.738271Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-28T22:52:45.158296Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 58237ede-0d1a-4d75-b01b-bb0c4bd791a9 · inbound

Grounded SAM: Assembling Open-World Models for Diverse Visual Tasks cites this paper.

Grounded SAM: Assembling Open-World Models for Diverse Visual Tasks GLIPv2: Unifying Localization and Vision-Language Understanding

Reference 81

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:20:16.184999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-11T06:20:15.656356Z digest=sha256:4f5ff6d357b3a28b138a6551383e51cd4ffa8fbe1599cb138c790b7aad0566e0

Observation cefe5213-571d-4075-85e9-fa4563dec278 · inbound

AnySynth: Harnessing the Power of Image Synthetic Data Generation for Generalized Vision-Language Tasks cites this paper.

AnySynth: Harnessing the Power of Image Synthetic Data Generation for Generalized Vision-Language Tasks GLIPv2: Unifying Localization and Vision-Language Understanding

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-12T14:03:28.738271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:03:28.738271Z digest=sha256:42a5ffe4c0e37494c014e9042d19d3da1f4071e4e08533dcca537d01ab57d4b5

Observation 4258e75c-6c85-4a3b-8cd9-29474c05c172 · inbound

Textual Inversion for Efficient Adaptation of Open-Vocabulary Object Detectors Without Forgetting cites this paper.

Textual Inversion for Efficient Adaptation of Open-Vocabulary Object Detectors Without Forgetting GLIPv2: Unifying Localization and Vision-Language Understanding

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T23:31:48.441669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:31:48.441669Z digest=sha256:3d0b5935a14565c565a47256d63ab83b98801b5014fa68d9177c7c6e9233d521

Observation 7e36ec06-2120-4b26-a830-aa9c57c85ec8 · inbound

LaV-CoT: Language-Aware Visual CoT with Multi-Aspect Reward Optimization for Real-World Multilingual VQA cites this paper.

LaV-CoT: Language-Aware Visual CoT with Multi-Aspect Reward Optimization for Real-World Multilingual VQA GLIPv2: Unifying Localization and Vision-Language Understanding

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-18T17:56:42.221129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-18T17:53:08.136677Z digest=sha256:c12c64a42e48b990ad1c93639cf38dd95b1e144d58bc655457dd414fd0a8af39

Observation 8ab7e9fb-dee8-4409-99fa-0b1b5541f649 · inbound

Vision Harnessing Agent for Open Ad-hoc Segmentation cites this paper.

Vision Harnessing Agent for Open Ad-hoc Segmentation GLIPv2: Unifying Localization and Vision-Language Understanding

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-20T05:53:04.454464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-20T05:52:40.429412Z digest=sha256:a7557f57f9254e8c4d76762001fc73aad5220d026cfd97e6fc2307d3626c2dc7

Observation 73fb1ece-de34-48bc-9792-0ddfd17d7a01 · inbound

FOCUS: Forcing In-Context Object Localization through Visual Support Constraints and Policy Optimization cites this paper.

FOCUS: Forcing In-Context Object Localization through Visual Support Constraints and Policy Optimization GLIPv2: Unifying Localization and Vision-Language Understanding

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T22:52:45.159900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-28T22:50:21.596865Z digest=sha256:2fae6ae7defd6ffcae42559905b0cef87d8bd0beb89a7c26ff580cca67a37a2d

Observation 93de2272-4edd-4896-bdb1-354bcfb0d231 · inbound

Hi-TOPS: Hierarchical Topology-aware Scoring Prior for 3D Part Decomposition cites this paper.

Hi-TOPS: Hierarchical Topology-aware Scoring Prior for 3D Part Decomposition GLIPv2: Unifying Localization and Vision-Language Understanding

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T00:25:54.332570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T00:25:54.332570Z digest=sha256:7ecff8ba72fc20580dd613031b8f495fa2b2237d5da871cad99fbcb5f6408618