Pith. sign in

Paper Citation Record · LEDGER

OV-DINO: Unified Open-Vocabulary Detection with Language-Aware Selective Fusion

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2407.07844.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.07844 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:25:43.919389Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-12T11:01:30.278582Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 46e99aa6-694f-4db8-ab1f-aac1b4d565c3 · inbound

VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion cites this paper.

VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion OV-DINO: Unified Open-Vocabulary Detection with Language-Aware Selective Fusion

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:43.919389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:43.919389Z digest=sha256:48362cfb928270c5a6b176ab68fd30e2eebdeea8ece5d3da2408d890ac48b5cf

Observation 98f4faab-7e40-497d-b19a-127eb7ff4cb4 · inbound

Advancing Visual Large Language Model for Multi-granular Versatile Perception cites this paper.

Advancing Visual Large Language Model for Multi-granular Versatile Perception OV-DINO: Unified Open-Vocabulary Detection with Language-Aware Selective Fusion

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-06T15:25:03.131454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:25:03.131454Z digest=sha256:5957c703980d4ac8f7dda5c954fe15c17d5d4b0f0b10e0fe86dea368ea8ce257

Observation 11e6a2d6-7d3d-4387-8737-21f65eda7f78 · inbound

Dynamic-DINO: Fine-Grained Mixture of Experts Tuning for Real-time Open-Vocabulary Object Detection cites this paper.

Dynamic-DINO: Fine-Grained Mixture of Experts Tuning for Real-time Open-Vocabulary Object Detection OV-DINO: Unified Open-Vocabulary Detection with Language-Aware Selective Fusion

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T14:54:36.715356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:54:36.715356Z digest=sha256:34617b32dd1073361249434d22a19057b65e957a237e77a5f930511a00abd496

Observation 9742e30f-4e89-4261-add6-acde697eeb25 · inbound

ROVI: A VLM-LLM Re-Captioned Dataset for Open-Vocabulary Instance-Grounded Text-to-Image Generation cites this paper.

ROVI: A VLM-LLM Re-Captioned Dataset for Open-Vocabulary Instance-Grounded Text-to-Image Generation OV-DINO: Unified Open-Vocabulary Detection with Language-Aware Selective Fusion

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T05:59:15.615039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:59:15.615039Z digest=sha256:72d5f7d839caebb9449916a3a77e9f43dea12af05c792f3d5f72d26fb1507868

Observation b7bfff33-ee27-4915-bbca-c45c3a5ea562 · inbound

FG-CLIP 2: A Bilingual Fine-grained Vision-Language Alignment Model cites this paper.

FG-CLIP 2: A Bilingual Fine-grained Vision-Language Alignment Model OV-DINO: Unified Open-Vocabulary Detection with Language-Aware Selective Fusion

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T10:18:44.133419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T10:18:44.133419Z digest=sha256:61f17461c38a310c127aa01e1ec1533b42b17182d936a4c80ab0f774e3b83c2b

Observation 9ffa1cd3-1c7a-4f2a-8b55-aad5ad78509a · inbound

Bridging Time and Space: Decoupled Spatio-Temporal Alignment for Video Grounding cites this paper.

Bridging Time and Space: Decoupled Spatio-Temporal Alignment for Video Grounding OV-DINO: Unified Open-Vocabulary Detection with Language-Aware Selective Fusion

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:15:52.714098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T18:38:16.204012Z digest=sha256:adb24eb3d0e6ffc6d4f43d4a280cf4c1b156a1c98c9a843d1d513d3bf07bc8dd

Observation 55c0b928-ff47-4129-897b-666db6c1e72d · inbound

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection cites this paper.

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection OV-DINO: Unified Open-Vocabulary Detection with Language-Aware Selective Fusion

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-12T11:01:30.280840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T01:24:56.236650Z digest=sha256:1a92ade18dbf6971cf19dfa0e36f454d4292637beefb06fdabb8fb6ff2b8a459

Observation aaec8475-5878-4758-b6aa-33f753f3d931 · inbound

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection cites this paper.

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection OV-DINO: Unified Open-Vocabulary Detection with Language-Aware Selective Fusion

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:00:35.786329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T19:14:12.092478Z digest=sha256:5e978ae539566c685805a4f6839c4d6406fbaa8b456b583cc9b8108eeecb203f

Observation 5a6c34d4-c1f2-460d-82a2-a42632974133 · inbound

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection cites this paper.

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection OV-DINO: Unified Open-Vocabulary Detection with Language-Aware Selective Fusion

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:51:48.406458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T01:33:55.317107Z digest=sha256:22b87dfae365b729e51f78a3ccb757df7f05e20be82a700e75fd8227c239204b

Observation a0643a01-b863-4231-b322-c3311c8b21cd · inbound

Open-Vocabulary Gaze Object Prediction: Benchmark and Method cites this paper.

Open-Vocabulary Gaze Object Prediction: Benchmark and Method OV-DINO: Unified Open-Vocabulary Detection with Language-Aware Selective Fusion

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-01T14:18:38.052966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:18:38.052966Z digest=sha256:70c9318616a2e0b264c9f76d4ac8375d44b987562ce8618220f8a0c78fec4fca

Observation 043f68e6-4dbd-4137-a18f-8f2421d77bb4 · inbound

Teaching MLLMs to Say No: Generalized Referring Expression Comprehension via Refusal Calibrated GRPO cites this paper.

Teaching MLLMs to Say No: Generalized Referring Expression Comprehension via Refusal Calibrated GRPO OV-DINO: Unified Open-Vocabulary Detection with Language-Aware Selective Fusion

Reference 101

Resolution
unresolved
no resolver link, observed 2026-08-06T18:44:51.279963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:44:51.279963Z digest=sha256:df2bd555809e68fc0940ee21cca14214509572cb5568e5b83b9343efd1a71c27