Pith. sign in

Paper Citation Record · LEDGER

OV-DINO: Unified Open-Vocabulary Detection with Language-Aware Selective Fusion

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2407.07844.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.07844 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:25:43.919389Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-12T11:01:30.278582Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 46e99aa6-694f-4db8-ab1f-aac1b4d565c3 · inbound

VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion cites this paper.

VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion OV-DINO: Unified Open-Vocabulary Detection with Language-Aware Selective Fusion

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:43.919389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:43.919389Z digest=sha256:2a321213bc41010db2d7185cdd96825d63bb5c60dc8ca2a768df022a07050101

Observation 98f4faab-7e40-497d-b19a-127eb7ff4cb4 · inbound

Advancing Visual Large Language Model for Multi-granular Versatile Perception cites this paper.

Advancing Visual Large Language Model for Multi-granular Versatile Perception OV-DINO: Unified Open-Vocabulary Detection with Language-Aware Selective Fusion

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-06T15:25:03.131454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:25:03.131454Z digest=sha256:5957c703980d4ac8f7dda5c954fe15c17d5d4b0f0b10e0fe86dea368ea8ce257

Observation 11e6a2d6-7d3d-4387-8737-21f65eda7f78 · inbound

Dynamic-DINO: Fine-Grained Mixture of Experts Tuning for Real-time Open-Vocabulary Object Detection cites this paper.

Dynamic-DINO: Fine-Grained Mixture of Experts Tuning for Real-time Open-Vocabulary Object Detection OV-DINO: Unified Open-Vocabulary Detection with Language-Aware Selective Fusion

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T14:54:36.715356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:54:36.715356Z digest=sha256:c3728a0ef78c412498adf060af59aef835136b67a95e42b031e632746f1023bb

Observation 9742e30f-4e89-4261-add6-acde697eeb25 · inbound

ROVI: A VLM-LLM Re-Captioned Dataset for Open-Vocabulary Instance-Grounded Text-to-Image Generation cites this paper.

ROVI: A VLM-LLM Re-Captioned Dataset for Open-Vocabulary Instance-Grounded Text-to-Image Generation OV-DINO: Unified Open-Vocabulary Detection with Language-Aware Selective Fusion

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T05:59:15.615039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:59:15.615039Z digest=sha256:72d5f7d839caebb9449916a3a77e9f43dea12af05c792f3d5f72d26fb1507868

Observation b7bfff33-ee27-4915-bbca-c45c3a5ea562 · inbound

FG-CLIP 2: A Bilingual Fine-grained Vision-Language Alignment Model cites this paper.

FG-CLIP 2: A Bilingual Fine-grained Vision-Language Alignment Model OV-DINO: Unified Open-Vocabulary Detection with Language-Aware Selective Fusion

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T10:18:44.133419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T10:18:44.133419Z digest=sha256:61f17461c38a310c127aa01e1ec1533b42b17182d936a4c80ab0f774e3b83c2b

Observation 9ffa1cd3-1c7a-4f2a-8b55-aad5ad78509a · inbound

Bridging Time and Space: Decoupled Spatio-Temporal Alignment for Video Grounding cites this paper.

Bridging Time and Space: Decoupled Spatio-Temporal Alignment for Video Grounding OV-DINO: Unified Open-Vocabulary Detection with Language-Aware Selective Fusion

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:15:52.714098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T18:38:16.204012Z digest=sha256:729b2ecd1807f8989b4f9a72b77b9b3d975b5d401a3cf89dfd310073294ea757

Observation 55c0b928-ff47-4129-897b-666db6c1e72d · inbound

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection cites this paper.

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection OV-DINO: Unified Open-Vocabulary Detection with Language-Aware Selective Fusion

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-12T11:01:30.280840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T01:24:56.236650Z digest=sha256:67bb559b862b85a177e52656b8b88c41d4c45a5dcdfb3bf3f1b0072863a4f2dc

Observation aaec8475-5878-4758-b6aa-33f753f3d931 · inbound

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection cites this paper.

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection OV-DINO: Unified Open-Vocabulary Detection with Language-Aware Selective Fusion

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:00:35.786329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T19:14:12.092478Z digest=sha256:bba411270e2921cb4553ab780870105791617cdc93cf02591cf5aeaa56af768a

Observation 5a6c34d4-c1f2-460d-82a2-a42632974133 · inbound

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection cites this paper.

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection OV-DINO: Unified Open-Vocabulary Detection with Language-Aware Selective Fusion

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:51:48.406458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T01:33:55.317107Z digest=sha256:02285575378f4b53688251941724c081b38ef51bcc1ef26e8957392b1848ce42

Observation a0643a01-b863-4231-b322-c3311c8b21cd · inbound

Open-Vocabulary Gaze Object Prediction: Benchmark and Method cites this paper.

Open-Vocabulary Gaze Object Prediction: Benchmark and Method OV-DINO: Unified Open-Vocabulary Detection with Language-Aware Selective Fusion

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-01T14:18:38.052966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:18:38.052966Z digest=sha256:70c9318616a2e0b264c9f76d4ac8375d44b987562ce8618220f8a0c78fec4fca

Observation 043f68e6-4dbd-4137-a18f-8f2421d77bb4 · inbound

Teaching MLLMs to Say No: Generalized Referring Expression Comprehension via Refusal Calibrated GRPO cites this paper.

Teaching MLLMs to Say No: Generalized Referring Expression Comprehension via Refusal Calibrated GRPO OV-DINO: Unified Open-Vocabulary Detection with Language-Aware Selective Fusion

Reference 101

Resolution
unresolved
no resolver link, observed 2026-08-06T18:44:51.279963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:44:51.279963Z digest=sha256:cf5a4540f8f77ebd72ae1bfbd00740c48a3b95e9b60a474f0cf98acb1b583c06