Pith. sign in

Paper Citation Record · LEDGER

An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 21 inbound Pith citation observations for arXiv:2401.02361.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2401.02361 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 21 of 21 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:41:15.950528Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T10:27:02.358494Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 4b86998f-e0b9-4839-8036-bb0f69105d49 · inbound

SemSegBench & DetecBench: Benchmarking Reliability and Generalization Beyond Classification cites this paper.

SemSegBench & DetecBench: Benchmarking Reliability and Generalization Beyond Classification An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:15.950528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:41:15.950528Z digest=sha256:1dfc43ebe585545fada03ed160f7e141334d367798990bcdf42fb95c280bae10

Observation 1efe85d8-c42e-4f8a-a403-3af50ae9eedf · inbound

DINO-R1: Incentivizing Reasoning Capability in Vision Foundation Models cites this paper.

DINO-R1: Incentivizing Reasoning Capability in Vision Foundation Models An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T12:43:41.692788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:43:41.692788Z digest=sha256:e4dd2f98cb176f09a851b9ea298dbca126b6f2bbed1fb43d8cc215e5aa634bf8

Observation 45e8d754-781f-409c-999e-6debc3a739d7 · inbound

3D-MOOD: Lifting 2D to 3D for Monocular Open-Set Object Detection cites this paper.

3D-MOOD: Lifting 2D to 3D for Monocular Open-Set Object Detection An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T10:42:01.065564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:42:01.065564Z digest=sha256:b5f012dadc1e204e005f5b3021f62f00f862567ce69dd50067363ac3f34dbe03

Observation e99a1f33-3852-4387-84a7-c1421d2a9899 · inbound

FG-CLIP 2: A Bilingual Fine-grained Vision-Language Alignment Model cites this paper.

FG-CLIP 2: A Bilingual Fine-grained Vision-Language Alignment Model An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T10:18:44.680887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T10:18:44.680887Z digest=sha256:ca62c704680efdd26a0c8923381e9f2637d67ff7ed40e4faa09bab47ca1f0756

Observation 5c190997-ec5f-41ba-b72f-d5d5304883e6 · inbound

Detector-Empowered Video Large Language Model for Efficient Spatio-Temporal Grounding cites this paper.

Detector-Empowered Video Large Language Model for Efficient Spatio-Temporal Grounding An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 80

Resolution
verified exact
arxiv_id, observed 2026-05-17T00:58:46.594864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T00:54:53.789523Z digest=sha256:bf8f58974b2a88f2b0cc815f3df5328622d6419a3cf0068b3bbe576afd2c2ab3

Observation e88c9d94-71db-4dde-bb93-b8fa46b96bbd · inbound

Focus on What Really Matters in Low-Altitude Governance: A Management-Centric Multi-Modal Benchmark with Implicitly Coordinated Vision-Language Reasoning Framework cites this paper.

Focus on What Really Matters in Low-Altitude Governance: A Management-Centric Multi-Modal Benchmark with Implicitly Coordinated Vision-Language Reasoning Framework An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:07:48.345220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T11:03:37.176829Z digest=sha256:5a21384dab37f55df58063e7a7f8bb6cc14d6e6352621cfb199f93b342e7e131

Observation 7f335ee2-8af0-41d3-bc57-160cecb213a3 · inbound

PET-DINO: Unifying Visual Cues into Grounding DINO with Prompt-Enriched Training cites this paper.

PET-DINO: Unifying Visual Cues into Grounding DINO with Prompt-Enriched Training An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:23:21.416722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T22:19:21.723963Z digest=sha256:0d18c2a7a731b26f16cebc0e19a558c8aab4081bc3be115a2fc1bbadda3a5a8b

Observation 540a292e-b921-46dc-a206-39bcb5603324 · inbound

Training a Student Expert via Semi-Supervised Foundation Model Distillation cites this paper.

Training a Student Expert via Semi-Supervised Foundation Model Distillation An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-13T16:53:00.102253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T16:50:57.376622Z digest=sha256:c4e6ff32cfe668c2d5644c8e0bb8659ecec2efc64e20f6b4029dac87eb0dcc32

Observation 209b0ead-cc99-482c-8f98-a1d7eaf71716 · inbound

Beyond Semantic Search: Towards Referential Anchoring in Composed Image Retrieval cites this paper.

Beyond Semantic Search: Towards Referential Anchoring in Composed Image Retrieval An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:35:50.534296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T19:42:23.224746Z digest=sha256:b4493120e4fa1bc5c166dc28a245e859839d25d1609f8e4d8109467522bfab6c

Observation 463a44b0-acb2-4e75-b519-aeb67b593d8c · inbound

The Second Challenge on Cross-Domain Few-Shot Object Detection at NTIRE 2026: Methods and Results cites this paper.

The Second Challenge on Cross-Domain Few-Shot Object Detection at NTIRE 2026: Methods and Results An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 127

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:21:00.707452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T16:05:16.070144Z digest=sha256:338aa0bb32a383be9461a24519d9f7f9d24d9e80ad5811d300b8ef67acd646bd

Observation 2f0e812a-c15e-440b-9b25-eca914c50af2 · inbound

DETR-ViP: Detection Transformer with Robust Discriminative Visual Prompts cites this paper.

DETR-ViP: Detection Transformer with Robust Discriminative Visual Prompts An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T12:10:21.951732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T12:07:21.203513Z digest=sha256:466a66d6a0d186fdaf20f5d68953917aed8ef3847e1d90c59afe62d2b4fcd33a

Observation f9b45074-bd74-46d8-9c0b-962acfd60426 · inbound

DETR-ViP: Detection Transformer with Robust Discriminative Visual Prompts cites this paper.

DETR-ViP: Detection Transformer with Robust Discriminative Visual Prompts An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-12T20:05:20.702092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T20:05:20.702092Z digest=sha256:35ccf14959bf11ed3074213fc191573205cfb0fbd78c15365fac30e8c95ad97e

Observation 1cd0ed13-1760-4394-8ae9-705145babb7e · inbound

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection cites this paper.

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-12T11:01:30.276802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T01:24:56.236650Z digest=sha256:923331ea16e317dbfaa93cdf81cb73a525b299228aac90a19af30ad0a6f1c4c7

Observation 9d01a686-8cd3-4429-8355-412294974194 · inbound

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection cites this paper.

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:00:35.797463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T19:14:12.092478Z digest=sha256:a023732f6f4339314ef54cb74f5f78118c0c9e6816e20ffc0cae2c7de316f9ed

Observation baa1efb4-1ade-4b71-a36d-6a1bf977a9e0 · inbound

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection cites this paper.

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:51:48.454982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T01:33:55.317107Z digest=sha256:60e294856cdad9b8b028bb29bb4074ecfb99b9fb8e3e1dd01071d00121fc225f

Observation 61f9333f-e049-465d-bad7-866ed712caeb · inbound

DetRefiner: Model-Agnostic Detection Refinement with Feature Fusion Transformer cites this paper.

DetRefiner: Model-Agnostic Detection Refinement with Feature Fusion Transformer An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:26:18.879905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T03:25:49.511875Z digest=sha256:be0ac9fb131ddf72bff65bf2cbd834e9ce6a2d08a96395c2863fae5b20373792

Observation 9ddef33d-510d-4ed1-9cb8-6cc45a0324a5 · inbound

Robust Onion: Peeling Open Vocab Object Detectors Under Noise cites this paper.

Robust Onion: Peeling Open Vocab Object Detectors Under Noise An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 69

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T12:59:53.215985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T05:33:25.373918Z digest=sha256:ce63a5528c5ea364273764ebc905f65ff33e0405be4a81b5a994ddd232c8e817

Observation 439bc8e2-3a67-4be8-8230-69ae5c0e53a5 · inbound

Robust Onion: Peeling Open Vocab Object Detectors Under Noise cites this paper.

Robust Onion: Peeling Open Vocab Object Detectors Under Noise An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 75

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T16:15:49.989638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T00:42:36.605869Z digest=sha256:c8c6fb8818fea8991a989803d3159a9042f8cd0a5b010a844828264646e9c3a1

Observation 509f46f5-f5b3-43df-bde3-c13a696c041e · inbound

DroneFINE: Domain-Aware Parameter-Efficient Fine-Tuning of Vision-Language Detectors for Drone Images cites this paper.

DroneFINE: Domain-Aware Parameter-Efficient Fine-Tuning of Vision-Language Detectors for Drone Images An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 46

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T15:27:04.695200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-02T15:23:42.245556Z digest=sha256:7f534594dd8c4a2dd731b6e6befe77ed4c39b43010ba552486a75b4872746476

Observation 5fc73722-0554-4add-84c8-b5d95d9b9f21 · inbound

UniRef-UAV: A Multimodal Benchmark for Universal Referring in UAV Imagery cites this paper.

UniRef-UAV: A Multimodal Benchmark for Universal Referring in UAV Imagery An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-07-10T10:27:02.359976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T10:22:14.308951Z digest=sha256:2085494cd49d6ea723b839161a325278412b63c2a168c0e272fb08ba0ccc85b9

Observation c4342e8e-9501-4f06-ab95-cf5278aac247 · inbound

Teaching MLLMs to Say No: Generalized Referring Expression Comprehension via Refusal Calibrated GRPO cites this paper.

Teaching MLLMs to Say No: Generalized Referring Expression Comprehension via Refusal Calibrated GRPO An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 107

Resolution
unresolved
no resolver link, observed 2026-08-06T18:44:51.853525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:44:51.853525Z digest=sha256:d819ca3a4bdf336c6ffefa26a792516f19c0e9476a11c256dadcfdad55fa48a7