Pith. sign in

Paper Citation Record · LEDGER

An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 22 inbound Pith citation observations for arXiv:2401.02361.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2401.02361 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 22 of 22 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T21:55:23.351984Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T10:27:02.358494Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 2a559df2-72fb-4529-82c1-3936457f60a1 · inbound

LLMDet: Learning Strong Open-Vocabulary Object Detectors under the Supervision of Large Language Models cites this paper.

LLMDet: Learning Strong Open-Vocabulary Object Detectors under the Supervision of Large Language Models An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-09T21:55:23.351984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T21:55:23.351984Z digest=sha256:0fa83ca06467312a45d6f5f32a670369226d932ccd4be160d96acbfd0d594c86

Observation 4b86998f-e0b9-4839-8036-bb0f69105d49 · inbound

SemSegBench & DetecBench: Benchmarking Reliability and Generalization Beyond Classification cites this paper.

SemSegBench & DetecBench: Benchmarking Reliability and Generalization Beyond Classification An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:15.950528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:41:15.950528Z digest=sha256:e772a408ad03448105cec5d4bf678c90d389ebc4906d5ffe89affea69048ae59

Observation 1efe85d8-c42e-4f8a-a403-3af50ae9eedf · inbound

DINO-R1: Incentivizing Reasoning Capability in Vision Foundation Models cites this paper.

DINO-R1: Incentivizing Reasoning Capability in Vision Foundation Models An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T12:43:41.692788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:43:41.692788Z digest=sha256:bbb2f8f5e25f3f5637eccb474fd27e697402bb6976a27a8305bcf7782ddc3895

Observation 45e8d754-781f-409c-999e-6debc3a739d7 · inbound

3D-MOOD: Lifting 2D to 3D for Monocular Open-Set Object Detection cites this paper.

3D-MOOD: Lifting 2D to 3D for Monocular Open-Set Object Detection An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T10:42:01.065564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:42:01.065564Z digest=sha256:e0ac9992fa98f570a6dd8aff436b9fe5dae2579435a297867b14ee333efe4daf

Observation e99a1f33-3852-4387-84a7-c1421d2a9899 · inbound

FG-CLIP 2: A Bilingual Fine-grained Vision-Language Alignment Model cites this paper.

FG-CLIP 2: A Bilingual Fine-grained Vision-Language Alignment Model An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T10:18:44.680887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T10:18:44.680887Z digest=sha256:ed9bce57dd8b6895993db3ae93feb7553d20cfbd2500d830aa45904f2416c808

Observation 5c190997-ec5f-41ba-b72f-d5d5304883e6 · inbound

Detector-Empowered Video Large Language Model for Efficient Spatio-Temporal Grounding cites this paper.

Detector-Empowered Video Large Language Model for Efficient Spatio-Temporal Grounding An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 80

Resolution
verified exact
arxiv_id, observed 2026-05-17T00:58:46.594864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-17T00:54:53.789523Z digest=sha256:a7380d9cb80585ed3fec24662bab581fea91b3d0831c04a1f8fbb206c31af6cf

Observation e88c9d94-71db-4dde-bb93-b8fa46b96bbd · inbound

Focus on What Really Matters in Low-Altitude Governance: A Management-Centric Multi-Modal Benchmark with Implicitly Coordinated Vision-Language Reasoning Framework cites this paper.

Focus on What Really Matters in Low-Altitude Governance: A Management-Centric Multi-Modal Benchmark with Implicitly Coordinated Vision-Language Reasoning Framework An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:07:48.345220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-16T11:03:37.176829Z digest=sha256:84377dc8fec5b133ca4f83c8070a598f69f77a3f316bb69efed1e6242615bd32

Observation 7f335ee2-8af0-41d3-bc57-160cecb213a3 · inbound

PET-DINO: Unifying Visual Cues into Grounding DINO with Prompt-Enriched Training cites this paper.

PET-DINO: Unifying Visual Cues into Grounding DINO with Prompt-Enriched Training An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:23:21.416722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T22:19:21.723963Z digest=sha256:66f117b202a2f761db5a3574a233177aafd28c229003977e30c6d4d2f6604976

Observation 540a292e-b921-46dc-a206-39bcb5603324 · inbound

Training a Student Expert via Semi-Supervised Foundation Model Distillation cites this paper.

Training a Student Expert via Semi-Supervised Foundation Model Distillation An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-13T16:53:00.102253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T16:50:57.376622Z digest=sha256:ee5b6e251f930188d5fba981001d3b3de7bf878d3637a92d1cf60e3ad691fb21

Observation 209b0ead-cc99-482c-8f98-a1d7eaf71716 · inbound

Beyond Semantic Search: Towards Referential Anchoring in Composed Image Retrieval cites this paper.

Beyond Semantic Search: Towards Referential Anchoring in Composed Image Retrieval An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:35:50.534296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T19:42:23.224746Z digest=sha256:aa69c12d3f254d57638bd8ab573b37a5a568bd1a62940fe87686b1d273b99c2b

Observation 463a44b0-acb2-4e75-b519-aeb67b593d8c · inbound

The Second Challenge on Cross-Domain Few-Shot Object Detection at NTIRE 2026: Methods and Results cites this paper.

The Second Challenge on Cross-Domain Few-Shot Object Detection at NTIRE 2026: Methods and Results An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 127

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:21:00.707452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T16:05:16.070144Z digest=sha256:be7f26ca3589929cd79526d7c876f76b77ef88b6cea3d098f739f72bd9d3c7aa

Observation 2f0e812a-c15e-440b-9b25-eca914c50af2 · inbound

DETR-ViP: Detection Transformer with Robust Discriminative Visual Prompts cites this paper.

DETR-ViP: Detection Transformer with Robust Discriminative Visual Prompts An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T12:10:21.951732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T12:07:21.203513Z digest=sha256:d35350797f32381334f2db98d3819dba127cb23a432c69c0117bc9245313cde4

Observation f9b45074-bd74-46d8-9c0b-962acfd60426 · inbound

DETR-ViP: Detection Transformer with Robust Discriminative Visual Prompts cites this paper.

DETR-ViP: Detection Transformer with Robust Discriminative Visual Prompts An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-12T20:05:20.702092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T20:05:20.702092Z digest=sha256:f2917c6c2ecf3767812eae270b101b3517b863a644cec9d3a3a23305b41937e3

Observation 1cd0ed13-1760-4394-8ae9-705145babb7e · inbound

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection cites this paper.

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-12T11:01:30.276802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T01:24:56.236650Z digest=sha256:f63b71fb0fcfe0407a8fed17cd9226d22ca31c9cd42faa309cdd818f8c86a221

Observation 9d01a686-8cd3-4429-8355-412294974194 · inbound

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection cites this paper.

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:00:35.797463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T19:14:12.092478Z digest=sha256:8d3622d7d1735c057b241facdd4fe21c129d7aba4cb4ed89d76a9d658e5b2873

Observation baa1efb4-1ade-4b71-a36d-6a1bf977a9e0 · inbound

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection cites this paper.

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:51:48.454982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-12T01:33:55.317107Z digest=sha256:feeaa728143a9ae575abdc947f65ae3e6d71e46bd6c868649fd053d3118e8def

Observation 61f9333f-e049-465d-bad7-866ed712caeb · inbound

DetRefiner: Model-Agnostic Detection Refinement with Feature Fusion Transformer cites this paper.

DetRefiner: Model-Agnostic Detection Refinement with Feature Fusion Transformer An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:26:18.879905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-12T03:25:49.511875Z digest=sha256:941b9d32787244d4fb430037b25149cd05d349c93c12e5ac535530ad967685e2

Observation 9ddef33d-510d-4ed1-9cb8-6cc45a0324a5 · inbound

Robust Onion: Peeling Open Vocab Object Detectors Under Noise cites this paper.

Robust Onion: Peeling Open Vocab Object Detectors Under Noise An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 69

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T12:59:53.215985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T05:33:25.373918Z digest=sha256:832c4aa54d079d8b2821701d6090796b1ad302511b48683466449ea647dc14b3

Observation 439bc8e2-3a67-4be8-8230-69ae5c0e53a5 · inbound

Robust Onion: Peeling Open Vocab Object Detectors Under Noise cites this paper.

Robust Onion: Peeling Open Vocab Object Detectors Under Noise An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 75

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T16:15:49.989638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T00:42:36.605869Z digest=sha256:a81d4f40200e5915c7744b5186128afff150d65c7d004e4257947493754c1c64

Observation 509f46f5-f5b3-43df-bde3-c13a696c041e · inbound

DroneFINE: Domain-Aware Parameter-Efficient Fine-Tuning of Vision-Language Detectors for Drone Images cites this paper.

DroneFINE: Domain-Aware Parameter-Efficient Fine-Tuning of Vision-Language Detectors for Drone Images An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 46

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T15:27:04.695200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-02T15:23:42.245556Z digest=sha256:529cb0003711ab98c56d08aaccecb348ae985e4b8955e9cefcc9943bbeeba436

Observation 5fc73722-0554-4add-84c8-b5d95d9b9f21 · inbound

UniRef-UAV: A Multimodal Benchmark for Universal Referring in UAV Imagery cites this paper.

UniRef-UAV: A Multimodal Benchmark for Universal Referring in UAV Imagery An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-07-10T10:27:02.359976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T10:22:14.308951Z digest=sha256:555faffbbe31452d2c2fe6f48ed2429c08b96559c14fc38486f998026700329d

Observation c4342e8e-9501-4f06-ab95-cf5278aac247 · inbound

Teaching MLLMs to Say No: Generalized Referring Expression Comprehension via Refusal Calibrated GRPO cites this paper.

Teaching MLLMs to Say No: Generalized Referring Expression Comprehension via Refusal Calibrated GRPO An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 107

Resolution
unresolved
no resolver link, observed 2026-08-06T18:44:51.853525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:44:51.853525Z digest=sha256:af6f0a27b7cd2496d9015bc12dc03b10c06741c4a66dd1ff80e57f8dfc603c46