Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:25:43.959038Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 7 inbound Pith citation observations for arXiv:2505.18986.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:25:43.959038Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-29T07:07:59.925795Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T10:17:58.075590Z
43 of 43 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 8031d5a1-ae22-493d-b571-74ad3918e70f · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Lang, Sourabh V ora, Venice Erin Liong, Qiang Xu, Anush Krishnan, Yu Pan, Giancarlo Baldan, and Oscar Beijbom
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 80bd6b94-22d0-4e78-98cc-3191d61c86c3 · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion MMDetection: Open MMLab Detection Toolbox and Benchmark
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6815a255-5f92-403c-a406-0daa57900b10 · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion A simple framework for contrastive learning of visual representations
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7695381b-dbd6-445b-acb7-586bd352201e · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Pix2seq: A Language Modeling Framework for Object Detection
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33c63cf9-3e2e-4215-ada8-85d55d388162 · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48a071c6-3d4b-4e2f-bffb-25d64aea5eb1 · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 288d0ccd-d231-4159-be59-242400cd12f7 · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Yolo-world: Real- time open-vocabulary object detection
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 83de4954-8d7b-44f0-9aa2-636c8924faf8 · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Evaluating Large-Vocabulary Object Detectors: The Devil is in the Details
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b01f4f06-47e3-4fb3-ae57-d385e3b4b817 · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Reducing network agnostophobia
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9db2f7de-1b34-4bd2-9a0c-f139f406c28a · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion An image is worth 16x16 words: Transformers for image recognition at scale
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f63d6b4c-fd85-4bc0-ad99-27228d92063a · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion LLMDet: Learning Strong Open-Vocabulary Object Detectors under the Supervision of Large Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19dd24c6-c3e8-4896-9945-e7cf4b572175 · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Recent advances in open set recognition: A survey
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d1c5f081-6694-4ff2-8139-7c0bcc87d935 · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Lvis: A dataset for large vocabulary instance segmentation
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 129904e7-f0fc-417d-a952-164a0f83f016 · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Ow- detr: Open-world detection transformer
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b5ca2b6e-0f50-4553-a3d8-05fb99cb394f · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Momentum contrast for unsupervised visual representation learning
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1c8338aa-67bb-4a63-9bc2-2445730ab42b · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Mask r-cnn
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e2ea59b2-12b6-42dc-9e04-128fea24ba2f · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion T-rex2: Towards generic object detection via text-visual prompt synergy
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d3a8e7ce-0b56-4ead-b1f4-6a1d7596a753 · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Segment anything
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c0219f15-f889-4c94-a70b-64a2b3560a8c · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Lisa: Reasoning segmentation via large language model
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 95ca5c2d-c1c2-4d79-807c-bc76acb402a1 · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion LLaVA-OneVision: Easy Visual Task Transfer
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a9ca5e0-6b27-41a1-bd80-8c35644b54c7 · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Dn-detr: Accelerate detr training by introducing query denoising
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a9ccb8e4-1767-4134-806f-97bc6c3f423b · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Coda: A real-world road corner case dataset for object detection in autonomous driving
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1767c512-b56c-4a29-8389-15275d85d357 · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Desco: Learning object recognition with rich language descriptions
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6188faa5-3e31-4a7e-a0f4-045ff99fa574 · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Grounded language-image pre-training
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1ea8ff82-22a0-4319-a530-fb6425cc897f · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Generative region-language pretraining for open-ended object detection
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b1cd2f1d-0e0a-4292-b2cf-c793afca37fb · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Training-free open-ended object detection and segmentation via attention as prompts
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e459fbe4-2f7e-433e-943f-ce29b7572115 · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Visual instruction tuning.Neural Information Processing Systems (NeurIPS), 2023
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 65e03801-1900-45f7-a0fb-91aff4a1bc80 · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Grounding dino: Marrying dino with grounded pre-training for open-set object detection
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 66878a2e-4dc8-44c5-a84a-3a94ab7d4a43 · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Swin transformer: Hierarchical vision transformer using shifted windows
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 53e798e5-d546-4307-b90d-0fb5d8dc97ad · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Capdet: Unifying dense captioning and open-world detection pretraining
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 86bf87f0-ecd7-4e52-9836-8ac65fd7ce75 · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Scaling open-vocabulary object detection.Neural Information Processing Systems (NeurIPS), 2023
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d0e3482c-010f-4af8-8e77-014c190f77be · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Learning transferable visual models from natural language supervision
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1cf861bd-aa92-4732-b410-788eb9ba0f3d · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Faster r-cnn: Towards real-time object detection with region proposal networks
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 952f7a13-222b-4092-b791-fa62baf447d8 · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Grounded SAM: Assembling Open-World Models for Diverse Visual Tasks
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc5ded50-7edf-402d-a22f-6193f73cc3b1 · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Scalability in perception for autonomous driving: Waymo open dataset
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 46e99aa6-694f-4db8-ab1f-aac1b4d565c3 · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion OV-DINO: Unified Open-Vocabulary Detection with Language-Aware Selective Fusion
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c90909b2-a67e-4f66-8a7d-cde1ca2d1edb · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Cogvlm: Visual expert for pretrained language models
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f802fc6a-8f2c-467f-998a-c86d33c79d40 · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Detclipv2: Scalable open-vocabulary object detection pre-training via word-region alignment
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a048e2b5-a136-4b19-ae26-7e841210f4b8 · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Detclip: Dictionary-enriched visual-concept paralleled pre-training for open-world detection
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0cfd0a67-5e40-40cb-ad5a-0454b65051c3 · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Detclipv3: Towards versatile generative open-vocabulary object detection
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e33a2214-f781-4e90-92e1-5313eccf1022 · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Ni, and Heung-Yeung Shum
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bf19aa99-7d41-4b33-830f-9a967b0586c9 · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Llava-grounding: Grounded visual chat with large multimodal models
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a910885d-f952-4053-a270-93fbdc12f077 · outbound
VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion Glipv2: Unifying localization and vision-language understanding
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f945f4ea-7141-4be0-a3e6-3f2553c5d7f9 · inbound
Visual Prompt Based Reasoning for Offroad Mapping using Multimodal LLMs VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fdd5e044-6917-4c84-b422-71aaa53799e8 · inbound
VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d0f78be8-c593-4474-959b-3a09d910a640 · inbound
VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 36545036-448a-4b16-a2fa-2703d6456869 · inbound
VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0e043359-7b35-4c85-b811-1538a58c95d3 · inbound
FoodCHA: Multi-Modal LLM Agent for Fine-Grained Food Analysis VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3bae7f73-e0ef-4b89-8e37-602e64d41309 · inbound
3DVLA: Enhancing Vision-Language-Action Models via 3D Spatial and Instance Understanding VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0aeb88e6-1235-4112-9af7-b80b89eb9aba · inbound
DrivingAgent: Design and Scheduling Agents for Autonomous Driving Systems VL-SAM-V2: Open-World Object Detection with General and Specific Query Fusion
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.