Pith. sign in

Paper Citation Record · LEDGER

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios

As of 23 August 2026, this Paper Citation Record lists 70 of 70 outbound references and 4 inbound Pith citation observations for arXiv:2512.24561.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2512.24561 v2

Coverage vector

measured 70 of 70 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-03T13:22:13.267737Z

measured 74 of 74 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T15:01:03.556212Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-13T23:18:25.494068Z

Reference resolution

70 of 70 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved69
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1407b649-e64b-40d9-8988-2a1232010af3 · outbound

This paper cites Qwen-vl: A versatile vision-language model for un- derstanding, localization, text reading, and beyond, 2023.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Qwen-vl: A versatile vision-language model for un- derstanding, localization, text reading, and beyond, 2023

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:07.475422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:07.475422Z digest=sha256:3fb485a7a37a6afa89847aafba199f10ef8ba7c08c03523fd870e7b2c82bd0fb

Observation bb100858-17d2-40bd-8059-6adb8e7479c1 · outbound

This paper cites Uniter: Universal image-text representation learning.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Uniter: Universal image-text representation learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:07.532784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:07.532784Z digest=sha256:380cb9b40faee32b94460abbbbc8a66b6a792ef4cd50e18e96289002121933c5

Observation 1ee842d9-9938-44b3-8875-9f63b5e079d2 · outbound

This paper cites Cops-ref: A new dataset and task on composi- tional referring expression comprehension.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Cops-ref: A new dataset and task on composi- tional referring expression comprehension

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:07.598268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:07.598268Z digest=sha256:f374e9fa17476f7adc8fd8ca87cb0084c4ac3d274a136c958fc70ea8b1c3133f

Observation 4f01e7a2-af88-41f5-81fe-71e9cd1c16a8 · outbound

This paper cites Unit3d: A unified transformer for 3d dense captioning and visual grounding.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unit3d: A unified transformer for 3d dense captioning and visual grounding

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:07.673878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:07.673878Z digest=sha256:3988a3ae1d3ae2ea777a650514583c1a927f497d965a9fb63d06403a23268ebe

Observation 7e262af5-d149-42b1-92a0-23ca710517fe · outbound

This paper cites Advancing visual grounding with scene knowl- edge: Benchmark and method.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Advancing visual grounding with scene knowl- edge: Benchmark and method

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:07.741282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:07.741282Z digest=sha256:0c8278163bf000f832c1a414801eb0b0302080fc33a5fb809581c0d01e0fd057

Observation c45022d3-1c20-4f82-9414-230b44a1a92e · outbound

This paper cites Transvg: End-to-end visual ground- ing with transformers.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Transvg: End-to-end visual ground- ing with transformers

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:07.814739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:07.814739Z digest=sha256:d0d73788251a89c028f43161360b44e4a86dace4f8c48d61d1747f601ca95818

Observation 309e8538-795e-47ea-a566-b1d1f7993058 · outbound

This paper cites Bert: Pre-training of deep bidirectional trans- formers for language understanding.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Bert: Pre-training of deep bidirectional trans- formers for language understanding

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:07.875627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:07.875627Z digest=sha256:5baa9361e5a695fbe44a2d3042f05ad57b6c6c61300e76477764b3ba0f65e9c5

Observation 49a85298-a2b6-4d7a-99f7-a5e211ef41ab · outbound

This paper cites D3t: Distinctive dual-domain teacher zigzagging across rgb- thermal gap for domain-adaptive object detection.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios D3t: Distinctive dual-domain teacher zigzagging across rgb- thermal gap for domain-adaptive object detection

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:07.982432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:07.982432Z digest=sha256:da2c0a26e0d4bd6f9100beec2b611fce1908cfbce6edaf9a5500ae4d2fc07e22

Observation 0af4a103-7809-46c1-9dc7-8b4af125131b · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:08.081744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:08.081744Z digest=sha256:498e6e56e1cf59b449d03953942cf721d30295445f785c891a06a01e3a7c54ff

Observation bdd8dfea-ecb2-4268-8898-5ba49f84d6a4 · outbound

This paper cites Large-scale adversarial training for vision- and-language representation learning.NeurIPS, 33:6616– 6628, 2020.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Large-scale adversarial training for vision- and-language representation learning.NeurIPS, 33:6616– 6628, 2020

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:08.183519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:08.183519Z digest=sha256:39db136fd968d53c74afd6d531d3721958b10940c90466ec65421af30f7d649c

Observation fec1444d-f117-40fc-b782-b787c38ec76d · outbound

This paper cites Room-and-object aware knowledge reasoning for remote embodied referring expression.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Room-and-object aware knowledge reasoning for remote embodied referring expression

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:08.253760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:08.253760Z digest=sha256:fa01aaaaec3989e044aee04de80e184ccec74b8491728dbdcecb2bc3c53a65fa

Observation b0f35396-7226-4bdd-bae3-aeb7dbf0eff2 · outbound

This paper cites The iapr tc-12 benchmark: A new eval- uation resource for visual information systems.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios The iapr tc-12 benchmark: A new eval- uation resource for visual information systems

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:08.349314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:08.349314Z digest=sha256:1454da1adba6b80ac501ea9944265302d099df8ab43c1342242302330d30faab

Observation cc6a1b54-228b-4f93-a57e-b5970a8e6004 · outbound

This paper cites Deep residual learning for image recognition.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Deep residual learning for image recognition

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:08.447263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:08.447263Z digest=sha256:ef1e2f02510b3abefa49291fbf4f389f76da20be5f5718f6552806b44c48e164

Observation a4bddb92-2038-49c6-b0a5-b5ccd64e6ac0 · outbound

This paper cites Lora: Low-rank adaptation of large language models.ICLR, 1(2):3, 2022.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Lora: Low-rank adaptation of large language models.ICLR, 1(2):3, 2022

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:08.559249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:08.559249Z digest=sha256:525c329b04bc5f9b5510a4cc132f6528b92a3234e68d6e65ba2cce0308dfe904

Observation 7db8483a-6139-4464-94e5-d1b3dd79514a · outbound

This paper cites Ei 2 det: Edge-guided illumination-aware interactive learning for visible-infrared object detection.TCSVT, 2025.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Ei 2 det: Edge-guided illumination-aware interactive learning for visible-infrared object detection.TCSVT, 2025

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:08.727448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:08.727448Z digest=sha256:4f1ea3235bcfc342ed61ddfb398fdbbc76fd45c938e18963b6c7b5ace3c42e62

Observation d41f7c38-41e1-478f-9bea-ac274ef82114 · outbound

This paper cites Beyond one-to-one: Re- thinking the referring image segmentation.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Beyond one-to-one: Re- thinking the referring image segmentation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:08.836138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:08.836138Z digest=sha256:c3fc9d3bad9a281aa36c3947faae7e36e5817c97a7dee83bbf82273c65a990a1

Observation 2e61ec8c-4d4e-468c-8cd5-b5de161dd86f · outbound

This paper cites Referitgame: Referring to objects in pho- tographs of natural scenes.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Referitgame: Referring to objects in pho- tographs of natural scenes

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:08.924987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:08.924987Z digest=sha256:767d3951ab811f51c8a492b4a7acc3b65d16f29cf29b10747716ae1ee4abd12c

Observation 4f621396-d7f2-4d0c-857b-b445360ee236 · outbound

This paper cites mplug: Effective and efficient vision-language learning by cross-modal skip-connections.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios mplug: Effective and efficient vision-language learning by cross-modal skip-connections

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:08.997393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:08.997393Z digest=sha256:1a9b00f12ff1477feb9c8e52bddd6bded608efc142bf68dd316b7291c573f768

Observation 567a5f88-d237-42e1-a838-7af3bed1a193 · outbound

This paper cites Rgb-t semantic segmentation with location, activation, and sharpening.TCSVT, 33(3):1223–1235, 2022.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Rgb-t semantic segmentation with location, activation, and sharpening.TCSVT, 33(3):1223–1235, 2022

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:09.187025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:09.187025Z digest=sha256:ca37dd4082ac08098efb18d5ec7bcb0ec69ead210c39d1c13cb06010f34f075b

Observation da724d24-8ca5-4d5b-af96-fad6aa94cef9 · outbound

This paper cites Microsoft coco: Common objects in context.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Microsoft coco: Common objects in context

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:09.301713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:09.301713Z digest=sha256:f2289a471f4a7f60de21655f5a8fbb0f9f00312c964706512a2e2d6a8f864393

Observation 0f038089-9ff8-44a7-93b0-98721e65a4cd · outbound

This paper cites Gres: Gen- eralized referring expression segmentation.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Gres: Gen- eralized referring expression segmentation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:09.356679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:09.356679Z digest=sha256:3da3c8ef798ca40b66ae7cbeced078fbc0aca096a6c87af3e68bfd348e7b31dd

Observation 91129e57-d05f-4298-9e37-598bc2d60a90 · outbound

This paper cites Refer-it-in-rgbd: A bottom-up ap- proach for 3d visual grounding in rgbd images.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Refer-it-in-rgbd: A bottom-up ap- proach for 3d visual grounding in rgbd images

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:09.469324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:09.469324Z digest=sha256:08125de843c1149d2ca3feea731068597fe4294332bfe8fd2911308bbab0e15e

Observation 805d9a1c-b4cc-4c9d-a0bb-5c60bdf049ce · outbound

This paper cites Visual instruction tuning.NeurIPS, 36:34892–34916, 2023.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Visual instruction tuning.NeurIPS, 36:34892–34916, 2023

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:09.627654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:09.627654Z digest=sha256:fabf40eb8a4df69f24837b3c82999fb5974c1c046c46bbe26798b2fec8ca14cf

Observation ef9c6c3c-4cfd-49bc-acf8-1c69cd8fed7d · outbound

This paper cites Target-aware dual adversarial learning and a multi-scenario multi-modality benchmark to fuse infrared and visible for object detection.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Target-aware dual adversarial learning and a multi-scenario multi-modality benchmark to fuse infrared and visible for object detection

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:09.685054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:09.685054Z digest=sha256:41d31d948f178f207631e15d6f9c9faa583009ec99a4be8b7a392cbc8e35f246

Observation 2edb6309-ac5c-4d98-8c91-1603d2351ad4 · outbound

This paper cites Cross3dvg: Cross-dataset 3d visual grounding on different rgb-d scans.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Cross3dvg: Cross-dataset 3d visual grounding on different rgb-d scans

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:09.797495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:09.797495Z digest=sha256:b708f62aa92ef216156513cb009df90ef1b56760ce17a13931001f47b4c283ae

Observation 959bbfc9-1512-42d2-9499-e1385b44ed0f · outbound

This paper cites Mod- eling context between objects for referring expression under- standing.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Mod- eling context between objects for referring expression under- standing

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:09.957036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:09.957036Z digest=sha256:1006f8f157afa95a3f37faee5f900382de7bbcc7a3696f92bbf6be70580bc4d9

Observation a25a0121-4167-462c-b42c-b046a3549897 · outbound

This paper cites Flickr30k entities: Collecting region-to-phrase corre- spondences for richer image-to-sentence models.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Flickr30k entities: Collecting region-to-phrase corre- spondences for richer image-to-sentence models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:10.034272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:10.034272Z digest=sha256:614d69f0fe78995622fa6a06753d77cf01035e882a1c947a3570f908ba0c7d36

Observation 031b6ad1-fc67-4964-bacc-bea7a1ee9cf7 · outbound

This paper cites Learn- ing transferable visual models from natural language super- vision.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Learn- ing transferable visual models from natural language super- vision

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:10.056944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:10.056944Z digest=sha256:25e6009b006ab6ecc8ce2dd06796b00b5b6afc39e4712a2eb5ba208cc994ebe1

Observation eb5116eb-af16-4e42-937b-d9932eec4ee4 · outbound

This paper cites Dynamic mdetr: A dynamic multimodal transformer decoder for visual grounding.IEEE Transactions on Pattern Analysis and Machine Intelligence, 46(2):1181–1198, 2023.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Dynamic mdetr: A dynamic multimodal transformer decoder for visual grounding.IEEE Transactions on Pattern Analysis and Machine Intelligence, 46(2):1181–1198, 2023

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:10.132196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:10.132196Z digest=sha256:f4b8462f4163c0a2382cd2ada3ca7d94a282e5c9e16eea5efee2d417c6143068

Observation 869f0c3c-cdca-453f-9644-63d37f1169f5 · outbound

This paper cites Attention is all you need.NeurIPS, 30, 2017.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Attention is all you need.NeurIPS, 30, 2017

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:10.165003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:10.165003Z digest=sha256:0dcf42e0621c894acecbff53a5c612ac650f7911ec599e7ab1de67761d9f6bee

Observation e8659ac2-f1dd-4ebf-a7c2-4e7b6a7a8794 · outbound

This paper cites Ofa: Unifying architectures, tasks, and modalities through a simple sequence-to-sequence learning framework.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Ofa: Unifying architectures, tasks, and modalities through a simple sequence-to-sequence learning framework

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:10.222723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:10.222723Z digest=sha256:7f5a6a3ae20333048e506bf61f15b71dedb28f01e3e1eeb1b2fcc85fe0522519

Observation aa35b809-d83d-43ca-a718-0ed735563433 · outbound

This paper cites Image as a foreign language: Beit pretraining for vision and vision- language tasks.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Image as a foreign language: Beit pretraining for vision and vision- language tasks

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:10.291114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:10.291114Z digest=sha256:08a8fc36e7307c2f94fcba25191c34ffdbbb52a56279e0798d17c9eb7819f6a3

Observation a43b6454-1678-48d8-b9d4-e4e4a696c5cb · outbound

This paper cites Clip-vg: Self-paced curriculum adapting of clip for visual grounding.TMM, 26:4334–4347,.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Clip-vg: Self-paced curriculum adapting of clip for visual grounding.TMM, 26:4334–4347,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:10.359561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:10.359561Z digest=sha256:639b231bf23638acd1e45f8e7c8a576fcef8affc66b444435edc648af3cd1267

Observation ceb9017f-6716-48c9-bdeb-10dcb348b875 · outbound

This paper cites Towards visual grounding: A survey.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Towards visual grounding: A survey

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:10.534984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:10.534984Z digest=sha256:898f5d85a3baf257aa64d20e53e4c422bed0ec7d570c8288868216dfe5022de6

Observation 9cbcbc35-a951-4641-b7ab-2c7b948aa364 · outbound

This paper cites Hivg: Hierarchical multimodal fine- grained modulation for visual grounding.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Hivg: Hierarchical multimodal fine- grained modulation for visual grounding

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:10.608296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:10.608296Z digest=sha256:1a2f9545999d5dabd826666e9fd1f1bd92e112a7c56483974e19b85433bb3539

Observation 1a5d5a3c-e6fa-4de4-81e3-4fa0ff4433bf · outbound

This paper cites Oneref: Unified one-tower expression grounding and segmentation with mask referring modeling.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Oneref: Unified one-tower expression grounding and segmentation with mask referring modeling

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:10.690297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:10.690297Z digest=sha256:deee5aee6230fa243557d19f75507efb725a15dfc7c4704b22a276009e856969

Observation 894c0ae2-e442-40e5-bf4b-3d0ed8e4c595 · outbound

This paper cites Described object detection: Liberating ob- ject detection with flexible expressions.NeurIPS, 36:79095– 79107, 2023.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Described object detection: Liberating ob- ject detection with flexible expressions.NeurIPS, 36:79095– 79107, 2023

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:10.751759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:10.751759Z digest=sha256:cd15f83f33b1fa37c10803e9e0545fcf43e2b376a7afe47e0174d4a6102915e1

Observation 12bd4718-e362-4862-b9b4-8c237dc051b9 · outbound

This paper cites Mc-bench: A bench- mark for multi-context visual grounding in the era of mllms.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Mc-bench: A bench- mark for multi-context visual grounding in the era of mllms

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:10.859018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:10.859018Z digest=sha256:3480cd9d012358f4f9ab45837e3603b94185e8977f583544f4717c83ddef9c20

Observation a5317b8b-90d2-4892-949b-1340eea2232c · outbound

This paper cites Improving one-stage visual grounding by recursive sub-query construction.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Improving one-stage visual grounding by recursive sub-query construction

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:10.963077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:10.963077Z digest=sha256:86d620ae04d6c4d3a76a7e90f19f43c241ce4bff6e38bed7d84bc1471d11b267

Observation 0e63b9eb-7172-4b40-80d8-6b4f06347bfc · outbound

This paper cites Unitab: Unifying text and box outputs for grounded vision- language modeling.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unitab: Unifying text and box outputs for grounded vision- language modeling

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:11.074241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:11.074241Z digest=sha256:143e5a27a9fd2540a40117ea3bfdf741240ac70cf6d5c535bad597a3883193a1

Observation 1cb76a83-6046-4b61-a18b-eb439b5d1d27 · outbound

This paper cites Vi- sual grounding with multi-modal conditional adaptation.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Vi- sual grounding with multi-modal conditional adaptation

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:11.136609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:11.136609Z digest=sha256:66e2a83717bd1ae66c4ef6b0f7313c537b3b7fe2092a686629a3b3b669ab9bdd

Observation 2c7be307-2a62-4605-a87a-d3324ed593d0 · outbound

This paper cites Shifting more attention to visual backbone: Query-modulated refinement networks for end-to-end visual grounding.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Shifting more attention to visual backbone: Query-modulated refinement networks for end-to-end visual grounding

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:11.208551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:11.208551Z digest=sha256:6826c3b91396de837f2e15c0999d1aa580831bde68a7d3dcd42632fda9d0b29d

Observation 183bb79e-93ef-4873-9853-bf4a36d29ada · outbound

This paper cites Modeling context in referring expres- sions.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Modeling context in referring expres- sions

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:11.270069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:11.270069Z digest=sha256:fe19fd7eba0547734c96a2716b5710c3071f63220ff1130c698f74883005f9cc

Observation ac0055fa-fcb5-4ab3-915c-2aedee084a49 · outbound

This paper cites C 2former: Calibrated and complementary transformer for rgb-infrared object de- tection.TGRS, 62:1–12, 2024.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios C 2former: Calibrated and complementary transformer for rgb-infrared object de- tection.TGRS, 62:1–12, 2024

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:11.310614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:11.310614Z digest=sha256:0ca53b39bffc558889dbe5bc1dd60f9fecfd6d59f1503d583562e6ad8d36e439

Observation f90310fa-9edf-45f6-8e1c-d43a87e061ba · outbound

This paper cites Trans- lation, scale and rotation: Cross-modal alignment meets rgb-infrared vehicle detection.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Trans- lation, scale and rotation: Cross-modal alignment meets rgb-infrared vehicle detection

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:11.392019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:11.392019Z digest=sha256:45052cd3e38f2e8f4d34e69df5fd53650eedbd92a6f9a7b6fcf74e22f767a4b4

Observation 019f9ec1-96dc-4537-91fa-742ced008088 · outbound

This paper cites Improving rgb-infrared object detection with cascade alignment-guided transformer.Information Fusion, 105:102246, 2024.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Improving rgb-infrared object detection with cascade alignment-guided transformer.Information Fusion, 105:102246, 2024

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:11.462234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:11.462234Z digest=sha256:427bf5c7fed95a8e894acd0f5220a92fd68f67c379b9127f3034f7237d19b26a

Observation 7933f4bf-8206-414e-9d2f-76a2fef01119 · outbound

This paper cites Unirgb-ir: A unified frame- work for visible-infrared semantic tasks via adapter tuning.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unirgb-ir: A unified frame- work for visible-infrared semantic tasks via adapter tuning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:11.542644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:11.542644Z digest=sha256:fc90179ab5cba7fb68161af74cb2072e151fd1a9df9828783e211f71f041b6b0

Observation 5dc45d55-c7c5-452b-95b5-0ee8441e4abf · outbound

This paper cites Multispectral fusion for object detection with cyclic fuse-and-refine blocks.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Multispectral fusion for object detection with cyclic fuse-and-refine blocks

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:11.610798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:11.610798Z digest=sha256:79f2752824ece491a745e976a4d2b8abf92f4013bac10b7be9c17bb9dc3db82f

Observation ee7be17c-9ea4-40a8-88c4-0ed0c8c9468e · outbound

This paper cites Abmdrnet: Adaptive-weighted bi-directional modality difference reduc- tion network for rgb-t semantic segmentation.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Abmdrnet: Adaptive-weighted bi-directional modality difference reduc- tion network for rgb-t semantic segmentation

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:11.644623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:11.644623Z digest=sha256:86dc11f3b8b175c450ee23bcb261166057e1c0590bda6471bf8d20249c703045

Observation 81ffd4b3-5722-4dc6-a3e6-57e85c43e6d4 · outbound

This paper cites Removal then selection: A coarse-to-fine fusion perspective for rgb-infrared object detection.arXiv e-prints, pages arXiv–2401, 2024.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Removal then selection: A coarse-to-fine fusion perspective for rgb-infrared object detection.arXiv e-prints, pages arXiv–2401, 2024

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:11.726883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:11.726883Z digest=sha256:ace78eb512faff68def481069bb0cdbea1ea6d9727bc914eb7e6d8c8e9fcdfbf

Observation e354113d-a2f4-4527-8820-a55cb53aab71 · outbound

This paper cites be- hind the truck.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios be- hind the truck

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:11.831105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:11.831105Z digest=sha256:38739fee853805e0eb1605c0ca88d8fb73a4dd52316856ac4aa40eb5ff968285

Observation c1623795-31fe-4d43-b53b-225543485a47 · outbound

This paper cites an unresolved cited work.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unresolved cited work

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:11.914135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:11.914135Z digest=sha256:9bbb5694133b353e77cd20872d84d3a3cf4efb1f05ddbc159bc1af851f6b9375

Observation edfd31a8-d30e-41d2-b6fc-ff48332fbe89 · outbound

This paper cites an unresolved cited work.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unresolved cited work

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:12.001270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:12.001270Z digest=sha256:ba3fd44a6ad81c52e8905395c31c8f753a0292416a2a285ab21d9ca74db67e08

Observation c2fb9113-c4ab-4fc6-b8f0-796a559df62c · outbound

This paper cites an unresolved cited work.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unresolved cited work

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:12.187388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:12.187388Z digest=sha256:837a25d0bacfad0afddb8c8dfe2babfe529c2195e22ade79121898fd912a9d89

Observation f96a5e63-4701-43df-b6b6-4fd64bd69e3e · outbound

This paper cites an unresolved cited work.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unresolved cited work

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:12.218174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:12.218174Z digest=sha256:9cb251a0a31449b52a131db207203e2d23afdb9adc8091348935be9608128d05

Observation f32a7b6f-b590-4184-867f-954a7f47a86c · outbound

This paper cites an unresolved cited work.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unresolved cited work

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:12.304908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:12.304908Z digest=sha256:a352335605c00b42333de988067978207a5ee1d198ac54849e93008598606362

Observation 0a7c6681-4264-499d-bf3a-75c1b9f15d4d · outbound

This paper cites an unresolved cited work.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unresolved cited work

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:12.488625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:12.488625Z digest=sha256:d46f3cb342400e28aa2b5137b76f6d9ac60c36349b5ca0a77b3234462350ecf7

Observation 38d3df62-5325-4b4b-a79a-586a1f9b87c8 · outbound

This paper cites an unresolved cited work.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unresolved cited work

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:12.615969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:12.615969Z digest=sha256:6af374d33e65f25d65d67bb15ab74e2d2ff8648c001dcb8dc4e9ed61ad019ac9

Observation 8540625b-e574-43fd-ac62-002f82383d36 · outbound

This paper cites an unresolved cited work.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unresolved cited work

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:12.686063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:12.686063Z digest=sha256:d5898d715a836ade0a00bd0575448aa3a65c4c8b07bf8fd323ebfa94e5491a36

Observation 400a225d-e54d-46b3-838d-7cec33fbae08 · outbound

This paper cites an unresolved cited work.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unresolved cited work

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:12.756090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:12.756090Z digest=sha256:c9fd2c7a4760aa79bafe069b928ff5e6520c7dfdb75a8a390074dac6b14a164c

Observation 7beeb4e8-13a5-4879-bb26-19bbae99ffcc · outbound

This paper cites near the stairs.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios near the stairs

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:12.820399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:12.820399Z digest=sha256:cc42975c26931313ecec82c9b38a41fd726b2820c1a1ef13ddc00bbf08f60f58

Observation 386a503d-75d2-4be4-93ec-f822478cd406 · outbound

This paper cites an unresolved cited work.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unresolved cited work

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:12.879990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:12.879990Z digest=sha256:723a7fbf8214f4dfd250bf2a1c119d7a27df4f6133022f6281fe08d6e02b55ed

Observation 4919d6e8-abb9-4622-94ab-c7e6a8db808f · outbound

This paper cites an unresolved cited work.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unresolved cited work

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:12.957929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:12.957929Z digest=sha256:6653cd1367dbfb094e69028091dee221e6b26d1d1411067af1dec106d3f36742

Observation bf3a12b4-79e7-4ca7-9f1a-43aa71c047d2 · outbound

This paper cites an unresolved cited work.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unresolved cited work

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:12.979078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:12.979078Z digest=sha256:17a817a1f2918d475ddfb4029caf273cc8ab3aedf743e472c4aa89067269c289

Observation 9c91cae0-6661-4ba5-a87e-7c2c9c824b6b · outbound

This paper cites Please return only one number corresponding to the lighting condition: 0 (very_weak_light), 1 (weak_light), 2 (normal_light), or 3 ( strong_light).

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Please return only one number corresponding to the lighting condition: 0 (very_weak_light), 1 (weak_light), 2 (normal_light), or 3 ( strong_light)

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:13.007925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:13.007925Z digest=sha256:07bbb4c4b240bafa570078b26d74947905383388881ae3ee50b79765cafa4ca1

Observation f67d87e2-450c-474c-ae53-981ab63e3435 · outbound

This paper cites an unresolved cited work.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unresolved cited work

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:13.051235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:13.051235Z digest=sha256:995c7750eef4fdfaf66eaab6f4335009260e3361d5cd69d943efb139ade2806a

Observation 6b6846cd-bf47-4eb4-8db7-40c503b7c1e3 · outbound

This paper cites an unresolved cited work.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unresolved cited work

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:13.131462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:13.131462Z digest=sha256:f7572c2f8e6a6d5aa683d8f56b4e70a011d7f0269589fd03e0a2a397948a38d6

Observation 540c878a-3b0f-4679-b930-fbd5c2801678 · outbound

This paper cites an unresolved cited work.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unresolved cited work

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:13.221032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:13.221032Z digest=sha256:10291fc41151cd425e95f11281ce8f0151b3c4c80ecd418fcb3eb94bb98f5867

Observation 806563eb-1840-4fbb-8b28-9cef8c06c8b0 · outbound

This paper cites small"ifsize_ratio < 0.01else.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios small"ifsize_ratio < 0.01else

Reference 75

Resolution
malformed identifier
no resolver link, observed 2026-08-03T13:22:13.267737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:13.267737Z digest=sha256:3b5fe3e23beafd4fbd1405a291e721f0c991efeabe01230fb405ff9b1c102370

Observation ef0166ab-e496-4d55-88d5-ec756ac1c30e · outbound

This paper cites an unresolved cited work.

RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios Unresolved cited work

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-03T13:22:10.429051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:22:10.429051Z digest=sha256:d00793b4d17783806f4991951e75c33cda87e15ebeaa3008e558e9d1cb87c099

Pith citing papers

Observation 627a73fa-a568-4ebf-95f0-97a79ed31635 · inbound

OmniSch: A Multimodal PCB Schematic Benchmark For Structured Diagram Visual Reasoning cites this paper.

OmniSch: A Multimodal PCB Schematic Benchmark For Structured Diagram Visual Reasoning RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-07-01T02:17:18.600387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-13T23:17:45.005518Z digest=sha256:be169d8d1d9ef55c965220a305400d6a6d509db967dcf1f55a8f15a6ef053852

Observation 88a6f25c-1961-4add-97de-c2a9524ed012 · inbound

OmniSch: A Multimodal PCB Schematic Benchmark For Structured Diagram Visual Reasoning cites this paper.

OmniSch: A Multimodal PCB Schematic Benchmark For Structured Diagram Visual Reasoning RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios

Reference 40

Resolution
unresolved
no resolver link, observed 2026-07-13T15:15:25.977993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T15:15:25.977993Z digest=sha256:10c4068320c7a6d6c9c66638f27439a35d27b4a065e20bf7496706f05bc4ee2f

Observation bc4d7ac6-4525-4f87-9402-b4fb9e1bb77a · inbound

SpectraDINO: Modality-Conditioned Adaptation of RGB Vision Foundation Models Across Infrared Bands cites this paper.

SpectraDINO: Modality-Conditioned Adaptation of RGB Vision Foundation Models Across Infrared Bands RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios

Reference 80

Resolution
verified exact
arxiv_id, observed 2026-07-01T02:17:18.600387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-09T16:10:23.516635Z digest=sha256:e767f3cac578366380e12542deb836eb03c309ee0995f7e2741d06c85d0465b9

Observation 5199272c-3e9f-4c74-9573-f11b653460ee · inbound

SpectraDINO: Modality-Conditioned Adaptation of RGB Vision Foundation Models Across Infrared Bands cites this paper.

SpectraDINO: Modality-Conditioned Adaptation of RGB Vision Foundation Models Across Infrared Bands RGBT-GroundBench: Visual Grounding Beyond RGB in Complex Real-World Scenarios

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-02T15:01:03.556212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T15:01:03.556212Z digest=sha256:2ec5642d78631f1df95132be6e0674917e68b10c1f19a0b638c7c9bc495390b7