Pith. sign in

Paper Citation Record · LEDGER

DiffV2IR: Visible-to-Infrared Diffusion Model via Vision-Language Understanding

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2503.19012.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.19012 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T04:45:30.384285Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-08T02:14:26.171030Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation d1ef1bd4-096f-4047-9367-630f465ad483 · inbound

SpectraDINO: Modality-Conditioned Adaptation of RGB Vision Foundation Models Across Infrared Bands cites this paper.

SpectraDINO: Modality-Conditioned Adaptation of RGB Vision Foundation Models Across Infrared Bands DiffV2IR: Visible-to-Infrared Diffusion Model via Vision-Language Understanding

Reference 56

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T16:36:06.358678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-09T16:10:23.516635Z digest=sha256:a902d24bf9939f340c384612bbfb1c87fb180809608cfb002cf56d6a0c5a390d

Observation d7b2a223-e258-48ac-a46f-af0888cfaf8d · inbound

SpectraDINO: Modality-Conditioned Adaptation of RGB Vision Foundation Models Across Infrared Bands cites this paper.

SpectraDINO: Modality-Conditioned Adaptation of RGB Vision Foundation Models Across Infrared Bands DiffV2IR: Visible-to-Infrared Diffusion Model via Vision-Language Understanding

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-02T15:01:00.534257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T15:01:00.534257Z digest=sha256:ebe289b1f153ce7219101156cab2ce69ce5c07468ae4244eef3f50bf79e221dc

Observation 7d88b748-8505-4f8b-9a25-b096683d2887 · inbound

UniTriGen: Unified Triplet Generation of Aligned Visible-Infrared-Label for Few-Shot RGB-T Semantic Segmentation cites this paper.

UniTriGen: Unified Triplet Generation of Aligned Visible-Infrared-Label for Few-Shot RGB-T Semantic Segmentation DiffV2IR: Visible-to-Infrared Diffusion Model via Vision-Language Understanding

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-15T05:55:05.299642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-15T05:52:02.748092Z digest=sha256:9a4a77715da9a3ad81beba3b61274297d9cd70f5108252a0e2dad3448da73b85

Observation 45b9a66a-b282-4a5c-a9b0-b8d95fd99822 · inbound

T-CLIP: Enabling Thermal Perception for Contrastive Language-Image Pretraining cites this paper.

T-CLIP: Enabling Thermal Perception for Contrastive Language-Image Pretraining DiffV2IR: Visible-to-Infrared Diffusion Model via Vision-Language Understanding

Reference 139

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T19:42:36.280336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-28T18:53:10.978631Z digest=sha256:0c75cb1a7917eac6ddb6b8e7b2297295f8b71401fb2d85ed3a7783f7087124e2

Observation b01884d0-867c-47ef-b4a6-b4d6e1defe6c · inbound

FusionRS: A Large-Scale RGB-Infrared-Style Remote Sensing Dataset for Cross-Modal Vision-Language Learning cites this paper.

FusionRS: A Large-Scale RGB-Infrared-Style Remote Sensing Dataset for Cross-Modal Vision-Language Learning DiffV2IR: Visible-to-Infrared Diffusion Model via Vision-Language Understanding

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T04:45:30.384285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T04:45:30.384285Z digest=sha256:60aefd10716ed4acd9011037c7856067fb1bc8fb59e081bb91160e8cb9397cb7

Observation 3f3f287b-2821-4b7e-a32a-6a702d48f3dc · inbound

AnyMatch: Supercharging Universal Multi-Modal Image Matching with Large-Scale Single-View Images cites this paper.

AnyMatch: Supercharging Universal Multi-Modal Image Matching with Large-Scale Single-View Images DiffV2IR: Visible-to-Infrared Diffusion Model via Vision-Language Understanding

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T09:25:41.369412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-07-01T06:36:25.904345Z digest=sha256:5864dae9d79cb2de1b00ef579860a0dfa439594973bc4b0697e0a4201489daeb

Observation 50cd2179-8ff1-4b8a-bd74-c151d6fcda81 · inbound

AnyMatch: Supercharging Universal Multi-Modal Image Matching with Large-Scale Single-View Images cites this paper.

AnyMatch: Supercharging Universal Multi-Modal Image Matching with Large-Scale Single-View Images DiffV2IR: Visible-to-Infrared Diffusion Model via Vision-Language Understanding

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:17:21.132991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-07-02T20:12:25.636045Z digest=sha256:c66774392f4a70ac8c28d3d97a7f963c29780eada407bde1b28162774494217c

Observation 8f810cbb-b1f8-4895-bc21-fc31f4faa148 · inbound

MonoIR-RS: Infrared Remote Sensing Vision-Language Learning with CLIP and VLM Adaptation cites this paper.

MonoIR-RS: Infrared Remote Sensing Vision-Language Learning with CLIP and VLM Adaptation DiffV2IR: Visible-to-Infrared Diffusion Model via Vision-Language Understanding

Reference 25

Resolution
metadata mismatch
local_arxiv, observed 2026-07-08T02:14:26.174795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-07-08T02:06:27.741680Z digest=sha256:dfef5b171a8ab9e25978c39d5be4523f0c7247709d24aaf6f559286c35a8a6fe

Observation 4a74ba66-e586-44d9-9343-4aaa66e53ae6 · inbound

MTVDiff: Multimodal Conditional Latent Diffusion for Enhanced Thermal-to-Visible Face Translation cites this paper.

MTVDiff: Multimodal Conditional Latent Diffusion for Enhanced Thermal-to-Visible Face Translation DiffV2IR: Visible-to-Infrared Diffusion Model via Vision-Language Understanding

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T11:28:39.116403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:28:39.116403Z digest=sha256:ac5448b4f1cc107551fafb30646c3f354adec6671034c848262fda5b72654baf