Pith. sign in

Paper Citation Record · LEDGER

DiffV2IR: Visible-to-Infrared Diffusion Model via Vision-Language Understanding

As of 12 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2503.19012.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.19012 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T04:45:30.384285Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-08T02:14:26.171030Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation d1ef1bd4-096f-4047-9367-630f465ad483 · inbound

SpectraDINO: Modality-Conditioned Adaptation of RGB Vision Foundation Models Across Infrared Bands cites this paper.

SpectraDINO: Modality-Conditioned Adaptation of RGB Vision Foundation Models Across Infrared Bands DiffV2IR: Visible-to-Infrared Diffusion Model via Vision-Language Understanding

Reference 56

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T16:36:06.358678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-09T16:10:23.516635Z digest=sha256:13c88c8da30185ce9390d4040e07e10ddac60b6ac3a8a916c6dd231eff0aa1f8

Observation d7b2a223-e258-48ac-a46f-af0888cfaf8d · inbound

SpectraDINO: Modality-Conditioned Adaptation of RGB Vision Foundation Models Across Infrared Bands cites this paper.

SpectraDINO: Modality-Conditioned Adaptation of RGB Vision Foundation Models Across Infrared Bands DiffV2IR: Visible-to-Infrared Diffusion Model via Vision-Language Understanding

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-02T15:01:00.534257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T15:01:00.534257Z digest=sha256:ebe289b1f153ce7219101156cab2ce69ce5c07468ae4244eef3f50bf79e221dc

Observation 7d88b748-8505-4f8b-9a25-b096683d2887 · inbound

UniTriGen: Unified Triplet Generation of Aligned Visible-Infrared-Label for Few-Shot RGB-T Semantic Segmentation cites this paper.

UniTriGen: Unified Triplet Generation of Aligned Visible-Infrared-Label for Few-Shot RGB-T Semantic Segmentation DiffV2IR: Visible-to-Infrared Diffusion Model via Vision-Language Understanding

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-15T05:55:05.299642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-15T05:52:02.748092Z digest=sha256:06530766b5bb858d368003d4a1ba4ac37df15ed88b4d4baeb405d7e93bd1fb51

Observation 45b9a66a-b282-4a5c-a9b0-b8d95fd99822 · inbound

T-CLIP: Enabling Thermal Perception for Contrastive Language-Image Pretraining cites this paper.

T-CLIP: Enabling Thermal Perception for Contrastive Language-Image Pretraining DiffV2IR: Visible-to-Infrared Diffusion Model via Vision-Language Understanding

Reference 139

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T19:42:36.280336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-06-28T18:53:10.978631Z digest=sha256:0ea94da506bce5565a7a96d3efb2553a4f6642c0b259e575a8d0717379f6a29c

Observation b01884d0-867c-47ef-b4a6-b4d6e1defe6c · inbound

FusionRS: A Large-Scale RGB-Infrared-Style Remote Sensing Dataset for Cross-Modal Vision-Language Learning cites this paper.

FusionRS: A Large-Scale RGB-Infrared-Style Remote Sensing Dataset for Cross-Modal Vision-Language Learning DiffV2IR: Visible-to-Infrared Diffusion Model via Vision-Language Understanding

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T04:45:30.384285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T04:45:30.384285Z digest=sha256:60aefd10716ed4acd9011037c7856067fb1bc8fb59e081bb91160e8cb9397cb7

Observation 3f3f287b-2821-4b7e-a32a-6a702d48f3dc · inbound

AnyMatch: Supercharging Universal Multi-Modal Image Matching with Large-Scale Single-View Images cites this paper.

AnyMatch: Supercharging Universal Multi-Modal Image Matching with Large-Scale Single-View Images DiffV2IR: Visible-to-Infrared Diffusion Model via Vision-Language Understanding

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T09:25:41.369412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-07-01T06:36:25.904345Z digest=sha256:84a98f238096e48f3b6716180dc00d40fc7ec5e843755a1cc02b948c9a8f7b96

Observation 50cd2179-8ff1-4b8a-bd74-c151d6fcda81 · inbound

AnyMatch: Supercharging Universal Multi-Modal Image Matching with Large-Scale Single-View Images cites this paper.

AnyMatch: Supercharging Universal Multi-Modal Image Matching with Large-Scale Single-View Images DiffV2IR: Visible-to-Infrared Diffusion Model via Vision-Language Understanding

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:17:21.132991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-07-02T20:12:25.636045Z digest=sha256:39b1b0f52214467ff74deb2a376e2b860bfd757eccbcf8a55dd6c9e6a7ac0509

Observation 8f810cbb-b1f8-4895-bc21-fc31f4faa148 · inbound

MonoIR-RS: Infrared Remote Sensing Vision-Language Learning with CLIP and VLM Adaptation cites this paper.

MonoIR-RS: Infrared Remote Sensing Vision-Language Learning with CLIP and VLM Adaptation DiffV2IR: Visible-to-Infrared Diffusion Model via Vision-Language Understanding

Reference 25

Resolution
metadata mismatch
local_arxiv, observed 2026-07-08T02:14:26.174795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-07-08T02:06:27.741680Z digest=sha256:025719dbeae781f3d18e96ca081913917245541f9838e6aaa3d15ad589e3f64c

Observation 4a74ba66-e586-44d9-9343-4aaa66e53ae6 · inbound

MTVDiff: Multimodal Conditional Latent Diffusion for Enhanced Thermal-to-Visible Face Translation cites this paper.

MTVDiff: Multimodal Conditional Latent Diffusion for Enhanced Thermal-to-Visible Face Translation DiffV2IR: Visible-to-Infrared Diffusion Model via Vision-Language Understanding

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T11:28:39.116403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:28:39.116403Z digest=sha256:ac5448b4f1cc107551fafb30646c3f354adec6671034c848262fda5b72654baf