Pith. sign in

Paper Citation Record · LEDGER

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization

As of 12 August 2026, this Paper Citation Record lists 65 of 65 outbound references and 1 inbound Pith citation observation for arXiv:2501.09499.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.09499 v1

Coverage vector

measured 65 of 65 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T20:01:25.565723Z

measured 66 of 66 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-01T06:34:15.790154Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T06:35:29.356653Z

Reference resolution

65 of 65 outbound references displayed

  • verified exact0
  • verified fuzzy49
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2cf6c9f1-e94c-4157-ad84-7f0866f58598 · outbound

This paper cites https : / / huggingface.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization https : / / huggingface

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.367058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.319906Z digest=sha256:ccc7f934bd99a60bb2485b6ea77e5306c09fd2e8e704b14070011940fb2f18d2

Observation 0defbe13-5a0b-4049-bebd-8c47e167281e · outbound

This paper cites Semantic-sparse coloriza- tion network for deep exemplar-based colorization.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Semantic-sparse coloriza- tion network for deep exemplar-based colorization

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.354170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.324304Z digest=sha256:393c249253a7b15b707b1a5dda9f9568470eb3fca2cd302f1307714e0f56e635

Observation 35af48cc-26e8-4fee-abe5-f1c60cf775fe · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T20:01:25.328376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:01:25.328376Z digest=sha256:cd82d8c217c7c8662c71bb806f6ef54051f00d8c7b47bacad6d4313d0542f52f

Observation ef50c5d9-4a13-4ed4-8bb0-ebfc8487ff06 · outbound

This paper cites Versatile Vision Foundation Model for Image and Video Colorization.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Versatile Vision Foundation Model for Image and Video Colorization

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.342326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.333161Z digest=sha256:28ca8868d47f42c08264f08883c8664aaffab079abaf0612f207211026b82b0f

Observation e41be00a-72a6-4014-a8d0-6a6b6f506c40 · outbound

This paper cites L-CoIns: Language-based colorization with instance awareness.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization L-CoIns: Language-based colorization with instance awareness

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.330446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.337065Z digest=sha256:69156ddbabd2601458aa14982bad9675f48ee57c33dc13d59e87324c8a5f9859

Observation cb042cab-4f17-4eee-81d6-1f72e4427904 · outbound

This paper cites Deep colorization.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Deep colorization

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.318969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.341105Z digest=sha256:d574710389601ea76736a94739171848a934c8ec9aded268d326cdb3a011f9c3

Observation 1e720e3b-7687-476f-82ec-35f808e7ee35 · outbound

This paper cites Automatic Controllable Colorization via Imagination.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Automatic Controllable Colorization via Imagination

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.307352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.345154Z digest=sha256:3fb78f7099d4443f2e1aa2a3d13fbae5bbaa91f549bcf0dd76eaf9a9a898106f

Observation c29deb4b-d48e-4b96-af14-188eba18a429 · outbound

This paper cites Learn- ing large-scale automatic image colorization.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Learn- ing large-scale automatic image colorization

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.294037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.349063Z digest=sha256:3de6dc3c0d0d39efced3c1d2b97ec62fa41703ca6a100590edb1644d3a94b4a8

Observation 8911e4fc-1da8-4be8-890d-99ae3bbf452b · outbound

This paper cites Learning diverse image colorization.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Learning diverse image colorization

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.279690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.352987Z digest=sha256:5dc805d6cc4b05b407063af2a9a7fb632b747879c3e1e0ba3abf5f59aafc8c11

Observation bbeecf3a-4c02-4927-988f-ae2ea62b1705 · outbound

This paper cites A superpixel-based variational model for image colorization.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization A superpixel-based variational model for image colorization

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.265713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.356740Z digest=sha256:16ade387b43682399bb1fb34b6097bca94d270189cfed8ad50ce06c5dcc366ed

Observation 369c9030-57dd-4d76-93fe-dfcdd2eda53f · outbound

This paper cites A Neural Algorithm of Artistic Style.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization A Neural Algorithm of Artistic Style

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T20:01:25.360322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:01:25.360322Z digest=sha256:c188dbc3326c111e62ddce8fd1acf72f320919e0d81f1f3d9350999022864a01

Observation 47427e31-075f-49fd-8aed-f88df2e8b1cc · outbound

This paper cites Long video generation with time-agnostic vqgan and time- sensitive transformer.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Long video generation with time-agnostic vqgan and time- sensitive transformer

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.252578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.364481Z digest=sha256:45a8080c84cc757973ded8490a2b6e2628c02b73c0e85a3931fca55dac23e27b

Observation 92159ea6-549c-4bd6-8b45-4bfa26f82aa3 · outbound

This paper cites Measuring color- fulness in natural images.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Measuring color- fulness in natural images

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.241362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.368338Z digest=sha256:d8be06fc6de6762eeeb7eb5cd73b834b52f0e53c33f51af2879203d10ccc2a05

Observation f2ce9f71-bf9c-45eb-806e-117ed837ad2b · outbound

This paper cites CLIPScore: A Reference-free Evaluation Metric for Image Captioning.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization CLIPScore: A Reference-free Evaluation Metric for Image Captioning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T20:01:25.372325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:01:25.372325Z digest=sha256:39f39c214ac9f8b6da271102f9fe577659fc746574ecc39333c61ed7e87cd6f8

Observation 7559ee70-5e3e-4a8c-affb-707247e5b3dc · outbound

This paper cites Unicolor: A unified framework for multi-modal colorization with trans- former.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Unicolor: A unified framework for multi-modal colorization with trans- former

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.228621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.376480Z digest=sha256:4f157f5e3b2e23e2387830f4885340ae86422bb166dc9f221fcc7ef8a4c7cd83

Observation 54c89dd9-0d91-4f4a-be86-d8802b8745db · outbound

This paper cites Deepremaster: tem- poral source-reference attention networks for comprehensive video enhancement.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Deepremaster: tem- poral source-reference attention networks for comprehensive video enhancement

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.214775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.380275Z digest=sha256:b014718a0767dc85cba597f9d256a441a073aa7a6bacfb864126d0894ed221ac

Observation 76630b9a-6b05-4d63-9e48-99a20d7b6894 · outbound

This paper cites Colorformer: Image colorization via color memory assisted hybrid-attention transformer.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Colorformer: Image colorization via color memory assisted hybrid-attention transformer

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.202687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.383941Z digest=sha256:5af5dbd990c091c8ff9e178cec15cd4c3fb2e3ac82f34342e0559c2a531c4d27

Observation fc6597ea-dcce-49c8-a37d-52a8ae97aca2 · outbound

This paper cites DDColor: Towards photo- realistic image colorization via dual decoders.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization DDColor: Towards photo- realistic image colorization via dual decoders

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.190891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.387649Z digest=sha256:681c9ec4457b5a6db31921eeb61ef98fc0d2618e9b1e35f7907ff317f91f7d50

Observation 5c5223c8-8578-42ff-bac8-e377d3e45e38 · outbound

This paper cites CoTracker: It is Better to Track Together.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization CoTracker: It is Better to Track Together

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T20:01:25.391228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:01:25.391228Z digest=sha256:648d4c8473d97fd54733981ef932a22850b20f6cd6d5d2ceb3d151c64764ae3d

Observation 53f4ef50-a638-480a-8f16-efdc4f9daf86 · outbound

This paper cites Neural preset for color style transfer.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Neural preset for color style transfer

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.179665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.395266Z digest=sha256:9d9710cb5095b6d287fae65815719d245c5776ea81b97357cafa5cddfc18f9eb

Observation f7aeaa9e-ecc1-4d36-8542-936274d4425a · outbound

This paper cites Slic: Self-supervised learning with iterative clus- tering for human action videos.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Slic: Self-supervised learning with iterative clus- tering for human action videos

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.168146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.399115Z digest=sha256:e0a208b8cf53313f2ab9c9846c0277a3765a493a60eb4a6bf9f574f15dcfee63

Observation a57226e8-7698-4397-822c-9fa0550c09d4 · outbound

This paper cites Auto-Encoding Variational Bayes.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Auto-Encoding Variational Bayes

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T20:01:25.403000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:01:25.403000Z digest=sha256:5015985a42be1f0f94b53d27b06930c1e8ad8c565c0073b03a25a608ca4f0fe7

Observation d4cb9daa-83b7-4e20-b2c2-4d735cb79bf0 · outbound

This paper cites Learning representations for automatic colorization.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Learning representations for automatic colorization

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.156797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.406792Z digest=sha256:4fbebb9dda8887dc3c740d6ccf717402439e9692aa2683f38de477a5527fde70

Observation c255e9e9-5f9d-4d46-b25b-0602a2b07f17 · outbound

This paper cites Fully automatic video colorization with self-regularization and diversity.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Fully automatic video colorization with self-regularization and diversity

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.144862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.410571Z digest=sha256:e6a665d7f9edf5989ce0deeb9f785107e1ac2c9aff99db742c612ec1263740d2

Observation cd77d1e5-8fe3-4363-8db5-17e7495a1943 · outbound

This paper cites Deep video prior for video consistency and propagation.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Deep video prior for video consistency and propagation

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.133504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.414237Z digest=sha256:2ad836a5172cf810091b4e954b95b12bbd373276f94503cc0dc9bd4aed081fd3

Observation 61b6fc24-2450-48df-abb9-2877b5246150 · outbound

This paper cites Automatic example-based image colorization using location- aware cross-scale matching.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Automatic example-based image colorization using location- aware cross-scale matching

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.122301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.418209Z digest=sha256:48672e3b0cb0aa4730908a93efc5667fdeb294287a2ecc7bf7480c2414538bd2

Observation 97d45809-bee5-461f-91d4-fbff11b68fc0 · outbound

This paper cites Towards Photorealistic Video Colorization via Gated Color- Guided Image Diffusion Models.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Towards Photorealistic Video Colorization via Gated Color- Guided Image Diffusion Models

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.110244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.422118Z digest=sha256:ff76cc401ff6a724ac911c9ff6dc90cab2c4d823039ccb17165f67be59c57d91

Observation 114490dc-56e7-4775-8248-8b2104fff21b · outbound

This paper cites VideoGen: A Reference-Guided Latent Diffusion Approach for High Definition Text-to-Video Generation.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization VideoGen: A Reference-Guided Latent Diffusion Approach for High Definition Text-to-Video Generation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T20:01:25.425868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:01:25.425868Z digest=sha256:77c72ac8f64982355937207a8ab84bced101b71dbe31d4e21d8df3cac6eb0deb

Observation 25c0c84d-69b8-453f-9071-5e9499f68d6e · outbound

This paper cites Control Color: Multimodal Diffusion-based Interactive Image Colorization.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Control Color: Multimodal Diffusion-based Interactive Image Colorization

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T20:01:25.430224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:01:25.430224Z digest=sha256:bb7415703064444c5dca21bab27f64dbf086381dd909e1495abdbe9f07396355

Observation 3991614a-0534-48a2-854d-a06095f6f7e9 · outbound

This paper cites Video Colorization with Pre-trained Text-to-Image Diffusion Models.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Video Colorization with Pre-trained Text-to-Image Diffusion Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T20:01:25.434190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:01:25.434190Z digest=sha256:d89a910872ed616f5af6ba27510c6eb31306ca60c7716ed3be516b7b129e39a1

Observation 737a304f-2cb7-441f-ba72-2f1e6b7eaccf · outbound

This paper cites Fr\'echet Video Motion Distance: A Metric for Evaluating Motion Consistency in Videos.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Fr\'echet Video Motion Distance: A Metric for Evaluating Motion Consistency in Videos

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T20:01:25.438600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:01:25.438600Z digest=sha256:c79d85c55d191220b2fe3dbe69c0618ce86c4cd76e71c18015fee609bc313d47

Observation 1403326e-c205-4173-9225-e646984ce772 · outbound

This paper cites Switch- able temporal propagation network.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Switch- able temporal propagation network

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.096879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.442389Z digest=sha256:82709f262829e5bbf9cae8afc5e09b924acd85b3c32772434b52f953696be31b

Observation d783abc3-2f47-40ce-b067-00e78b21c75f · outbound

This paper cites Decoupled Weight Decay Regularization.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Decoupled Weight Decay Regularization

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T20:01:25.446011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:01:25.446011Z digest=sha256:4e47e5eb845221bc337a8e9a422e1693ae579ec701f546b64911eb33a9f5398d

Observation 31337d8c-0c10-41b7-93fa-40a76ef6951b · outbound

This paper cites Unpaired cartoon image synthesis via gated cycle mapping.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Unpaired cartoon image synthesis via gated cycle mapping

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.084046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.449520Z digest=sha256:4c82051a6ff736b123bca3c74243587e4fc1d0abfa1af51fbc4e30e32e4ff731

Observation f5d91190-eed0-4bdf-b4d4-6540428b7c6b · outbound

This paper cites OpenVid-1M: A Large-Scale High-Quality Dataset for Text-to-video Generation.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization OpenVid-1M: A Large-Scale High-Quality Dataset for Text-to-video Generation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T20:01:25.452931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:01:25.452931Z digest=sha256:5794e08652fb498eca1f583f12d1908808d7884cb915a1f82c13a61554890e4b

Observation 0cbdf23f-8792-44e5-93de-166f76cd1c7a · outbound

This paper cites Perazzi, J.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Perazzi, J

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.072099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.457096Z digest=sha256:0fd50f17eb8999160bc4f8a9c5cb6f8b648fcab78336426fe49fb31ad3d2f5dd

Observation d6c1de74-f70e-4371-9ed4-0928fd953e82 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Learning transferable visual models from natural language supervi- sion

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.059842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.460738Z digest=sha256:f6b8abe3dc8ec8fed268f144fc9f123cf6195aae2cf3eb61345883a9967334bb

Observation 55b54c06-5b82-429f-b0ec-989fd1db683b · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization High-resolution image synthesis with latent diffusion models

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.047020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.464952Z digest=sha256:85d6126129d35cb9a0db94c48004978d2e5db7912bf560ace6096029e594692e

Observation 13290b23-9792-4281-b8f2-7c765cee721a · outbound

This paper cites Instance- aware image colorization.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Instance- aware image colorization

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.034202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.468663Z digest=sha256:c2ee8bca9040c0bb842282fe700ab8116741f4264d7895f8c9fb69774fe2a6db

Observation ddce61d5-7d3b-4259-89d6-cd8f025ba7a3 · outbound

This paper cites Raft: Recurrent all-pairs field transforms for optical flow.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Raft: Recurrent all-pairs field transforms for optical flow

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T20:01:25.472126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:01:25.472126Z digest=sha256:ca80c4618ee3a1b8a40a9c75d34aa12a6bd596325da960d0bf19995fdea19743

Observation fe3a301a-5b98-4d6a-8cee-1647e491f147 · outbound

This paper cites Towards Accurate Generative Models of Video: A New Metric & Challenges.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Towards Accurate Generative Models of Video: A New Metric & Challenges

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-10T20:01:25.475947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:01:25.475947Z digest=sha256:e8ab84367210ce39996ee3448b9be41673f41515c8afc853ed4eb08e76c02c48

Observation 9aebdcde-6545-4267-b1ee-dbd48902a3ba · outbound

This paper cites Tracking emerges by colorizing videos.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Tracking emerges by colorizing videos

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.014261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.480116Z digest=sha256:397aeb20466899a474a22ee7c6f71e757b777053e03810c83bb2214319aa6ee0

Observation f80168d4-0842-4ab2-a99e-52ef8de6a2fb · outbound

This paper cites Unsupervised deep exemplar colorization via pyramid dual non-local attention.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Unsupervised deep exemplar colorization via pyramid dual non-local attention

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.002287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.483699Z digest=sha256:2fd530dd2f7925de524a6fd93eb4926e771c9a9d18301727d0e293d185ebf815

Observation c2875ce4-f248-4705-a0e3-c3be5cf0a4df · outbound

This paper cites CT 2: Colorization transformer via color tokens.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization CT 2: Colorization transformer via color tokens

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.990705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.487512Z digest=sha256:bbc17609182881ace036e4b124a7ca0810bcdd16d5aa5e071d054501f3c0ae20

Observation 24f8cfd8-890e-46cb-9cf3-eaad0c4ebe02 · outbound

This paper cites L-CAD: Language-based colorization with any-level descriptions using diffusion priors.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization L-CAD: Language-based colorization with any-level descriptions using diffusion priors

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.978652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.491193Z digest=sha256:a9412d37d2fff460cdfd26ac3a4a597c883aecd5a9ba6d90311d03553baca0fa

Observation 1063e8bb-059f-430f-9a94-bb378e3087f2 · outbound

This paper cites Dynamicrafter: Animating open-domain images with video diffusion priors.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Dynamicrafter: Animating open-domain images with video diffusion priors

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.965993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.494932Z digest=sha256:bfcf7fdd1a46617d2782ed8b02cfa1f85f06cf82ed55b612e75293bd22525fa7

Observation 48335aa9-11e0-48c5-8870-4b7fc4731521 · outbound

This paper cites GMFlow: Learning optical flow via global matching.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization GMFlow: Learning optical flow via global matching

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.954378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.498445Z digest=sha256:ee1eb3b8516d64c43dcadd7315702663ddee6b77e89453619cf1671eb1d385cd

Observation f534be7a-0a08-44de-9d0b-2324dc2f58ef · outbound

This paper cites Stylization-based architecture for fast deep exemplar colorization.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Stylization-based architecture for fast deep exemplar colorization

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.941558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.502299Z digest=sha256:b10f53929584d755c3136e5723380130bbe0972901723edc5f86c0650c349c9c

Observation c8851c77-f7bc-4dea-9dc6-30d28e4442e8 · outbound

This paper cites Depth Anything V2.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Depth Anything V2

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-10T20:01:25.505980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:01:25.505980Z digest=sha256:bb7a71e8bd48837a757be256fbfb645132b42690a73c80d0cfcf692307c15137

Observation fb5de070-c5a8-4e26-8870-22def420f921 · outbound

This paper cites Bistnet: Semantic image prior guided bidirectional temporal feature fusion for deep exemplar-based video colorization.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Bistnet: Semantic image prior guided bidirectional temporal feature fusion for deep exemplar-based video colorization

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.927343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.509980Z digest=sha256:88940324fdb7b2b99c690493f6bdd629a67cbe293f8491c22c2659e3ebea166b

Observation ff4fed0d-6403-4bc9-bf6c-74b13df051bf · outbound

This paper cites ColorMNet: A Memory-based Deep Spatial-Temporal Fea- ture Propagation Network for Video Colorization.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization ColorMNet: A Memory-based Deep Spatial-Temporal Fea- ture Propagation Network for Video Colorization

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.914472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.513458Z digest=sha256:f5309a810a4aa3909d7749e49d8433ca8cd59f7141d70da07ddb847636975614

Observation c1331a13-adac-412d-b4ad-8dd4d95aad7b · outbound

This paper cites Photorealistic style transfer via wavelet transforms.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Photorealistic style transfer via wavelet transforms

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.902709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.517297Z digest=sha256:4820086dd9727129629de37c86066058a07db0dbc7c725ae8535b2e1febc0707

Observation 5a3aadf5-b088-40ee-8208-5a93f9f3bf4f · outbound

This paper cites iColoriT: Towards propagating local hints to the right region in interactive colorization by leveraging vision transformer.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization iColoriT: Towards propagating local hints to the right region in interactive colorization by leveraging vision transformer

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.890878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.520884Z digest=sha256:1d43384aae58393e7e14f046533d84b4ee38fb24cb4821eab4b715b7c57ddaec

Observation 47e8cfc4-a4d5-47a4-9c69-8058860f4a77 · outbound

This paper cites Diffusing Colors: Image Colorization with Text Guided Diffusion.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Diffusing Colors: Image Colorization with Text Guided Diffusion

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.878739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.524500Z digest=sha256:013bdc3886d9b2faecbbd59357c27dd3e6bd491212b830bb392286f1c3eace6e

Observation 3185ac72-34a7-4303-8bf4-ee26b712a7bb · outbound

This paper cites Deep exemplar- based video colorization.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Deep exemplar- based video colorization

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.866550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.528364Z digest=sha256:8f964216845b375aae78c98790e47b9c6ce457cf98299e073e449a4096ecd123

Observation 6ed7009e-f098-45dc-b3f9-807c7b7f848d · outbound

This paper cites A tale of two features: Stable diffusion complements dino for zero-shot semantic correspondence.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization A tale of two features: Stable diffusion complements dino for zero-shot semantic correspondence

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.855327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.532316Z digest=sha256:eb95329fd7b5fa1dccfad3cd0a0e12f37edfadaf2cdf5ac58a5d51f1cf6c86f9

Observation 1e4cd868-329b-4576-9fee-7cec81a5724d · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Adding conditional control to text-to-image diffusion models

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.843920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.535749Z digest=sha256:5c38922e5427f7ee45d4d59e0d4e68f87512698a934cecec9bd1c4d22222c5da

Observation 060ce9f1-b07c-4578-afbe-78d49257c12e · outbound

This paper cites Colorful image colorization.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Colorful image colorization

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.832275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.539312Z digest=sha256:47388371e7182fbbdd0151d650a25ee01587f20788bfe447f57083ddc0227bd1

Observation 89f9657c-725a-40ae-ac93-747eaa4f2af2 · outbound

This paper cites CV-VAE: A Compatible Video VAE for Latent Generative Video Models.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization CV-VAE: A Compatible Video VAE for Latent Generative Video Models

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-10T20:01:25.543114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:01:25.543114Z digest=sha256:a6845071a0b53fc088fd3c0e6f4c9357015a0d5a3be8f8f58887fe8222eb13d2

Observation 49b53f7d-be82-4956-ad40-62a57be8f08f · outbound

This paper cites VC- GAN: Video colorization with hybrid generative adversarial network.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization VC- GAN: Video colorization with hybrid generative adversarial network

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.819249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.547005Z digest=sha256:07a2e69a3f81bec42303fce47e6756c4e1e042d8f7a98d4456c63c5212d52afb

Observation 5f41c89a-a5d7-4277-a0c2-68663515f21f · outbound

This paper cites SVCNet: Scribble-based video colorization network with temporal aggregation.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization SVCNet: Scribble-based video colorization network with temporal aggregation

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.807116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.550664Z digest=sha256:6689d7147bf35b0452c8c77713168ab394f6e5a0efbe9b964a9aab1c2fcdfe81

Observation 773da9c5-1166-422c-a53b-2ef4115014f9 · outbound

This paper cites an unresolved cited work.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-08-10T20:01:25.781763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.558212Z digest=sha256:8bcc186d369a35a3dc5cb70e7fdd3feaf0cb80029647771f7e4dd3b0cb10549f

Observation fcd39e48-89a7-450e-95d8-6c9fd724b124 · outbound

This paper cites Negative values indicate green, and positive values indicate red.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Negative values indicate green, and positive values indicate red

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.769655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.562051Z digest=sha256:b0d33e16b315c468f6b4535aaf30adb0854c3810a3e132d00b8738b721998591

Observation c490a63c-23cf-46e7-b0eb-f2869b98b324 · outbound

This paper cites Red ducks.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Red ducks

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.756254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.565723Z digest=sha256:ff645c188246f59fb687d570b5e888f814f0f2aacc5be8492e405b1f8ff9f24e

Observation 292f0c44-6ce5-48ee-a291-ea92f228c75a · outbound

This paper cites A, including the parameter settings of the network during training and the analysis of luma channel replacement.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization A, including the parameter settings of the network during training and the analysis of luma channel replacement

Reference 2023

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.794282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:01:25.554501Z digest=sha256:a3147b21d1992a7fbabc7f8cc26b81162d68520917099f36896ff527be3ce8b4

Pith citing papers

Observation 333dd639-f753-46aa-af36-f6450e2484d2 · inbound

RaBitQCache: Rotated Binary Quantization for KVCache in Long Context LLM Inference cites this paper.

RaBitQCache: Rotated Binary Quantization for KVCache in Long Context LLM Inference VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-01T06:35:29.359106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-07-01T06:34:15.790154Z digest=sha256:abebe43ae5d9dd4283ef655783db46aee0dddebfe3d34fcfb8efbc8ea14f24a7