Pith. sign in

Paper Citation Record · LEDGER

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization

As of 12 August 2026, this Paper Citation Record lists 65 of 65 outbound references and 1 inbound Pith citation observation for arXiv:2501.09499.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.09499 v1

Coverage vector

measured 65 of 65 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T20:01:25.565723Z

measured 66 of 66 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-01T06:34:15.790154Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T06:35:29.356653Z

Reference resolution

65 of 65 outbound references displayed

  • verified exact0
  • verified fuzzy49
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2cf6c9f1-e94c-4157-ad84-7f0866f58598 · outbound

This paper cites https : / / huggingface.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization https : / / huggingface

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.367058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.319906Z digest=sha256:cfae4792eb685f37998bf570a4d4f480c62aa29b0518ace392f32e044f6c9ffa

Observation 0defbe13-5a0b-4049-bebd-8c47e167281e · outbound

This paper cites Semantic-sparse coloriza- tion network for deep exemplar-based colorization.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Semantic-sparse coloriza- tion network for deep exemplar-based colorization

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.354170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.324304Z digest=sha256:ff76d0689777bf4e31ede42ab8f66d2661c704c09295c8003fd7eee865611a1b

Observation 35af48cc-26e8-4fee-abe5-f1c60cf775fe · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T20:01:25.328376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:01:25.328376Z digest=sha256:cd82d8c217c7c8662c71bb806f6ef54051f00d8c7b47bacad6d4313d0542f52f

Observation ef50c5d9-4a13-4ed4-8bb0-ebfc8487ff06 · outbound

This paper cites Versatile Vision Foundation Model for Image and Video Colorization.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Versatile Vision Foundation Model for Image and Video Colorization

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.342326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.333161Z digest=sha256:e5430884af8ef530c363f2407f1a269a782b0121af3290a831180bdf6befc599

Observation e41be00a-72a6-4014-a8d0-6a6b6f506c40 · outbound

This paper cites L-CoIns: Language-based colorization with instance awareness.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization L-CoIns: Language-based colorization with instance awareness

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.330446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.337065Z digest=sha256:1395acd0b0f01129dfa4b8601de9d2dd73b2bfd5790880ab8b9696ea8175244f

Observation cb042cab-4f17-4eee-81d6-1f72e4427904 · outbound

This paper cites Deep colorization.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Deep colorization

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.318969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.341105Z digest=sha256:5e2f32b8d2827b3b9bc336f13c738b67597627be7921f3649d78a5b55dd323ea

Observation 1e720e3b-7687-476f-82ec-35f808e7ee35 · outbound

This paper cites Automatic Controllable Colorization via Imagination.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Automatic Controllable Colorization via Imagination

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.307352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.345154Z digest=sha256:3fc343bd509d12ee8482a18166ce3414d64db8a8829bcdf9f1a68e5fa7a7143a

Observation c29deb4b-d48e-4b96-af14-188eba18a429 · outbound

This paper cites Learn- ing large-scale automatic image colorization.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Learn- ing large-scale automatic image colorization

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.294037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.349063Z digest=sha256:bf6132f2ad03cbd51e5b24ffb6c9dc594bf17fc4ed0be7e7f0738483b4e58ab8

Observation 8911e4fc-1da8-4be8-890d-99ae3bbf452b · outbound

This paper cites Learning diverse image colorization.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Learning diverse image colorization

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.279690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.352987Z digest=sha256:557be648e319d1e5a2b4fd71e8ef830a9ab5445e700bbf5a577b1a5139f4882a

Observation bbeecf3a-4c02-4927-988f-ae2ea62b1705 · outbound

This paper cites A superpixel-based variational model for image colorization.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization A superpixel-based variational model for image colorization

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.265713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.356740Z digest=sha256:747d75040867d9b0d1fb09cb740b10b78e9942b8ba0cc20dd3a0e9baca2fb019

Observation 369c9030-57dd-4d76-93fe-dfcdd2eda53f · outbound

This paper cites A Neural Algorithm of Artistic Style.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization A Neural Algorithm of Artistic Style

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T20:01:25.360322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:01:25.360322Z digest=sha256:c188dbc3326c111e62ddce8fd1acf72f320919e0d81f1f3d9350999022864a01

Observation 47427e31-075f-49fd-8aed-f88df2e8b1cc · outbound

This paper cites Long video generation with time-agnostic vqgan and time- sensitive transformer.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Long video generation with time-agnostic vqgan and time- sensitive transformer

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.252578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.364481Z digest=sha256:a50261b1601890ef4fcbe2d4aadca729eb9db5827cd7a78ba9a82ebeeb9b67b7

Observation 92159ea6-549c-4bd6-8b45-4bfa26f82aa3 · outbound

This paper cites Measuring color- fulness in natural images.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Measuring color- fulness in natural images

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.241362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.368338Z digest=sha256:c477341f5f2197a98b01756b5ad894a532e682f2a3579dcdf0b83ee9b05caf73

Observation f2ce9f71-bf9c-45eb-806e-117ed837ad2b · outbound

This paper cites CLIPScore: A Reference-free Evaluation Metric for Image Captioning.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization CLIPScore: A Reference-free Evaluation Metric for Image Captioning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T20:01:25.372325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:01:25.372325Z digest=sha256:39f39c214ac9f8b6da271102f9fe577659fc746574ecc39333c61ed7e87cd6f8

Observation 7559ee70-5e3e-4a8c-affb-707247e5b3dc · outbound

This paper cites Unicolor: A unified framework for multi-modal colorization with trans- former.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Unicolor: A unified framework for multi-modal colorization with trans- former

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.228621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.376480Z digest=sha256:3cd992327134656922726ff2d1c6024a517719d4f1c5b736e54a369df866a99b

Observation 54c89dd9-0d91-4f4a-be86-d8802b8745db · outbound

This paper cites Deepremaster: tem- poral source-reference attention networks for comprehensive video enhancement.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Deepremaster: tem- poral source-reference attention networks for comprehensive video enhancement

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.214775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.380275Z digest=sha256:fd8e3df5c1d5e64300c05bb22b7d47795e45fa4559722be162d665291c1c8f35

Observation 76630b9a-6b05-4d63-9e48-99a20d7b6894 · outbound

This paper cites Colorformer: Image colorization via color memory assisted hybrid-attention transformer.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Colorformer: Image colorization via color memory assisted hybrid-attention transformer

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.202687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.383941Z digest=sha256:6e70d8a65a3793e7d938d159aac3013b16477a1c7747a30f3acef9b60282f357

Observation fc6597ea-dcce-49c8-a37d-52a8ae97aca2 · outbound

This paper cites DDColor: Towards photo- realistic image colorization via dual decoders.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization DDColor: Towards photo- realistic image colorization via dual decoders

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.190891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.387649Z digest=sha256:b70737e50e7d1e5df289a75be75fff6fa79e75e7f30d80805de3100fb07731a0

Observation 5c5223c8-8578-42ff-bac8-e377d3e45e38 · outbound

This paper cites CoTracker: It is Better to Track Together.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization CoTracker: It is Better to Track Together

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T20:01:25.391228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:01:25.391228Z digest=sha256:1199f552b2af9acfdc84f3b45a1330ef568174dfa6846f9c3fcab7abad6da3eb

Observation 53f4ef50-a638-480a-8f16-efdc4f9daf86 · outbound

This paper cites Neural preset for color style transfer.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Neural preset for color style transfer

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.179665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.395266Z digest=sha256:8c004d0c5142e4d0c941ed4898e50dea0c423bdca978d1c6ddc882bb5a218026

Observation f7aeaa9e-ecc1-4d36-8542-936274d4425a · outbound

This paper cites Slic: Self-supervised learning with iterative clus- tering for human action videos.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Slic: Self-supervised learning with iterative clus- tering for human action videos

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.168146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.399115Z digest=sha256:f3c3b980e14d97b146bb686cef06142427edeeefef42a5d2867996ba3d87c6b3

Observation a57226e8-7698-4397-822c-9fa0550c09d4 · outbound

This paper cites Auto-Encoding Variational Bayes.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Auto-Encoding Variational Bayes

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T20:01:25.403000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:01:25.403000Z digest=sha256:5015985a42be1f0f94b53d27b06930c1e8ad8c565c0073b03a25a608ca4f0fe7

Observation d4cb9daa-83b7-4e20-b2c2-4d735cb79bf0 · outbound

This paper cites Learning representations for automatic colorization.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Learning representations for automatic colorization

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.156797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.406792Z digest=sha256:1ae380bb929f38e0fc3a62e55a2d793884c028aee1f66fdce5d20bf47f67e254

Observation c255e9e9-5f9d-4d46-b25b-0602a2b07f17 · outbound

This paper cites Fully automatic video colorization with self-regularization and diversity.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Fully automatic video colorization with self-regularization and diversity

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.144862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.410571Z digest=sha256:832228a65dab92837dbb7597656040c12e8a04d92c53958bd6fab7eda2b19b90

Observation cd77d1e5-8fe3-4363-8db5-17e7495a1943 · outbound

This paper cites Deep video prior for video consistency and propagation.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Deep video prior for video consistency and propagation

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.133504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.414237Z digest=sha256:6856db58fb53327551c46496f2b6aadd3ea80945ef456f616f0914743d193c8c

Observation 61b6fc24-2450-48df-abb9-2877b5246150 · outbound

This paper cites Automatic example-based image colorization using location- aware cross-scale matching.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Automatic example-based image colorization using location- aware cross-scale matching

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.122301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.418209Z digest=sha256:2153506bace7455b1b72b457bb1444a3cc942b05b9d8afdddd5ae07e8525a253

Observation 97d45809-bee5-461f-91d4-fbff11b68fc0 · outbound

This paper cites Towards Photorealistic Video Colorization via Gated Color- Guided Image Diffusion Models.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Towards Photorealistic Video Colorization via Gated Color- Guided Image Diffusion Models

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.110244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.422118Z digest=sha256:e08faa12661cc523075718f13d52935b9d9acb070b3aaf743e3ce8630a5af306

Observation 114490dc-56e7-4775-8248-8b2104fff21b · outbound

This paper cites VideoGen: A Reference-Guided Latent Diffusion Approach for High Definition Text-to-Video Generation.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization VideoGen: A Reference-Guided Latent Diffusion Approach for High Definition Text-to-Video Generation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T20:01:25.425868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:01:25.425868Z digest=sha256:77c72ac8f64982355937207a8ab84bced101b71dbe31d4e21d8df3cac6eb0deb

Observation 25c0c84d-69b8-453f-9071-5e9499f68d6e · outbound

This paper cites Control Color: Multimodal Diffusion-based Interactive Image Colorization.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Control Color: Multimodal Diffusion-based Interactive Image Colorization

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T20:01:25.430224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:01:25.430224Z digest=sha256:bb7415703064444c5dca21bab27f64dbf086381dd909e1495abdbe9f07396355

Observation 3991614a-0534-48a2-854d-a06095f6f7e9 · outbound

This paper cites Video Colorization with Pre-trained Text-to-Image Diffusion Models.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Video Colorization with Pre-trained Text-to-Image Diffusion Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T20:01:25.434190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:01:25.434190Z digest=sha256:da9cb202063fa05e7f8f21deb7010723f7b6173b17dd8e7c551f32f83e8a4428

Observation 737a304f-2cb7-441f-ba72-2f1e6b7eaccf · outbound

This paper cites Fr\'echet Video Motion Distance: A Metric for Evaluating Motion Consistency in Videos.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Fr\'echet Video Motion Distance: A Metric for Evaluating Motion Consistency in Videos

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T20:01:25.438600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:01:25.438600Z digest=sha256:c79d85c55d191220b2fe3dbe69c0618ce86c4cd76e71c18015fee609bc313d47

Observation 1403326e-c205-4173-9225-e646984ce772 · outbound

This paper cites Switch- able temporal propagation network.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Switch- able temporal propagation network

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.096879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.442389Z digest=sha256:71bb70e95372a1e6226f8b33364eded674242aecfff2cb63be634a1dddc40ecb

Observation d783abc3-2f47-40ce-b067-00e78b21c75f · outbound

This paper cites Decoupled Weight Decay Regularization.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Decoupled Weight Decay Regularization

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T20:01:25.446011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:01:25.446011Z digest=sha256:4e47e5eb845221bc337a8e9a422e1693ae579ec701f546b64911eb33a9f5398d

Observation 31337d8c-0c10-41b7-93fa-40a76ef6951b · outbound

This paper cites Unpaired cartoon image synthesis via gated cycle mapping.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Unpaired cartoon image synthesis via gated cycle mapping

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.084046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.449520Z digest=sha256:f275bf63ae6fcd1d3e3d71e5f58409126a6ae73ee93e66e619acbda488a4d3ba

Observation f5d91190-eed0-4bdf-b4d4-6540428b7c6b · outbound

This paper cites OpenVid-1M: A Large-Scale High-Quality Dataset for Text-to-video Generation.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization OpenVid-1M: A Large-Scale High-Quality Dataset for Text-to-video Generation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T20:01:25.452931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:01:25.452931Z digest=sha256:5794e08652fb498eca1f583f12d1908808d7884cb915a1f82c13a61554890e4b

Observation 0cbdf23f-8792-44e5-93de-166f76cd1c7a · outbound

This paper cites Perazzi, J.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Perazzi, J

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.072099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.457096Z digest=sha256:cd1e4acec10a4181dbf29915d582de159e4e53c2c9eba5f2214da4876bc84f73

Observation d6c1de74-f70e-4371-9ed4-0928fd953e82 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Learning transferable visual models from natural language supervi- sion

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.059842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.460738Z digest=sha256:d4eeceb890fe42db82ff1ab26105ce29f95b40bfa07181f100b2670e02cb145d

Observation 55b54c06-5b82-429f-b0ec-989fd1db683b · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization High-resolution image synthesis with latent diffusion models

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.047020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.464952Z digest=sha256:297547efbaa888d6c745232149904a5942d49bc523895030a320338e5917f7b5

Observation 13290b23-9792-4281-b8f2-7c765cee721a · outbound

This paper cites Instance- aware image colorization.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Instance- aware image colorization

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.034202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.468663Z digest=sha256:9a1de2d374e945998e2d3c5174aae35cc6ec6a7a25122298a095ef8ac7041246

Observation ddce61d5-7d3b-4259-89d6-cd8f025ba7a3 · outbound

This paper cites Raft: Recurrent all-pairs field transforms for optical flow.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Raft: Recurrent all-pairs field transforms for optical flow

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T20:01:25.472126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:01:25.472126Z digest=sha256:ca80c4618ee3a1b8a40a9c75d34aa12a6bd596325da960d0bf19995fdea19743

Observation fe3a301a-5b98-4d6a-8cee-1647e491f147 · outbound

This paper cites Towards Accurate Generative Models of Video: A New Metric & Challenges.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Towards Accurate Generative Models of Video: A New Metric & Challenges

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-10T20:01:25.475947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:01:25.475947Z digest=sha256:e8ab84367210ce39996ee3448b9be41673f41515c8afc853ed4eb08e76c02c48

Observation 9aebdcde-6545-4267-b1ee-dbd48902a3ba · outbound

This paper cites Tracking emerges by colorizing videos.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Tracking emerges by colorizing videos

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.014261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.480116Z digest=sha256:7bc2b10f2cf598ca518c63d51f696e3511b145b3cc416df382cc5396874df477

Observation f80168d4-0842-4ab2-a99e-52ef8de6a2fb · outbound

This paper cites Unsupervised deep exemplar colorization via pyramid dual non-local attention.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Unsupervised deep exemplar colorization via pyramid dual non-local attention

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:26.002287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.483699Z digest=sha256:72fe51a99b57dd774679a0f725a5becef67a48a6294ca017ec149fe5e8445d68

Observation c2875ce4-f248-4705-a0e3-c3be5cf0a4df · outbound

This paper cites CT 2: Colorization transformer via color tokens.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization CT 2: Colorization transformer via color tokens

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.990705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.487512Z digest=sha256:a15b97685b3ad845108c94381bd7beeeafb1a087ca0537bd65baae90e8ca4264

Observation 24f8cfd8-890e-46cb-9cf3-eaad0c4ebe02 · outbound

This paper cites L-CAD: Language-based colorization with any-level descriptions using diffusion priors.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization L-CAD: Language-based colorization with any-level descriptions using diffusion priors

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.978652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.491193Z digest=sha256:ae0050e007d2f8214c61d0edfd235556227416c1701351710d5fcc7424abb124

Observation 1063e8bb-059f-430f-9a94-bb378e3087f2 · outbound

This paper cites Dynamicrafter: Animating open-domain images with video diffusion priors.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Dynamicrafter: Animating open-domain images with video diffusion priors

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.965993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.494932Z digest=sha256:2bceddb3af61500bf3b060c4314b9e41d90c662764baa89f46eab23a4a854ce0

Observation 48335aa9-11e0-48c5-8870-4b7fc4731521 · outbound

This paper cites GMFlow: Learning optical flow via global matching.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization GMFlow: Learning optical flow via global matching

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.954378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.498445Z digest=sha256:bee721a8dc85d89f5c927e5fd7ba3bc72de86d3d6458c5add1c76b38d34e5b5b

Observation f534be7a-0a08-44de-9d0b-2324dc2f58ef · outbound

This paper cites Stylization-based architecture for fast deep exemplar colorization.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Stylization-based architecture for fast deep exemplar colorization

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.941558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.502299Z digest=sha256:0820d0841d017820f74ff35ac8448afba639534f920727540e8d020c46f1c3c9

Observation c8851c77-f7bc-4dea-9dc6-30d28e4442e8 · outbound

This paper cites Depth Anything V2.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Depth Anything V2

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-10T20:01:25.505980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:01:25.505980Z digest=sha256:bb7a71e8bd48837a757be256fbfb645132b42690a73c80d0cfcf692307c15137

Observation fb5de070-c5a8-4e26-8870-22def420f921 · outbound

This paper cites Bistnet: Semantic image prior guided bidirectional temporal feature fusion for deep exemplar-based video colorization.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Bistnet: Semantic image prior guided bidirectional temporal feature fusion for deep exemplar-based video colorization

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.927343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.509980Z digest=sha256:d5668472e14700cb3f62cabdaad22f72b92fab243cc89d62ad6cdb851e6b83ce

Observation ff4fed0d-6403-4bc9-bf6c-74b13df051bf · outbound

This paper cites ColorMNet: A Memory-based Deep Spatial-Temporal Fea- ture Propagation Network for Video Colorization.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization ColorMNet: A Memory-based Deep Spatial-Temporal Fea- ture Propagation Network for Video Colorization

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.914472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.513458Z digest=sha256:2b6114b63d15ec9e8c6e8cc0483f4707950b8405128c6b27d1e985c7f32d74db

Observation c1331a13-adac-412d-b4ad-8dd4d95aad7b · outbound

This paper cites Photorealistic style transfer via wavelet transforms.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Photorealistic style transfer via wavelet transforms

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.902709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.517297Z digest=sha256:97800acdd8dc7133b99f2d0198bbb112c52ab030967dde6bb5fd774bd1ddee76

Observation 5a3aadf5-b088-40ee-8208-5a93f9f3bf4f · outbound

This paper cites iColoriT: Towards propagating local hints to the right region in interactive colorization by leveraging vision transformer.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization iColoriT: Towards propagating local hints to the right region in interactive colorization by leveraging vision transformer

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.890878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.520884Z digest=sha256:2f799725c7554c35400ddc54cb6f902532ed14d07e0a7a42bdbec064da5522be

Observation 47e8cfc4-a4d5-47a4-9c69-8058860f4a77 · outbound

This paper cites Diffusing Colors: Image Colorization with Text Guided Diffusion.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Diffusing Colors: Image Colorization with Text Guided Diffusion

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.878739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.524500Z digest=sha256:d01df9d32e2927be56ede1cc7a9c9cb462874333428cc77659bd560a771eedfe

Observation 3185ac72-34a7-4303-8bf4-ee26b712a7bb · outbound

This paper cites Deep exemplar- based video colorization.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Deep exemplar- based video colorization

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.866550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.528364Z digest=sha256:7b1780d78a6ad6ad72efd35191e9be8180d024b34f285c686b8bcd4a2eb97394

Observation 6ed7009e-f098-45dc-b3f9-807c7b7f848d · outbound

This paper cites A tale of two features: Stable diffusion complements dino for zero-shot semantic correspondence.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization A tale of two features: Stable diffusion complements dino for zero-shot semantic correspondence

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.855327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.532316Z digest=sha256:95b5babfa1011147aef339adab0169bea24a940fcd02489be402d09df52ff7b3

Observation 1e4cd868-329b-4576-9fee-7cec81a5724d · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Adding conditional control to text-to-image diffusion models

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.843920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.535749Z digest=sha256:8e308134e7e58a5b7440084049b270156165a389b3acae0f9f4cf628186cc8ce

Observation 060ce9f1-b07c-4578-afbe-78d49257c12e · outbound

This paper cites Colorful image colorization.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Colorful image colorization

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.832275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.539312Z digest=sha256:e6fe33509cdb062b13aa5a03f90d1683763e4afda1bbfdcd6cfdd97e9db5c5ea

Observation 89f9657c-725a-40ae-ac93-747eaa4f2af2 · outbound

This paper cites CV-VAE: A Compatible Video VAE for Latent Generative Video Models.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization CV-VAE: A Compatible Video VAE for Latent Generative Video Models

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-10T20:01:25.543114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:01:25.543114Z digest=sha256:a6845071a0b53fc088fd3c0e6f4c9357015a0d5a3be8f8f58887fe8222eb13d2

Observation 49b53f7d-be82-4956-ad40-62a57be8f08f · outbound

This paper cites VC- GAN: Video colorization with hybrid generative adversarial network.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization VC- GAN: Video colorization with hybrid generative adversarial network

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.819249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.547005Z digest=sha256:90d7e7fe335d6d336c2ad50d3b3c7995bf264e151ef8881d680e6042378e311b

Observation 5f41c89a-a5d7-4277-a0c2-68663515f21f · outbound

This paper cites SVCNet: Scribble-based video colorization network with temporal aggregation.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization SVCNet: Scribble-based video colorization network with temporal aggregation

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.807116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.550664Z digest=sha256:9d3a92d212aaf24c32cfbda7235ca25c84da2e7956ed75f0f75923b5a07ec128

Observation 773da9c5-1166-422c-a53b-2ef4115014f9 · outbound

This paper cites an unresolved cited work.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-08-10T20:01:25.781763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.558212Z digest=sha256:8e42462b6771eb92cacf19a4c3aa174db9468ba09400c1bc2a1fed0a9b47111b

Observation fcd39e48-89a7-450e-95d8-6c9fd724b124 · outbound

This paper cites Negative values indicate green, and positive values indicate red.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Negative values indicate green, and positive values indicate red

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.769655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.562051Z digest=sha256:e343d233c44275ee855f98d3feaeede7bffcc07a19a6422d5922ce48b8e4d302

Observation c490a63c-23cf-46e7-b0eb-f2869b98b324 · outbound

This paper cites Red ducks.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization Red ducks

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.756254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.565723Z digest=sha256:5ae9fd8bd4a77d8268c78b922d679cdc346995c08b49edea97f7bba9ba168c54

Observation 292f0c44-6ce5-48ee-a291-ea92f228c75a · outbound

This paper cites A, including the parameter settings of the network during training and the analysis of luma channel replacement.

VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization A, including the parameter settings of the network during training and the analysis of luma channel replacement

Reference 2023

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:01:25.794282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:01:25.554501Z digest=sha256:c27566bc613f7ce2da56e8788c1cdfa4d272cc0633b72e935c4ef04f6c97e1f4

Pith citing papers

Observation 333dd639-f753-46aa-af36-f6450e2484d2 · inbound

RaBitQCache: Rotated Binary Quantization for KVCache in Long Context LLM Inference cites this paper.

RaBitQCache: Rotated Binary Quantization for KVCache in Long Context LLM Inference VanGogh: A Unified Multimodal Diffusion-based Framework for Video Colorization

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-01T06:35:29.359106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-07-01T06:34:15.790154Z digest=sha256:0143ae72149eaf9d1c27071668fa7cb49edf8f88972111ae22fa120069e3f7b1