Pith. sign in

Paper Citation Record · LEDGER

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention

As of 14 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 0 inbound Pith citation observations for arXiv:2412.03756.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.03756 v1

Coverage vector

measured 32 of 32 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T22:12:35.666619Z

measured 32 of 32 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

32 of 32 outbound references displayed

  • verified exact1
  • verified fuzzy18
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 39d3f69b-7061-4ea6-8533-05f76bddce74 · outbound

This paper cites Multidiffusion: Fusing diffusion paths for controlled image generation.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Multidiffusion: Fusing diffusion paths for controlled image generation

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:36.288303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:12:35.474549Z digest=sha256:bb657b450b8e11ac83311ed7b9cf6d28c738c714c6528ff7af7909cda7281900

Observation 20753695-8249-4216-98ef-5338d103ade9 · outbound

This paper cites Align your latents: High-resolution video synthesis with latent diffusion models.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Align your latents: High-resolution video synthesis with latent diffusion models

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:36.270573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:12:35.481308Z digest=sha256:874c4618ef0744e58f56d2a8dd43c562f9bfd4bc3eec671e6c3cc704a139210e

Observation cbc3f534-3694-4acf-89d0-f34c5117f017 · outbound

This paper cites Matterport3D: Learning from RGB-D Data in Indoor Environments.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Matterport3D: Learning from RGB-D Data in Indoor Environments

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T22:12:35.487268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:12:35.487268Z digest=sha256:b6482b309bb47065ca833ac49229bebea0cea085044cf7d7338bdb444aeabde8

Observation 148561cc-ef3d-4251-b88d-39e182aa6bc9 · outbound

This paper cites Attend-and-excite: Attention-based semantic guidance for text-to-image diffusion models.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Attend-and-excite: Attention-based semantic guidance for text-to-image diffusion models

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:36.252913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:12:35.493080Z digest=sha256:95b2f0f00400be295048d82f52a041704aa762843dc36dbbfd9480cecc4ebc55

Observation e9984c13-1734-44cb-9ce8-3a3f8f438455 · outbound

This paper cites Scannet: Richly-annotated 3d reconstructions of indoor scenes.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Scannet: Richly-annotated 3d reconstructions of indoor scenes

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:36.234905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:12:35.498195Z digest=sha256:46cf4ea757664b41d33667b093aaec154812c9826a06087bccb261ed50a073e7

Observation c18e978d-16ee-495e-9201-8fb602cafbad · outbound

This paper cites Preserve your own correlation: A noise prior for video diffusion models.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Preserve your own correlation: A noise prior for video diffusion models

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:36.215626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:12:35.504771Z digest=sha256:29e13f4b1da8e3443deb533b58149e7c1efe2f97e70ed85eb0039d176c22d530

Observation d26a3603-52ba-4a12-875b-104d59ca959e · outbound

This paper cites TokenFlow: Consistent Diffusion Features for Consistent Video Editing.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention TokenFlow: Consistent Diffusion Features for Consistent Video Editing

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T22:12:35.510659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:12:35.510659Z digest=sha256:38632cbab4a9acd464bd7982e5db30acff2cb9e7bb945bed72c5cb56fd96f7d9

Observation 75c44599-4033-4d1c-b3c7-278434a95748 · outbound

This paper cites Reuse and Diffuse: Iterative Denoising for Text-to-Video Generation.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Reuse and Diffuse: Iterative Denoising for Text-to-Video Generation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T22:12:35.516441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:12:35.516441Z digest=sha256:ddf606c2f9433716bb035fd06dcc216843ce124c273707c43ea22da4a168f003

Observation 31c1c96c-aaf3-42a9-b9e5-dbb831685a71 · outbound

This paper cites Prompt-to-prompt image editing with cross-attention control.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Prompt-to-prompt image editing with cross-attention control

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:36.197889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:12:35.521913Z digest=sha256:67eb7b371f01a4fbd33cbb5239dce9f39bb238e8575992e160b49c98b473d8b4

Observation ee30a7e3-2974-413b-a8c4-50a9b912f2c4 · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilibrium.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Gans trained by a two time-scale update rule converge to a local nash equilibrium

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:36.180447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:12:35.536137Z digest=sha256:a364b2abb029bc62293fb90ec20850b26b39eb00d77ef033ae4d73bd71a9473c

Observation 84b4858a-f0a8-4ca3-b55b-da51df3b0991 · outbound

This paper cites Text2Room: Extracting Textured 3D Meshes from 2D Text-to-Image Models.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Text2Room: Extracting Textured 3D Meshes from 2D Text-to-Image Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T22:12:35.541973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:12:35.541973Z digest=sha256:dc15d2ea01061ddf9e3165edd7e4e99ce0d70ccb52b93cf3be4e8494907cb227

Observation b66cda66-ffad-4dad-8780-aa0af6d4e361 · outbound

This paper cites SyncDiffusion: Coherent Montage via Synchronized Joint Diffusions.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention SyncDiffusion: Coherent Montage via Synchronized Joint Diffusions

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T22:12:35.547538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:12:35.547538Z digest=sha256:ae9d03d7a26046b46e277e48bf61b5d17469d49813c32cfc645335bc4c15f8ab

Observation 3d670d82-7b36-4d8b-8a21-072ef5c4b73e · outbound

This paper cites Common diffusion noise schedules and sample steps are flawed.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Common diffusion noise schedules and sample steps are flawed

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:36.162984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:12:35.554333Z digest=sha256:899b0f915102836f5c1544d2b919dd9b0df7dbcae053c3ce356c972c77b474de

Observation 89b6d7db-44ef-4b55-8b69-f227506b59ae · outbound

This paper cites GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T22:12:35.559616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:12:35.559616Z digest=sha256:8b5e8ddf02049d2edbfd81ed79b57082f6e9a090e808a35d104629eabf583a1b

Observation dd1e31fc-0856-4626-8d38-57889c0450f8 · outbound

This paper cites Glide: Towards photorealistic image generation and editing with text-guided diffusion models.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Glide: Towards photorealistic image generation and editing with text-guided diffusion models

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:36.142724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:12:35.565890Z digest=sha256:13b41121fbb08437f302e1daa1f8da72bc93caf3fdc5691daf0b855ab9b91c6c

Observation 64262177-2fbf-4342-8f83-526e0daa1d7e · outbound

This paper cites Zero-shot image-to-image translation.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Zero-shot image-to-image translation

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:36.120847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:12:35.571475Z digest=sha256:a798cafacbc93f593123f543749c60a955a8242e79a3aa9d4c7387a9866ff529

Observation a27b086c-d8d9-4b33-8ef3-bfe600815192 · outbound

This paper cites FreeNoise: Tuning-Free Longer Video Diffusion via Noise Rescheduling.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention FreeNoise: Tuning-Free Longer Video Diffusion via Noise Rescheduling

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T22:12:35.576188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:12:35.576188Z digest=sha256:d26b8f77d87d0f27028df85c41bdaf419da481059ed752bdc952cd37c8d4e68c

Observation 1b4dedec-6a4d-43f6-b53d-5584fd3d3d90 · outbound

This paper cites Learning transferable visual models from natural language supervision.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Learning transferable visual models from natural language supervision

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:36.101972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:12:35.581407Z digest=sha256:7963e43b0149662490301f8e91e8aa557d2039af82ef35f32ae1977abf890856

Observation 772e6e4c-179f-4e3b-b84e-67610d05d568 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T22:12:35.586714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:12:35.586714Z digest=sha256:68fc7bb8c3b4569cd2736de0f02c69de7b2386af7c8b78d7aae83930a4d27473

Observation 87cd3522-8cfd-4da0-b788-34e2ff75fd81 · outbound

This paper cites ConsistI2V: Enhancing Visual Consistency for Image-to-Video Generation.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention ConsistI2V: Enhancing Visual Consistency for Image-to-Video Generation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T22:12:35.592721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:12:35.592721Z digest=sha256:e2fddf58f0a7b97e0c13ba85b0e6b692c256dc4d951ec0068df26b974c34abe3

Observation 74fbb41d-58bf-457f-bef1-bc8d7ce28070 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention High-resolution image synthesis with latent diffusion models

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:36.082395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:12:35.598177Z digest=sha256:3a33906e74a1e6f8932b0fe7a6f70f67501e2736e10e3c37f6dd23b8339a20e9

Observation 45bcd2a0-035d-4185-8fe7-2cf23b120b84 · outbound

This paper cites U-net: Convolutional networks for biomedical image segmentation.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention U-net: Convolutional networks for biomedical image segmentation

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:36.064762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:12:35.603873Z digest=sha256:69d711ae4bf3106bfcdace32b9578c19b9ab857e3979a9c82221afa59d72270e

Observation eacd56c9-fd3f-428c-9f54-30203b667688 · outbound

This paper cites Photorealistic text-to-image diffusion models with deep language understand- ing.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Photorealistic text-to-image diffusion models with deep language understand- ing

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:36.046023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:12:35.609718Z digest=sha256:c183ab2089f14a00afddd49a6c1d0da8e44e8ed1aad5da20946f092df3c72d38

Observation deeba151-656f-4c7b-a922-97b32852d700 · outbound

This paper cites MVDiffusion: Enabling Holistic Multi-view Image Generation with Correspondence-Aware Diffusion.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention MVDiffusion: Enabling Holistic Multi-view Image Generation with Correspondence-Aware Diffusion

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T22:12:35.614836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:12:35.614836Z digest=sha256:16e274eb4a17f381d4579d411125e536d58eaad43419ed32b59594ffc31cb4f1

Observation a3b4e943-c53e-4723-b434-998baaf67da0 · outbound

This paper cites Attention is all you need.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Attention is all you need

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T22:12:35.621290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:12:35.621290Z digest=sha256:57f7e64e6a419baf7f5d09be1bb5b4cf2e88d4db028c0c1847e81a54cdce6ba3

Observation 84d09a35-daaf-4de3-8419-37e3430e031b · outbound

This paper cites Diffusers: State-of-the-art diffusion models.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Diffusers: State-of-the-art diffusion models

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:36.016032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:12:35.628239Z digest=sha256:0c907a9590c16a67eb02f0eaf9e90066a234d631a583fec4aa0396868f3a8f96

Observation 70fa4d08-60f6-491f-ab08-577ef7cc2513 · outbound

This paper cites FreeInit: Bridging Initialization Gap in Video Diffusion Models.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention FreeInit: Bridging Initialization Gap in Video Diffusion Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T22:12:35.634946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:12:35.634946Z digest=sha256:5ad949f65b6f9662e9fd3b7cc02ab2a5e59fa37ed780476031b282014aff797f

Observation 470ca0e1-6415-4b6b-9731-4df782bfd3a3 · outbound

This paper cites Freestyle layout-to-image synthesis.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Freestyle layout-to-image synthesis

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:35.998158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:12:35.644795Z digest=sha256:2ca7072919a565bb0a6d57d173d144e9fc8801344adf6216b2d7a1673f3a34bd

Observation 5268a08c-3e08-42de-9f0a-baf32f2e8282 · outbound

This paper cites Preserving image properties through initializations in diffusion models.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Preserving image properties through initializations in diffusion models

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:35.979928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:12:35.650107Z digest=sha256:13e47a23855be73512833d1439df68378f60038cc40a4029636aef0b834334c8

Observation 5f4c8e9a-e8c1-4e23-bcbe-51b0ff19d1d6 · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Adding conditional control to text-to-image diffusion models

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:35.961042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:12:35.655501Z digest=sha256:1b3dee04c3d4eafde4a44d5c35099aaa9e95d4449395b625c77f85bdce4b4852

Observation 3381c663-2fcc-4a53-996b-2d7611e8075c · outbound

This paper cites DiffCollage: Parallel Generation of Large Content with Diffusion Models.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention DiffCollage: Parallel Generation of Large Content with Diffusion Models

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-08-11T22:12:35.718932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T22:12:35.661069Z digest=sha256:6a2290c98c71a538caef0056ed16753e90fd75ee8e0778db3b2b061a1f17f92c

Observation ddbc1c79-3a4d-4fb4-a32c-2dd18097e10d · outbound

This paper cites The unreasonable effectiveness of deep features as a perceptual metric.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention The unreasonable effectiveness of deep features as a perceptual metric

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T22:12:35.666619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:12:35.666619Z digest=sha256:07bea0fa419625fbc10219088b7396657f9a2274364fab25df5b004caf3f806f

Pith citing papers

No inbound Pith citation observations are available.