Pith. sign in

Paper Citation Record · LEDGER

Unified Multi-Modal Latent Diffusion for Joint Subject and Text Conditional Image Generation

As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2303.09319.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2303.09319 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T17:33:24.402742Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T22:36:17.742600Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a5d99e01-48cf-4088-a13b-c0c89d107815 · inbound

PersonaCraft: Personalized and Controllable Full-Body Multi-Human Scene Generation Using Occlusion-Aware 3D-Conditioned Diffusion cites this paper.

PersonaCraft: Personalized and Controllable Full-Body Multi-Human Scene Generation Using Occlusion-Aware 3D-Conditioned Diffusion Unified Multi-Modal Latent Diffusion for Joint Subject and Text Conditional Image Generation

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-12T11:36:17.238796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:36:17.238796Z digest=sha256:9a630962d52cca575b76c4fcf287cba2553895e8d8f9d42dfb9db6332da8e508

Observation f6e471b6-1425-48f7-b38e-054d90d11343 · inbound

DreamBlend: Advancing Personalized Fine-tuning of Text-to-Image Diffusion Models cites this paper.

DreamBlend: Advancing Personalized Fine-tuning of Text-to-Image Diffusion Models Unified Multi-Modal Latent Diffusion for Joint Subject and Text Conditional Image Generation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T10:17:11.814348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:17:11.814348Z digest=sha256:3d1f3b772a01df28a6a2aebd2f8256a9f158dc237ea7ddb377d67e94f6f9ff9d

Observation 24a3f34d-8c30-4bfc-a45e-9e084766f7cc · inbound

Numerical Study of Oblique Detonation Initiation Assisted by Local Energy Deposition cites this paper.

Numerical Study of Oblique Detonation Initiation Assisted by Local Energy Deposition Unified Multi-Modal Latent Diffusion for Joint Subject and Text Conditional Image Generation

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T17:32:40.016408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:32:40.016408Z digest=sha256:6322b58334ea5bec107ca1d7209e12c95a8c07e8016771333d7635dc96f06f40

Observation 644cf613-1b0b-48b4-b6bf-5832cb843d57 · inbound

Lay2Story: Extending Diffusion Transformers for Layout-Togglable Story Generation cites this paper.

Lay2Story: Extending Diffusion Transformers for Layout-Togglable Story Generation Unified Multi-Modal Latent Diffusion for Joint Subject and Text Conditional Image Generation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T17:33:24.402742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:33:24.402742Z digest=sha256:c81ad0cf29edee508117aaf6dbc9c7d95fefde169a00776c31a095eb74cde5ae

Observation 9fd020f2-265d-4aa5-8e34-447f3d42f4ff · inbound

Per-Query Visual Concept Learning cites this paper.

Per-Query Visual Concept Learning Unified Multi-Modal Latent Diffusion for Joint Subject and Text Conditional Image Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T21:18:51.407931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:18:51.407931Z digest=sha256:b9a6ebe463349ec7dfae66c7f2e638818251e18ca861e7e130917556e27dccd6

Observation a1e344d6-d5dd-4349-9efd-9c5b58bbda1a · inbound

Adversarial Concept Distillation for One-Step Diffusion Personalization cites this paper.

Adversarial Concept Distillation for One-Step Diffusion Personalization Unified Multi-Modal Latent Diffusion for Joint Subject and Text Conditional Image Generation

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-18T04:50:53.616261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-18T04:50:01.364000Z digest=sha256:3b42abb689cbe0cca5dd44e1d840e149c0ca8dfe318cf9a72e4fda2d791453aa

Observation cf782946-a286-4bbe-854f-cf6c448d418b · inbound

Intrinsic Concept Extraction Based on Compositional Interpretability cites this paper.

Intrinsic Concept Extraction Based on Compositional Interpretability Unified Multi-Modal Latent Diffusion for Joint Subject and Text Conditional Image Generation

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T12:15:34.356784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-15T12:15:08.559058Z digest=sha256:20f0288bac33136c55da00a5ce6555b778c2e1e601b9eff5541c1a028f52f66b

Observation 8e053777-22b8-4abf-964b-ad09d4cfe36f · inbound

Equilibrated Diffusion: Frequency-aware Textual Embedding for Equilibrated Image Customization cites this paper.

Equilibrated Diffusion: Frequency-aware Textual Embedding for Equilibrated Image Customization Unified Multi-Modal Latent Diffusion for Joint Subject and Text Conditional Image Generation

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:36:17.744401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-28T15:11:20.359551Z digest=sha256:6fcdbf985d093802cdaf5ba33facdd4f217c58fa9e93bd420e7f6c774028e17f

Observation ff7a76c8-911a-424a-9428-a543670821b3 · inbound

Text-to-Image Generation for Projector-Camera System Registration cites this paper.

Text-to-Image Generation for Projector-Camera System Registration Unified Multi-Modal Latent Diffusion for Joint Subject and Text Conditional Image Generation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-12T05:14:28.636488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T05:14:28.636488Z digest=sha256:c480c163c3964761a5f6270ad046280fd347f2b5e92d74f35fc8bf0b8fb712e5