Pith. sign in

Paper Citation Record · LEDGER

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation

As of 8 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 15 inbound Pith citation observations for arXiv:2506.21416.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.21416 v1

Coverage vector

measured 37 of 37 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:31:39.642465Z

measured 52 of 52 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 15 of 15 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T12:53:38.411642Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T13:39:50.272756Z

Reference resolution

37 of 37 outbound references displayed

  • verified exact0
  • verified fuzzy11
  • unresolved26
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e4570a16-3e24-490d-a1fc-597e0f9d9ac3 · outbound

This paper cites Generative adversarial nets.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Generative adversarial nets

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:31:42.127277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:31:36.471687Z digest=sha256:d49d56eab1214e51079408e4a5a3b08bcae7b5c01094606584737f09d70952ae

Observation 8fb18fec-6931-42f9-9291-8562fecf16f8 · outbound

This paper cites Auto-Encoding Variational Bayes.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Auto-Encoding Variational Bayes

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:36.583900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:36.583900Z digest=sha256:7f07acf707ab3b4b188969765b5054ad66d18f3bf382b94062d1bbc59f5268d7

Observation bb5f2fd5-e91d-40eb-806e-f144cf46e618 · outbound

This paper cites Deep unsupervised learning using nonequilibrium thermodynamics.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Deep unsupervised learning using nonequilibrium thermodynamics

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:31:41.948860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:31:36.680516Z digest=sha256:b48e3dee3af66204c8db8e83f8f4064f2d06aad84b6713c41a012a5ff118c97a

Observation 6c70bbc0-3f6c-4b18-8bca-2d0f1c84a552 · outbound

This paper cites Denoising diffusion probabilistic models.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Denoising diffusion probabilistic models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:36.793772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:36.793772Z digest=sha256:471d3ad401ee13b7ab8c3f3a7e41584d823bf4531d86bd3ec79873ffce7577af

Observation e3e54b6f-977d-43d9-a66e-d58748817776 · outbound

This paper cites Scaling rectified flow trans- formers for high-resolution image synthesis.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Scaling rectified flow trans- formers for high-resolution image synthesis

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:36.899491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:36.899491Z digest=sha256:568f4b71c3c4257f94c02fa7a1f006e599824529ee081c68af67a7e783d2ef4e

Observation 4c5114f4-4c04-4c39-9bf6-7ec36c2b7f89 · outbound

This paper cites Flux: Official inference repository for flux.1 models, 2024.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Flux: Official inference repository for flux.1 models, 2024

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:36.961397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:36.961397Z digest=sha256:a6a69e7977f762e8dedce42f0a538eb6b8999ce0414808ffd73f3365938f5b9f

Observation 2fbe35fa-cafc-45ea-97df-c4692d0e7227 · outbound

This paper cites An image is worth one word: Personalizing text-to-image generation using textual inversion.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation An image is worth one word: Personalizing text-to-image generation using textual inversion

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:31:41.738609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:31:37.016227Z digest=sha256:6f22e6737c328740821eb48a81a0e8b44fb0a073e75b9bce027efd45646a10a7

Observation 159babfa-0675-49ae-8b39-125559d34e9e · outbound

This paper cites Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:37.112591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:37.112591Z digest=sha256:38ab16a9588d128215bd8f672230fda187de75a246f66fad70ee589ef733a4be

Observation f87ff64d-ad1d-4052-9bea-704555dad77c · outbound

This paper cites IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:37.221483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:37.221483Z digest=sha256:5e7296e29913ccc0a6c76a1db2a4f6898fb7ab0b2e7f971fd0127c3315b5e12d

Observation 6511219b-c77d-4d58-8478-2534065bdb50 · outbound

This paper cites PhotoVerse: Tuning-Free Image Customization with Text-to-Image Diffusion Models.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation PhotoVerse: Tuning-Free Image Customization with Text-to-Image Diffusion Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:37.284527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:37.284527Z digest=sha256:dc59c611ffe64bc340f82abe6273f5be8ad09fd4143d0e775ab8d7ad53b1d3b0

Observation c8f0840b-cc97-4efc-9975-c8ca26b8a315 · outbound

This paper cites OminiControl: Minimal and Universal Control for Diffusion Transformer.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation OminiControl: Minimal and Universal Control for Diffusion Transformer

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:37.369398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:37.369398Z digest=sha256:6b1ef632348a80735b20107a4de35422c5c4706d925bf8fe66bedc66f63ab5cb

Observation 014d59ee-aa01-4d8d-ab3c-afd2cc834df1 · outbound

This paper cites UniReal: Universal Image Generation and Editing via Learning Real-world Dynamics.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation UniReal: Universal Image Generation and Editing via Learning Real-world Dynamics

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:37.447843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:37.447843Z digest=sha256:47388ec0d17cce21ebea9b83752a24fd9dd3d6b50cff7ea1c8d14e8f4b4f100c

Observation f96b9a0a-e68f-47f6-ae2c-98ae8c6a048d · outbound

This paper cites Less-to-More Generalization: Unlocking More Controllability by In-Context Generation.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Less-to-More Generalization: Unlocking More Controllability by In-Context Generation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:37.542461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:37.542461Z digest=sha256:2ade0e3198f21528aa535f860414b8a5727894b02560239ed8cfeaaa01658d81

Observation 49598c4f-28c7-42dc-9b1b-eb3c294c4036 · outbound

This paper cites Dreamo: A unified framework for image customization.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Dreamo: A unified framework for image customization

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:37.601156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:37.601156Z digest=sha256:92a3737f2b89f718e1650ea94bfa4b314e073b05c40432b14192330d6670b1f5

Observation 26b22ae6-6a67-4a54-b659-2e7e3f0773df · outbound

This paper cites Scalable diffusion models with transformers.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Scalable diffusion models with transformers

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:37.643202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:37.643202Z digest=sha256:999af25ba4627365df3792ccfee53962c2a62b4b3353b74cab3631f5f4dac261

Observation d829d934-7f9a-4273-ad6c-6d7012e20903 · outbound

This paper cites A style-based generator architecture for generative adversarial networks.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation A style-based generator architecture for generative adversarial networks

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:37.696393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:37.696393Z digest=sha256:992629647011109ebb5a01b52f6f27bac57faabcb6eb2a4993de9f08c580c1e6

Observation 232c08e2-2e84-487c-ab27-c5aa14d0faa5 · outbound

This paper cites Designing an encoder for stylegan image manipulation.ACM Transactions on Graphics (TOG), 40(4):1–14, 2021.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Designing an encoder for stylegan image manipulation.ACM Transactions on Graphics (TOG), 40(4):1–14, 2021

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:31:41.531963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:31:37.767651Z digest=sha256:b45e933dc7d933e770d91d28d1fecd14cb0931b070db9960484f0bdeceb02460

Observation 8ea6dd3d-a355-4dc3-a39b-3a54ec5f80d5 · outbound

This paper cites Ganspace: Discovering interpretable gan controls.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Ganspace: Discovering interpretable gan controls

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:31:41.322231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:31:37.826734Z digest=sha256:8bda7962c20b6fe43a669cfbc61bec0d80484c7b58d99db8dd4f18896ffef120

Observation 52db1cd3-ed70-4cc6-bf8c-b375d71af0f7 · outbound

This paper cites Encoding in style: a stylegan encoder for image-to-image translation.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Encoding in style: a stylegan encoder for image-to-image translation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:37.904851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:37.904851Z digest=sha256:5f2acc1fe3c7cff26119117a9e3a33178ca946be88cb13fced3ab1e395c343d1

Observation 63192005-370f-4988-bde6-26998597e266 · outbound

This paper cites Pivotal tuning for latent-based editing of real images.ACM Transactions on graphics (TOG), 42(1):1–13, 2022.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Pivotal tuning for latent-based editing of real images.ACM Transactions on graphics (TOG), 42(1):1–13, 2022

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:31:41.081831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:31:38.074184Z digest=sha256:ead5d66a5538ea2829cf97d030f4001d9c5b2818b3e953f214a481e173e56d92

Observation b4da9b7a-6c5f-4b42-b36f-7097e7b81195 · outbound

This paper cites In-domain gan inversion for real image editing.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation In-domain gan inversion for real image editing

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:31:40.885707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:31:38.218140Z digest=sha256:205c9d4c9de98fe7454a27f4f74eab9f31846ca81def7b0a131e500bec196474

Observation 3dab4c66-e729-46c7-87a8-dd4deb3fe3af · outbound

This paper cites Film: Visual reasoning with a general conditioning layer.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Film: Visual reasoning with a general conditioning layer

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:38.296097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:38.296097Z digest=sha256:29541f2c4f29ebbe5e0aa3af1bd53eb6aee0b1ce14548a0313e1284b0db3c868

Observation 3c872b78-359f-4561-8a1b-f038e971f4b6 · outbound

This paper cites TokenVerse: Versatile Multi-concept Personalization in Token Modulation Space.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation TokenVerse: Versatile Multi-concept Personalization in Token Modulation Space

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:38.434050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:38.434050Z digest=sha256:d046d154444ae1b77541f9d023f74263ad19e1b6affec0065da534df066e752d

Observation 2d18dd4d-e6a2-4c3f-9b5b-fa3871114d64 · outbound

This paper cites Learning transferable visual models from natural language supervision.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Learning transferable visual models from natural language supervision

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:38.586255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:38.586255Z digest=sha256:02c554b50d47bd320fc72d901aa0eab4d13a6b0bd7eb198cf38fd2faf33c44b1

Observation c8faf2a2-c291-4ff2-aac2-50afb7b9af38 · outbound

This paper cites Flamingo: a visual language model for few-shot learning.Advances in neural information processing systems, 35: 23716–23736, 2022.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Flamingo: a visual language model for few-shot learning.Advances in neural information processing systems, 35: 23716–23736, 2022

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:38.748794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:38.748794Z digest=sha256:b1354853441bb7b443344df546bd1cd8d4e0bf726406613477ee42ef2fa29f38

Observation d6ccbcd1-7eec-4f9f-9816-f38436d43d8f · outbound

This paper cites Florence-2: Advancing a unified representation for a variety of vision tasks.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Florence-2: Advancing a unified representation for a variety of vision tasks

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:38.845612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:38.845612Z digest=sha256:d5eb368b3727feed47313b0785d855427d0489d5fde31c050bd03dc50cb7814e

Observation aa057ec5-667f-42dd-8281-30edb7310c8f · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation SAM 2: Segment Anything in Images and Videos

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:38.955842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:38.955842Z digest=sha256:d5cd8b6afe21833f2911bcf98a13748de178cfefa0400b0c75937a8ac04b6f96

Observation bef5d56a-ffae-4352-a18e-16f927a2d838 · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation DINOv2: Learning Robust Visual Features without Supervision

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:39.025304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:39.025304Z digest=sha256:e1c74d973c40e806412d94b39a402ca18293ed8d85a617deb71961687505645f

Observation ad9308bd-0d90-4c40-9088-81af9a9850b5 · outbound

This paper cites Lora: Low-rank adaptation of large language models.ICLR, 1 (2):3, 2022.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Lora: Low-rank adaptation of large language models.ICLR, 1 (2):3, 2022

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:39.091950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:39.091950Z digest=sha256:3d737739589cc7138111b0d2eeb329703aea7049c05b5d618f107f4190a90062

Observation a08f6626-0509-4dec-b39f-f9948e1c75b7 · outbound

This paper cites Dreambench++: A human-aligned benchmark for personalized image generation.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Dreambench++: A human-aligned benchmark for personalized image generation

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:31:40.710770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:31:39.150402Z digest=sha256:bc3858e1e238ab20b07e1745af0bcd6e8e80a234a03be23220c2578ef93ae416

Observation 382fe658-2f9f-42a0-8064-1f434393b4aa · outbound

This paper cites ELLA: Equip Diffusion Models with LLM for Enhanced Semantic Alignment.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation ELLA: Equip Diffusion Models with LLM for Enhanced Semantic Alignment

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:39.198887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:39.198887Z digest=sha256:f72e787e5f3a07cc16dd53f61e6f59dc27a1ebbb0f829a6cbba2ba01fd919d06

Observation 3678e2e0-4235-406e-9d52-98dcd8a0673c · outbound

This paper cites Arcface: Additive angular margin loss for deep face recognition.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Arcface: Additive angular margin loss for deep face recognition

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:31:40.514199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:31:39.249183Z digest=sha256:f0b6b978be667bf8c70ce71f69d760383864ec6b1a6695cb7f0b54cae2da375c

Observation 7bec32a5-3960-47ef-9241-ca029922e028 · outbound

This paper cites Aesthetic predictor v2.5: Siglip-based aesthetic score predictor.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Aesthetic predictor v2.5: Siglip-based aesthetic score predictor

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:31:40.329806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:31:39.318972Z digest=sha256:75a65a850104fcedfa8e2ae6d0fe52c49d4da1a538112505f4481a410be25ed9

Observation cd942f24-4cfc-4ea0-995f-1cb464a3073f · outbound

This paper cites Ms-diffusion: Multi- subject zero-shot image personalization with layout guidance.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Ms-diffusion: Multi- subject zero-shot image personalization with layout guidance

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:39.412245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:39.412245Z digest=sha256:efbe10849c7a91936a780d5ee1caaef8c3845ce4429a1df7b71824f3f9bea5b0

Observation f4a49b79-f5dd-4167-954b-0cedaf21bb2d · outbound

This paper cites Resolving multi- condition confusion for finetuning-free personalized image generation.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Resolving multi- condition confusion for finetuning-free personalized image generation

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:31:40.140747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T22:31:39.479358Z digest=sha256:21078ac9ebc8eacfcd74a917c1884a6228b7d2757ed8d440aefd2c31c0dfa9ba

Observation d52336b2-704b-48e2-9341-70681ce8c7e4 · outbound

This paper cites OmniGen: Unified Image Generation.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation OmniGen: Unified Image Generation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:39.566138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:39.566138Z digest=sha256:6239cf7603ac5ae1e03a06e147ed4eba84aa0742bb20bb8ffba087dbdccfe6a9

Observation c6695ebe-3967-4611-900a-7da26f6548a4 · outbound

This paper cites OmniGen2: Towards Instruction-Aligned Multimodal Generation.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation OmniGen2: Towards Instruction-Aligned Multimodal Generation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:39.642465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:39.642465Z digest=sha256:d5e58c5960d369bf2644a074bb3a9d9ff0defadb1118cdd943bed29b421374b9

Pith citing papers

Observation 0cf400a6-efad-4964-a0b9-38b718e0c123 · inbound

FocusDPO: Dynamic Preference Optimization for Multi-Subject Personalized Image Generation via Adaptive Focus cites this paper.

FocusDPO: Dynamic Preference Optimization for Multi-Subject Personalized Image Generation via Adaptive Focus XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T12:53:38.411642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:53:38.411642Z digest=sha256:0f95741044f3571748e63540aaec87e4a13508bd6a23e091c3015040e0894292

Observation 9fd33aeb-613c-48d9-b8e6-543cfa57ce5f · inbound

MOSAIC: Multi-Subject Personalized Generation via Correspondence-Aware Alignment and Disentanglement cites this paper.

MOSAIC: Multi-Subject Personalized Generation via Correspondence-Aware Alignment and Disentanglement XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T12:04:28.875711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:04:28.875711Z digest=sha256:3d842d25e965d2d0632ef85a879646faa4e3effa89cc222cee8e107aa9f247c8

Observation 19927f7f-b8c2-4d30-b728-7f6573860d98 · inbound

EditIDv2: Editable ID Customization with Data-Lubricated ID Feature Integration for Text-to-Image Generation cites this paper.

EditIDv2: Editable ID Customization with Data-Lubricated ID Feature Integration for Text-to-Image Generation XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T05:19:59.576677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:19:59.576677Z digest=sha256:26ee71a9305d760ebac83fae528cefa35a466b61e95bfecfc7a3fc16193b0ca2

Observation 65abac06-99d6-4a96-99e5-8bbec1b5019b · inbound

UMO: Scaling Multi-Identity Consistency for Image Customization via Matching Reward cites this paper.

UMO: Scaling Multi-Identity Consistency for Image Customization via Matching Reward XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T23:06:09.033237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:06:09.033237Z digest=sha256:0115d08cc78719986f5c6a4570dfe9134734289a59712d3e544b3362b3728913

Observation 52d119ce-1930-42b7-b741-8ea8dd5fcf5d · inbound

Adversarial Concept Distillation for One-Step Diffusion Personalization cites this paper.

Adversarial Concept Distillation for One-Step Diffusion Personalization XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-18T04:50:53.769526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T04:50:01.364000Z digest=sha256:7343bc9ca86d997dcb492373e624da8c3d26aeda13be98241277b8bfbaea9695

Observation 39f2e4af-cbc7-4c2a-b997-d543aab3f810 · inbound

PSR: Scaling Multi-Subject Personalized Image Generation with Pairwise Subject-Consistency Rewards cites this paper.

PSR: Scaling Multi-Subject Personalized Image Generation with Pairwise Subject-Consistency Rewards XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:51:29.393701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T03:49:05.489626Z digest=sha256:786b007e82cbf1325f486857cff5974b68d69c4c4480d25c0a296da983101897

Observation 9d60235d-6d44-4dd0-a4d0-d9c40dc4487b · inbound

Scone: Bridging Composition and Distinction in Subject-Driven Image Generation via Unified Understanding-Generation Modeling cites this paper.

Scone: Bridging Composition and Distinction in Subject-Driven Image Generation via Unified Understanding-Generation Modeling XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-16T22:41:19.250140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T22:39:32.955779Z digest=sha256:0ecb4006d8c83b2ce0736131e9b3723a72e82744fed875676ba5f132904989f9

Observation 77d8101a-9bd5-4e66-97ac-aefe3bfd821c · inbound

Scone: Bridging Composition and Distinction in Subject-Driven Image Generation via Unified Understanding-Generation Modeling cites this paper.

Scone: Bridging Composition and Distinction in Subject-Driven Image Generation via Unified Understanding-Generation Modeling XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T16:37:04.207705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:37:04.207705Z digest=sha256:684b57b281ff54f4054c36c9568056af5cd36d59c432bfa70a5fccd5580014b1

Observation 21c8b8d1-1359-4227-9cce-028b245b611d · inbound

InstructMoLE: Instruction-Guided Mixture of Low-rank Experts for Multi-Conditional Image Generation cites this paper.

InstructMoLE: Instruction-Guided Mixture of Low-rank Experts for Multi-Conditional Image Generation XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-16T19:11:11.492007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T19:10:47.425041Z digest=sha256:8bbd9670753a32502c065f9c2a90b23d28711046a5370c290a594859a0fa78e1

Observation 4b8f698b-0c51-40e3-98f9-4e44c4be5dc1 · inbound

Hierarchical Concept-to-Appearance Guidance for Multi-Subject Image Generation cites this paper.

Hierarchical Concept-to-Appearance Guidance for Multi-Subject Image Generation XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-03T05:01:41.368190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:01:41.368190Z digest=sha256:f9890419d644f98a6ca36051a7f7bcd6e79418e7bf3b587c2844adb7b53ae870

Observation bdb99f14-0731-46ed-a3c3-425c3c4247cd · inbound

Premier: Personalized Preference Modulation with Learnable User Embedding in Text-to-Image Generation cites this paper.

Premier: Personalized Preference Modulation with Learnable User Embedding in Text-to-Image Generation XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-15T07:19:49.464581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T07:19:16.609341Z digest=sha256:6d96ff94fb9c0629809994674c4dd5909b474d5c832a015a3e3de071c2c18ca2

Observation 0a8a5e6e-c8fa-4f6f-bb05-2a5876052da7 · inbound

Training-Free Image Editing with Visual Context Integration and Concept Alignment cites this paper.

Training-Free Image Editing with Visual Context Integration and Concept Alignment XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:05:47.713675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T20:17:34.383852Z digest=sha256:0695dd0f186a4f2d7bc52d35052852c5fedbee54a9c4268e614d80134136b9df

Observation 4768bda1-a6a7-4c3a-b08e-267986581324 · inbound

UniVerse: A Unified Modulation Framework for Segmentation-Free,Disentangled Multi-Concept Personalization cites this paper.

UniVerse: A Unified Modulation Framework for Segmentation-Free,Disentangled Multi-Concept Personalization XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:26:00.759415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T22:34:05.043186Z digest=sha256:94d778727beb0cbdac08b149f2f1180392bcb9f79b5734d16d4bac13a70fab9d

Observation 72f3ce7d-6bfe-4b89-87ca-9a5d9fbfc7fb · inbound

Scaling Multi-Reference Image Generation with Dynamic Reward Optimization cites this paper.

Scaling Multi-Reference Image Generation with Dynamic Reward Optimization XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:39:50.274596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T05:06:12.721122Z digest=sha256:14df5129d631ede8abf632f5f193a23ad0531fb171d08f0727431c869ed19b0e

Observation 494b16af-ea7a-4c42-a8ad-556cc7f39c60 · inbound

MIBE: Multi-subject Interaction Benchmark and Evaluator for Personalized Image Generation cites this paper.

MIBE: Multi-subject Interaction Benchmark and Evaluator for Personalized Image Generation XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-03T21:08:57.537513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-03T21:05:29.882088Z digest=sha256:3e1c931a0ba522ecae8c43e2e3342f2faa2069b64ecade3be2a94e4237a8aad7