Pith. sign in

Paper Citation Record · LEDGER

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation

As of 20 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 15 inbound Pith citation observations for arXiv:2506.21416.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.21416 v1

Coverage vector

measured 37 of 37 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:31:39.642465Z

measured 52 of 52 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 15 of 15 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T12:53:38.411642Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T13:39:50.272756Z

Reference resolution

37 of 37 outbound references displayed

  • verified exact0
  • verified fuzzy11
  • unresolved26
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e4570a16-3e24-490d-a1fc-597e0f9d9ac3 · outbound

This paper cites Generative adversarial nets.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Generative adversarial nets

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:31:42.127277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T22:31:36.471687Z digest=sha256:de859c8e64afa433a73460c3ee6de740b4535e3bf01494de64ebfb39e024e258

Observation 8fb18fec-6931-42f9-9291-8562fecf16f8 · outbound

This paper cites Auto-Encoding Variational Bayes.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Auto-Encoding Variational Bayes

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:36.583900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:36.583900Z digest=sha256:3d16208dc09083d7b6fd737d0402d30a14bf152432338eb6d8af93c8e9332f82

Observation bb5f2fd5-e91d-40eb-806e-f144cf46e618 · outbound

This paper cites Deep unsupervised learning using nonequilibrium thermodynamics.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Deep unsupervised learning using nonequilibrium thermodynamics

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:31:41.948860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T22:31:36.680516Z digest=sha256:b64ef3ee2654e60d92f8234f0c8c85ce7d9fcbfced846360a1cd60dab1ab9c24

Observation 6c70bbc0-3f6c-4b18-8bca-2d0f1c84a552 · outbound

This paper cites Denoising diffusion probabilistic models.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Denoising diffusion probabilistic models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:36.793772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:36.793772Z digest=sha256:10a2e83197d75472c8091c0f2be3d28b86bdef89a0e01a0f2def758e3c9454cc

Observation e3e54b6f-977d-43d9-a66e-d58748817776 · outbound

This paper cites Scaling rectified flow trans- formers for high-resolution image synthesis.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Scaling rectified flow trans- formers for high-resolution image synthesis

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:36.899491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:36.899491Z digest=sha256:6f0aaee14dd943899e5c92df10308847d42b990d8ad10fc51d205f5bff3164f7

Observation 4c5114f4-4c04-4c39-9bf6-7ec36c2b7f89 · outbound

This paper cites Flux: Official inference repository for flux.1 models, 2024.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Flux: Official inference repository for flux.1 models, 2024

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:36.961397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:36.961397Z digest=sha256:7a1c74f50ca68d692bf991fdd9a4cca0aa1b55d68f65039e30964634dfbc1d39

Observation 2fbe35fa-cafc-45ea-97df-c4692d0e7227 · outbound

This paper cites An image is worth one word: Personalizing text-to-image generation using textual inversion.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation An image is worth one word: Personalizing text-to-image generation using textual inversion

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:31:41.738609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T22:31:37.016227Z digest=sha256:8dba029619666fe82ef7b117374abe01a3821696d5d565b948fba5fe15d3c0d6

Observation 159babfa-0675-49ae-8b39-125559d34e9e · outbound

This paper cites Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:37.112591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:37.112591Z digest=sha256:eb6fc9715a694266cf58fcfd7763a7d57620f79db671ce16968ccb0d6593e89f

Observation f87ff64d-ad1d-4052-9bea-704555dad77c · outbound

This paper cites IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:37.221483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:37.221483Z digest=sha256:b986686270f960da0632b4738815b40a71952b8ad495e276240988cfb0e102c9

Observation 6511219b-c77d-4d58-8478-2534065bdb50 · outbound

This paper cites PhotoVerse: Tuning-Free Image Customization with Text-to-Image Diffusion Models.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation PhotoVerse: Tuning-Free Image Customization with Text-to-Image Diffusion Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:37.284527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:37.284527Z digest=sha256:483266c210d73020ca38a36886ac67ae812f538cc42ed6cb7d87d18e07cdb7aa

Observation c8f0840b-cc97-4efc-9975-c8ca26b8a315 · outbound

This paper cites OminiControl: Minimal and Universal Control for Diffusion Transformer.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation OminiControl: Minimal and Universal Control for Diffusion Transformer

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:37.369398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:37.369398Z digest=sha256:a34e252160482d29c950004da600fe5a948a3dbaac73904a741a07aee75fe371

Observation 014d59ee-aa01-4d8d-ab3c-afd2cc834df1 · outbound

This paper cites UniReal: Universal Image Generation and Editing via Learning Real-world Dynamics.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation UniReal: Universal Image Generation and Editing via Learning Real-world Dynamics

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:37.447843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:37.447843Z digest=sha256:ac6d0f310c5e8ee96d6d5c4431af5763139bd1aee3da84d521def5634a1ba23a

Observation f96b9a0a-e68f-47f6-ae2c-98ae8c6a048d · outbound

This paper cites Less-to-More Generalization: Unlocking More Controllability by In-Context Generation.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Less-to-More Generalization: Unlocking More Controllability by In-Context Generation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:37.542461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:37.542461Z digest=sha256:715c58afe9f7dff914bc5e87b2eb046e35be415fa97dfbe7a18a0c61d2757738

Observation 49598c4f-28c7-42dc-9b1b-eb3c294c4036 · outbound

This paper cites Dreamo: A unified framework for image customization.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Dreamo: A unified framework for image customization

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:37.601156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:37.601156Z digest=sha256:36fd3cb8f86ef15ff892a7a8593d0619f6c5121b8221191c40e6a451117fb625

Observation 26b22ae6-6a67-4a54-b659-2e7e3f0773df · outbound

This paper cites Scalable diffusion models with transformers.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Scalable diffusion models with transformers

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:37.643202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:37.643202Z digest=sha256:028fb37ef0e1341f49ea37f5eda24416c8b42c4d1956182d97d262af5d2b49b5

Observation d829d934-7f9a-4273-ad6c-6d7012e20903 · outbound

This paper cites A style-based generator architecture for generative adversarial networks.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation A style-based generator architecture for generative adversarial networks

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:37.696393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:37.696393Z digest=sha256:15d5593f58edc43ca353bf44463f7234ed3f174e2a5240082d66016b3285d3fb

Observation 232c08e2-2e84-487c-ab27-c5aa14d0faa5 · outbound

This paper cites Designing an encoder for stylegan image manipulation.ACM Transactions on Graphics (TOG), 40(4):1–14, 2021.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Designing an encoder for stylegan image manipulation.ACM Transactions on Graphics (TOG), 40(4):1–14, 2021

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:31:41.531963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T22:31:37.767651Z digest=sha256:1b1e798761af5f93b19cb6a30ef5368b27c0991a1d18d9f65c51e71fa937c61a

Observation 8ea6dd3d-a355-4dc3-a39b-3a54ec5f80d5 · outbound

This paper cites Ganspace: Discovering interpretable gan controls.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Ganspace: Discovering interpretable gan controls

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:31:41.322231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T22:31:37.826734Z digest=sha256:644de91cde60f7032f44f80ae976288c868e1849a71a1f113c4f12b5d3162de9

Observation 52db1cd3-ed70-4cc6-bf8c-b375d71af0f7 · outbound

This paper cites Encoding in style: a stylegan encoder for image-to-image translation.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Encoding in style: a stylegan encoder for image-to-image translation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:37.904851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:37.904851Z digest=sha256:203a7f1e5867325735d19994e49fe4dbcf07b7d06b3b12f6184719f8e3eacc39

Observation 63192005-370f-4988-bde6-26998597e266 · outbound

This paper cites Pivotal tuning for latent-based editing of real images.ACM Transactions on graphics (TOG), 42(1):1–13, 2022.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Pivotal tuning for latent-based editing of real images.ACM Transactions on graphics (TOG), 42(1):1–13, 2022

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:31:41.081831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T22:31:38.074184Z digest=sha256:3cd533288d717420dd98b7427c57d217a5e4a8183f054cf7af9eea827f30ddbf

Observation b4da9b7a-6c5f-4b42-b36f-7097e7b81195 · outbound

This paper cites In-domain gan inversion for real image editing.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation In-domain gan inversion for real image editing

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:31:40.885707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T22:31:38.218140Z digest=sha256:f556a0985be627e574e2a74e4eda828f55813c157fcaf9d7155644a5502119cf

Observation 3dab4c66-e729-46c7-87a8-dd4deb3fe3af · outbound

This paper cites Film: Visual reasoning with a general conditioning layer.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Film: Visual reasoning with a general conditioning layer

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:38.296097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:38.296097Z digest=sha256:588a266bc27728d70be3a01290d1720bbf9092ea31dd383ef3c3927e9f654d92

Observation 3c872b78-359f-4561-8a1b-f038e971f4b6 · outbound

This paper cites TokenVerse: Versatile Multi-concept Personalization in Token Modulation Space.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation TokenVerse: Versatile Multi-concept Personalization in Token Modulation Space

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:38.434050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:38.434050Z digest=sha256:f82caedc4a4b281a4f363d03ff9d3c896bda7c67132b905e03520bf34af8d627

Observation 2d18dd4d-e6a2-4c3f-9b5b-fa3871114d64 · outbound

This paper cites Learning transferable visual models from natural language supervision.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Learning transferable visual models from natural language supervision

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:38.586255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:38.586255Z digest=sha256:bf5552f408413ce2e928cbd6056d4b0fda52a9271fba160ebe44d3a4e1aeab9e

Observation c8faf2a2-c291-4ff2-aac2-50afb7b9af38 · outbound

This paper cites Flamingo: a visual language model for few-shot learning.Advances in neural information processing systems, 35: 23716–23736, 2022.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Flamingo: a visual language model for few-shot learning.Advances in neural information processing systems, 35: 23716–23736, 2022

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:38.748794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:38.748794Z digest=sha256:1a8dea803bfed12a9013676b57b04720b0dd97bda0983f10ad2c6ffea52a441b

Observation d6ccbcd1-7eec-4f9f-9816-f38436d43d8f · outbound

This paper cites Florence-2: Advancing a unified representation for a variety of vision tasks.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Florence-2: Advancing a unified representation for a variety of vision tasks

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:38.845612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:38.845612Z digest=sha256:66af2ee561550f1ac823599c8e2a27cae27e75a9450765ff5853f99c0e940ec3

Observation aa057ec5-667f-42dd-8281-30edb7310c8f · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation SAM 2: Segment Anything in Images and Videos

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:38.955842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:38.955842Z digest=sha256:0fdd032be2b3b73ed79ba3c9a6d51391ab23a5675abd797ec96bfb2243709cb2

Observation bef5d56a-ffae-4352-a18e-16f927a2d838 · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation DINOv2: Learning Robust Visual Features without Supervision

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:39.025304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:39.025304Z digest=sha256:d5a9b719129f37d4beca3d7d8fbad9a90874fbfa3e36a44a279c53782c2e3367

Observation ad9308bd-0d90-4c40-9088-81af9a9850b5 · outbound

This paper cites Lora: Low-rank adaptation of large language models.ICLR, 1 (2):3, 2022.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Lora: Low-rank adaptation of large language models.ICLR, 1 (2):3, 2022

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:39.091950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:39.091950Z digest=sha256:233f6357faf9d175740d40b2067531c08549b15c05ceaf3f4180b3269b17aed8

Observation a08f6626-0509-4dec-b39f-f9948e1c75b7 · outbound

This paper cites Dreambench++: A human-aligned benchmark for personalized image generation.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Dreambench++: A human-aligned benchmark for personalized image generation

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:31:40.710770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T22:31:39.150402Z digest=sha256:99ae3793fdb575df9e821305f9ecf51630651d601f3538bee130eef4321f7137

Observation 382fe658-2f9f-42a0-8064-1f434393b4aa · outbound

This paper cites ELLA: Equip Diffusion Models with LLM for Enhanced Semantic Alignment.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation ELLA: Equip Diffusion Models with LLM for Enhanced Semantic Alignment

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:39.198887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:39.198887Z digest=sha256:8239831b85420a63a85f7a38110aeda00e10e05ec30f3e8d5e7072b2791e615b

Observation 3678e2e0-4235-406e-9d52-98dcd8a0673c · outbound

This paper cites Arcface: Additive angular margin loss for deep face recognition.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Arcface: Additive angular margin loss for deep face recognition

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:31:40.514199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T22:31:39.249183Z digest=sha256:6d698825a614b0a7f3ac340203a533aec0f89f70af1439a1e4fb425e4c8330b1

Observation 7bec32a5-3960-47ef-9241-ca029922e028 · outbound

This paper cites Aesthetic predictor v2.5: Siglip-based aesthetic score predictor.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Aesthetic predictor v2.5: Siglip-based aesthetic score predictor

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:31:40.329806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T22:31:39.318972Z digest=sha256:7f836b871c06fa99107a4f2b27246240f2ab5aec5e9e6b5c9047f3c496216e9d

Observation cd942f24-4cfc-4ea0-995f-1cb464a3073f · outbound

This paper cites Ms-diffusion: Multi- subject zero-shot image personalization with layout guidance.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Ms-diffusion: Multi- subject zero-shot image personalization with layout guidance

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:39.412245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:39.412245Z digest=sha256:2ea577173f210ddddb243d06af2df1f932012139d62c93884d2bfc66c95c7433

Observation f4a49b79-f5dd-4167-954b-0cedaf21bb2d · outbound

This paper cites Resolving multi- condition confusion for finetuning-free personalized image generation.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation Resolving multi- condition confusion for finetuning-free personalized image generation

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:31:40.140747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T22:31:39.479358Z digest=sha256:c41b500109e6d61294bfcf33a4e2fc50724569c66d1aeec215bd3834760efbb8

Observation d52336b2-704b-48e2-9341-70681ce8c7e4 · outbound

This paper cites OmniGen: Unified Image Generation.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation OmniGen: Unified Image Generation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:39.566138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:39.566138Z digest=sha256:31a04c98cce0b3f42ae8373614fbdc739601ebe7e506bbae2c5a5b6b0a0d944a

Observation c6695ebe-3967-4611-900a-7da26f6548a4 · outbound

This paper cites OmniGen2: Towards Instruction-Aligned Multimodal Generation.

XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation OmniGen2: Towards Instruction-Aligned Multimodal Generation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:39.642465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:39.642465Z digest=sha256:ec1fd0767eada67455cecacc1c625bfae316148baa4574f3985439b6327b058b

Pith citing papers

Observation 0cf400a6-efad-4964-a0b9-38b718e0c123 · inbound

FocusDPO: Dynamic Preference Optimization for Multi-Subject Personalized Image Generation via Adaptive Focus cites this paper.

FocusDPO: Dynamic Preference Optimization for Multi-Subject Personalized Image Generation via Adaptive Focus XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T12:53:38.411642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:53:38.411642Z digest=sha256:c9080f611a709c762cf1700096735f99018253d06c9865cd6901fe495a10e0f6

Observation 9fd33aeb-613c-48d9-b8e6-543cfa57ce5f · inbound

MOSAIC: Multi-Subject Personalized Generation via Correspondence-Aware Alignment and Disentanglement cites this paper.

MOSAIC: Multi-Subject Personalized Generation via Correspondence-Aware Alignment and Disentanglement XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T12:04:28.875711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:04:28.875711Z digest=sha256:3fb95ba27902dec11435919ca2265d3d064d93d570876c94141f198553a8b33c

Observation 19927f7f-b8c2-4d30-b728-7f6573860d98 · inbound

EditIDv2: Editable ID Customization with Data-Lubricated ID Feature Integration for Text-to-Image Generation cites this paper.

EditIDv2: Editable ID Customization with Data-Lubricated ID Feature Integration for Text-to-Image Generation XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T05:19:59.576677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:19:59.576677Z digest=sha256:8ed0aeeee784533b60b65d356e52d5092d9a316dae73bbb37e23f93004e3cf7d

Observation 65abac06-99d6-4a96-99e5-8bbec1b5019b · inbound

UMO: Scaling Multi-Identity Consistency for Image Customization via Matching Reward cites this paper.

UMO: Scaling Multi-Identity Consistency for Image Customization via Matching Reward XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T23:06:09.033237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:06:09.033237Z digest=sha256:92e0b6d602716240017308b8cab4ce7bb2ae74a120f38e7b22dcd53b123b8f01

Observation 52d119ce-1930-42b7-b741-8ea8dd5fcf5d · inbound

Adversarial Concept Distillation for One-Step Diffusion Personalization cites this paper.

Adversarial Concept Distillation for One-Step Diffusion Personalization XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-18T04:50:53.769526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-18T04:50:01.364000Z digest=sha256:ec975dd624f335e962b4e9ec28fb05ce2ac27ae7fd9761bdc4392e78accccf7b

Observation 39f2e4af-cbc7-4c2a-b997-d543aab3f810 · inbound

PSR: Scaling Multi-Subject Personalized Image Generation with Pairwise Subject-Consistency Rewards cites this paper.

PSR: Scaling Multi-Subject Personalized Image Generation with Pairwise Subject-Consistency Rewards XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:51:29.393701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-17T03:49:05.489626Z digest=sha256:e5da705e2e51dc9096da0ffdfd8a97e44291ea9ab7be34925bd4d8e7d1d245d6

Observation 9d60235d-6d44-4dd0-a4d0-d9c40dc4487b · inbound

Scone: Bridging Composition and Distinction in Subject-Driven Image Generation via Unified Understanding-Generation Modeling cites this paper.

Scone: Bridging Composition and Distinction in Subject-Driven Image Generation via Unified Understanding-Generation Modeling XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-16T22:41:19.250140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-16T22:39:32.955779Z digest=sha256:d50f5b91f9fdc04b0b9efb3df0ac7fa35c2625dec16cfb87d72f624ae700c7e7

Observation 77d8101a-9bd5-4e66-97ac-aefe3bfd821c · inbound

Scone: Bridging Composition and Distinction in Subject-Driven Image Generation via Unified Understanding-Generation Modeling cites this paper.

Scone: Bridging Composition and Distinction in Subject-Driven Image Generation via Unified Understanding-Generation Modeling XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T16:37:04.207705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:37:04.207705Z digest=sha256:4b1f7993752f6e202234418ec5fff4ea169fca22d45adbcacccbeabfba1f9c54

Observation 21c8b8d1-1359-4227-9cce-028b245b611d · inbound

InstructMoLE: Instruction-Guided Mixture of Low-rank Experts for Multi-Conditional Image Generation cites this paper.

InstructMoLE: Instruction-Guided Mixture of Low-rank Experts for Multi-Conditional Image Generation XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-16T19:11:11.492007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-16T19:10:47.425041Z digest=sha256:e1022d2eadea19050f162a70e62a4b9f769fdd3c4fd9738f1ea6234b900f45a0

Observation 4b8f698b-0c51-40e3-98f9-4e44c4be5dc1 · inbound

Hierarchical Concept-to-Appearance Guidance for Multi-Subject Image Generation cites this paper.

Hierarchical Concept-to-Appearance Guidance for Multi-Subject Image Generation XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-03T05:01:41.368190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:01:41.368190Z digest=sha256:f45272fc99c1751154961aec5ce6a0fa1a00462749fb6bcbf9b52c9787da9ddb

Observation bdb99f14-0731-46ed-a3c3-425c3c4247cd · inbound

Premier: Personalized Preference Modulation with Learnable User Embedding in Text-to-Image Generation cites this paper.

Premier: Personalized Preference Modulation with Learnable User Embedding in Text-to-Image Generation XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-15T07:19:49.464581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-15T07:19:16.609341Z digest=sha256:9f35b976c440155526d46d84b18372083c3c5fa668a38a8f050e8d572d813040

Observation 0a8a5e6e-c8fa-4f6f-bb05-2a5876052da7 · inbound

Training-Free Image Editing with Visual Context Integration and Concept Alignment cites this paper.

Training-Free Image Editing with Visual Context Integration and Concept Alignment XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:05:47.713675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T20:17:34.383852Z digest=sha256:39f2bf948e7dc264c1e099b452319a8609162655307af782a5666fba3a681ddf

Observation 4768bda1-a6a7-4c3a-b08e-267986581324 · inbound

UniVerse: A Unified Modulation Framework for Segmentation-Free,Disentangled Multi-Concept Personalization cites this paper.

UniVerse: A Unified Modulation Framework for Segmentation-Free,Disentangled Multi-Concept Personalization XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:26:00.759415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T22:34:05.043186Z digest=sha256:7fbc331fcac48149d384d287ffc246a45fb34551641e3ce629bb1ddb282cb45d

Observation 72f3ce7d-6bfe-4b89-87ca-9a5d9fbfc7fb · inbound

Scaling Multi-Reference Image Generation with Dynamic Reward Optimization cites this paper.

Scaling Multi-Reference Image Generation with Dynamic Reward Optimization XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:39:50.274596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-26T05:06:12.721122Z digest=sha256:986370149da8b1a7282148df3c14dded9bce21b2d5a33e30b76571044ea1394d

Observation 494b16af-ea7a-4c42-a8ad-556cc7f39c60 · inbound

MIBE: Multi-subject Interaction Benchmark and Evaluator for Personalized Image Generation cites this paper.

MIBE: Multi-subject Interaction Benchmark and Evaluator for Personalized Image Generation XVerse: Consistent Multi-Subject Control of Identity and Semantic Attributes via DiT Modulation

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-03T21:08:57.537513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-03T21:05:29.882088Z digest=sha256:2592c5f397b70274f4f10937c281dd95727f2a185591dd6df880637d8f05509f