Omni-Attribute is a new open-vocabulary image attribute encoder trained on semantically linked pairs with dual objectives to produce disentangled representations for personalization and compositional generation.
Denoising diffu- sion probabilistic models
4 Pith papers cite this work. Polarity classification is still indexing.
fields
cs.CV 4representative citing papers
AIA loss teaches unified multimodal models task-specific cross-modal attention patterns to reduce conflicts between image understanding and generation without architecture decoupling.
TES applies early global alignment then iterative CLIP-guided refinement to text embeddings in Stable Diffusion to mitigate bias while preserving quality.
citing papers explorer
-
Omni-Attribute: Open-vocabulary Attribute Encoder for Visual Concept Personalization
Omni-Attribute is a new open-vocabulary image attribute encoder trained on semantically linked pairs with dual objectives to produce disentangled representations for personalization and compositional generation.
-
AIA: Rethinking Architecture Decoupling Strategy In Unified Multimodal Model
AIA loss teaches unified multimodal models task-specific cross-modal attention patterns to reduce conflicts between image understanding and generation without architecture decoupling.
-
Training-Free Debiasing of Diffusion Models via CLIP-Guided Denoising Optimization
TES applies early global alignment then iterative CLIP-guided refinement to text embeddings in Stable Diffusion to mitigate bias while preserving quality.
- Optimization-Guided Diffusion for Interactive Scene Generation