Pith. sign in

Paper Citation Record · LEDGER

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment

As of 13 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 0 inbound Pith citation observations for arXiv:2412.00306.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.00306 v1

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T05:35:49.359376Z

measured 30 of 30 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

30 of 30 outbound references displayed

  • verified exact2
  • verified fuzzy5
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 45a79a92-a61d-41d8-90b9-e7e618e61fe4 · outbound

This paper cites We train the model with a batch size of 192 and drop the image embedding at a rate of 0.1.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment We train the model with a batch size of 192 and drop the image embedding at a rate of 0.1

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:49.929428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T05:35:49.354097Z digest=sha256:f14fcb2d6cf15ee80246ff324aeaaee221115f71516705f9c8c5d5e9699e24fe

Observation 9f02b43c-6e70-40d3-a3e8-9d95f5b1e620 · outbound

This paper cites Vision Transformers Need Registers.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment Vision Transformers Need Registers

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.234411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.234411Z digest=sha256:fc91306df35ae801fc12c688545e138ad288dab6fe79c15053d2a47c9b30d71b

Observation ca0d8515-820f-4219-99e3-62c4b54ed134 · outbound

This paper cites An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual Inversion.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual Inversion

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.240485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.240485Z digest=sha256:a915b0bcdb30d9b9a1e9bf934cae779747a18487d4a3ef5c371e45850c6550af

Observation 09952882-1a03-4fb1-b46e-0d2cd9fd92a2 · outbound

This paper cites SwapAnything: Enabling Arbitrary Object Swapping in Personalized Visual Editing.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment SwapAnything: Enabling Arbitrary Object Swapping in Personalized Visual Editing

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.252814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.252814Z digest=sha256:70286cc1c301abdefbb280e2134bcf78140807db9615d18f94fc4f245f52d212

Observation b5cbb33e-cbc6-4312-834c-4c7e6d372533 · outbound

This paper cites COHO: Context-Sensitive City-Scale Hierarchical Urban Layout Generation.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment COHO: Context-Sensitive City-Scale Hierarchical Urban Layout Generation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.258400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.258400Z digest=sha256:e365e67b453ba93ba61afa0ba2ada844ae61407c06ec125dc1132166de5cc55c

Observation ea7f2463-4780-4487-9ebb-11fc2e68efce · outbound

This paper cites Kubrick: Multimodal Agent Collaborations for Synthetic Video Generation.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment Kubrick: Multimodal Agent Collaborations for Synthetic Video Generation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.263716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.263716Z digest=sha256:ce7d543c5190ea3349fed80f17922444f854d69fc75722c8ee2dbeaffda2829d

Observation ecdbac20-165d-4508-b0c1-0b94cdcde5d1 · outbound

This paper cites Prompt-to-Prompt Image Editing with Cross Attention Control.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment Prompt-to-Prompt Image Editing with Cross Attention Control

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.270326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.270326Z digest=sha256:ceb4b790eca7afa4beda612f069fb85c5661935029ed05c389384db1b57700b8

Observation 69d610d4-e174-4bed-9d6f-423da042533e · outbound

This paper cites Multi-concept customization of text-to-image diffusion.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment Multi-concept customization of text-to-image diffusion

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:49.983476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T05:35:49.281681Z digest=sha256:c5b336d0387f08a0c054dcf8539021de2e271a590a2fbaf0be83d5056b7df87f

Observation dda2c42b-1a2d-4700-a094-67ce62b2702b · outbound

This paper cites BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.287588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.287588Z digest=sha256:210d9cf47e4afe45383d39f8ebfe60f1e36a7ebbcb0af4982fcb3fed52877cfc

Observation a337f01a-bfa9-44ef-9e2e-db61bc6bc27e · outbound

This paper cites Unihuman: A unified model for editing human images in the wild.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment Unihuman: A unified model for editing human images in the wild

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:49.964355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T05:35:49.292725Z digest=sha256:8299b5996933af42cadc36089b5be932bd8b23a62dc1bd768663f66db1f9e508

Observation 701b4eaf-ac43-4827-a4f1-f753e81ef7db · outbound

This paper cites CliqueParcel: An Approach For Batching LLM Prompts That Jointly Optimizes Efficiency And Faithfulness.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment CliqueParcel: An Approach For Batching LLM Prompts That Jointly Optimizes Efficiency And Faithfulness

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-12T05:35:49.665549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T05:35:49.297777Z digest=sha256:b029fcf0ca90af00aa5a0b08eb060b8f7268234fb8dd77414133b7dce5d05f6b

Observation 95aa426a-84bb-4ecc-b7f6-f3b4bab6c186 · outbound

This paper cites Latent Consistency Models: Synthesizing High-Resolution Images with Few-Step Inference.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment Latent Consistency Models: Synthesizing High-Resolution Images with Few-Step Inference

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.302850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.302850Z digest=sha256:4c5943725c7a25d402ceed44b4fa7740d6f586fb9b08771c7123cfd3344b9998

Observation b19f6622-8eb3-48cf-869a-8f78ede00d03 · outbound

This paper cites Kosmos-G: Generating Images in Context with Multimodal Large Language Models.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment Kosmos-G: Generating Images in Context with Multimodal Large Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.307584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.307584Z digest=sha256:69598bdbe2c0c4866e094b9e80474f906db9f7402214ccc5f45fe3d78e9ba900

Observation 7bd69111-d137-4f0d-8427-83c770c72f45 · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.313063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.313063Z digest=sha256:cd8a7c54c72ff5ff6e2a7771dc6d92298d660123162019f40b4ad1c59d426008

Observation f3218787-7d44-4fad-9acf-36a2f85c2ecb · outbound

This paper cites Orb: An efficient alternative to sift or surf.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment Orb: An efficient alternative to sift or surf

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:49.946537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T05:35:49.317747Z digest=sha256:181c8fa86a5d3d1b59435d10052d1c9b751f2186f00376d94613b3cc8f9ed573

Observation ea202a3a-e097-472c-b5e6-7492ba2629e4 · outbound

This paper cites InstantBooth: Personalized Text-to-Image Generation without Test-Time Finetuning.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment InstantBooth: Personalized Text-to-Image Generation without Test-Time Finetuning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.327527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.327527Z digest=sha256:385b2d8a0a98c2f96b345b35460f48c81f8b87f22ae7fa99ec0e82ae4edd01a8

Observation 3593f28f-97c4-45b6-aee3-b6e3716cace5 · outbound

This paper cites Empower- ing llms with pseudo-untrimmed videos for audio-visual temporal understanding, 2024b.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment Empower- ing llms with pseudo-untrimmed videos for audio-visual temporal understanding, 2024b

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.332573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.332573Z digest=sha256:c7b2f67a4e9d7838fffb0aac8de5a4eb4a84a6be7dfbce56e98d76ef7d043202

Observation 0af59515-f522-40d6-9c5a-0ab4f95a5bbb · outbound

This paper cites GroundingBooth: Grounding Text-to-Image Customization.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment GroundingBooth: Grounding Text-to-Image Customization

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.338092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.338092Z digest=sha256:98fed182c4590dc03a1bf7df5a733a09d0d01b61719b3039f6a9564818d82163

Observation a799112a-eb4c-4033-8e53-411439055035 · outbound

This paper cites PromptFix: You Prompt and We Fix the Photo.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment PromptFix: You Prompt and We Fix the Photo

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.343195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.343195Z digest=sha256:0b448d89d85fda7aecceb10da280265283ea22310e507df32d039a7120d3947b

Observation c8779f89-33b1-4f36-bef6-f6c49fd042a0 · outbound

This paper cites LLMExplainer: Large Language Model based Bayesian Inference for Graph Explanation Generation.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment LLMExplainer: Large Language Model based Bayesian Inference for Graph Explanation Generation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.348093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.348093Z digest=sha256:01d9c3020c655067f745b06ef6ae8a37d0f9a7c9ce8b342fd5db54316ab684a5

Observation 0301323f-f1b7-455c-8eca-c6ceea8c7ff5 · outbound

This paper cites an unresolved cited work.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-12T05:35:49.912416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T05:35:49.359376Z digest=sha256:49a6f120ba891004050e3a9296e5619e4a924c987e364ff83390ba4b6f182f1d

Observation dff6126b-5035-4045-830f-27558ee014a0 · outbound

This paper cites SynArtifact: Classifying and Alleviating Artifacts in Synthetic Images via Vision-Language Model.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment SynArtifact: Classifying and Alleviating Artifacts in Synthetic Images via Vision-Language Model

Reference 2006

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.212116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.212116Z digest=sha256:3d59eb4c8d5268efa74ce11ee705e2ed9878cf13dc288d7bddaeb25c7aa63db5

Observation bf1480ec-7946-4984-ac1c-670d7bfe6848 · outbound

This paper cites HyperDreamBooth: HyperNetworks for Fast Personalization of Text-to-Image Models.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment HyperDreamBooth: HyperNetworks for Fast Personalization of Text-to-Image Models

Reference 2011

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.322696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.322696Z digest=sha256:d36c0cddc011ac4dee135ff3ecbb70823629eb65aeee09d9277dd501e3afdbe8

Observation 85720df2-7b3b-4219-968a-73fb03e7488e · outbound

This paper cites Photoswap: Personalized Subject Swapping in Images.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment Photoswap: Personalized Subject Swapping in Images

Reference 2014

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.246983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.246983Z digest=sha256:6555d8912c9547cc3a04a1ccd2d2ce7ff83deba6956e2c0381822c033754c552

Observation ea0c4b71-066b-441c-aa6e-0139cb5ccc8a · outbound

This paper cites Cross- image attention for zero-shot appearance transfer.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment Cross- image attention for zero-shot appearance transfer

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.200848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.200848Z digest=sha256:c826987aaf8b56547afa54a9d1c46ab496c6e6b7465154fc2aba2563102712f5

Observation e95f8177-7c36-4217-9b7c-e29fea831f5c · outbound

This paper cites FINEMATCH: Aspect-based Fine-grained Image and Text Mismatch Detection and Correction.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment FINEMATCH: Aspect-based Fine-grained Image and Text Mismatch Detection and Correction

Reference 2020

Resolution
verified exact
local_arxiv, observed 2026-08-12T05:35:49.707068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T05:35:49.276116Z digest=sha256:822bce378ced4fb4601aefbcbb9eae722f8c866a7b7e7321610746bcbfae399f

Observation 215b6e92-e1ad-44a2-b4fd-42ad8416616f · outbound

This paper cites Improving Diffusion Models for Authentic Virtual Try-on in the Wild.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment Improving Diffusion Models for Authentic Virtual Try-on in the Wild

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.228583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.228583Z digest=sha256:f2c2371363059d0cf926160337c9e9e6ec745c2ffae445a6d80c9b81a2ec4d78

Observation 983caa76-1d01-4e5b-9705-87edcc664404 · outbound

This paper cites Surf: Speeded up robust features.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment Surf: Speeded up robust features

Reference 2022

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:35:50.001997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T05:35:49.207143Z digest=sha256:7c1df9eccd0ef6cd50cc49a18da8a18a2c677b9547e078ef995df1334d3fa0c1

Observation 1cccc3fc-e6d1-4f31-ba60-a96a4f2463c1 · outbound

This paper cites Zero-shot Image Editing with Reference Imitation.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment Zero-shot Image Editing with Reference Imitation

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.222660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.222660Z digest=sha256:4269d6a856e2110c9b1e529120286ba7a4c16962d6c94c2e8772f2d93a1c5521

Observation eade9098-e1f0-4e26-8019-b16d720e4f17 · outbound

This paper cites AnyDoor: Zero-shot Object-level Image Customization.

Refine-by-Align: Reference-Guided Artifacts Refinement through Semantic Alignment AnyDoor: Zero-shot Object-level Image Customization

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-12T05:35:49.217312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:35:49.217312Z digest=sha256:0e00669e4fe66f4076644dc3754a3a9d93e0d536fed221114c4b0229bbd13d32

Pith citing papers

No inbound Pith citation observations are available.