Pith. sign in

Paper Citation Record · LEDGER

Rethink Sparse Signals for Pose-guided Text-to-image Generation

As of 18 August 2026, this Paper Citation Record lists 59 of 59 outbound references and 0 inbound Pith citation observations for arXiv:2506.20983.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.20983 v1

Coverage vector

measured 59 of 59 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:40:23.917814Z

measured 59 of 59 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

59 of 59 outbound references displayed

  • verified exact3
  • verified fuzzy44
  • unresolved11
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 833bc21f-dc56-4240-95b8-ce59b7836de8 · outbound

This paper cites Spatext: Spatio-textual representation for con- trollable image generation.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Spatext: Spatio-textual representation for con- trollable image generation

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:32.798550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:18.387547Z digest=sha256:ce6b51e37dbef37fdaf68a1efdcfd400a203dc7ccc11d6f859b2918ce5d063f8

Observation 45547b5a-3975-438f-8a80-e554d2d9599b · outbound

This paper cites eDiff-I: Text-to-Image Diffusion Models with an Ensemble of Expert Denoisers.

Rethink Sparse Signals for Pose-guided Text-to-image Generation eDiff-I: Text-to-Image Diffusion Models with an Ensemble of Expert Denoisers

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T22:40:18.447091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:40:18.447091Z digest=sha256:afd1770164ca23b196f916235d1169eddcf3d313b798c361bd2e36d9a519b9c1

Observation 44218b58-17a2-47d5-b3f3-1ea80d3510a3 · outbound

This paper cites ZoeDepth: Zero-shot Transfer by Combining Relative and Metric Depth.

Rethink Sparse Signals for Pose-guided Text-to-image Generation ZoeDepth: Zero-shot Transfer by Combining Relative and Metric Depth

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T22:40:18.534281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:40:18.534281Z digest=sha256:0ef54f47066d9023a94b0aed00ee99d656258868f109a8c9d83603be71449bad

Observation 5628c2ab-30af-4872-888c-e7383a3d72a1 · outbound

This paper cites Person image synthesis via de- noising diffusion model.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Person image synthesis via de- noising diffusion model

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:32.639476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:18.624771Z digest=sha256:39fee65e6f501e8f1706e32d26d5645e1ff66497408ccbc551e7d67bcae4ec30

Observation 5ab2491b-4378-4387-9606-075d955b3969 · outbound

This paper cites Openpose: Realtime multi-person 2d pose estimation using part affinity fields.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Openpose: Realtime multi-person 2d pose estimation using part affinity fields

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:32.444330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:18.685457Z digest=sha256:e5b8a1b60ed27f9c029bcd60fe7247cb38415f04fce78a46daf3adfbeb62f29f

Observation f0ee8c9b-2a7e-4cf6-9751-6724cb5b10a3 · outbound

This paper cites Magicpose: Realistic human poses and facial expressions retargeting with identity-aware diffu- sion.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Magicpose: Realistic human poses and facial expressions retargeting with identity-aware diffu- sion

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:32.248922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:18.778458Z digest=sha256:c5e743711a8bf52d8cba7c856b18f4d89c91bccdf0ba7f039f7fa8e77f0b57b6

Observation 607853e3-52f3-44ea-b3c7-8e5ac184cf04 · outbound

This paper cites Up- gpt: Universal diffusion model for person image generation, editing and pose transfer.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Up- gpt: Universal diffusion model for person image generation, editing and pose transfer

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:32.064430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:18.884292Z digest=sha256:7e9bc22ad874d9800e73b79906312d17a5c3264a38b5a005258e54568b22125d

Observation fe2847af-3a21-4c59-a7b5-e3523d7f771d · outbound

This paper cites Openmmlab pose estimation tool- box and benchmark.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Openmmlab pose estimation tool- box and benchmark

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:31.812432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:18.965482Z digest=sha256:9a054c9fcf909259faed40a482df0f4077c411588c2f420d7b39fb3fcde8b802

Observation c22768ff-8cb5-466d-a26d-0bd141b8e2cd · outbound

This paper cites Make-a-scene: Scene- based text-to-image generation with human priors.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Make-a-scene: Scene- based text-to-image generation with human priors

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T22:40:19.108485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:40:19.108485Z digest=sha256:df3003d390336ce087cd93bdda680253cec3b4dfe749048e98c3895779fab4e9

Observation 2c05867f-a099-4f8a-9a47-ee4101208c0a · outbound

This paper cites An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual Inversion.

Rethink Sparse Signals for Pose-guided Text-to-image Generation An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual Inversion

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T22:40:19.223666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:40:19.223666Z digest=sha256:931866a5eb5fda54889feeaa2fdbafd612f8a4b6b15d0ae477cc92402846908a

Observation acac9e1e-f3f4-4ff5-b258-00547f502898 · outbound

This paper cites Controllable person image synthesis with pose-constrained latent diffusion.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Controllable person image synthesis with pose-constrained latent diffusion

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:31.564562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:19.334652Z digest=sha256:01c5b90a39d4449b0da6fa076c6354077f0520cb39c3ed48697308c92e9aa0ad

Observation 7d530aa7-e0e8-4bda-bc64-6701f675d284 · outbound

This paper cites Prompt-to-prompt image editing with cross-attention control.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Prompt-to-prompt image editing with cross-attention control

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:31.391334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:19.479717Z digest=sha256:1986e43abe304746bcd87ed1a438e8848581de26c3f749d2246aa82c2d603cf4

Observation e5d999f1-1c13-4f02-9527-39893fc807d6 · outbound

This paper cites Clipscore: A reference-free evaluation met- ric for image captioning.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Clipscore: A reference-free evaluation met- ric for image captioning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T22:40:19.644611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:40:19.644611Z digest=sha256:e09f34497e3aea0ba21602708405e1943efc2b204f50f449a6b0ef7c0915bb1b

Observation a01b8f56-54a9-47fb-8b03-71f48e7721fc · outbound

This paper cites Animate anyone: Consistent and controllable image- to-video synthesis for character animation.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Animate anyone: Consistent and controllable image- to-video synthesis for character animation

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:31.130883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:19.795100Z digest=sha256:fe07a5b1901a33400f79899662d665f44a7cb38be37abd96624ec6f6668d2cb8

Observation 6a1bccc9-3dc9-4d47-93f4-1a8cb391bf44 · outbound

This paper cites Composer: Creative and controllable im- age synthesis with composable conditions.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Composer: Creative and controllable im- age synthesis with composable conditions

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:30.925992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:19.888999Z digest=sha256:c8f49513473eb509f1173c16e638f45368015568896fd6e115618dc2f28aaade

Observation f941104b-625f-4ac8-ac21-f92395e651d6 · outbound

This paper cites Stable-pose: Leveraging transformers for pose-guided text-to-image generation.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Stable-pose: Leveraging transformers for pose-guided text-to-image generation

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:30.703108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:19.965654Z digest=sha256:17b0198338163b5ccdcc0ccc5517c15bdd71dd782117169ad84a7c3e6c1c1ea6

Observation 3cf98ae1-f372-48e2-b8d6-6706dcd805a4 · outbound

This paper cites SPAC-Net: Synthetic Pose-aware Animal ControlNet for Enhanced Pose Estimation.

Rethink Sparse Signals for Pose-guided Text-to-image Generation SPAC-Net: Synthetic Pose-aware Animal ControlNet for Enhanced Pose Estimation

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-08-06T22:40:24.462808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:20.049697Z digest=sha256:5686e90d1fd939d6174a09b2ea9ba2e746f2cbc2066d4feb5fc7f46f94187fac

Observation af2d24ef-0310-4618-a707-be84a04873cf · outbound

This paper cites Skip-and-Play: Depth-Driven Pose-Preserved Image Generation for Any Objects.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Skip-and-Play: Depth-Driven Pose-Preserved Image Generation for Any Objects

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-06T22:40:24.274129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:20.131528Z digest=sha256:64460a30f907379cf338d32839177d3fd17a5c3080ec048112c25d5a620745ba

Observation e090c3f2-84a6-4ede-b58e-7564d8eada9c · outbound

This paper cites Human-art: A versatile human-centric dataset bridg- ing natural and artificial scenes.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Human-art: A versatile human-centric dataset bridg- ing natural and artificial scenes

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:30.449937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:20.219216Z digest=sha256:051648ccccd216e1ce50a1db53f4306ea127fd373161c34fa34c61e2816266b7

Observation 2d79e33b-121f-47f2-94b5-b25e00a34248 · outbound

This paper cites Humansd: A native skeleton-guided diffusion model for human image generation.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Humansd: A native skeleton-guided diffusion model for human image generation

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:30.256036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:20.293266Z digest=sha256:73dddad50cc0619f2f3a5d4d577d4a3a8720318a27746b4d6e62dd83a511f274

Observation 25cb4b2a-f888-4c65-8186-54b21222ff5c · outbound

This paper cites Dreampose: Fashion image-to-video synthesis via stable diffusion.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Dreampose: Fashion image-to-video synthesis via stable diffusion

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:30.116455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:20.396502Z digest=sha256:ebcc9c0c8b032a9413999e675d725fed9585a2c90c0390b4a0ac585020fdbdb0

Observation 7184493b-22ca-44a7-b19d-82dd367fcff4 · outbound

This paper cites Segment anything in high quality.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Segment anything in high quality

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:29.972257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:20.512885Z digest=sha256:bb01a853179c69c8b9d0fd9143f1e39afe2e74d290e8f0669b48487c837dd687

Observation 1ec9def2-159a-4df0-891f-0413aa5e4cb8 · outbound

This paper cites Dense text-to-image generation with attention modulation.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Dense text-to-image generation with attention modulation

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:29.804186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:20.595026Z digest=sha256:17ffe36c5a2a4242dacb028979d885bf81920e6aeeeac99041ba440362201c43

Observation 607a66ca-c9bd-42e4-9e61-6732cd769cce · outbound

This paper cites DisPose: Disentangling Pose Guidance for Controllable Human Image Animation.

Rethink Sparse Signals for Pose-guided Text-to-image Generation DisPose: Disentangling Pose Guidance for Controllable Human Image Animation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T22:40:20.708669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:40:20.708669Z digest=sha256:e658db243083645cb22841b100f01f54bc73a14d0533aeee9afe26204e271af9

Observation 783e2f72-2b44-4d2e-a2bd-49810a47c9c5 · outbound

This paper cites Controlnet++: Improving conditional controls with efficient consistency feedback.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Controlnet++: Improving conditional controls with efficient consistency feedback

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:29.666948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:20.793234Z digest=sha256:1a10f3d912a2af8dda220a1e9924ff224fbdb7f02224f4c7de08e988c8a62d38

Observation f69f1314-6a64-4010-a217-9626108da35b · outbound

This paper cites ECNet: Effective Controllable Text-to-Image Diffusion Models.

Rethink Sparse Signals for Pose-guided Text-to-image Generation ECNet: Effective Controllable Text-to-Image Diffusion Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T22:40:20.850877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:40:20.850877Z digest=sha256:7b7643b4276090eea915f35c8898527562c6cb5b44cc08045fbcce2134deffdd

Observation 52b2b2fb-567d-4835-a095-42c980c5b3f2 · outbound

This paper cites Gligen: Open-set grounded text-to-image generation.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Gligen: Open-set grounded text-to-image generation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T22:40:20.931244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:40:20.931244Z digest=sha256:95ec24071a4f87150961cf5253f9d423636cd431e7d79570f214fae7bcb78bc4

Observation 119d5b4a-1c81-4081-b23b-e238c38c97ce · outbound

This paper cites Controllable text-to-3d generation via surface-aligned gaus- sian splatting.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Controllable text-to-3d generation via surface-aligned gaus- sian splatting

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:29.491203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:21.050057Z digest=sha256:00da035f13296e07fa5e0f70b3e2d62a666437063590f4aeca0c6e684fe096c3

Observation d38beb51-20af-492f-8711-9b700076ecec · outbound

This paper cites Ctrl-x: Controlling structure and appear- ance for text-to-image generation without guidance.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Ctrl-x: Controlling structure and appear- ance for text-to-image generation without guidance

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:29.354096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:21.171516Z digest=sha256:9154ded305114c7bf20452c0f29faf9439dac7fa00bd1325f6485ed31cce214f

Observation ee0095ed-0ed8-43f5-9242-aec745325c01 · outbound

This paper cites Grounding dino: Marrying dino with grounded pre-training for open-set object detection.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Grounding dino: Marrying dino with grounded pre-training for open-set object detection

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T22:40:21.255648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:40:21.255648Z digest=sha256:f47beacc284ced599f490d5620c4a9f0d61d62a519598dce8a250b4d0ad43535

Observation 13543dd2-06d3-4ca0-92a3-4747a533ad52 · outbound

This paper cites Hyperhuman: Hyper-realistic human generation with latent structural diffusion.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Hyperhuman: Hyper-realistic human generation with latent structural diffusion

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:29.191061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:21.369782Z digest=sha256:ef70df6cb64f04fdb38c1acaf3a532b4f5eea6fc83855aa18b7b6fe53472ba19

Observation f1dd4690-39e6-4a9c-9caa-9cf447c62710 · outbound

This paper cites Smartcontrol: Enhancing controlnet for handling rough visual conditions.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Smartcontrol: Enhancing controlnet for handling rough visual conditions

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:29.015061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:21.456151Z digest=sha256:4ca12d4d8001c8b27bcf767e53361b5e878822c904e06da32b60faa26e0b7d2c

Observation 993a21a4-7d32-43e1-a379-0d171d8b410b · outbound

This paper cites Controllable person image synthesis with attribute-decomposed gan.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Controllable person image synthesis with attribute-decomposed gan

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:28.843139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:21.568663Z digest=sha256:298a05453f2902b4932ee6ecaa850957f0a680a666fe53e3ed1362e1e6d05cfd

Observation 91adaafd-c272-4938-8cf1-6dde34bff1de · outbound

This paper cites Freecontrol: Training-free spatial control of any text-to-image diffusion model with any condition.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Freecontrol: Training-free spatial control of any text-to-image diffusion model with any condition

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:28.645243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:21.644746Z digest=sha256:e5e80665ad9b0ec5295d7f96102c974765234ef517ab5f3ade0c6de26ba9b512

Observation 808ad8c8-ba86-4a9d-a047-274bd831b626 · outbound

This paper cites T2i-adapter: Learning adapters to dig out more controllable ability for text-to-image diffusion models.

Rethink Sparse Signals for Pose-guided Text-to-image Generation T2i-adapter: Learning adapters to dig out more controllable ability for text-to-image diffusion models

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:28.506372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:21.732972Z digest=sha256:404e1c27e162ae076719f0aacd3dc6324644ae57afecec4a947359eec6a198cf

Observation 1d451bb1-a1f3-4a38-bf7f-618b1ac63fa6 · outbound

This paper cites Consolidating attention features for multi-view image editing.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Consolidating attention features for multi-view image editing

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:28.334312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:21.818579Z digest=sha256:84efcf4f4aacd8f4b0d68da214790b6ceb8a1b2ede4af4b917d358abd0d111ba

Observation 9c39c07c-2e35-41cf-86bd-dd14baf73fc2 · outbound

This paper cites ControlNeXt: Powerful and Efficient Control for Image and Video Generation.

Rethink Sparse Signals for Pose-guided Text-to-image Generation ControlNeXt: Powerful and Efficient Control for Image and Video Generation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T22:40:21.916192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:40:21.916192Z digest=sha256:46f645bc878124c3cfd2d5971fbdec68142569fe1b05867759014173a79e0184

Observation 0d88ba47-61e9-41bd-b7e9-1e56f189f0c0 · outbound

This paper cites Learn, imagine and create: Text-to-image generation from prior knowledge.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Learn, imagine and create: Text-to-image generation from prior knowledge

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:28.132901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:22.006626Z digest=sha256:cc2a9a30afca8c2cbda40a3ee35f18c403d85a79b8ee5c62ad14b1a0955fd819

Observation 05a64b06-8e81-48fc-85a5-3ca9b9b271ad · outbound

This paper cites Mirrorgan: Learning text-to-image generation by re- description.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Mirrorgan: Learning text-to-image generation by re- description

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:27.933999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:22.076431Z digest=sha256:c774175ad8b906f0ffe839326a3091d9ad8813ec3923e9d0c5ba964a96d1f25e

Observation 72703871-4aa1-4f6a-8c2f-94d34b17d84a · outbound

This paper cites Unicontrol: A unified diffu- sion model for controllable visual generation in the wild.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Unicontrol: A unified diffu- sion model for controllable visual generation in the wild

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:27.738273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:22.160510Z digest=sha256:a873e9343439ca874a3458f3787715b2261cd7236abd46d56672ea5f1a9c5722

Observation 1a0eead5-185a-4469-b772-310b5493210b · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Rethink Sparse Signals for Pose-guided Text-to-image Generation High-resolution image synthesis with latent diffusion models

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:27.558803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:22.267932Z digest=sha256:aa67a0753520c3790c1e006e79dd9603095001f1fc381744cdbd3e47bbf5a85f

Observation 83d0645b-7bff-4c3e-9d2a-5eccb02dc216 · outbound

This paper cites Benchmarking and error diagnosis in multi-instance pose estimation.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Benchmarking and error diagnosis in multi-instance pose estimation

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:27.374856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:22.354649Z digest=sha256:ccd5f6ff5230ff6027be83b5be222b946ba3062b13b48521203690f4d0176c10

Observation c87315b9-434f-461f-8cff-34b5f42a6f15 · outbound

This paper cites Stable Diffusion v1-5, 2022.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Stable Diffusion v1-5, 2022

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:27.151560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:22.437441Z digest=sha256:06601cb966625ec3e7a1ae46d835659b8b562adeed969c7b00f36bf01722ed75

Observation 86d5a380-274a-47dd-9cef-f2acde020253 · outbound

This paper cites Advancing pose-guided image synthesis with pro- gressive conditional diffusion models.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Advancing pose-guided image synthesis with pro- gressive conditional diffusion models

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:26.932845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:22.559546Z digest=sha256:634804f6b7553bdb267516b3e4cbe776f258617e8715b3cd81c5d98684138510

Observation 1445cae7-bdb8-4e1e-b56d-87fe324abe08 · outbound

This paper cites Deep high-resolution representation learning for human pose es- timation.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Deep high-resolution representation learning for human pose es- timation

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:26.729070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:22.670807Z digest=sha256:d166766abaf5cd714957dc6b65dc25c76adf93a2844a64b93cc08f5df4dedd5d

Observation 272e4a01-41a1-4d12-b231-ca0ffe1d8095 · outbound

This paper cites What the DAAM: Interpreting stable diffu- sion using cross attention.

Rethink Sparse Signals for Pose-guided Text-to-image Generation What the DAAM: Interpreting stable diffu- sion using cross attention

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:26.481776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:22.795614Z digest=sha256:7c23d8b25e2a4af33fc0325769607d97d39dff24c6f5747a190e70ce468d83d3

Observation 3af9e700-3e09-4412-bd5e-0140522cdec1 · outbound

This paper cites Visualizing data using t-sne.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Visualizing data using t-sne

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:26.260908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:22.889767Z digest=sha256:a3ebbd3b744fcc8be7adcd33449a1e1f9296551fbe1fc4a7a102a71d13128446

Observation b36ee586-7416-44ad-a2bf-8227b61bd322 · outbound

This paper cites AnimateZoo: Zero-shot Video Generation of Cross-Species Animation via Subject Alignment.

Rethink Sparse Signals for Pose-guided Text-to-image Generation AnimateZoo: Zero-shot Video Generation of Cross-Species Animation via Subject Alignment

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-06T22:40:24.089793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:22.979387Z digest=sha256:1f14d87d615a12e80f67b4d76a677513df62913bcf57c21585cd72a0643a2c76

Observation 8fcca519-5e3b-4860-9e55-37269ad6398a · outbound

This paper cites Vit- pose++: Vision transformer for generic body pose estima- tion.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Vit- pose++: Vision transformer for generic body pose estima- tion

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:26.127713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:23.064502Z digest=sha256:4adfefa0a25375c2421fef1da1a6722238f80efce22867bb81b356b8b8b265b4

Observation 34235a47-1d52-499f-9d9d-7f07ea86d56b · outbound

This paper cites Magicanimate: Temporally consistent human im- age animation using diffusion model.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Magicanimate: Temporally consistent human im- age animation using diffusion model

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T22:40:23.153655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:40:23.153655Z digest=sha256:8a36fb3bd988b0c770cf881a9c4a5e0874f3eae128f1f4d87515d5ca3b990ef2

Observation 6358fd66-da64-4fd1-876c-ee32e3ac8255 · outbound

This paper cites When controlnet meets inexplicit masks: A case study of controlnet on its contour- following ability.

Rethink Sparse Signals for Pose-guided Text-to-image Generation When controlnet meets inexplicit masks: A case study of controlnet on its contour- following ability

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:25.958791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:23.205566Z digest=sha256:2974fe64f2ab23f6c49fcdaaf7a353d1e762a3bf9d04628804c111e0aa781a79

Observation a06cd6da-07c0-4043-bfaf-f54265992b12 · outbound

This paper cites Grpose: Learning graph relations for human image generation with pose priors.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Grpose: Learning graph relations for human image generation with pose priors

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:25.776788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:23.282647Z digest=sha256:d41a7cb92cd3c1f872bedffb75d8612091ffd737ebcd53510891ead50b650f49

Observation 1ac9b915-db30-4958-8005-df02b151b425 · outbound

This paper cites Ap-10k: A benchmark for animal pose estima- tion in the wild.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Ap-10k: A benchmark for animal pose estima- tion in the wild

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:25.616665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:23.382631Z digest=sha256:e295e71485c80ddfa58c580e62c5217cec900085f30ddd18c6c71e30fb21ff18

Observation 9207dcfd-c897-4005-b695-294835838505 · outbound

This paper cites Pise: Person image synthesis and editing with decoupled gan.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Pise: Person image synthesis and editing with decoupled gan

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:25.458888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:23.483847Z digest=sha256:b246e0d6c795080c0d360b7ac4e3176ffae3b07b25b3429e6d3b782215e8a9cc

Observation 1ccaa88c-f7b5-463c-9e73-d35c2edba244 · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Adding conditional control to text-to-image diffusion models

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:25.283427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:23.564585Z digest=sha256:5c86e4849d7babda6aa0fb66fe071a8eeb72e52b16e0f21fbb9139130e05e216

Observation 9d89d071-f55c-404f-828f-2ad97bd5aa08 · outbound

This paper cites Uni-controlnet: All-in-one control to text-to-image diffusion models.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Uni-controlnet: All-in-one control to text-to-image diffusion models

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:25.105082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:23.662160Z digest=sha256:62bb7f95f0a11f53fe0c069df06a88c5a3daa505971b040b06894f0dc974e7eb

Observation 5208c7f9-7def-4072-8b94-6218f2966c44 · outbound

This paper cites Unipc: A unified predictor-corrector framework for fast sampling of diffusion models.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Unipc: A unified predictor-corrector framework for fast sampling of diffusion models

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:24.924657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:23.776671Z digest=sha256:30cb3d859ad6b970614466833185c3631baba9b10f530b6178b7e8ed9100394e

Observation 647fed9d-f1d4-4794-ab9e-ee636f9fe1ea · outbound

This paper cites Cross attention based style distribution for controllable person image synthesis.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Cross attention based style distribution for controllable person image synthesis

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:40:24.769862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:23.848578Z digest=sha256:8ffde6d913c4d6e13b0cedc4f30020096df37ff3160d4205985ff27e297a2828

Observation 9475ba99-a852-441d-a906-d2c102ee2fdb · outbound

This paper cites Champ: Controllable and consistent human image an- imation with 3d parametric guidance.

Rethink Sparse Signals for Pose-guided Text-to-image Generation Champ: Controllable and consistent human image an- imation with 3d parametric guidance

Reference 59

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T22:40:24.624338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T22:40:23.917814Z digest=sha256:9ada1f505a8f4a3e2328b7ff8e3bd0f063ac83fdac562098dcec3a147042041d

Pith citing papers

No inbound Pith citation observations are available.