Pith. sign in

Paper Citation Record · LEDGER

Stable Diffusion Models are Secretly Good at Visual In-Context Learning

As of 10 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 0 inbound Pith citation observations for arXiv:2508.09949.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.09949 v1

Coverage vector

measured 52 of 52 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T20:45:11.810244Z

measured 52 of 52 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

52 of 52 outbound references displayed

  • verified exact0
  • verified fuzzy39
  • unresolved12
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ea0e9fe5-1447-49a4-81bf-e1c267cba23c · outbound

This paper cites Cross-image attention for zero- shot appearance transfer.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Cross-image attention for zero- shot appearance transfer

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:19.260996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:07.514302Z digest=sha256:e5cda2be461b2e7193b4c89752e5936e29617d6f1e9f40daeedcbd3ae2a988cb

Observation 4ea1ad37-8362-43e2-b517-aca312ee46cb · outbound

This paper cites Sequential modeling enables scalable learn- ing for large vision models.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Sequential modeling enables scalable learn- ing for large vision models

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:19.059029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:07.596299Z digest=sha256:ffb0d310894782216e027e85cdae7a3ea1528e1b2c09e982106402a9c8d3e194

Observation 0c78c0da-cb8a-450d-9ceb-40e3db54a21e · outbound

This paper cites Visual prompting via image inpaint- ing.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Visual prompting via image inpaint- ing

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:18.879570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:07.653433Z digest=sha256:1c30b7451e557b88ae40acf76dc01031177ddd4206d611ae1b23274dedb82b2e

Observation 279f86f4-f708-4a88-972a-89722b9a1430 · outbound

This paper cites In- structpix2pix: Learning to follow image editing instructions.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning In- structpix2pix: Learning to follow image editing instructions

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:18.714494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:07.727032Z digest=sha256:66889b7ddc5d763eedab6db2e7e77a2cbfd7029fd7d19ac48b3f323fe8481400

Observation 1e441169-98af-43af-926c-38ff8468b4f1 · outbound

This paper cites Lan- guage models are few-shot learners.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Lan- guage models are few-shot learners

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:18.540437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:07.783552Z digest=sha256:4492956b4d9f3b328c810ab645b3254f600019755f78bb48d33d50c4d62b418f

Observation ab96f64c-7e4f-47b6-bbcb-ef99c54e959f · outbound

This paper cites Masactrl: Tuning-free mu- tual self-attention control for consistent image synthesis and editing.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Masactrl: Tuning-free mu- tual self-attention control for consistent image synthesis and editing

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:18.392573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:07.843136Z digest=sha256:490f2da3b711e60747daeff8b8923048dd68dd83377f88e0061d68532e32c89e

Observation 20da4964-ea57-4fca-96e3-d14d20015449 · outbound

This paper cites Palm: Scaling language modeling with pathways.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Palm: Scaling language modeling with pathways

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:18.194866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:07.920859Z digest=sha256:39be7a038704af0411e4658e5fa8f7f9528868136e14c73e7038f26e661da842

Observation a6019d4b-ae55-475a-a205-e86bb6897692 · outbound

This paper cites The cityscapes dataset for semantic urban scene understanding.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning The cityscapes dataset for semantic urban scene understanding

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:18.005280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:07.986102Z digest=sha256:1efc09bd44563e514f3d510913d7b706e5af3d717b6565865b4ab2e43a825469

Observation a37b9586-9823-497a-b435-ef0d4c998b6f · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning An image is worth 16x16 words: Transformers for image recognition at scale

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:17.840449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:08.078082Z digest=sha256:ca559da64c32f7cf63b206cbf1d812eb8a244d7273c7528febbe9d88760fa21b

Observation fe7acafe-8409-4c76-a38a-6fefa71c9c5c · outbound

This paper cites Taming transformers for high-resolution image synthesis.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Taming transformers for high-resolution image synthesis

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:17.711926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:08.157184Z digest=sha256:b7d0ca6365c8727bb830d21307cd596e610ccd3ea52cddfe7b02273790809855

Observation 92968a93-2f4c-4d58-866e-094b743fbdac · outbound

This paper cites Explore in-context learning for 3d point cloud understanding.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Explore in-context learning for 3d point cloud understanding

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:17.523814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:08.221200Z digest=sha256:f03c957480bc3092e8d650c248e71228209aeb734dec801e072f8d8ee2beb435

Observation bdb00826-7da4-4ba0-9f21-afa5a1913ce3 · outbound

This paper cites Openllama: An open reproduc- tion of llama, 2023.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Openllama: An open reproduc- tion of llama, 2023

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:17.333844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:08.272090Z digest=sha256:10d3d11675a6c9f6fab4bce3ff818326c842dac0dfe216b3e0bb027c2aeecd18

Observation 0a76cdf3-950e-41c2-86d4-2c52527f4881 · outbound

This paper cites Language Models are General-Purpose Interfaces.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Language Models are General-Purpose Interfaces

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T20:45:08.328312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:45:08.328312Z digest=sha256:0aa9f0e9fd3d89c2a1ece1b4e869f08a413cf24d62724ee1c886033c93db0ea8

Observation 3f09f22a-d9ee-4c88-abfb-07858b1c03c5 · outbound

This paper cites Masked autoencoders are scalable vision learners.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Masked autoencoders are scalable vision learners

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:17.143042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:08.399731Z digest=sha256:f4b2261a435862fa5a4c0965e053e4d7076c10e469df3c552257dc23a4221539

Observation e22ed390-084b-44ea-91cb-b7913bfe02e1 · outbound

This paper cites Unsupervised keypoints from pretrained diffusion models.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Unsupervised keypoints from pretrained diffusion models

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:16.981249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:08.476205Z digest=sha256:502cb7c9a3aa2b6482f053e79bddb4472c3ee3039f95fa89d33008086bbff428

Observation e14efd6c-fdb3-4f5a-881e-5ddbcbf6248d · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilib- rium.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Gans trained by a two time-scale update rule converge to a local nash equilib- rium

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:16.797606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:08.585457Z digest=sha256:4e0a924efad480799a1f438332aebed9515cd636e33023c7eb6220bfec6abb01

Observation 21b0b5c6-4357-4d7c-87a8-2d63a06ef7dc · outbound

This paper cites Classifier-Free Diffusion Guidance.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Classifier-Free Diffusion Guidance

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T20:45:08.698082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:45:08.698082Z digest=sha256:494289eb03234361e07213a15a99c11822812599798758574ac9537e80af00f7

Observation f1032768-867f-4539-a491-87c2d55c1be7 · outbound

This paper cites Arbitrary style transfer in real-time with adaptive instance normalization.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Arbitrary style transfer in real-time with adaptive instance normalization

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:16.700582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:08.762078Z digest=sha256:f1a4f003e3b41ad5dcca7831582cc1a9126e3864276b347932f91d9c25faabb4

Observation 097e27bf-7126-4538-8307-7a9c084f15fc · outbound

This paper cites An edit friendly ddpm noise space: Inversion and manipulations.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning An edit friendly ddpm noise space: Inversion and manipulations

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:16.517854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:08.820187Z digest=sha256:42398b9e30acb546e3965adf3aa6e5a1d326ca37a6b33720e17538fcab767c08

Observation e65fa974-f178-4267-b48a-ee822bf7af11 · outbound

This paper cites Fss-1000: A 1000-class dataset for few- shot segmentation.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Fss-1000: A 1000-class dataset for few- shot segmentation

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:16.356034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:08.873462Z digest=sha256:53390b963cf7212e37d580c4bca1cade443d65208b143a04fc8aaefa0163aa87

Observation ff2ed0fe-3963-4cfa-acfc-fb9a47cfdbb5 · outbound

This paper cites Microsoft coco: Common objects in context.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Microsoft coco: Common objects in context

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:16.187764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:08.938454Z digest=sha256:b265dc323ed7db4a702bfb9cb4d60645290a0874f999fe748af19ee565ff7735

Observation 2d64a71a-81f3-4847-83dd-4cf2f1126cd6 · outbound

This paper cites Explicit visual prompting for low-level structure segmenta- tions.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Explicit visual prompting for low-level structure segmenta- tions

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:16.037748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:09.036719Z digest=sha256:25e558bd4c09dedf40593ca97944c84ae146ad17f859895a66ce080915536f19

Observation 616c6738-5689-4846-bde1-91441c31a463 · outbound

This paper cites Instaflow: One step is enough for high-quality diffusion- based text-to-image generation.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Instaflow: One step is enough for high-quality diffusion- based text-to-image generation

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:15.817633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:09.124066Z digest=sha256:68ed0f8bcdfbbec0100277f392bd273aecd32966d75730b7a4756c25cee5266a

Observation 49ae3865-0d90-47ca-8dd0-929f6ed843fa · outbound

This paper cites Deepfashion: Powering robust clothes recognition and retrieval with rich annotations.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Deepfashion: Powering robust clothes recognition and retrieval with rich annotations

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:15.617910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:09.199623Z digest=sha256:93cf199a0701428823782ac31fadedf1f2ad7608ad0a75e56a3dfcd578719338

Observation a7669b71-d91c-4ddb-b930-23a48fd25abb · outbound

This paper cites Localizing object-level shape variations with text-to-image diffusion models.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Localizing object-level shape variations with text-to-image diffusion models

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:15.473594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:09.275040Z digest=sha256:3506130b6a6717d34bd0b1e1580c6ac7284568e8a798dc21b3cab7e342f567f7

Observation 26bb486c-c1c8-4cf1-870b-acb41d9279cc · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Learning transferable visual models from natural language supervi- sion

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T20:45:09.388248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:45:09.388248Z digest=sha256:c2887b292d0d3caf80bbd9f10ed45fcc1797e66a9a5587dbf3e3529482e3380b

Observation 971880ba-b9bd-4a29-8caf-3a2ec646574e · outbound

This paper cites Scaling Language Models: Methods, Analysis & Insights from Training Gopher.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Scaling Language Models: Methods, Analysis & Insights from Training Gopher

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T20:45:09.443522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:45:09.443522Z digest=sha256:75a17d51da2f01b0aa712c25894479c7c198eea0edc962390e1cda218c4ab7c2

Observation 241695fd-c785-4c00-8cc9-d96f30560e88 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning High-resolution image synthesis with latent diffusion models

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:15.284617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:09.528641Z digest=sha256:f2ef84d37784bdcb8acfeef6afc247ec9194cd2e7489273d8f0c75a0a056a635

Observation d3ca2731-a12e-4c2d-8967-8821137393a8 · outbound

This paper cites Imagenet large scale visual recognition challenge.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Imagenet large scale visual recognition challenge

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:15.132089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:09.586460Z digest=sha256:d444c86cae2e874b6ee7adb1513b194e0211b8fbd31e9ba6742be4ea25d6126e

Observation fdc02c7e-29c3-4394-8e00-7d38d771fc7d · outbound

This paper cites Laion-5b: An open large-scale dataset for training next generation image-text models.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Laion-5b: An open large-scale dataset for training next generation image-text models

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:14.965749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:09.630964Z digest=sha256:0e15a17b437f2af028026851bb2ea8454280de24f2f78d5b1b7d48f9b59b4858

Observation beb072df-3064-489d-a413-a39b42bd99bb · outbound

This paper cites One-Shot Learning for Semantic Segmentation.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning One-Shot Learning for Semantic Segmentation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T20:45:09.687821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:45:09.687821Z digest=sha256:1521ab9f90e3adac6669fe2e301bd0e0acd7a79a95e07dbf54ae8047f4c0c82b

Observation c8f21ec4-1925-4b96-8948-9823c16e68d0 · outbound

This paper cites Indoor segmentation and support inference from rgbd images.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Indoor segmentation and support inference from rgbd images

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:14.831772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:09.772103Z digest=sha256:f1de92c48f7037c3eafd97912f55e41eac0a765f2066f61759ff4198709af102

Observation 2347fc21-f0c2-432f-aaba-ee081af82a43 · outbound

This paper cites LaMDA: Language Models for Dialog Applications.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning LaMDA: Language Models for Dialog Applications

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T20:45:09.834456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:45:09.834456Z digest=sha256:5ad6d9d8ffa4b8478d34376e138856df016bd1dd8614f5de068912a994af0f7d

Observation 23bf833b-62e0-4108-a269-a7dd2ae6e9c3 · outbound

This paper cites Diffuse attend and segment: Un- supervised zero-shot segmentation using stable diffusion.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Diffuse attend and segment: Un- supervised zero-shot segmentation using stable diffusion

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:14.658644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:09.900414Z digest=sha256:4cfd12e085672fa9a5a0d0060149f4ec1a6839161076ee522e9bc3ff05418411

Observation 923b9bb1-eafc-4abe-b358-85ada635f178 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning LLaMA: Open and Efficient Foundation Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-05T20:45:09.978506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:45:09.978506Z digest=sha256:47ed59b5a258efb4175614717ab6e50557b246edf0c09cbded6661c94ebadd07

Observation cb9f60f3-d965-4aa2-bb28-8fcc1c32aabf · outbound

This paper cites Plug-and-play diffusion features for text-driven image-to-image translation.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Plug-and-play diffusion features for text-driven image-to-image translation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-05T20:45:10.059580Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:45:10.059580Z digest=sha256:734eb0e88ba01ff131a6c58686a74d8d317e0a5fba836eae6c08ab506fee13b7

Observation a0d20d56-4c9f-4965-a019-5c9ec8260765 · outbound

This paper cites Attention is all you need.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Attention is all you need

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T20:45:10.129900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:45:10.129900Z digest=sha256:aa218939d8d126073e62ec74719cd46b8b2becc371367aabfa97441a61644da9

Observation 9de3235e-7d0e-4f19-8c58-7bede78fabf1 · outbound

This paper cites Images speak in images: A generalist painter for in-context visual learning.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Images speak in images: A generalist painter for in-context visual learning

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:14.460737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:10.206858Z digest=sha256:9d37c456ad1c6075fadcf56698534d6ea3075138afe5d5da1aed24f10f79cfd5

Observation d6743e03-b4d1-4a9f-ad12-21eb31c0a679 · outbound

This paper cites SegGPT: Segmenting Everything In Context.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning SegGPT: Segmenting Everything In Context

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-05T20:45:10.267368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:45:10.267368Z digest=sha256:47a67e2b70b7f4141a3771311c6764ea08f412fb9148837f3c4e9e990596a160

Observation ea9b798f-32c4-4a78-b75e-c0e1fd9b19cb · outbound

This paper cites Skeleton-in-context: Unified skeleton sequence modeling with in-context learning.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Skeleton-in-context: Unified skeleton sequence modeling with in-context learning

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:14.185823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:10.361851Z digest=sha256:baccee2334815f82902ef8bf6cc19dd1f34a24d564423dbe9d54fb2d44fb31d8

Observation e35bfc9e-0305-4407-a16f-21226894854e · outbound

This paper cites In- context learning unlocked for diffusion models.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning In- context learning unlocked for diffusion models

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:13.929754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:10.441571Z digest=sha256:dc6c2dbfd36e5f887b7d097068c68f1d55d4f1f07217281b0293f3c70083a008

Observation 01c703b2-bea8-4173-874e-c1ffd15056a0 · outbound

This paper cites Emergent abilities of large language models.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Emergent abilities of large language models

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:13.481069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:10.625430Z digest=sha256:09fe559ea436a54acac7e77efd15bd510be3acec448b381bece26529f9c3371d

Observation baa751ca-64cb-49a5-bd05-07d8de37512f · outbound

This paper cites Holistically-nested edge de- tection.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Holistically-nested edge de- tection

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:13.288223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:10.743864Z digest=sha256:fe941eb03f667a2b32e855838dd498b7b2cc641c2c34fc463d9dea730c449c95

Observation f585997b-c4f4-4bff-a58f-7d6ab1a7d75f · outbound

This paper cites IMProv: Inpainting-based Multimodal Prompting for Computer Vision Tasks.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning IMProv: Inpainting-based Multimodal Prompting for Computer Vision Tasks

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-05T20:45:10.874525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:45:10.874525Z digest=sha256:f04f6b0f5d99541dd46eab360ba9c93b61123aac4a9d79eaca8faa10df84eb83

Observation d8c56f9a-ba14-4fde-8add-e172b04bc932 · outbound

This paper cites Improved Distribution Matching Distillation for Fast Image Synthesis.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Improved Distribution Matching Distillation for Fast Image Synthesis

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-05T20:45:11.013444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:45:11.013444Z digest=sha256:771c61d13245e9a4582e3ce7d1c1c55e5b3d09ee622a6f5f2fc2a7af263c6f86

Observation 46c9168b-3274-4100-ab56-7067f4fc9e32 · outbound

This paper cites The unreasonable effectiveness of deep features as a perceptual metric.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning The unreasonable effectiveness of deep features as a perceptual metric

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:13.143015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:11.193266Z digest=sha256:50a45d0a9672f2750279e23a8aa9a5590e747daabdf7ddc9569a59d51d1610e1

Observation 8a303289-20e1-4d0a-b51c-1dc3719dbed8 · outbound

This paper cites What makes good examples for visual in-context learning? Advances in Neural Information Processing Systems, 36:17773–17794,.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning What makes good examples for visual in-context learning? Advances in Neural Information Processing Systems, 36:17773–17794,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:12.966932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:11.248935Z digest=sha256:9207e41da8339edf81994d64a0a50f52441a8c6b3265555e3cf8a029ae838423

Observation bf780141-304e-4389-9986-410db0428342 · outbound

This paper cites Semantic under- standing of scenes through the ade20k dataset.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Semantic under- standing of scenes through the ade20k dataset

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:12.841015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:11.402783Z digest=sha256:5eefa12733d711a386a886dd303bf52288763631a5b81a150d7f2bebd570ba8d

Observation 1a9829e6-0b1f-4825-9430-44fad156dbf2 · outbound

This paper cites To accommo- date the different spatial scales, we apply Gaussians with smaller variance for facial keypoints, which are relatively finer, and larger variance for body keypoints.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning To accommo- date the different spatial scales, we apply Gaussians with smaller variance for facial keypoints, which are relatively finer, and larger variance for body keypoints

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:12.718357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:11.506721Z digest=sha256:60231f8853f118a51b03a230a7b822c631ca8f2c8c047bfce24eb55f2fcb729a

Observation d4e52e8c-3d23-41ad-9ca7-58450e1771b1 · outbound

This paper cites We compute the LPIPS loss and the FID score [16] between the original colored image and the colorized prediction to evaluate the perceptual simi- larity.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning We compute the LPIPS loss and the FID score [16] between the original colored image and the colorized prediction to evaluate the perceptual simi- larity

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:12.512074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:11.654332Z digest=sha256:67f5d8be5c3ff73ab2667213ea5cfc86564ef04de7539e5b2d31b0ecafbd3e7e

Observation 405de464-d3f0-4a4b-aa9b-a4de4ef4bbe3 · outbound

This paper cites LAION5B [30]) that span annotated, unannotated, and sequence images.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning LAION5B [30]) that span annotated, unannotated, and sequence images

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:45:12.222950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:11.810244Z digest=sha256:13e761afe696874da24fdd7c5cff983c52556bf09496dfa348d67d463b3b133b

Observation baf02844-f133-45ce-9239-b9f829f1ea4e · outbound

This paper cites an unresolved cited work.

Stable Diffusion Models are Secretly Good at Visual In-Context Learning Unresolved cited work

Reference 2023

Resolution
parse uncertain
raw_fallback, observed 2026-08-05T20:45:13.707769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-05T20:45:10.502447Z digest=sha256:76e0c797078d97a63fcec04e952d9b1f81445bf36f068f0c4091352d73131420

Pith citing papers

No inbound Pith citation observations are available.