Pith. sign in

Paper Citation Record · LEDGER

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC

As of 16 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 0 inbound Pith citation observations for arXiv:2412.05619.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.05619 v1

Coverage vector

measured 55 of 55 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T20:35:35.506601Z

measured 55 of 55 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

55 of 55 outbound references displayed

  • verified exact2
  • verified fuzzy29
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation db3a6920-fcbc-4a6f-a2c8-c679f3d23b69 · outbound

This paper cites InstructPix2Pix: Learning to Follow Image Editing Instructions.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC InstructPix2Pix: Learning to Follow Image Editing Instructions

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T20:35:35.170866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:35:35.170866Z digest=sha256:e7025e24c740da3c6a0f46041caeaa8ba2fe4e3f39539bc185c7c1ee86f4d25b

Observation 5cfbc285-3400-4de8-a09c-fd3450bd1ba7 · outbound

This paper cites In- structpix2pix: Learning to follow image editing instructions.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC In- structpix2pix: Learning to follow image editing instructions

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T20:35:35.177165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:35:35.177165Z digest=sha256:abfcfaa65917a77a9b0b2419dc95c6ae41b941d1a959ff5a01d5b3a51180e7c1

Observation 0e918aec-f3a5-4939-813e-1ae817d7f511 · outbound

This paper cites Lan- guage models are few-shot learners.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Lan- guage models are few-shot learners

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:35:36.880607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T20:35:35.182726Z digest=sha256:83b6d4daa5a40aaf138fc228ed011a5cb189af258b16cbf47d853bb98f0b6115

Observation 78e2705d-4e39-4de0-bbc3-eab48dccc3ab · outbound

This paper cites Viton-hd: High-resolution virtual try-on via misalignment-aware normalization.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Viton-hd: High-resolution virtual try-on via misalignment-aware normalization

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:35:36.860548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T20:35:35.187902Z digest=sha256:96b85e8654cddaf5c16034fb1f53c5e63d0c717ff5a54075c70916bc2cd2d857

Observation fa7efd9e-8ddd-40f9-a6c8-062afb7570c6 · outbound

This paper cites Diffusion models beat gans on image synthesis.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Diffusion models beat gans on image synthesis

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T20:35:35.194037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:35:35.194037Z digest=sha256:646cc0aaae5c8e9105adf460fc29984d174f5396a62e753f6bb6d79850f507df

Observation 8181d36a-a2cf-47f6-aefc-63c455ff987a · outbound

This paper cites Make-a-scene: Scene- based text-to-image generation with human priors.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Make-a-scene: Scene- based text-to-image generation with human priors

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:35:36.821511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T20:35:35.200150Z digest=sha256:6bc5e235b75b5279796bc823e8f4e15fcc2ef1745ab1fa80fce0227ca46a6667

Observation 1066ad90-e5cc-4fda-bb70-1555ee2d27dc · outbound

This paper cites An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual Inversion.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual Inversion

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T20:35:35.206891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:35:35.206891Z digest=sha256:828d70ec5986bd14a59beff4d84557e55df8880ac5e081a0c5cbe4ad15aa7b37

Observation 9fc890e7-ec6a-4ad4-af64-cbda6ccdc07a · outbound

This paper cites Making Pre-trained Language Models Better Few-shot Learners.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Making Pre-trained Language Models Better Few-shot Learners

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T20:35:35.214142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:35:35.214142Z digest=sha256:6dd54a47f4da6deb9f3dba55a42fe9fec939ed99a3fb1eea54a30958902d64bb

Observation c2db7efc-fb81-4999-a38b-16b74239e343 · outbound

This paper cites TaleCrafter: Interactive Story Visualization with Multiple Characters.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC TaleCrafter: Interactive Story Visualization with Multiple Characters

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T20:35:35.220631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:35:35.220631Z digest=sha256:7ce1b1388fe3f08ff6e94c29e52df5e312001d70b381dff4ee02c77552ceec2f

Observation 570f3637-1474-4382-adf5-9f924b2a5cac · outbound

This paper cites Vec- tor quantized diffusion model for text-to-image synthesis.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Vec- tor quantized diffusion model for text-to-image synthesis

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:35:36.791412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T20:35:35.226937Z digest=sha256:7a633614b7ffc4f6bc001a7a8a3bc5bee58bfac27c219266459d05ea53287bd6

Observation 1eb910d4-1ee0-4777-ad73-90975900b8a9 · outbound

This paper cites Denoising diffu- sion probabilistic models.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Denoising diffu- sion probabilistic models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T20:35:35.232247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:35:35.232247Z digest=sha256:6e23fd952315bf7a0692663d7fcd7e38a8d3146423e0a4b49ec67ca7df84a431

Observation 8983ecd6-9bc5-4a23-9f8f-4f36bd453748 · outbound

This paper cites Cascaded diffusion models for high fidelity image generation.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Cascaded diffusion models for high fidelity image generation

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:35:36.749374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T20:35:35.237708Z digest=sha256:d7628e4abea1ed7d8bd8c7fc2577b48ec38426db2c7b87d4ea4c3f5bce07d7af

Observation 6fdd1597-6d2f-419f-99a1-d5d5f746a8e8 · outbound

This paper cites Stableviton: Learning semantic correspon- dence with latent diffusion model for virtual try-on.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Stableviton: Learning semantic correspon- dence with latent diffusion model for virtual try-on

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:35:36.722161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T20:35:35.242936Z digest=sha256:0901ab5b4eaffb603b6edaa74d9fe185a0cc5bfff8d7f0bc4a6065c7215240c3

Observation c8061cfa-58a6-441f-be25-8f91ee251d19 · outbound

This paper cites High-resolution virtual try-on with misalignment and occlusion-handled conditions.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC High-resolution virtual try-on with misalignment and occlusion-handled conditions

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:35:36.681877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T20:35:35.248016Z digest=sha256:857d8f715bfbc377478cb78657f7fe5730846cbc0143d1625933bfaf12a6fce7

Observation 443c4852-c067-4a8c-aedd-70023d53bdc0 · outbound

This paper cites Blip-diffusion: Pre- trained subject representation for controllable text-to-image generation and editing.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Blip-diffusion: Pre- trained subject representation for controllable text-to-image generation and editing

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:35:36.659395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T20:35:35.252837Z digest=sha256:09d4a94202bdcc303e3a31f41f5ed7b0c30ced12ad4b80672103e9124fa93e19

Observation cc40dc17-c0d6-462a-a6bd-85ff35a889de · outbound

This paper cites Controlnet++: Improving conditional controls with efficient consistency feedback, 2024.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Controlnet++: Improving conditional controls with efficient consistency feedback, 2024

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:35:36.639649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T20:35:35.257664Z digest=sha256:503d8e76994eeb1a221aa35c2a16bf9a0547e6bff79ea5c159bb963aca88e5bb

Observation dc8c4438-0ff3-4bda-9ddb-222d4ad7fdf9 · outbound

This paper cites Gligen: Open-set grounded text-to-image generation.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Gligen: Open-set grounded text-to-image generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T20:35:35.263266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:35:35.263266Z digest=sha256:6148aa557c34bdd31a67269a444b70eabbc11550167a8f1f4082e118868daaea

Observation 414b54d1-87c4-4335-9ad9-17a04e92d277 · outbound

This paper cites Intelligent Grimm -- Open-ended Visual Storytelling via Latent Diffusion Models.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Intelligent Grimm -- Open-ended Visual Storytelling via Latent Diffusion Models

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-11T20:35:35.841274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T20:35:35.268470Z digest=sha256:1bbcf8c34bcb9f64ea7a20db5c8262f85d8718f303b93738df44dc447771f1c4

Observation bd8406eb-4c27-42f8-a811-0df26713b607 · outbound

This paper cites Open-edit: Open- domain image manipulation with open-vocabulary instruc- tions.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Open-edit: Open- domain image manipulation with open-vocabulary instruc- tions

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:35:36.599250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T20:35:35.274231Z digest=sha256:87511834ec42053239d289b9f262696dd41548c5d2d51725973f4faeeeeff4dc

Observation 5c0f5fbe-260e-4794-bf51-4e768859cd57 · outbound

This paper cites P-tuning: Prompt tuning can be comparable to fine-tuning across scales and tasks.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC P-tuning: Prompt tuning can be comparable to fine-tuning across scales and tasks

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:35:36.573846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T20:35:35.279924Z digest=sha256:890a2ef1218de128a942cbe81aa8c3d9729654721a652fadae9ab5600bf6550f

Observation 490f10b2-03fd-42ec-89ca-b8e29a6470d7 · outbound

This paper cites Repaint: Inpainting using denoising diffusion probabilistic models.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Repaint: Inpainting using denoising diffusion probabilistic models

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:35:36.552194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T20:35:35.287160Z digest=sha256:8ee552249c0be75f12c1b82d9592905bbe72c3c40813946f632f185cf0299215

Observation 2b64e61d-240a-45ea-ad19-e45cd8c3d944 · outbound

This paper cites Peft: State-of-the-art parameter-efficient fine-tuning meth- ods.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Peft: State-of-the-art parameter-efficient fine-tuning meth- ods

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:35:36.528109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T20:35:35.292932Z digest=sha256:57e8d7cbecc3fac4bcc2afd0e0b6a00ecf0a91b901b80f07791d94d660048521

Observation dd5935d4-721d-460a-8a42-86ea1840a614 · outbound

This paper cites SDEdit: Guided image synthesis and editing with stochastic differential equa- tions.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC SDEdit: Guided image synthesis and editing with stochastic differential equa- tions

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T20:35:35.299419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:35:35.299419Z digest=sha256:8791673a5c75f766f811c92c254344b1ed7ff299210b4c2872593c883c3f2e84

Observation e542d542-0413-4290-9c1b-3f4de17638b6 · outbound

This paper cites LaDI-VTON: Latent Diffusion Textual-Inversion Enhanced Virtual Try-On.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC LaDI-VTON: Latent Diffusion Textual-Inversion Enhanced Virtual Try-On

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T20:35:35.306833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:35:35.306833Z digest=sha256:773828f8cf8a787f37365eb3a15c352353b6de9d1b948fd9e4e47f14387e5377

Observation cb4f1fd0-2bd1-4fa4-b993-27cf956491cf · outbound

This paper cites T2i- adapter: Learning adapters to dig out more controllable abil- ity for text-to-image diffusion models, 2023.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC T2i- adapter: Learning adapters to dig out more controllable abil- ity for text-to-image diffusion models, 2023

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T20:35:35.312489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:35:35.312489Z digest=sha256:9b067001cd974492c829fa503405a5b3b052dabaa72183a00ab1b4bab580fc02

Observation 83ce0e42-9717-44bc-9abd-8eebf94655ef · outbound

This paper cites Improved denoising diffusion probabilistic models.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Improved denoising diffusion probabilistic models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T20:35:35.317462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:35:35.317462Z digest=sha256:5980b92618d3c719f703cb439740dcbd15318ba2f917530839d0020a3d4b07d3

Observation dcb7ec94-c1ba-4a02-b6bc-371278c7d98b · outbound

This paper cites GLIDE: towards photorealis- tic image generation and editing with text-guided diffusion models.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC GLIDE: towards photorealis- tic image generation and editing with text-guided diffusion models

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:35:36.447938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T20:35:35.322891Z digest=sha256:4bc93ab71c7b8751c92aeb9672ae1012912ff958cdd58c3115083afac4d2632e

Observation 4b30c706-484f-4164-8021-5c7e49977767 · outbound

This paper cites Glide: Towards photorealis- tic image generation and editing with text-guided diffusion models.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Glide: Towards photorealis- tic image generation and editing with text-guided diffusion models

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:35:36.423041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T20:35:35.329447Z digest=sha256:2150b7f3859b3a2d1b6b91781912d1642a0772e6194c423ddf5125996678a9c1

Observation a1fdaf0e-362b-467c-839f-14822c612260 · outbound

This paper cites Scalable diffusion models with transformers.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Scalable diffusion models with transformers

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T20:35:35.335010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:35:35.335010Z digest=sha256:c1d9ab059b959cc7139ab18ec984de2fc3fa61f1f68bb7193d1dd4b485b4d0af

Observation d78d5fda-2843-4b40-b779-a247cf0727d9 · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T20:35:35.340357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:35:35.340357Z digest=sha256:432039de1b8ee860130163e498195437f297acaece9cddc6ce0906ddbe65734d

Observation 84ee5af5-af2f-4d68-b5d7-1b3afc3d14e4 · outbound

This paper cites UniControl: A Unified Diffusion Model for Controllable Visual Generation In the Wild.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC UniControl: A Unified Diffusion Model for Controllable Visual Generation In the Wild

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T20:35:35.345745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:35:35.345745Z digest=sha256:4cb4a22aa4a7e1117b00155ad4d0fc2fecf19b92ff399f9a6f2ae37b320ed6ee

Observation 28e76c5f-c6c4-4ef3-9c5c-a07e491ca861 · outbound

This paper cites Language models are unsu- pervised multitask learners.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Language models are unsu- pervised multitask learners

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:35:36.367113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T20:35:35.352931Z digest=sha256:d5fb12ec206232c9a5dcfdfa83c71f089cce99578a4986df947e926e875dfc05

Observation 9c374a86-9f24-49a7-8dcb-1952cd346f73 · outbound

This paper cites Learn- ing transferable visual models from natural language super- vision.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Learn- ing transferable visual models from natural language super- vision

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:35:36.342608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T20:35:35.359170Z digest=sha256:4549d472abaf04e364400a57bef2f57097c740969b7b34678484422b9f670f20

Observation 8d3f1035-c9cd-41a3-86b9-7760e14257b5 · outbound

This paper cites Exploring the limits of transfer learning with a unified text-to-text transformer.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Exploring the limits of transfer learning with a unified text-to-text transformer

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:35:36.314913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T20:35:35.365296Z digest=sha256:85fbce22716eb42c848ae49427ff63d28ea0bf7bb4cf5bb696eb95ed2356431c

Observation 8717a37c-cea4-4831-adbe-b33510f593be · outbound

This paper cites Make-a-story: Visual memory conditioned consistent story generation.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Make-a-story: Visual memory conditioned consistent story generation

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:35:36.291134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T20:35:35.370663Z digest=sha256:da5191e31725300eb4276910d342808213e3fe61b8ae00f92a461a2f26d58129

Observation 96142e50-8619-4a2b-8c64-5133b4cbdcf8 · outbound

This paper cites Zero-shot text-to-image generation.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Zero-shot text-to-image generation

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:35:36.269491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T20:35:35.376709Z digest=sha256:afa438e717e574e8c9109b53e5a4b8f74699fc7f0e4d010c78103a8a5f90053e

Observation 6e69a5e8-1d97-45bd-b043-c7a153e31105 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T20:35:35.381762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:35:35.381762Z digest=sha256:a35534bfd7c44f1a1cc563666baeef68146c2b9882a1e3d88d33815c002d808d

Observation 3273c662-0101-427d-bd6f-2773284318cc · outbound

This paper cites High-resolution image syn- thesis with latent diffusion models.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC High-resolution image syn- thesis with latent diffusion models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T20:35:35.387127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:35:35.387127Z digest=sha256:205d5c8684c2a40ea50d345851936e0745237fa0655053288e7014501ff2d590

Observation 3fb80e19-7a73-4aad-bc22-5ad67d67c169 · outbound

This paper cites Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:35:36.229178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T20:35:35.392317Z digest=sha256:964b0fd25d4c3deed293b42bc2edda6a5445e16c82f98fc687497d9662858b08

Observation 98b2f119-1d52-4f3e-abce-15a2bc0585fa · outbound

This paper cites Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T20:35:35.398099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:35:35.398099Z digest=sha256:9cb902e7e28e7ffb8eb6a594ac1413c5cc7c3059a84e8d830b051b45e15d82e9

Observation 8c3036b4-42dd-4e0f-946e-cda6a825d48a · outbound

This paper cites Deep unsupervised learning using nonequilibrium thermodynamics.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Deep unsupervised learning using nonequilibrium thermodynamics

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-11T20:35:35.403032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:35:35.403032Z digest=sha256:6afb2fc3ed08384c4b628976e5b4ac9b800899bc7cd4f8f8ccd3b07d81237361

Observation 381fbe4e-91bd-45df-aa3d-c735364fdc83 · outbound

This paper cites Df-gan: A simple and effective baseline for text-to-image synthesis.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Df-gan: A simple and effective baseline for text-to-image synthesis

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:35:36.186465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T20:35:35.409532Z digest=sha256:9cd92bb9d71fa7ba2e2cbad65f18f03285d16689350abdadde6d7d5374f9c718

Observation 817ff40c-92d1-4c9a-9966-c94e2b68ccf4 · outbound

This paper cites GALIP: Generative Adversarial CLIPs for Text-to-Image Synthesis.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC GALIP: Generative Adversarial CLIPs for Text-to-Image Synthesis

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-08-11T20:35:35.663300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T20:35:35.415709Z digest=sha256:40b1f094cdb0edc90fcb19b7599d87463c75ca3bccec0b9249a51b2104f9405a

Observation 4564569a-9564-414a-8169-afb469f22f5c · outbound

This paper cites Storyimager: A unified and efficient frame- work for coherent story visualization and completion.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Storyimager: A unified and efficient frame- work for coherent story visualization and completion

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:35:36.164572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T20:35:35.422550Z digest=sha256:7b74455f8486af2e92172950b12fe2a066e8d941d2d993c9cadad8b9187a1920

Observation b33e32e6-7e3b-41cc-8d51-f424ea37b6ed · outbound

This paper cites In- context learning unlocked for diffusion models.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC In- context learning unlocked for diffusion models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T20:35:35.432201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:35:35.432201Z digest=sha256:6b1500b5749ac552e2ac207e08ebd75925b2f7e4145c69888feeaeb374612723

Observation b64a9855-fedc-47fb-a139-ab8619124d41 · outbound

This paper cites OmniGen: Unified Image Generation.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC OmniGen: Unified Image Generation

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-11T20:35:35.438936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:35:35.438936Z digest=sha256:02e94b58cfc64a752ceb5fa0a82e6d17a1161794cdd6d486dbbb275d0944385d

Observation dee15de7-871f-4109-83f6-04de902ec6b2 · outbound

This paper cites Gp- vton: Towards general purpose virtual try-on via collabo- rative local-flow global-parsing learning.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Gp- vton: Towards general purpose virtual try-on via collabo- rative local-flow global-parsing learning

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:35:36.121879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T20:35:35.444748Z digest=sha256:e7203297afaea2df1e090d3bef57a87ff606986bafe317f6e2b86e6a475ef385

Observation 4b820a67-c1a9-4e9d-83fb-60a656345199 · outbound

This paper cites Attngan: Fine- grained text to image generation with attentional generative adversarial networks.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Attngan: Fine- grained text to image generation with attentional generative adversarial networks

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:35:36.097451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T20:35:35.451396Z digest=sha256:72c98f1d1d0abfa6ccd06e37b13f75204554d9446a082e8da5406e935f86b7e9

Observation 4c1ba89d-9e1c-442b-9053-2add82363d2b · outbound

This paper cites IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-11T20:35:35.457387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:35:35.457387Z digest=sha256:fb4e0024533f8efc5afd93af5ec97147d8667d42908836bc03013f4e11d97943

Observation e9219d00-e942-4ce8-9e40-17ca8ca675a4 · outbound

This paper cites Scaling Autoregressive Models for Content-Rich Text-to-Image Generation.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Scaling Autoregressive Models for Content-Rich Text-to-Image Generation

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-11T20:35:35.465748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:35:35.465748Z digest=sha256:b1834ac0581c1222881292a98fa1a774defe4a21531a150252750d03e8d91446

Observation 4f84cdd9-e027-475d-a1e5-c3d5b68f35c1 · outbound

This paper cites Stack- gan: Text to photo-realistic image synthesis with stacked generative adversarial networks.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Stack- gan: Text to photo-realistic image synthesis with stacked generative adversarial networks

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:35:36.069884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T20:35:35.473322Z digest=sha256:4d4dfd9611a31236b8582d8114ecee27d11efb6143b912f9ac3cdb1de231a096

Observation 48f422f3-68b0-4d00-9aa5-c962bbaaaf68 · outbound

This paper cites Magicbrush: A manually annotated dataset for instruction- guided image editing.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Magicbrush: A manually annotated dataset for instruction- guided image editing

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:35:36.046849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T20:35:35.481362Z digest=sha256:7570d789a4734d66d98b3685e6135b566940fa4e3e30956739a47e16c9c9c935

Observation 83d67812-8854-48b8-84b0-cd85c1403254 · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Adding conditional control to text-to-image diffusion models

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:35:36.023776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T20:35:35.489152Z digest=sha256:f5892c2fdf150f21d316d9bf5159941d1829347f26f294bf243112475b4ea0c0

Observation c861ebfe-9976-4840-8d8f-edf2b7528a53 · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Adding conditional control to text-to-image diffusion models

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-11T20:35:35.499143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:35:35.499143Z digest=sha256:a6c95479f0a3668ced8be8889cee5682e87bd9e6d6acbfb8ff2f0004a9d861a9

Observation a02a1900-0cd5-43e2-bef6-0a2eac227f01 · outbound

This paper cites Dm- gan: Dynamic memory generative adversarial networks for text-to-image synthesis.

Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC Dm- gan: Dynamic memory generative adversarial networks for text-to-image synthesis

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T20:35:35.981003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T20:35:35.506601Z digest=sha256:eecaaa6ea0a48aec121f7e6874d7823693e2567d9fa9b28b8105399b2efae5ba

Pith citing papers

No inbound Pith citation observations are available.