Pith. sign in

Paper Citation Record · LEDGER

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation

As of 9 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 0 inbound Pith citation observations for arXiv:2506.14015.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.14015 v1

Coverage vector

measured 61 of 61 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:31:08.145034Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

61 of 61 outbound references displayed

  • verified exact5
  • verified fuzzy41
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 735566f0-9610-4d9f-b7e3-458d0f21bd36 · outbound

This paper cites Clipface: Text-guided editing of textured 3d mor- phable models.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Clipface: Text-guided editing of textured 3d mor- phable models

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.907134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:07.898389Z digest=sha256:92a840292f58facea192e70982980d417578ed16779f524f8f2ba502cfd5e80d

Observation 7d03369b-3ed5-4970-b7a1-47346e794324 · outbound

This paper cites Bergman, Petr Kellnhofer, Yifan Wang, Eric R.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Bergman, Petr Kellnhofer, Yifan Wang, Eric R

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.896548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:07.903530Z digest=sha256:62ab8fa0dd12a27a384a71053e4c2b5dcc3a47c7932e1e8b0f2f9b4b3d1a193b

Observation 2bb60aaf-6068-4582-9d0a-fe92ffa4c3a5 · outbound

This paper cites Demystifying MMD GANs.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Demystifying MMD GANs

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:07.907862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:07.907862Z digest=sha256:a7078ec545fcc0ca3c607c859a19de254aa03f835a604cfa2b28dcb258214ef8

Observation f037af5b-db65-4d06-8022-db41774bdd87 · outbound

This paper cites Text and image guided 3d avatar generation and ma- nipulation.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Text and image guided 3d avatar generation and ma- nipulation

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.885244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:07.912939Z digest=sha256:77ad282bd8bad991b0851bdaa910b007c05f552a33980a2aa692fef52859bab7

Observation 7a8249a7-fc52-4ccb-9b14-8941913789a8 · outbound

This paper cites Chan, Connor Z.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Chan, Connor Z

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.873696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:07.917859Z digest=sha256:cadd48d038e4fbe24732d42da9992b94ef5bc75ca34b0d42fb1db90e9915264c

Observation 0658603d-2022-4351-8dae-9c5b743717c6 · outbound

This paper cites Efficient Text-Guided 3D-Aware Portrait Generation with Score Distillation Sampling on Distribution.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Efficient Text-Guided 3D-Aware Portrait Generation with Score Distillation Sampling on Distribution

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:31:08.345738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:07.921887Z digest=sha256:654578fc05c0b5284e615a8a8dc14a898d368cebe8b1f281bd27bc6fb5e07e9d

Observation 7f4c8178-9253-419a-8611-114b9894bb22 · outbound

This paper cites Generalizable and Animatable Gaussian Head Avatar.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Generalizable and Animatable Gaussian Head Avatar

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:07.926317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:07.926317Z digest=sha256:e7d0c10b4fb6f0e6288403bedbf81fccefc970678d2533ef97ceaec93d4f6829

Observation b433735a-71aa-4617-84cf-1f5cb4b6179e · outbound

This paper cites Gen- erative adversarial networks: An overview.IEEE signal processing magazine, 35(1):53–65, 2018.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Gen- erative adversarial networks: An overview.IEEE signal processing magazine, 35(1):53–65, 2018

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.862171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:07.931478Z digest=sha256:436c96b9e466c5098e972a951ea5b562825a88bea399ca8e78a6a53eace81667

Observation 86fc0d01-1ae1-420b-b519-1d0a7d9e4b19 · outbound

This paper cites Cogview: Mastering text-to-image generation via transformers.Advances in Neural Information Processing Systems, 34:19822–19835, 2021.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Cogview: Mastering text-to-image generation via transformers.Advances in Neural Information Processing Systems, 34:19822–19835, 2021

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.849292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:07.935502Z digest=sha256:30794498ef9e11463617a3443334a4b4f51ff21e749855b380dabb0ee037d5c9

Observation 53ea7cbf-1b06-482c-b969-82bc5854778a · outbound

This paper cites Cogview2: Faster and better text-to-image generation via hierarchical transformers.Advances in Neural Information Processing Systems, 35:16890–16902, 2022.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Cogview2: Faster and better text-to-image generation via hierarchical transformers.Advances in Neural Information Processing Systems, 35:16890–16902, 2022

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.837144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:07.939875Z digest=sha256:fc21a35a5240ebfd0ada2ef9ee03f32ad4a492c8b9b781c331c931c680474cd1

Observation 88001989-687a-45ae-9954-4b9776ae0453 · outbound

This paper cites Semantic image synthesis via adversarial learning.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Semantic image synthesis via adversarial learning

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.825624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:07.944122Z digest=sha256:ef6a6b370f0ff2a4660ba0cd84cef1a046a90445b4c2bbaf29a792bdbb4fe530

Observation 9ce23194-ef2d-434c-af26-a5a448a2fce7 · outbound

This paper cites Imagebart: Bidirectional context with multinomial diffusion for autoregressive image synthesis.Advances in neural information processing systems, 34:3518–3532, 2021.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Imagebart: Bidirectional context with multinomial diffusion for autoregressive image synthesis.Advances in neural information processing systems, 34:3518–3532, 2021

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.812341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:07.947588Z digest=sha256:7582f2f298b68e445f9574c72e641bf711fd01c4cb8e8a1f3d1711a510f400c0

Observation 2974ba24-a5c5-44a1-a96f-4e184953b605 · outbound

This paper cites Black, and Timo Bolkart.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Black, and Timo Bolkart

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.799938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:07.951370Z digest=sha256:7d57cba7869b3a42fc1e260489b06ddc89962a67b6d57c4f3b56754817f46059

Observation d2c986d5-db8e-4ecb-8199-9fb53d6fa1dc · outbound

This paper cites Generative adversarial nets.Advances in neural information processing systems, 27, 2014.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Generative adversarial nets.Advances in neural information processing systems, 27, 2014

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.788960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:07.954592Z digest=sha256:6612346f34942aad33fc07623f2ce57bf56d5a472ff41ec4bbbe4101e675bafd

Observation ffe053af-9ceb-4890-b3b8-97d97d978c03 · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilibrium.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Gans trained by a two time-scale update rule converge to a local nash equilibrium

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.777099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:07.958861Z digest=sha256:6c72b87cee572b32bbcd05d385bb7384c9264fd90745c8d8dafd2dbc333c43e5

Observation a7c2aaa2-f6c5-42f9-8adc-2dfc1ae1a485 · outbound

This paper cites Denoising diffu- sion probabilistic models.Advances in neural information processing systems, 33:6840–6851, 2020.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Denoising diffu- sion probabilistic models.Advances in neural information processing systems, 33:6840–6851, 2020

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.763621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:07.962580Z digest=sha256:68f81650972a1d9cb03b185f71de4339b55182d087b1cf6429bc1dfe2b5ab759

Observation 0241f0a9-ca3a-4eeb-8a02-2e28c7d011f0 · outbound

This paper cites Removing the quality tax in controllable face gener- ation.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Removing the quality tax in controllable face gener- ation

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.752237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:07.966318Z digest=sha256:cebd9d7f01aa9d2e8ee6d0b3a0eb13ce10384880843feb71243b9816a4c1aa83

Observation 74915187-9c66-4ddf-a0cb-fa4e1a59dd6e · outbound

This paper cites GSGAN: Adversarial Learning for Hierarchical Generation of 3D Gaussian Splats.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation GSGAN: Adversarial Learning for Hierarchical Generation of 3D Gaussian Splats

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:31:08.319241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:07.971098Z digest=sha256:44ad0dcd2b5d2c475560f888d8e3b0e7ca32932efd4211ad78e018eff44612b0

Observation 880bfad9-a36c-43dc-90ad-c07c88529523 · outbound

This paper cites ClipMatrix: Text-controlled Creation of 3D Textured Meshes.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation ClipMatrix: Text-controlled Creation of 3D Textured Meshes

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:07.975824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:07.975824Z digest=sha256:72bdad0e042a592a556a16871d74a9557bc7b175be82d3fae318b6dbabbd068f

Observation 3f66b315-86c1-4ce8-b2db-ff6ced427dae · outbound

This paper cites A style-based generator architecture for generative adversarial networks.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation A style-based generator architecture for generative adversarial networks

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.739919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:07.980228Z digest=sha256:1e81c316207f84e63904efa1ffafd0442c9cd68240ee095b2c020731f85bb62d

Observation 7cbadd74-90ea-40f3-8ee9-680ead2a340f · outbound

This paper cites Analyzing and improv- ing the image quality of stylegan.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Analyzing and improv- ing the image quality of stylegan

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.726573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:07.983506Z digest=sha256:249827b34838c8bee0760ecb40d3ac680837cdb2f9f65bd7148012518b56596a

Observation befcd4f9-07ef-44d1-8eff-5413d8aaf0d3 · outbound

This paper cites GGHead: Fast and Generalizable 3D Gaussian Heads.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation GGHead: Fast and Generalizable 3D Gaussian Heads

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:07.987075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:07.987075Z digest=sha256:39980c4d7be7f665ed0efd89d1df7bb35791a68fc28177334a3c43489e844452

Observation 44bd85d5-84b6-4b1d-8c60-4ca0ea6135b8 · outbound

This paper cites Gaus- sian3diff: 3d gaussian diffusion for 3d full head synthesis and editing.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Gaus- sian3diff: 3d gaussian diffusion for 3d full head synthesis and editing

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.712770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:07.991026Z digest=sha256:094573bf3b42ecf8b262b9bdad4033024043a9932f06df8c7138b8a208618921

Observation 19d0e1e1-2a9a-4b14-a6cd-102b4858f22d · outbound

This paper cites Autoregressive image generation using resid- ual quantization.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Autoregressive image generation using resid- ual quantization

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.696917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:07.994745Z digest=sha256:ce32497841a9dfbc292ceef02bbedfee0cc9d2749656a95e9367b8619d6ee436

Observation efe04c76-1438-4092-9792-aeac5f121e48 · outbound

This paper cites Controllable text-to-image generation.Advances in Neural Information Processing Systems, 32, 2019.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Controllable text-to-image generation.Advances in Neural Information Processing Systems, 32, 2019

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.682714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:07.998351Z digest=sha256:ec38454b2963e365b7840f2945ff642a7eb0e739442f9354631ca8157e0df091

Observation c91da531-567f-4987-9d93-521c96403568 · outbound

This paper cites an unresolved cited work.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:31:08.669685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:08.002098Z digest=sha256:47f112306e32a7758cf0e6307c216eaf304fcbe61b1fb69b1b3e384bcdff405c

Observation 16653aef-0c9f-43ae-a81a-624c66a64664 · outbound

This paper cites Mind the gap: Understanding the modality gap in multi-modal contrastive representation learning.Advances in Neural Information Processing Systems, 35:17612–17625, 2022.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Mind the gap: Understanding the modality gap in multi-modal contrastive representation learning.Advances in Neural Information Processing Systems, 35:17612–17625, 2022

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.655069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:08.006665Z digest=sha256:0b566a27a313a1d81eb6347693f7980b3675856b5a4f1b43666e00602d33a7ad

Observation 9f5d5daa-5666-4899-8ce2-de44176ffab5 · outbound

This paper cites Visual instruction tuning.Advances in neural information processing systems, 36, 2024.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Visual instruction tuning.Advances in neural information processing systems, 36, 2024

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:08.011103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:08.011103Z digest=sha256:8519ca1a771dd5ea77fb5d6dee06935edae119c17bf9538e6427f36207805c82

Observation ed161552-764b-494d-aee3-1dfca435fe4b · outbound

This paper cites Which training methods for gans do actually converge? In International conference on machine learning, pages 3481–.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Which training methods for gans do actually converge? In International conference on machine learning, pages 3481–

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:08.014647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:08.014647Z digest=sha256:885a3511dbd9a611ce82ae82408e11e4331c3f08d2f69bcb8c4c7a7995f9d2c0

Observation 55ba6465-0819-453d-beb0-c4fa7e1ac671 · outbound

This paper cites Text2mesh: Text-driven neural stylization for meshes.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Text2mesh: Text-driven neural stylization for meshes

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.625751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:08.018446Z digest=sha256:dfb0c3f26bb2e56c062aeb307c725c95cf04c11ebeeaf50c3c3d0ddb7155f030

Observation cdecfc0c-bb8f-40f3-b823-f10f2e6dd514 · outbound

This paper cites Text2facegan: Face generation from fine grained textual de- scriptions.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Text2facegan: Face generation from fine grained textual de- scriptions

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.613241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:08.022978Z digest=sha256:482268cbe65fd3117cc3b893911fb98ca15215a49454a4d740f8e1f429df7d84

Observation fb625bde-e6e2-4a41-a5ba-cd2506121b33 · outbound

This paper cites GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:08.027014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:08.027014Z digest=sha256:bb56c159b4759634c45d1c658f8271d9cbf88de2ec239bf54bac2b234d1ee53e

Observation 9696588e-eb09-44ed-8acb-ca15bb5c0a41 · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Representation Learning with Contrastive Predictive Coding

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:08.031156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:08.031156Z digest=sha256:88cec782ee061685f875a57c3d5029d7027680aeaed928d55d536405c6271502

Observation d3ce2d11-3fdd-40d2-a298-91c524589def · outbound

This paper cites Paysan, R.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Paysan, R

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.601451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:08.035456Z digest=sha256:91293471ff0c7094e3fa1a7e629df88d9432f13a43f68d74354927afb7026118

Observation a71babae-a84b-464d-be3e-6d794c104e67 · outbound

This paper cites Towards open-ended text-to-face generation, combination and manipulation.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Towards open-ended text-to-face generation, combination and manipulation

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.589174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:08.040150Z digest=sha256:f97b5eb91965273c5c21c4c9868d10b993b0071e4299e509cd4f119b64d86e18

Observation a3fc27ad-1fec-4a4a-9a2d-25e40644444d · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Learning transferable visual models from natural language supervi- sion

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:08.044522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:08.044522Z digest=sha256:3779b6113711c842d94370eb66f5cea89e65869256c2e190cd9cd4ea7a79c330

Observation 213a317a-79b8-4bae-8e43-e9043f6835df · outbound

This paper cites Zero-shot text-to-image generation.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Zero-shot text-to-image generation

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.569754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:08.048293Z digest=sha256:205f9c4d7219e480017ca10e70d1158a37de16038998d01fd396406529994361

Observation 26366c0e-fb3a-495b-b58e-64addc435256 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:08.052495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:08.052495Z digest=sha256:5c83831c82030a765bddea834e7f709d06d5dfadb20d1b6b1c16506d3056ebdb

Observation d42417b0-85cb-4b41-a860-62d8c34b4897 · outbound

This paper cites Generative adver- sarial text to image synthesis.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Generative adver- sarial text to image synthesis

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.558408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:08.057191Z digest=sha256:b38906543d503564251cba7c1f23788e42cfc709a95ec1a4b92f454470d16a32

Observation dd772fa9-2bd6-44fd-a59d-bae03601436c · outbound

This paper cites Higher order contractive auto-encoder.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Higher order contractive auto-encoder

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.545188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:08.061800Z digest=sha256:ec30f67aa0142b3ac06971490e8d1f03d656de93c8ab141f1ce8fbabbe5e8e7d

Observation c335105b-d031-4d89-a0ec-55b35709bc41 · outbound

This paper cites High-resolution image syn- thesis with latent diffusion models.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation High-resolution image syn- thesis with latent diffusion models

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.532044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:08.065870Z digest=sha256:4d505bcfc77955bad49da2908d08cca05c1f9b47f24b522a63d7831944f118d1

Observation d27c33ae-10cc-4752-86d2-2e213eeeb127 · outbound

This paper cites Pho- torealistic text-to-image diffusion models with deep language understanding.Advances in Neural Information Processing Systems, 35:36479–36494, 2022.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Pho- torealistic text-to-image diffusion models with deep language understanding.Advances in Neural Information Processing Systems, 35:36479–36494, 2022

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.519580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:08.069695Z digest=sha256:6aa4995a0d89080c96501a55062604d017fcf37a240459fbb4cdd8556762a3b0

Observation 7100a072-63c4-41a3-ba65-2e7a0008653f · outbound

This paper cites Conditional Image Generation and Manipulation for User-Specified Content.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Conditional Image Generation and Manipulation for User-Specified Content

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:31:08.251257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:08.073737Z digest=sha256:b0951dea346fbf5d205b471e59dc07b5cef26cc707cff543f43b385be97e1d10

Observation 0cbec195-ad96-4239-913f-17d8ca92bb32 · outbound

This paper cites Multi-caption text-to-face synthesis: Dataset and algo- rithm.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Multi-caption text-to-face synthesis: Dataset and algo- rithm

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.507844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:08.077667Z digest=sha256:9ae27aeca664fb582ad075529391a64382f26dcabcef3b67ec31c557a997dc47

Observation e8b7e879-7c43-44d6-9366-f1341cbc5b15 · outbound

This paper cites DF-GAN: A Simple and Effective Baseline for Text-to-Image Synthesis.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation DF-GAN: A Simple and Effective Baseline for Text-to-Image Synthesis

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:08.080980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:08.080980Z digest=sha256:d307a40f3d1b3147bf1de1518a695e2ba8f7567d771cc3809c22ee499c4aa29f

Observation 7059f785-b9ca-4257-893d-c43652b03ed4 · outbound

This paper cites Attention is all you need.Advances in neural information processing systems, 30, 2017.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Attention is all you need.Advances in neural information processing systems, 30, 2017

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.496591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:08.085430Z digest=sha256:7d412aef34cc28c9034cac0ae9af889cc92f8405c9e9252ccc5368db38917497

Observation dbd661aa-31c7-4aa9-b57d-72a9b2714125 · outbound

This paper cites Faces a la carte: Text-to-face generation via attribute disentanglement.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Faces a la carte: Text-to-face generation via attribute disentanglement

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.485757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:08.089341Z digest=sha256:3cc7253cccae9196c84a5d1b44f56dfa09d7cf45a036b2c7ca4a2da2240cffed

Observation c9486153-9104-45a6-a460-1bafc4d19fc0 · outbound

This paper cites High-fidelity 3d face genera- tion from natural language descriptions.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation High-fidelity 3d face genera- tion from natural language descriptions

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.473247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:08.093446Z digest=sha256:531b39f43456ab9566a0b15a6242cdaf5a1ee6579c682f4a21d215cfb7b55219

Observation 3950c2f3-7091-40ea-bc7f-7c152449f2a1 · outbound

This paper cites Tedigan: Text-guided diverse face image generation and ma- nipulation.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Tedigan: Text-guided diverse face image generation and ma- nipulation

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:08.097334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:08.097334Z digest=sha256:572bb51de9fed948e9fdbf215ee66b81427e5c739185408c1b8586611d5567c1

Observation fb10680f-332b-4266-b744-b0518f0a5121 · outbound

This paper cites Omniavatar: Geometry-guided controllable 3d head syn- thesis.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Omniavatar: Geometry-guided controllable 3d head syn- thesis

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.453055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:08.101556Z digest=sha256:09aecb7afedfbfb92440a9c2b217ae410883184bbadb493d69eadf94e11c19d3

Observation 0c14bdda-ff54-4a71-a0c6-dac0b2c93233 · outbound

This paper cites Attngan: Fine- grained text to image generation with attentional generative adversarial networks.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Attngan: Fine- grained text to image generation with attentional generative adversarial networks

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.441227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:08.105439Z digest=sha256:1dc30e866ef5613188966670a76cfdd3f3fb676655d572d67f59efe704d5bb98

Observation d43cd068-0b1e-4ce4-ad49-1798121f3570 · outbound

This paper cites Towards high-fidelity text-guided 3d face genera- tion and manipulation using only images.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Towards high-fidelity text-guided 3d face genera- tion and manipulation using only images

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.427768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:08.110294Z digest=sha256:cc7171e9742ff7a5eca2ae22048a2e3181a15a83442ca9813ffee5e598e997ae

Observation e979a7b2-90d9-4ec4-8946-4ae5a8d52129 · outbound

This paper cites Scaling Autoregressive Models for Content-Rich Text-to-Image Generation.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Scaling Autoregressive Models for Content-Rich Text-to-Image Generation

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:08.113817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:08.113817Z digest=sha256:89cd639f8ee8c300e9b483ea584560d3e3caa8deca984b653382fdf6dbdd4c0a

Observation 8893a9f8-19ac-4385-9cb5-284c1d8f5228 · outbound

This paper cites Stack- gan: Text to photo-realistic image synthesis with stacked generative adversarial networks.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Stack- gan: Text to photo-realistic image synthesis with stacked generative adversarial networks

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.415061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:08.117732Z digest=sha256:d0f57db5a0561bddc8bb9dfcf625edbd2fd40bc806c894121d9fa79248c85009

Observation 30d6e647-65ee-4e1d-88ea-2f4c1f240ebd · outbound

This paper cites Stack- gan++: Realistic image synthesis with stacked generative adversarial networks.IEEE transactions on pattern analysis and machine intelligence, 41(8):1947–1962, 2018.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Stack- gan++: Realistic image synthesis with stacked generative adversarial networks.IEEE transactions on pattern analysis and machine intelligence, 41(8):1947–1962, 2018

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.403186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:08.121195Z digest=sha256:e18133679e2080e9d6a19b0f9064f21723f4643681f098b4462e807199e4932e

Observation 7c4839b6-b5d2-44ce-b479-1494bdf7799b · outbound

This paper cites DreamFace: Progressive Generation of Animatable 3D Faces under Text Guidance.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation DreamFace: Progressive Generation of Animatable 3D Faces under Text Guidance

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:08.125192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:08.125192Z digest=sha256:1a39d1c472f5cb79887ed422a5d3ebd9bd4950765855eb8ad47e1530fb45bb2e

Observation 426d1768-4ea8-4558-9641-85788ae17f1b · outbound

This paper cites M6-UFC: Unifying Multi-Modal Controls for Conditional Image Synthesis via Non-Autoregressive Generative Transformers.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation M6-UFC: Unifying Multi-Modal Controls for Conditional Image Synthesis via Non-Autoregressive Generative Transformers

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:31:08.203740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:08.128900Z digest=sha256:222ed59c5c71f5e1926f1a9267d2ac75a655d9d0788bea8d1d89536df1248ee1

Observation 334e4b03-11e9-4b82-ba61-618f27db6aa0 · outbound

This paper cites DiffGS: Functional Gaussian Splatting Diffusion.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation DiffGS: Functional Gaussian Splatting Diffusion

Reference 58

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:31:08.184971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:08.132902Z digest=sha256:1305b6e475c74aa5804e5530f5e65c666e0581f2b74a326b78b6b12d6056d127

Observation 2cc47fac-b3db-41f9-b79e-1a1fd9f57f64 · outbound

This paper cites Generative adversarial network for text-to-face synthesis and manipulation.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Generative adversarial network for text-to-face synthesis and manipulation

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.390930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:08.136966Z digest=sha256:fdf522b80153851afbb4d75c3f19d45f76fa0d6eb52d9900402e97d009abeeb4

Observation 0881f328-e6c4-419e-bf31-7a9eaa0d5516 · outbound

This paper cites Generative adversar- ial network for text-to-face synthesis and manipulation with pretrained bert model.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Generative adversar- ial network for text-to-face synthesis and manipulation with pretrained bert model

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.379003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:08.141082Z digest=sha256:52d90d7f14c35d42b533e548df71a6a76f4c1c8305298c0b7ac013086bdc7c00

Observation 3f82aeca-0e7a-4eb1-86bc-93ebdcf242c7 · outbound

This paper cites blonde”, “blue eyes.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation blonde”, “blue eyes

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.367112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:31:08.145034Z digest=sha256:94dfa8bd30bbc51ce15604a6c837081a04de9e25c6afbc34fbaf45d676ac188c

Pith citing papers

No inbound Pith citation observations are available.