Pith. sign in

Paper Citation Record · LEDGER

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation

As of 16 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 0 inbound Pith citation observations for arXiv:2506.14015.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.14015 v1

Coverage vector

measured 61 of 61 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:31:08.145034Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

61 of 61 outbound references displayed

  • verified exact5
  • verified fuzzy41
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 735566f0-9610-4d9f-b7e3-458d0f21bd36 · outbound

This paper cites Clipface: Text-guided editing of textured 3d mor- phable models.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Clipface: Text-guided editing of textured 3d mor- phable models

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.907134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:07.898389Z digest=sha256:a32f96efafd13fbd57d36b77a6926c693b02553161c305632910c61167bf8248

Observation 7d03369b-3ed5-4970-b7a1-47346e794324 · outbound

This paper cites Bergman, Petr Kellnhofer, Yifan Wang, Eric R.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Bergman, Petr Kellnhofer, Yifan Wang, Eric R

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.896548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:07.903530Z digest=sha256:42a38c8eff1eb2f5fcf9fb0b73d24750222ebcef7fd829beb22253ba9982d05b

Observation 2bb60aaf-6068-4582-9d0a-fe92ffa4c3a5 · outbound

This paper cites Demystifying MMD GANs.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Demystifying MMD GANs

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:07.907862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:07.907862Z digest=sha256:533859a2d5941619f0675bb4fbb64abcbd9109a31a73e45a7816ed563d04676a

Observation f037af5b-db65-4d06-8022-db41774bdd87 · outbound

This paper cites Text and image guided 3d avatar generation and ma- nipulation.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Text and image guided 3d avatar generation and ma- nipulation

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.885244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:07.912939Z digest=sha256:d5301fbb5ad6fef7af620fcbab80720732e57c9289e3d9b2d9353b1e002889b0

Observation 7a8249a7-fc52-4ccb-9b14-8941913789a8 · outbound

This paper cites Chan, Connor Z.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Chan, Connor Z

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.873696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:07.917859Z digest=sha256:f3639cc50d3031ad1764836f733b28308385305f58bc5434810dc1eaff3d8a24

Observation 0658603d-2022-4351-8dae-9c5b743717c6 · outbound

This paper cites Efficient Text-Guided 3D-Aware Portrait Generation with Score Distillation Sampling on Distribution.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Efficient Text-Guided 3D-Aware Portrait Generation with Score Distillation Sampling on Distribution

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:31:08.345738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:07.921887Z digest=sha256:33942230210f0fbf9ccc93d52776256084ef7a180ba13b3548f32734a9c9e3b0

Observation 7f4c8178-9253-419a-8611-114b9894bb22 · outbound

This paper cites Generalizable and Animatable Gaussian Head Avatar.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Generalizable and Animatable Gaussian Head Avatar

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:07.926317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:07.926317Z digest=sha256:413a6e807e2cf7358c46fed0914b8c88553a92de30eec0faf62d0d6b59b7d62c

Observation b433735a-71aa-4617-84cf-1f5cb4b6179e · outbound

This paper cites Gen- erative adversarial networks: An overview.IEEE signal processing magazine, 35(1):53–65, 2018.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Gen- erative adversarial networks: An overview.IEEE signal processing magazine, 35(1):53–65, 2018

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.862171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:07.931478Z digest=sha256:9d759e8971d557522afa48d67791274ae81a3cdac1853fb70f1703d1a2f1bcdb

Observation 86fc0d01-1ae1-420b-b519-1d0a7d9e4b19 · outbound

This paper cites Cogview: Mastering text-to-image generation via transformers.Advances in Neural Information Processing Systems, 34:19822–19835, 2021.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Cogview: Mastering text-to-image generation via transformers.Advances in Neural Information Processing Systems, 34:19822–19835, 2021

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.849292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:07.935502Z digest=sha256:bbdea6d19b8eeee8f71e67d17581b3910f4cd7f52cf08b94e30f9c5eccc5d29a

Observation 53ea7cbf-1b06-482c-b969-82bc5854778a · outbound

This paper cites Cogview2: Faster and better text-to-image generation via hierarchical transformers.Advances in Neural Information Processing Systems, 35:16890–16902, 2022.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Cogview2: Faster and better text-to-image generation via hierarchical transformers.Advances in Neural Information Processing Systems, 35:16890–16902, 2022

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.837144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:07.939875Z digest=sha256:10310eb5e81c7e4dddeb65269515fadaf2e71fbc51a3c2aba24f13d0ab603997

Observation 88001989-687a-45ae-9954-4b9776ae0453 · outbound

This paper cites Semantic image synthesis via adversarial learning.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Semantic image synthesis via adversarial learning

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.825624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:07.944122Z digest=sha256:8091f8fe0329fbb9e6d1c46dc1916505646acc1a08f603dcca5866b5ded4bc54

Observation 9ce23194-ef2d-434c-af26-a5a448a2fce7 · outbound

This paper cites Imagebart: Bidirectional context with multinomial diffusion for autoregressive image synthesis.Advances in neural information processing systems, 34:3518–3532, 2021.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Imagebart: Bidirectional context with multinomial diffusion for autoregressive image synthesis.Advances in neural information processing systems, 34:3518–3532, 2021

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.812341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:07.947588Z digest=sha256:e1398bd457b2b928a86d5fc2afac5c8d953bac70a9497a62b42a2722009e2d58

Observation 2974ba24-a5c5-44a1-a96f-4e184953b605 · outbound

This paper cites Black, and Timo Bolkart.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Black, and Timo Bolkart

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.799938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:07.951370Z digest=sha256:2bd0f960021c3063bf4fb48c1366568d6323d82d3848d3ff95cfb18f17bd3dae

Observation d2c986d5-db8e-4ecb-8199-9fb53d6fa1dc · outbound

This paper cites Generative adversarial nets.Advances in neural information processing systems, 27, 2014.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Generative adversarial nets.Advances in neural information processing systems, 27, 2014

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.788960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:07.954592Z digest=sha256:c84887aeec1bf6fd5fcaada2b44ce63665a69bac2e309ebc5c4d12c5e797308b

Observation ffe053af-9ceb-4890-b3b8-97d97d978c03 · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilibrium.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Gans trained by a two time-scale update rule converge to a local nash equilibrium

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.777099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:07.958861Z digest=sha256:330552e958d6ac105bcc5c325c65d0da26438f590334fa04e0d536b1f7a712d3

Observation a7c2aaa2-f6c5-42f9-8adc-2dfc1ae1a485 · outbound

This paper cites Denoising diffu- sion probabilistic models.Advances in neural information processing systems, 33:6840–6851, 2020.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Denoising diffu- sion probabilistic models.Advances in neural information processing systems, 33:6840–6851, 2020

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.763621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:07.962580Z digest=sha256:fa474968f6e0cdb3f0c255831dc69d14a0425bdc7e5ef280b550bbca71e34fd3

Observation 0241f0a9-ca3a-4eeb-8a02-2e28c7d011f0 · outbound

This paper cites Removing the quality tax in controllable face gener- ation.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Removing the quality tax in controllable face gener- ation

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.752237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:07.966318Z digest=sha256:3a1098713d683e06d43695b682c30c92ed309ed100c44fa20e004f90362b155d

Observation 74915187-9c66-4ddf-a0cb-fa4e1a59dd6e · outbound

This paper cites GSGAN: Adversarial Learning for Hierarchical Generation of 3D Gaussian Splats.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation GSGAN: Adversarial Learning for Hierarchical Generation of 3D Gaussian Splats

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:31:08.319241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:07.971098Z digest=sha256:116972651118cb718a5bd9ed8b1d6ae8e26d25ecafc7c38ec31a7c656e1679b0

Observation 880bfad9-a36c-43dc-90ad-c07c88529523 · outbound

This paper cites ClipMatrix: Text-controlled Creation of 3D Textured Meshes.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation ClipMatrix: Text-controlled Creation of 3D Textured Meshes

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:07.975824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:07.975824Z digest=sha256:e8da66539df04b3f3f089c7cbdab661c9eced080abfe55bc8c641eb957115604

Observation 3f66b315-86c1-4ce8-b2db-ff6ced427dae · outbound

This paper cites A style-based generator architecture for generative adversarial networks.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation A style-based generator architecture for generative adversarial networks

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.739919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:07.980228Z digest=sha256:59437e342db23a97cf59fcc9f16304f7bc1232af09d3ea199014d3acdc793baa

Observation 7cbadd74-90ea-40f3-8ee9-680ead2a340f · outbound

This paper cites Analyzing and improv- ing the image quality of stylegan.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Analyzing and improv- ing the image quality of stylegan

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.726573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:07.983506Z digest=sha256:9e8b1ea922c74f4a2adff416e1d5c1286bdd2552d95f968d783c47f9167befe9

Observation befcd4f9-07ef-44d1-8eff-5413d8aaf0d3 · outbound

This paper cites GGHead: Fast and Generalizable 3D Gaussian Heads.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation GGHead: Fast and Generalizable 3D Gaussian Heads

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:07.987075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:07.987075Z digest=sha256:880ff884b8e787f605d9a6790050580b278ad742c6d98e6f92cdedc937a7b38d

Observation 44bd85d5-84b6-4b1d-8c60-4ca0ea6135b8 · outbound

This paper cites Gaus- sian3diff: 3d gaussian diffusion for 3d full head synthesis and editing.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Gaus- sian3diff: 3d gaussian diffusion for 3d full head synthesis and editing

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.712770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:07.991026Z digest=sha256:4bd9e51ebbdf2bd232bce8c5a43630729f49af9ad1bcbe84f6afd30f2badaca0

Observation 19d0e1e1-2a9a-4b14-a6cd-102b4858f22d · outbound

This paper cites Autoregressive image generation using resid- ual quantization.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Autoregressive image generation using resid- ual quantization

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.696917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:07.994745Z digest=sha256:85888c7ce00f18fda160665973e00808a21aed408550fe7b3c880b6559032e7e

Observation efe04c76-1438-4092-9792-aeac5f121e48 · outbound

This paper cites Controllable text-to-image generation.Advances in Neural Information Processing Systems, 32, 2019.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Controllable text-to-image generation.Advances in Neural Information Processing Systems, 32, 2019

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.682714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:07.998351Z digest=sha256:1ebf27341891b91a9e7329c08bd4abe9acf9569d3ef2745967f80063065b3c12

Observation c91da531-567f-4987-9d93-521c96403568 · outbound

This paper cites an unresolved cited work.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:31:08.669685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:08.002098Z digest=sha256:6a66c43ababe1c78206ded80ef404cb635a4b380cfec0b02ed70c3ef9dc55592

Observation 16653aef-0c9f-43ae-a81a-624c66a64664 · outbound

This paper cites Mind the gap: Understanding the modality gap in multi-modal contrastive representation learning.Advances in Neural Information Processing Systems, 35:17612–17625, 2022.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Mind the gap: Understanding the modality gap in multi-modal contrastive representation learning.Advances in Neural Information Processing Systems, 35:17612–17625, 2022

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.655069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:08.006665Z digest=sha256:c0386aff231f8d9fb64ca2c225a7c5794bc1a33b1795264f5ee4908c51d58ea6

Observation 9f5d5daa-5666-4899-8ce2-de44176ffab5 · outbound

This paper cites Visual instruction tuning.Advances in neural information processing systems, 36, 2024.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Visual instruction tuning.Advances in neural information processing systems, 36, 2024

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:08.011103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:08.011103Z digest=sha256:9f9d9159d33347da28513efde8198bbf2f4746a1c08ad1793078e9154234726d

Observation ed161552-764b-494d-aee3-1dfca435fe4b · outbound

This paper cites Which training methods for gans do actually converge? In International conference on machine learning, pages 3481–.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Which training methods for gans do actually converge? In International conference on machine learning, pages 3481–

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:08.014647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:08.014647Z digest=sha256:05961a6136594b7e92876d27176dc929d6ea983e182478f39686189b499c240d

Observation 55ba6465-0819-453d-beb0-c4fa7e1ac671 · outbound

This paper cites Text2mesh: Text-driven neural stylization for meshes.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Text2mesh: Text-driven neural stylization for meshes

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.625751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:08.018446Z digest=sha256:d3dccd6611cf08b7dd5c7fa3f52b2b307ff98adfce136c1f5d91d9ad91f21a65

Observation cdecfc0c-bb8f-40f3-b823-f10f2e6dd514 · outbound

This paper cites Text2facegan: Face generation from fine grained textual de- scriptions.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Text2facegan: Face generation from fine grained textual de- scriptions

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.613241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:08.022978Z digest=sha256:5f29bf80c0e7eeb344433792864823be4d1f26e223131710ce3ddfc3f4ed58fe

Observation fb625bde-e6e2-4a41-a5ba-cd2506121b33 · outbound

This paper cites GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:08.027014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:08.027014Z digest=sha256:f0ef24e1ea30768964f7d20ab4788487cb29e00106d4005bec7f827993b16216

Observation 9696588e-eb09-44ed-8acb-ca15bb5c0a41 · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Representation Learning with Contrastive Predictive Coding

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:08.031156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:08.031156Z digest=sha256:66282ec282100c66d927ff5f0dffb431befc5341e4faebfbe50ff6443e666d3f

Observation d3ce2d11-3fdd-40d2-a298-91c524589def · outbound

This paper cites Paysan, R.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Paysan, R

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.601451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:08.035456Z digest=sha256:428b8e11a334c22493d6c8fd3601d0295570e61b3be41484cb5486fa33f8f6ed

Observation a71babae-a84b-464d-be3e-6d794c104e67 · outbound

This paper cites Towards open-ended text-to-face generation, combination and manipulation.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Towards open-ended text-to-face generation, combination and manipulation

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.589174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:08.040150Z digest=sha256:4215adb001dff9d26bfdb7e192930103091a090a93b0a3022350986eef6cb53a

Observation a3fc27ad-1fec-4a4a-9a2d-25e40644444d · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Learning transferable visual models from natural language supervi- sion

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:08.044522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:08.044522Z digest=sha256:b252a30ffb14b027a605541df4c386ca20ddc79110044bdc674f30e6b8ddda05

Observation 213a317a-79b8-4bae-8e43-e9043f6835df · outbound

This paper cites Zero-shot text-to-image generation.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Zero-shot text-to-image generation

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.569754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:08.048293Z digest=sha256:62718ff77dd88ec7e684b2eb1c65f3efb39f6a921f8f2ad2581078df79909fa3

Observation 26366c0e-fb3a-495b-b58e-64addc435256 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:08.052495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:08.052495Z digest=sha256:463a693061ef48604ec018faff0b0c23eb57b79072ca86c65073d24fef6a92f7

Observation d42417b0-85cb-4b41-a860-62d8c34b4897 · outbound

This paper cites Generative adver- sarial text to image synthesis.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Generative adver- sarial text to image synthesis

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.558408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:08.057191Z digest=sha256:46e63cb46d132bfa990c5536652ac79b048bf3ebd351efcf7689f8f028c36400

Observation dd772fa9-2bd6-44fd-a59d-bae03601436c · outbound

This paper cites Higher order contractive auto-encoder.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Higher order contractive auto-encoder

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.545188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:08.061800Z digest=sha256:9682dd6c8028d184b3d6f6374e7580a9c82b2fc88df4c40c1248ec970bdc5161

Observation c335105b-d031-4d89-a0ec-55b35709bc41 · outbound

This paper cites High-resolution image syn- thesis with latent diffusion models.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation High-resolution image syn- thesis with latent diffusion models

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.532044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:08.065870Z digest=sha256:c81d7bb932d313f311935240c06d804f99c116ee38245c5d650caebdfbdfa531

Observation d27c33ae-10cc-4752-86d2-2e213eeeb127 · outbound

This paper cites Pho- torealistic text-to-image diffusion models with deep language understanding.Advances in Neural Information Processing Systems, 35:36479–36494, 2022.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Pho- torealistic text-to-image diffusion models with deep language understanding.Advances in Neural Information Processing Systems, 35:36479–36494, 2022

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.519580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:08.069695Z digest=sha256:1c4d2bdd3ce5d4abbba4bdeb04492b0aa99f954eef29da12c23979cd61ad75f3

Observation 7100a072-63c4-41a3-ba65-2e7a0008653f · outbound

This paper cites Conditional Image Generation and Manipulation for User-Specified Content.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Conditional Image Generation and Manipulation for User-Specified Content

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:31:08.251257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:08.073737Z digest=sha256:3f8e08194f37820beba49d3b5443d13aec42ed4f7de6131545775b19ab22d783

Observation 0cbec195-ad96-4239-913f-17d8ca92bb32 · outbound

This paper cites Multi-caption text-to-face synthesis: Dataset and algo- rithm.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Multi-caption text-to-face synthesis: Dataset and algo- rithm

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.507844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:08.077667Z digest=sha256:ab87c6e22c6fc6098dacf44c48246061bffb8ba3536a37130fd3720abb5296fe

Observation e8b7e879-7c43-44d6-9366-f1341cbc5b15 · outbound

This paper cites DF-GAN: A Simple and Effective Baseline for Text-to-Image Synthesis.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation DF-GAN: A Simple and Effective Baseline for Text-to-Image Synthesis

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:08.080980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:08.080980Z digest=sha256:61b368f25dfa71a2ce3c20f1be74d2af32e9664df599897a9ab87d6aa3120f69

Observation 7059f785-b9ca-4257-893d-c43652b03ed4 · outbound

This paper cites Attention is all you need.Advances in neural information processing systems, 30, 2017.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Attention is all you need.Advances in neural information processing systems, 30, 2017

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.496591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:08.085430Z digest=sha256:3857741607c5df9869717ae2e5af8a507af05d05bccb0007ac0c5cde0ae44541

Observation dbd661aa-31c7-4aa9-b57d-72a9b2714125 · outbound

This paper cites Faces a la carte: Text-to-face generation via attribute disentanglement.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Faces a la carte: Text-to-face generation via attribute disentanglement

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.485757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:08.089341Z digest=sha256:57fafbf2943763b17a5515a3ee7c27581989535d70bc71b1ee571fa30a5a8365

Observation c9486153-9104-45a6-a460-1bafc4d19fc0 · outbound

This paper cites High-fidelity 3d face genera- tion from natural language descriptions.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation High-fidelity 3d face genera- tion from natural language descriptions

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.473247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:08.093446Z digest=sha256:d422cabf15f7f78ef109964652abb4a8815c26773b3363728e4b383e2946afd1

Observation 3950c2f3-7091-40ea-bc7f-7c152449f2a1 · outbound

This paper cites Tedigan: Text-guided diverse face image generation and ma- nipulation.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Tedigan: Text-guided diverse face image generation and ma- nipulation

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:08.097334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:08.097334Z digest=sha256:2b4ea2af0ebafff297dec7ccfffa8fbf0f223c560146c3c54eb6bf2c69e884ff

Observation fb10680f-332b-4266-b744-b0518f0a5121 · outbound

This paper cites Omniavatar: Geometry-guided controllable 3d head syn- thesis.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Omniavatar: Geometry-guided controllable 3d head syn- thesis

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.453055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:08.101556Z digest=sha256:01ea73bed4330c5a1d78acd3461d5a35f143331ca653fedf85a2ffb7b2ffcc40

Observation 0c14bdda-ff54-4a71-a0c6-dac0b2c93233 · outbound

This paper cites Attngan: Fine- grained text to image generation with attentional generative adversarial networks.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Attngan: Fine- grained text to image generation with attentional generative adversarial networks

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.441227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:08.105439Z digest=sha256:e64764b3afb226681c91a142b40092e73aecf5e45fd1e8ffcad70d29057e8be6

Observation d43cd068-0b1e-4ce4-ad49-1798121f3570 · outbound

This paper cites Towards high-fidelity text-guided 3d face genera- tion and manipulation using only images.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Towards high-fidelity text-guided 3d face genera- tion and manipulation using only images

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.427768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:08.110294Z digest=sha256:9503ad273d813389488d4e21febd59be1208b49af28f1b91a7f6ee799b0c9610

Observation e979a7b2-90d9-4ec4-8946-4ae5a8d52129 · outbound

This paper cites Scaling Autoregressive Models for Content-Rich Text-to-Image Generation.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Scaling Autoregressive Models for Content-Rich Text-to-Image Generation

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:08.113817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:08.113817Z digest=sha256:3de12cbe2f0f34d6b66a6ff5b79eb7747f6cdebd2e1c8d0a7f331c5be88c4ae8

Observation 8893a9f8-19ac-4385-9cb5-284c1d8f5228 · outbound

This paper cites Stack- gan: Text to photo-realistic image synthesis with stacked generative adversarial networks.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Stack- gan: Text to photo-realistic image synthesis with stacked generative adversarial networks

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.415061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:08.117732Z digest=sha256:b0c512edf461561209bcfb5aa930b27aa2b786d6d1308b46d8b17ba0848c3aa9

Observation 30d6e647-65ee-4e1d-88ea-2f4c1f240ebd · outbound

This paper cites Stack- gan++: Realistic image synthesis with stacked generative adversarial networks.IEEE transactions on pattern analysis and machine intelligence, 41(8):1947–1962, 2018.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Stack- gan++: Realistic image synthesis with stacked generative adversarial networks.IEEE transactions on pattern analysis and machine intelligence, 41(8):1947–1962, 2018

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.403186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:08.121195Z digest=sha256:266244d947c731cdc82e6b193c5a017a63617f997645cfd4b88c180b8ee1f8f5

Observation 7c4839b6-b5d2-44ce-b479-1494bdf7799b · outbound

This paper cites DreamFace: Progressive Generation of Animatable 3D Faces under Text Guidance.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation DreamFace: Progressive Generation of Animatable 3D Faces under Text Guidance

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:08.125192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:08.125192Z digest=sha256:adc66c251dd3c93b7b0e78941817e30ee92a2d624472c8489acc1a7a05f1da1a

Observation 426d1768-4ea8-4558-9641-85788ae17f1b · outbound

This paper cites M6-UFC: Unifying Multi-Modal Controls for Conditional Image Synthesis via Non-Autoregressive Generative Transformers.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation M6-UFC: Unifying Multi-Modal Controls for Conditional Image Synthesis via Non-Autoregressive Generative Transformers

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:31:08.203740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:08.128900Z digest=sha256:26d884ef11106163fb82b87ad735c452dd8a881a36739351c68f094985a521fe

Observation 334e4b03-11e9-4b82-ba61-618f27db6aa0 · outbound

This paper cites DiffGS: Functional Gaussian Splatting Diffusion.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation DiffGS: Functional Gaussian Splatting Diffusion

Reference 58

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:31:08.184971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:08.132902Z digest=sha256:eead6f28f084095eb5117bd1a839582fa2af8c0159c7cf9b3cee37310c3d2a5b

Observation 2cc47fac-b3db-41f9-b79e-1a1fd9f57f64 · outbound

This paper cites Generative adversarial network for text-to-face synthesis and manipulation.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Generative adversarial network for text-to-face synthesis and manipulation

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.390930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:08.136966Z digest=sha256:e4c6da63620495f1e1f16cf586aefdfa780c582c55d49d57629d15c260aabeb9

Observation 0881f328-e6c4-419e-bf31-7a9eaa0d5516 · outbound

This paper cites Generative adversar- ial network for text-to-face synthesis and manipulation with pretrained bert model.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Generative adversar- ial network for text-to-face synthesis and manipulation with pretrained bert model

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.379003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:08.141082Z digest=sha256:d968fda80d5eab90533ddb3b9275d297d72b6881acb1fc03cf6e6770aecc2207

Observation 3f82aeca-0e7a-4eb1-86bc-93ebdcf242c7 · outbound

This paper cites blonde”, “blue eyes.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation blonde”, “blue eyes

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.367112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T00:31:08.145034Z digest=sha256:62a8bb057e01476ddc9c34f4058d2ef0e88d2e58e46bcf3923d7afcc422c8284

Pith citing papers

No inbound Pith citation observations are available.