Pith. sign in

Paper Citation Record · LEDGER

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation

As of 8 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 0 inbound Pith citation observations for arXiv:2506.14015.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.14015 v1

Coverage vector

measured 61 of 61 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:31:08.145034Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

61 of 61 outbound references displayed

  • verified exact5
  • verified fuzzy41
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 735566f0-9610-4d9f-b7e3-458d0f21bd36 · outbound

This paper cites Clipface: Text-guided editing of textured 3d mor- phable models.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Clipface: Text-guided editing of textured 3d mor- phable models

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.907134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:07.898389Z digest=sha256:807553cc92a88514922939162799e973f0776e2739888d4370c2de7985cba12e

Observation 7d03369b-3ed5-4970-b7a1-47346e794324 · outbound

This paper cites Bergman, Petr Kellnhofer, Yifan Wang, Eric R.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Bergman, Petr Kellnhofer, Yifan Wang, Eric R

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.896548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:07.903530Z digest=sha256:86e1fb8c138235eb02c54f906aa74c51f281e2bc31c77b3a338953407f2c24e5

Observation 2bb60aaf-6068-4582-9d0a-fe92ffa4c3a5 · outbound

This paper cites Demystifying MMD GANs.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Demystifying MMD GANs

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:07.907862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:07.907862Z digest=sha256:983974cd463ef7f5f70b778eae8113e5c7335747e9619225c769aa369b9238c4

Observation f037af5b-db65-4d06-8022-db41774bdd87 · outbound

This paper cites Text and image guided 3d avatar generation and ma- nipulation.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Text and image guided 3d avatar generation and ma- nipulation

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.885244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:07.912939Z digest=sha256:82d8b52e9822d0a7455974d7cbfc53ef3053d8558beecc324f34c79d484b571c

Observation 7a8249a7-fc52-4ccb-9b14-8941913789a8 · outbound

This paper cites Chan, Connor Z.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Chan, Connor Z

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.873696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:07.917859Z digest=sha256:d2771cfacf366c606cfa7b7228444016184ddf45b841b39f3abf82b425137557

Observation 0658603d-2022-4351-8dae-9c5b743717c6 · outbound

This paper cites Efficient Text-Guided 3D-Aware Portrait Generation with Score Distillation Sampling on Distribution.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Efficient Text-Guided 3D-Aware Portrait Generation with Score Distillation Sampling on Distribution

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:31:08.345738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:07.921887Z digest=sha256:8d94c8b4c49ce37dfbf270f837d625e6b16f992ec7435c7c7ad3abf3f456b289

Observation 7f4c8178-9253-419a-8611-114b9894bb22 · outbound

This paper cites Generalizable and Animatable Gaussian Head Avatar.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Generalizable and Animatable Gaussian Head Avatar

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:07.926317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:07.926317Z digest=sha256:d6bea5da30f8f9e9160b8717b43d5d2050070494cf47f27aa07b08708107aa38

Observation b433735a-71aa-4617-84cf-1f5cb4b6179e · outbound

This paper cites Gen- erative adversarial networks: An overview.IEEE signal processing magazine, 35(1):53–65, 2018.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Gen- erative adversarial networks: An overview.IEEE signal processing magazine, 35(1):53–65, 2018

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.862171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:07.931478Z digest=sha256:3634cf09dcf3c456c398483a1c96e7bc5c58427dc21d7f3af219c695a37f1535

Observation 86fc0d01-1ae1-420b-b519-1d0a7d9e4b19 · outbound

This paper cites Cogview: Mastering text-to-image generation via transformers.Advances in Neural Information Processing Systems, 34:19822–19835, 2021.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Cogview: Mastering text-to-image generation via transformers.Advances in Neural Information Processing Systems, 34:19822–19835, 2021

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.849292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:07.935502Z digest=sha256:432760b382a1e685c37db6f217ba946caf295af72aadf356b73f723ad66a2b74

Observation 53ea7cbf-1b06-482c-b969-82bc5854778a · outbound

This paper cites Cogview2: Faster and better text-to-image generation via hierarchical transformers.Advances in Neural Information Processing Systems, 35:16890–16902, 2022.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Cogview2: Faster and better text-to-image generation via hierarchical transformers.Advances in Neural Information Processing Systems, 35:16890–16902, 2022

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.837144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:07.939875Z digest=sha256:1a155624083990a8052f33bcf69e035f7fd6d11d5d91e00b7e384b9c4641fc25

Observation 88001989-687a-45ae-9954-4b9776ae0453 · outbound

This paper cites Semantic image synthesis via adversarial learning.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Semantic image synthesis via adversarial learning

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.825624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:07.944122Z digest=sha256:0552e8d6b92a2f8a043f84aacec3bcac89e81716d34f86813e12039b9efb38ed

Observation 9ce23194-ef2d-434c-af26-a5a448a2fce7 · outbound

This paper cites Imagebart: Bidirectional context with multinomial diffusion for autoregressive image synthesis.Advances in neural information processing systems, 34:3518–3532, 2021.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Imagebart: Bidirectional context with multinomial diffusion for autoregressive image synthesis.Advances in neural information processing systems, 34:3518–3532, 2021

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.812341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:07.947588Z digest=sha256:ed8349b389e82adf90ccc7e0adfa75fab712f45ee938f264323a5984f806a841

Observation 2974ba24-a5c5-44a1-a96f-4e184953b605 · outbound

This paper cites Black, and Timo Bolkart.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Black, and Timo Bolkart

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.799938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:07.951370Z digest=sha256:1649d2c26734f6204931eb0e940c96377332ab530fa1e83a584180ee45779466

Observation d2c986d5-db8e-4ecb-8199-9fb53d6fa1dc · outbound

This paper cites Generative adversarial nets.Advances in neural information processing systems, 27, 2014.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Generative adversarial nets.Advances in neural information processing systems, 27, 2014

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.788960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:07.954592Z digest=sha256:9a7c291ace00c5cd4b3ad56f5a890c026edb74431239c57f688dd777b8d319bb

Observation ffe053af-9ceb-4890-b3b8-97d97d978c03 · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilibrium.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Gans trained by a two time-scale update rule converge to a local nash equilibrium

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.777099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:07.958861Z digest=sha256:019a26f680c47d0738bbfdc7c2f5d4fa37bad29618ccb9ab47fe40bfeeeb0c8c

Observation a7c2aaa2-f6c5-42f9-8adc-2dfc1ae1a485 · outbound

This paper cites Denoising diffu- sion probabilistic models.Advances in neural information processing systems, 33:6840–6851, 2020.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Denoising diffu- sion probabilistic models.Advances in neural information processing systems, 33:6840–6851, 2020

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.763621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:07.962580Z digest=sha256:4966f38ffe53b22b800103701a9f47d22d38108438dd8c9ff26fb44ee3b68c3f

Observation 0241f0a9-ca3a-4eeb-8a02-2e28c7d011f0 · outbound

This paper cites Removing the quality tax in controllable face gener- ation.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Removing the quality tax in controllable face gener- ation

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.752237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:07.966318Z digest=sha256:0761a8912257b6677f6d53c9f86b4f1fbc8d7bba9c586035ed903688f6710c2b

Observation 74915187-9c66-4ddf-a0cb-fa4e1a59dd6e · outbound

This paper cites GSGAN: Adversarial Learning for Hierarchical Generation of 3D Gaussian Splats.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation GSGAN: Adversarial Learning for Hierarchical Generation of 3D Gaussian Splats

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:31:08.319241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:07.971098Z digest=sha256:cde00e957089b49ff6c78a0298dfea64f403d9c627ddeb9e88882ac51a26976c

Observation 880bfad9-a36c-43dc-90ad-c07c88529523 · outbound

This paper cites ClipMatrix: Text-controlled Creation of 3D Textured Meshes.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation ClipMatrix: Text-controlled Creation of 3D Textured Meshes

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:07.975824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:07.975824Z digest=sha256:eb6c38200091e6d8417483c502764bb9fc3e63db968eb5c0c60e93cf1a67f9cf

Observation 3f66b315-86c1-4ce8-b2db-ff6ced427dae · outbound

This paper cites A style-based generator architecture for generative adversarial networks.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation A style-based generator architecture for generative adversarial networks

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.739919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:07.980228Z digest=sha256:1c620afd09ba341ee98d0fa955099f39bed8ef8716bb8a5c7407f279932784c6

Observation 7cbadd74-90ea-40f3-8ee9-680ead2a340f · outbound

This paper cites Analyzing and improv- ing the image quality of stylegan.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Analyzing and improv- ing the image quality of stylegan

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.726573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:07.983506Z digest=sha256:8aff1a9cc63d94f624e6f940519b60c4db9414a490d29e5410ab094255bca876

Observation befcd4f9-07ef-44d1-8eff-5413d8aaf0d3 · outbound

This paper cites GGHead: Fast and Generalizable 3D Gaussian Heads.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation GGHead: Fast and Generalizable 3D Gaussian Heads

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:07.987075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:07.987075Z digest=sha256:6ab024ff989fc520464417244f3f9bcfd4450144a6b6c2c9e9127765afc0ca5a

Observation 44bd85d5-84b6-4b1d-8c60-4ca0ea6135b8 · outbound

This paper cites Gaus- sian3diff: 3d gaussian diffusion for 3d full head synthesis and editing.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Gaus- sian3diff: 3d gaussian diffusion for 3d full head synthesis and editing

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.712770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:07.991026Z digest=sha256:d83d50cb8abccf660aea0805117eb2df8a433b7fb5f4111e444a00e2f6426cc2

Observation 19d0e1e1-2a9a-4b14-a6cd-102b4858f22d · outbound

This paper cites Autoregressive image generation using resid- ual quantization.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Autoregressive image generation using resid- ual quantization

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.696917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:07.994745Z digest=sha256:dbcd527b17ad5eab560da5f1433eae8150657c232843e008515079ce5694cec6

Observation efe04c76-1438-4092-9792-aeac5f121e48 · outbound

This paper cites Controllable text-to-image generation.Advances in Neural Information Processing Systems, 32, 2019.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Controllable text-to-image generation.Advances in Neural Information Processing Systems, 32, 2019

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.682714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:07.998351Z digest=sha256:97790dee92b165ed8ddbb47b895ce7452a210431ee42a05464980df4a934f8ac

Observation c91da531-567f-4987-9d93-521c96403568 · outbound

This paper cites an unresolved cited work.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:31:08.669685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:08.002098Z digest=sha256:a5a1afed5e50ab428c7fd84a12d7157c5d01a27c3e091f30e27e40c2435ca7d3

Observation 16653aef-0c9f-43ae-a81a-624c66a64664 · outbound

This paper cites Mind the gap: Understanding the modality gap in multi-modal contrastive representation learning.Advances in Neural Information Processing Systems, 35:17612–17625, 2022.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Mind the gap: Understanding the modality gap in multi-modal contrastive representation learning.Advances in Neural Information Processing Systems, 35:17612–17625, 2022

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.655069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:08.006665Z digest=sha256:4ad74a561a9f1af18e369c1b588cef7fe96dcf48dad87aa2b752816ba83cc8ea

Observation 9f5d5daa-5666-4899-8ce2-de44176ffab5 · outbound

This paper cites Visual instruction tuning.Advances in neural information processing systems, 36, 2024.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Visual instruction tuning.Advances in neural information processing systems, 36, 2024

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:08.011103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:08.011103Z digest=sha256:49f2e479dc7e24ec0763c5cbce8133ea3f290167515d37264a2eb23601391c01

Observation ed161552-764b-494d-aee3-1dfca435fe4b · outbound

This paper cites Which training methods for gans do actually converge? In International conference on machine learning, pages 3481–.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Which training methods for gans do actually converge? In International conference on machine learning, pages 3481–

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:08.014647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:08.014647Z digest=sha256:7852a3a62b2a262f9a6807e93d080db334e1635343d61d123964b5ed46228436

Observation 55ba6465-0819-453d-beb0-c4fa7e1ac671 · outbound

This paper cites Text2mesh: Text-driven neural stylization for meshes.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Text2mesh: Text-driven neural stylization for meshes

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.625751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:08.018446Z digest=sha256:ddd0027e0903e455f2ab112da400e06c343d7d0ff62e39da30ec5bbecbc42db7

Observation cdecfc0c-bb8f-40f3-b823-f10f2e6dd514 · outbound

This paper cites Text2facegan: Face generation from fine grained textual de- scriptions.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Text2facegan: Face generation from fine grained textual de- scriptions

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.613241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:08.022978Z digest=sha256:09cbaadc466e03bd4191330be47b138405d0b093129d4adf9190342a1d11919b

Observation fb625bde-e6e2-4a41-a5ba-cd2506121b33 · outbound

This paper cites GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:08.027014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:08.027014Z digest=sha256:15021ac07186dcf54ad09886f17cac87871d89138565640b512114090270ed70

Observation 9696588e-eb09-44ed-8acb-ca15bb5c0a41 · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Representation Learning with Contrastive Predictive Coding

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:08.031156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:08.031156Z digest=sha256:ea3a356920fb670019c0226d22c18ea7353ed2f0619ecb01e5c2258aa5c2d1f5

Observation d3ce2d11-3fdd-40d2-a298-91c524589def · outbound

This paper cites Paysan, R.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Paysan, R

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.601451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:08.035456Z digest=sha256:20b428841f9f5bcd7460aba808335a76c9037e42cac20f398dcb5953f4caac1b

Observation a71babae-a84b-464d-be3e-6d794c104e67 · outbound

This paper cites Towards open-ended text-to-face generation, combination and manipulation.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Towards open-ended text-to-face generation, combination and manipulation

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.589174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:08.040150Z digest=sha256:b98ea368e2a971231aa2d4ed5fdbff98afd5a4d83e869d0c370d16103b6ebfbd

Observation a3fc27ad-1fec-4a4a-9a2d-25e40644444d · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Learning transferable visual models from natural language supervi- sion

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:08.044522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:08.044522Z digest=sha256:ae5f39f12fd96c2bc2ca39ddf4668dadb743de082d38acdfd1b92f38804c6160

Observation 213a317a-79b8-4bae-8e43-e9043f6835df · outbound

This paper cites Zero-shot text-to-image generation.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Zero-shot text-to-image generation

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.569754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:08.048293Z digest=sha256:2b3246997abf04d7e63375c6495ea6c0be5b6fa403c9d4b743c2c4c7207b6c9d

Observation 26366c0e-fb3a-495b-b58e-64addc435256 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:08.052495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:08.052495Z digest=sha256:fc13ae0851a1717e797d19516afad905abc3eb509b5c1af4f1e3919cdc95e054

Observation d42417b0-85cb-4b41-a860-62d8c34b4897 · outbound

This paper cites Generative adver- sarial text to image synthesis.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Generative adver- sarial text to image synthesis

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.558408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:08.057191Z digest=sha256:b4aab12a736635c9e71b3899db1e244a86d088837b2ab7bb4128c8be1df82178

Observation dd772fa9-2bd6-44fd-a59d-bae03601436c · outbound

This paper cites Higher order contractive auto-encoder.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Higher order contractive auto-encoder

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.545188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:08.061800Z digest=sha256:98ae095583b97b255726ca93de142b1e028382b94db36a9b4b3b61a2092b64c9

Observation c335105b-d031-4d89-a0ec-55b35709bc41 · outbound

This paper cites High-resolution image syn- thesis with latent diffusion models.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation High-resolution image syn- thesis with latent diffusion models

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.532044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:08.065870Z digest=sha256:6117af6e5057fc0d7739ede7445dec6508c4ceab894893bbdef1579fe116f749

Observation d27c33ae-10cc-4752-86d2-2e213eeeb127 · outbound

This paper cites Pho- torealistic text-to-image diffusion models with deep language understanding.Advances in Neural Information Processing Systems, 35:36479–36494, 2022.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Pho- torealistic text-to-image diffusion models with deep language understanding.Advances in Neural Information Processing Systems, 35:36479–36494, 2022

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.519580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:08.069695Z digest=sha256:dfbe2b648945e649730f7b4da8c0fefd0c2fb131d10863415de1d3609f5b1829

Observation 7100a072-63c4-41a3-ba65-2e7a0008653f · outbound

This paper cites Conditional Image Generation and Manipulation for User-Specified Content.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Conditional Image Generation and Manipulation for User-Specified Content

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:31:08.251257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:08.073737Z digest=sha256:06ba07d6c2d86dfce13986c991ad262ea351727e029bb8f9c6501bc151bb452f

Observation 0cbec195-ad96-4239-913f-17d8ca92bb32 · outbound

This paper cites Multi-caption text-to-face synthesis: Dataset and algo- rithm.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Multi-caption text-to-face synthesis: Dataset and algo- rithm

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.507844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:08.077667Z digest=sha256:438012ae220636bea11ae44f12c67e35d16ab61a611ba1db136e42ee1e55f88c

Observation e8b7e879-7c43-44d6-9366-f1341cbc5b15 · outbound

This paper cites DF-GAN: A Simple and Effective Baseline for Text-to-Image Synthesis.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation DF-GAN: A Simple and Effective Baseline for Text-to-Image Synthesis

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:08.080980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:08.080980Z digest=sha256:f8fc612c0ee79e15838380d026f342234715a97e3e49a24c2dc166fe96355dd9

Observation 7059f785-b9ca-4257-893d-c43652b03ed4 · outbound

This paper cites Attention is all you need.Advances in neural information processing systems, 30, 2017.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Attention is all you need.Advances in neural information processing systems, 30, 2017

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.496591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:08.085430Z digest=sha256:0bb285c6169f91bf7e499b4c41251dd42b9025735801f5d23eb24945cadc174d

Observation dbd661aa-31c7-4aa9-b57d-72a9b2714125 · outbound

This paper cites Faces a la carte: Text-to-face generation via attribute disentanglement.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Faces a la carte: Text-to-face generation via attribute disentanglement

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.485757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:08.089341Z digest=sha256:35349266a43ffa3c07d19b5cd6a48cfb4570c9b7f8c6cc6120cd0eb3178b1b7f

Observation c9486153-9104-45a6-a460-1bafc4d19fc0 · outbound

This paper cites High-fidelity 3d face genera- tion from natural language descriptions.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation High-fidelity 3d face genera- tion from natural language descriptions

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.473247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:08.093446Z digest=sha256:e4bc3f77ce31932db7927e5078554f2f1514e4873f21ff6328e18b462c6e9124

Observation 3950c2f3-7091-40ea-bc7f-7c152449f2a1 · outbound

This paper cites Tedigan: Text-guided diverse face image generation and ma- nipulation.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Tedigan: Text-guided diverse face image generation and ma- nipulation

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:08.097334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:08.097334Z digest=sha256:a1b639e4a167791df4f09d970420eb053f01ec541c1fa8f6747a4ca86309b8f9

Observation fb10680f-332b-4266-b744-b0518f0a5121 · outbound

This paper cites Omniavatar: Geometry-guided controllable 3d head syn- thesis.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Omniavatar: Geometry-guided controllable 3d head syn- thesis

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.453055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:08.101556Z digest=sha256:cf3e28b713ecc3d49d69788a5141995e7098b97b61a60031e48b17f19ff5189f

Observation 0c14bdda-ff54-4a71-a0c6-dac0b2c93233 · outbound

This paper cites Attngan: Fine- grained text to image generation with attentional generative adversarial networks.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Attngan: Fine- grained text to image generation with attentional generative adversarial networks

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.441227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:08.105439Z digest=sha256:3710a732a4a2359049072cea2eb2186b14e3537136ad09d8ee4d2032685690d7

Observation d43cd068-0b1e-4ce4-ad49-1798121f3570 · outbound

This paper cites Towards high-fidelity text-guided 3d face genera- tion and manipulation using only images.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Towards high-fidelity text-guided 3d face genera- tion and manipulation using only images

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.427768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:08.110294Z digest=sha256:a25c5d37847f7c33e06206e21179d768e39e5e8470946ba8a25bdb4e0172a4f0

Observation e979a7b2-90d9-4ec4-8946-4ae5a8d52129 · outbound

This paper cites Scaling Autoregressive Models for Content-Rich Text-to-Image Generation.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Scaling Autoregressive Models for Content-Rich Text-to-Image Generation

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:08.113817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:08.113817Z digest=sha256:75cf39b5cca53071189b2cf651def000edfa31bd6cf9499611f64096c0f20f73

Observation 8893a9f8-19ac-4385-9cb5-284c1d8f5228 · outbound

This paper cites Stack- gan: Text to photo-realistic image synthesis with stacked generative adversarial networks.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Stack- gan: Text to photo-realistic image synthesis with stacked generative adversarial networks

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.415061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:08.117732Z digest=sha256:c5685b4c9c698003c3d2374e08175bb040afb2bf9b518979af643db37cca5f71

Observation 30d6e647-65ee-4e1d-88ea-2f4c1f240ebd · outbound

This paper cites Stack- gan++: Realistic image synthesis with stacked generative adversarial networks.IEEE transactions on pattern analysis and machine intelligence, 41(8):1947–1962, 2018.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Stack- gan++: Realistic image synthesis with stacked generative adversarial networks.IEEE transactions on pattern analysis and machine intelligence, 41(8):1947–1962, 2018

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.403186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:08.121195Z digest=sha256:da62cf1ae3974ccb241b2371eabd72aa707c2eb46f02ebf749e1061412d46aa5

Observation 7c4839b6-b5d2-44ce-b479-1494bdf7799b · outbound

This paper cites DreamFace: Progressive Generation of Animatable 3D Faces under Text Guidance.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation DreamFace: Progressive Generation of Animatable 3D Faces under Text Guidance

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T00:31:08.125192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:31:08.125192Z digest=sha256:8a7541255a328c2b8dccdf1454697621684403065b8767e11387844ae4ac054e

Observation 426d1768-4ea8-4558-9641-85788ae17f1b · outbound

This paper cites M6-UFC: Unifying Multi-Modal Controls for Conditional Image Synthesis via Non-Autoregressive Generative Transformers.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation M6-UFC: Unifying Multi-Modal Controls for Conditional Image Synthesis via Non-Autoregressive Generative Transformers

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:31:08.203740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:08.128900Z digest=sha256:a8b4f076341b530a9d293aeb30feb5bc71b3cba0452710ee2e191630bef017b7

Observation 334e4b03-11e9-4b82-ba61-618f27db6aa0 · outbound

This paper cites DiffGS: Functional Gaussian Splatting Diffusion.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation DiffGS: Functional Gaussian Splatting Diffusion

Reference 58

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:31:08.184971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:08.132902Z digest=sha256:5f0fa218b79991048cb46593c132386e18c695ab5d4444cee5d804b993e80848

Observation 2cc47fac-b3db-41f9-b79e-1a1fd9f57f64 · outbound

This paper cites Generative adversarial network for text-to-face synthesis and manipulation.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Generative adversarial network for text-to-face synthesis and manipulation

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.390930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:08.136966Z digest=sha256:b9f5a5b50dcac9938b5ee3a138cf73a8bd7d991bec27f3e16e9f3a64f5e920e5

Observation 0881f328-e6c4-419e-bf31-7a9eaa0d5516 · outbound

This paper cites Generative adversar- ial network for text-to-face synthesis and manipulation with pretrained bert model.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation Generative adversar- ial network for text-to-face synthesis and manipulation with pretrained bert model

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.379003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:08.141082Z digest=sha256:8144871994cdf8cf9e722060258b693758f247780b4ef4473c101ca7091b97e0

Observation 3f82aeca-0e7a-4eb1-86bc-93ebdcf242c7 · outbound

This paper cites blonde”, “blue eyes.

Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation blonde”, “blue eyes

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:31:08.367112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:31:08.145034Z digest=sha256:c0bf8aee4c2a834c325c18184a27b07564246823455bcaeb2bb8d91d1aa90d26

Pith citing papers

No inbound Pith citation observations are available.