Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T13:58:05.169969Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 1 inbound Pith citation observation for arXiv:2507.19939.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T13:58:05.169969Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T09:52:01.768852Z
A source-named dated measurement, never combined with another source.
Source: cited_works
58 of 58 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 07e9ae56-ddf1-447f-bc70-bf34fe543b0c · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Diffit: Diffusion vision transformers for im- age generation
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ba332586-5090-4171-81ed-75404db30ac2 · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Sketch-guided text-to-image diffusion models
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c4ee3e43-266a-467d-95fb-8239ef05321d · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Language Models are Few-Shot Learners
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7146fd2-fcec-4aa8-aa49-eb21f363e065 · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Emergent Abilities of Large Language Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 117e6c31-1f2c-4410-bbcc-8af04532b50f · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs LLaMA: Open and Efficient Foundation Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 084b781f-ebe9-42a4-8549-550deea7ee0c · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Instruction Tuning with GPT-4
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7c8796a-5a87-450c-9b6f-9242b661bb1f · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs UniControl: A Unified Diffusion Model for Controllable Visual Generation In the Wild
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 171efb84-11c5-48a5-b5d8-9d20abe572af · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs T2i-adapter: Learning adapters to dig out more controllable ability for text-to-image diffusion models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bcd86d95-aa16-4441-ae02-c68103fd6549 · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Migc: Multi-instance generation controller for text-to-image synthesis
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 52a95384-79c8-48c5-9f2d-ef7c6610dcd7 · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Minigpt-4: Enhancing vision-language understanding with advanced large language models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6eaf496b-d8b5-457e-bbea-6eca8fcde988 · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Cogview: Mastering text-to-image generation via transformers
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6900dd29-b12c-4755-a336-471b5f0335be · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13a8f93a-9671-44cd-9f4c-7d768096f52f · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Freeman, Fr ´edo Durand, and Song Han
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e372bee8-fdf3-427a-a54e-9430c798aad7 · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Freeman, Fr ´edo Durand, and Song Han
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a026925c-8916-426c-af89-1f7a18dfca60 · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Cross-modal contrastive learning for text-to- image generation
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d7ddf754-1a80-441a-9475-41b3333b530c · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c672f31d-3987-41dd-ac11-5cb1fdbadc9d · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f49346e4-5003-42e9-b883-139052201eca · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Scaling autoregressive models for content-rich text-to-image genera- tion
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49cc29f7-ae85-47f5-9d8d-d0354b7ef22d · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Denoising Diffusion Implicit Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 387a71a9-9a1e-4380-965d-b4a3df88279d · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Open-vocabulary panop- tic segmentation with text-to-image diffusion models
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bb990f3b-b75d-4482-b143-6cdd2f770cd2 · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Classifier-Free Diffusion Guidance
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 045cb68c-5a5e-4c06-8745-e09cd8738412 · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Perception priori- tized training of diffusion models
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 01ca8309-71cb-4dc4-9625-b066f51b68a3 · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Training Compute-Optimal Large Language Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba033c7d-4d55-45e4-9218-fd745383be60 · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs GPT-4 Technical Report
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96d40ce9-fc3e-47cc-8936-c5206b6cfb6e · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Blip: Bootstrapping language-image pre-training for uni- fied vision-language understanding and generation
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dec473ca-5274-436a-8276-a27034d4da20 · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation aedc824d-b6f6-442a-a7bd-ae989798ff91 · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Composer: Creative and Controllable Image Synthesis with Composable Conditions
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a970c1c-d1f8-4434-9689-93710fb5c2ff · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs LLM-grounded Diffusion: Enhancing Prompt Understanding of Text-to-Image Diffusion Models with Large Language Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd0473fa-31b1-4309-9aeb-72ca7fb8cc91 · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Adding conditional control to text-to-image diffusion models
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d0195274-c817-4e8b-a3ba-0f901497dd7d · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Fourier features let networks learn high frequency functions in low dimen- sional domains
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9d934d97-73ad-40df-b623-a83d283a0493 · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5dd8cdb2-bf75-499e-b609-e59e05b3f2c8 · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Hyperdreambooth: Hypernetworks for fast personalization of text-to-image models
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d3c4b848-ad70-490f-9bbc-fdcc948ddf11 · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Spatext: Spatio-textual representation for con- trollable image generation
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f00b8a3d-e113-493b-ae07-9bb2c96cdf12 · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Scaling Rectified Flow Transformers for High-Resolution Image Synthesis
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f9b0f5c-a565-44a1-91c0-e3870f289e9c · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Diffusion models beat gans on image synthesis
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a820c14e-3831-45c8-b988-edf429f9c255 · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual Inversion
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e18275d2-9c57-48f6-a3e7-aca56999271f · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs High-resolution image syn- thesis with latent diffusion models
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1590a588-bddb-4a99-9aa1-282788ff9afe · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Generative ad- versarial text to image synthesis
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2882007e-6483-443f-b211-b9d621daf465 · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Younes Mirinezhad
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation aa69289a-f977-41b8-8bce-40f4f1818b70 · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Unresolved cited work
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8b607d78-345b-402f-9019-477285db138b · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dff75b55-9d13-4164-81f5-79445e019411 · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Freecontrol: Training-free spatial control of any text-to-image diffusion model with any condition
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9f202afb-5277-4256-9229-d79520cce8a1 · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Deep unsupervised learning using nonequilibrium thermodynamics
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 365b3635-be53-442c-bcc9-5976687a29ed · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Score-based 9 generative modeling through stochastic differential equa- tions
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f4a4d41f-fe1c-4a7e-9c56-368377856526 · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs It’s all about your sketch: Democratising sketch control in diffusion models
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ad81617a-d9db-4dc4-867e-6a734b62bd54 · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Attngan: Fine- grained text to image generation with attentional generative adversarial networks
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 042213c5-e384-4891-81bd-13bd131496f3 · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Plug-and-play diffusion features for text-driven image-to-image translation
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 40174456-e7e8-4dc2-aaae-aec564f43fd0 · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Compositional Text-to-Image Generation with Dense Blob Representations
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85990893-f0a0-45e3-b0ac-bfd9e6f7ab99 · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Layoutgpt: Compositional visual plan- ning and generation with large language models
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6625bbde-6465-4073-9aa9-78dc70909b7a · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs ELLA: Equip Diffusion Models with LLM for Enhanced Semantic Alignment
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dac461c6-ad6a-48a2-86ce-d515f6dd62ed · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Humansd: A native skeleton-guided diffusion model for human image generation
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2be2634a-9867-456f-97c1-6ceb237ba7e2 · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs eDiff-I: Text-to-Image Diffusion Models with an Ensemble of Expert Denoisers
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b6efc69-2ffd-4ae7-bccb-0251a4f3104e · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Lafite2: Few-shot Text-to-Image Generation
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b1d736c3-2c9e-4aed-b529-7dc8e21c151e · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Gligen: Open-set grounded text-to-image generation
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 64e4b6a5-2c47-43dc-998c-bf7fb7cd7e00 · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Ranni: Taming text-to-image diffusion for accurate instruction following
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c05889d0-fa51-4631-a53c-603f36e5e8fb · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Smartedit: Ex- ploring complex instruction-based image editing with multi- modal large language models
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0fbbccd9-48a4-463e-90e8-709a42ec2a94 · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Layoutdiffusion: Controllable diffu- sion model for layout-to-image generation
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation aa617503-f3f4-4e57-a2f1-b75cbf59137a · outbound
LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs Reco: Region- controlled text-to-image generation
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 136e6b49-1b99-449a-b459-2878b5d978c5 · inbound
EventOD: Event-Aware OD Flow Generation via LLM-Guided Semantic Modulation LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.