Pith. sign in

Paper Citation Record · LEDGER

End-to-end Training for Text-to-Image Synthesis using Dual-Text Embeddings

As of 18 August 2026, this Paper Citation Record lists 27 of 27 outbound references and 0 inbound Pith citation observations for arXiv:2502.01507.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.01507 v1

Coverage vector

measured 27 of 27 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T15:09:40.133207Z

measured 27 of 27 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

27 of 27 outbound references displayed

  • verified exact1
  • verified fuzzy4
  • unresolved18
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch4

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f3215e4f-506e-4fc5-8f3c-05c1bc892410 · outbound

This paper cites On Self Modulation for Generative Adversarial Networks.

End-to-end Training for Text-to-Image Synthesis using Dual-Text Embeddings On Self Modulation for Generative Adversarial Networks

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-09T15:09:39.990669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:09:39.990669Z digest=sha256:7d384670d01e8febae89a375f33e6baf71717569d204c41565de73fb0dc2478b

Observation c8a1d4cf-c7cb-417c-9ffb-036654322eac · outbound

This paper cites Taming Transformers for High-Resolution Image Synthesis.

End-to-end Training for Text-to-Image Synthesis using Dual-Text Embeddings Taming Transformers for High-Resolution Image Synthesis

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-09T15:09:40.008399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:09:40.008399Z digest=sha256:f2db366e6a4763002a5bec4f50ec1e1dbdfda9da1889fb0464c690b76911b531

Observation 4a7c5f4d-5c57-4048-93dc-05b7f04dc878 · outbound

This paper cites Kingma and Jimmy Ba.

End-to-end Training for Text-to-Image Synthesis using Dual-Text Embeddings Kingma and Jimmy Ba

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:09:41.024593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-09T15:09:40.013198Z digest=sha256:e929e773a2a502f29d565727a3da332be13cf25c03977563617895eb8cb419cb

Observation 2e47f44b-ddd7-4da9-96b7-be14f2dd00b2 · outbound

This paper cites an unresolved cited work.

End-to-end Training for Text-to-Image Synthesis using Dual-Text Embeddings Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-09T15:09:41.011812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-09T15:09:40.025480Z digest=sha256:d637686d55df110ad402d53c0f535bcd2dfb8def5d3bb90d9af7ca8605d65702

Observation fa88d0c9-2cf5-448b-abe1-22a5be1f69e9 · outbound

This paper cites Improved Denoising Diffusion Probabilistic Models.

End-to-end Training for Text-to-Image Synthesis using Dual-Text Embeddings Improved Denoising Diffusion Probabilistic Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-09T15:09:40.036150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:09:40.036150Z digest=sha256:7adc68320b0ca1910c0d7afe549768838192d8979793b891cb1acad06aa81474

Observation 8055126c-7d78-451d-b5a6-0e94e67a1838 · outbound

This paper cites Learning Transferable Visual Models From Natural Language Supervision.

End-to-end Training for Text-to-Image Synthesis using Dual-Text Embeddings Learning Transferable Visual Models From Natural Language Supervision

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-09T15:09:40.041613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:09:40.041613Z digest=sha256:ec12994589f25e60ec4f2982deb7d51b5098c7203a014a022cf05ba800d2afe7

Observation 9e2e35ea-cd90-488c-9099-f03f4104af40 · outbound

This paper cites Zero-Shot Text-to-Image Generation.

End-to-end Training for Text-to-Image Synthesis using Dual-Text Embeddings Zero-Shot Text-to-Image Generation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-09T15:09:40.047361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:09:40.047361Z digest=sha256:65123520ec9817a93992b37de78a04cf289f4b6de209867cbc08f5a59db4c417

Observation 46db79fc-cf05-461a-bad3-675e83439521 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

End-to-end Training for Text-to-Image Synthesis using Dual-Text Embeddings Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-09T15:09:40.053774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:09:40.053774Z digest=sha256:b945df063b5b2c16aa8f6e588fefb14f1058339af455ee5e5a23038794164642

Observation 68f6f8c4-9123-4e5e-8a14-5023ddf4ba3d · outbound

This paper cites Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding.

End-to-end Training for Text-to-Image Synthesis using Dual-Text Embeddings Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-09T15:09:40.060578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:09:40.060578Z digest=sha256:52c20115f0040205eb8bca883a2f91c80f707c66c53fa0323d86bfd4fab5bbf9

Observation 8f344a35-0309-444e-be54-25cd62c70348 · outbound

This paper cites Hongchen Tan, Xiuping Liu, Baocai Yin, and Xin Li.

End-to-end Training for Text-to-Image Synthesis using Dual-Text Embeddings Hongchen Tan, Xiuping Liu, Baocai Yin, and Xin Li

Reference 16

Resolution
metadata mismatch
raw_fallback, observed 2026-08-09T15:09:40.660278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-09T15:09:40.070944Z digest=sha256:1847faa5f253ed1fcad73717ce7891d31878bef30d46fb708a67ed351082b188

Observation 5ae53ed8-16f6-448b-83c1-584f3eefff29 · outbound

This paper cites Ming Tao, Hao Tang, Fei Wu, Xiao-Yuan Jing, Bing-Kun Bao, and Changsheng Xu.

End-to-end Training for Text-to-Image Synthesis using Dual-Text Embeddings Ming Tao, Hao Tang, Fei Wu, Xiao-Yuan Jing, Bing-Kun Bao, and Changsheng Xu

Reference 17

Resolution
metadata mismatch
raw_fallback, observed 2026-08-09T15:09:40.470829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-09T15:09:40.075955Z digest=sha256:49d466ca07544533ef55e14cbdd95edc4d7ab93c1fe4d8e4edeee0ab4f033077

Observation 5f97429b-964d-4103-925f-ff79a3d192a7 · outbound

This paper cites Stackgan++: Realistic image synthesis with stacked generative adversarial networks.IEEE Transactions on Pattern Analysis and Machine Intelligence, PP, 10 2017a.

End-to-end Training for Text-to-Image Synthesis using Dual-Text Embeddings Stackgan++: Realistic image synthesis with stacked generative adversarial networks.IEEE Transactions on Pattern Analysis and Machine Intelligence, PP, 10 2017a

Reference 21

Resolution
metadata mismatch
raw_fallback, observed 2026-08-09T15:09:40.303796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-09T15:09:40.095194Z digest=sha256:a4fcb4e62d495d8133120bffa0d6e0defb8e0a8ec40fd3cfcfa8466d9a62b1a3

Observation 0388d8c0-82c0-4f68-b0af-c69aeca4e5b9 · outbound

This paper cites DTGAN: Dual Attention Generative Adversarial Networks for Text-to-Image Generation.

End-to-end Training for Text-to-Image Synthesis using Dual-Text Embeddings DTGAN: Dual Attention Generative Adversarial Networks for Text-to-Image Generation

Reference 23

Resolution
metadata mismatch
local_arxiv, observed 2026-08-09T15:09:40.199618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-09T15:09:40.104617Z digest=sha256:765e6f1d7577f7920be9123b0ef46783b9323e96387749b9dc8cfdaf202c30be

Observation 72174d50-2523-42e0-b648-b3f25a7799e9 · outbound

This paper cites an unresolved cited work.

End-to-end Training for Text-to-Image Synthesis using Dual-Text Embeddings Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-09T15:09:40.956566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-09T15:09:40.114240Z digest=sha256:97bd43d2cab023f4a1d5b08afacc8766cc51c2ab9a2cfe170d11457aa0f4ba47

Observation 016d377a-26ce-42c5-bb42-e930c41d6303 · outbound

This paper cites an unresolved cited work.

End-to-end Training for Text-to-Image Synthesis using Dual-Text Embeddings Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-09T15:09:40.942812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-09T15:09:40.119309Z digest=sha256:d1f59a377fc8a9296cd0d875d6d21144946d6c2c8323b079a8eb660df68d5a5a

Observation 24bbd123-69ac-4daa-ae47-bceeaa9c6edc · outbound

This paper cites (2017) in both the generator and discriminator text encoders, and we have reported the results in Table.

End-to-end Training for Text-to-Image Synthesis using Dual-Text Embeddings (2017) in both the generator and discriminator text encoders, and we have reported the results in Table

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:09:40.929814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-09T15:09:40.123539Z digest=sha256:f56c876da4df83559ed42e812735cbac2ab5c76871a0a49acc621d28b3f4375f

Observation 0c176c39-e431-42cf-a881-811d811fcbb8 · outbound

This paper cites DTE-GAN architecture consists of a dual text embedding setup (Section C.1), a single-stage Generator (Section C.2) and a Discriminator (Section C.3).

End-to-end Training for Text-to-Image Synthesis using Dual-Text Embeddings DTE-GAN architecture consists of a dual text embedding setup (Section C.1), a single-stage Generator (Section C.2) and a Discriminator (Section C.3)

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:09:40.916217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-09T15:09:40.128871Z digest=sha256:068a3b997ff3770f4f1228100f2963ce83cd4b17bd5eacebdc3c47443c34e1b6

Observation 7813a44a-2f30-46b4-ae49-753246fa49e1 · outbound

This paper cites an unresolved cited work.

End-to-end Training for Text-to-Image Synthesis using Dual-Text Embeddings Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-09T15:09:40.901812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-09T15:09:40.133207Z digest=sha256:9d46cf6d3b3fcbc45d269407d5365854340cd72556f90539ebc3f739bab58722

Observation d6ce2c94-f544-4b88-9642-aaefcdef79e8 · outbound

This paper cites an unresolved cited work.

End-to-end Training for Text-to-Image Synthesis using Dual-Text Embeddings Unresolved cited work

Reference 200

Resolution
unresolved
raw_fallback, observed 2026-08-09T15:09:40.971967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-09T15:09:40.085888Z digest=sha256:cd9746bac8dab651999a13add9f85e6d64c1bb102980bd0a45a2cc7da90f6f42

Observation 5a78daf9-2ed3-4f5c-becb-054bb803caf8 · outbound

This paper cites Jascha Sohl-Dickstein, Eric Weiss, Niru Maheswaranathan, and Surya Ganguli.

End-to-end Training for Text-to-Image Synthesis using Dual-Text Embeddings Jascha Sohl-Dickstein, Eric Weiss, Niru Maheswaranathan, and Surya Ganguli

Reference 1997

Resolution
unresolved
no resolver link, observed 2026-08-09T15:09:40.065823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:09:40.065823Z digest=sha256:f632de1b5ded112ea8b9306a75458d45ccc676cf433d2750868e808506ce4432

Observation f248ba16-c3a5-47f1-81dc-5ef761925251 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

End-to-end Training for Text-to-Image Synthesis using Dual-Text Embeddings Adam: A Method for Stochastic Optimization

Reference 2015

Resolution
unresolved
no resolver link, observed 2026-08-09T15:09:40.018768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:09:40.018768Z digest=sha256:f2621408217eb0d2988d289ae73623239d8b6fbeaee3b3359e3c16d8b5c9873d

Observation 8e454184-18d3-4f51-b140-d03d0c906f42 · outbound

This paper cites an unresolved cited work.

End-to-end Training for Text-to-Image Synthesis using Dual-Text Embeddings Unresolved cited work

Reference 2017

Resolution
unresolved
raw_fallback, observed 2026-08-09T15:09:40.985254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-09T15:09:40.081152Z digest=sha256:8f9bd83b4ce98b1ada42ffea09adf6227ff4ae813213ea80c030328e6b7f2de3

Observation 429c2cb4-ef7b-479d-bf4b-5d3d4a18532a · outbound

This paper cites LAFITE: Towards Language-Free Training for Text-to-Image Generation.

End-to-end Training for Text-to-Image Synthesis using Dual-Text Embeddings LAFITE: Towards Language-Free Training for Text-to-Image Generation

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-09T15:09:40.109640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:09:40.109640Z digest=sha256:727f373e937f92b4513f840cfddd3f3d17d0e5239a09690a454f18e9ac805812

Observation 0fdaf98d-cd19-4f03-a75d-af131a256f1b · outbound

This paper cites Ting Chen, Mario Lučić, Neil Houlsby, and Sylvain Gelly.

End-to-end Training for Text-to-Image Synthesis using Dual-Text Embeddings Ting Chen, Mario Lučić, Neil Houlsby, and Sylvain Gelly

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-09T15:09:39.985250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:09:39.985250Z digest=sha256:f70e1c722fd051e2141fe78053711d0598a7373fa1471f98d03b4cb83db7e685

Observation 6cf6328c-c1e6-4189-bd6f-27cf470bee94 · outbound

This paper cites Ayushman Dash, John Cristian Borges Gamboa, Sheraz Ahmed, Marcus Liwicki, and Muhammad Zeshan Afzal.

End-to-end Training for Text-to-Image Synthesis using Dual-Text Embeddings Ayushman Dash, John Cristian Borges Gamboa, Sheraz Ahmed, Marcus Liwicki, and Muhammad Zeshan Afzal

Reference 2020

Resolution
verified exact
arxiv_id_nonexistent, observed 2026-08-09T15:09:40.797703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-09T15:09:39.995772Z digest=sha256:27fd37724bad1c1a0fa5b91501515a65e269d10519a311b2ba9104d3f62b20a5

Observation bb41dd06-7491-4981-8c0e-9c88b216d72d · outbound

This paper cites Lawrence Zitnick.

End-to-end Training for Text-to-Image Synthesis using Dual-Text Embeddings Lawrence Zitnick

Reference 2022

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T15:09:40.998624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-09T15:09:40.030996Z digest=sha256:0f95969414006ee5f356d2638aed82cc2e5e404028d036ad6d17a1693bedfa04

Observation 7ddc4aff-d043-4aed-aff6-496dc141da1a · outbound

This paper cites 17 Guojun Yin, Bin Liu, Lu Sheng, Nenghai Yu, Xiaogang Wang, and Jing Shao.

End-to-end Training for Text-to-Image Synthesis using Dual-Text Embeddings 17 Guojun Yin, Bin Liu, Lu Sheng, Nenghai Yu, Xiaogang Wang, and Jing Shao

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-09T15:09:40.090251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:09:40.090251Z digest=sha256:2150cb789a707584e8e623aba36425bf0fe183d783abe31332f157f62eb4c044

Pith citing papers

No inbound Pith citation observations are available.