Pith. sign in

Paper Citation Record · LEDGER

Zero-Shot Text-to-Image Generation

As of 23 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 99 inbound Pith citation observations for arXiv:2102.12092.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2102.12092 v2

Coverage vector

measured 23 of 23 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-13T22:26:09.216328Z

measured 122 of 122 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 99 of 99 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:10:03.609971Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

23 of 23 outbound references displayed

  • verified exact0
  • verified fuzzy20
  • unresolved3
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

1135
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 2df1834a-007e-409f-bd4b-1adefed903bc · outbound

This paper cites Bowman et al.

Zero-Shot Text-to-Image Generation Bowman et al

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T22:26:09.229996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-13T22:26:09.216328Z digest=sha256:6d66891089483b9e09ba2cdec59d84ded400bc94071bc7786ea387f7a3400c50

Observation d09cce28-e648-4d67-9024-ad2ddfda28bd · outbound

This paper cites Using a linear annealing schedule for this typically led to divergence.

Zero-Shot Text-to-Image Generation Using a linear annealing schedule for this typically led to divergence

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T22:26:09.232693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-13T22:26:09.216328Z digest=sha256:aeaad6ba56f8ac0be63fc3e832aa428f782fb30f46d4bee306a14bf41cfe80d0

Observation 6baae6b4-2780-45f8-bd98-8069e679c6ef · outbound

This paper cites row, column, row, row.

Zero-Shot Text-to-Image Generation row, column, row, row

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T22:26:09.235067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-13T22:26:09.216328Z digest=sha256:cc13abe12d2af8e4c501c832ad2a3cb5556ea0596ef608f3b7e32d7ddd580e8a

Observation 40a3e61c-9bf8-4631-832b-f784773a0f72 · outbound

This paper cites Our model uses 128 gradient scales, one for each of its resblocks.

Zero-Shot Text-to-Image Generation Our model uses 128 gradient scales, one for each of its resblocks

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T22:26:09.237344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-13T22:26:09.216328Z digest=sha256:5339a17c6d73722ab825c8ce4f59491e5a50f6dd992ab205fb551d0f391d738c

Observation 9be93f41-86d9-4ecc-a7bc-2a89864e11f6 · outbound

This paper cites In particular, store all gains, biases, embeddings, and unembeddings in 32-bit precision, with 32-bit gradients (including for remote communication) and 32-bit Adam moments.

Zero-Shot Text-to-Image Generation In particular, store all gains, biases, embeddings, and unembeddings in 32-bit precision, with 32-bit gradients (including for remote communication) and 32-bit Adam moments

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T22:26:09.239872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-13T22:26:09.216328Z digest=sha256:46256341669e019283a0691ec2d3cf29755f29b37d1ecaf2cd229bd533a493e5

Observation 569541e8-230d-48bb-b3c3-6ea27e66e3cf · outbound

This paper cites For data-parallel training, we need to divide the gradients by the total number of data-parallel workers M.

Zero-Shot Text-to-Image Generation For data-parallel training, we need to divide the gradients by the total number of data-parallel workers M

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T22:26:09.242339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-13T22:26:09.216328Z digest=sha256:daec7bb89dc391b3177bfb80f3793b3c23641a43dcf823dc8349d4e3e3a38fc3

Observation 4f56d470-0934-4020-b7af-94124e1964d4 · outbound

This paper cites an unresolved cited work.

Zero-Shot Text-to-Image Generation Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-05-13T22:26:09.244289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-13T22:26:09.216328Z digest=sha256:b8199b7416f65da4a420619e26978e7288e077f06c178817f7c209774aabd158

Observation c86db7c2-cf22-410c-b415-2e040fd168dd · outbound

This paper cites Otherwise, we do nothing and proceed with backpropagation; a single nonfinite value in the gradient means that the entire update will be skipped, which happens about 5% of the time.

Zero-Shot Text-to-Image Generation Otherwise, we do nothing and proceed with backpropagation; a single nonfinite value in the gradient means that the entire update will be skipped, which happens about 5% of the time

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T22:26:09.246687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-13T22:26:09.216328Z digest=sha256:cc3ecfb2f00408774be26493513ff86d78903a15d42b1766131d375d6a7eae86

Observation 942081b7-73bb-453b-a472-14e0deb0af10 · outbound

This paper cites Both the P and Q matrices are stored in 1-6-9 format and have their values scaled by predetermined constants, as discussed in Section D.

Zero-Shot Text-to-Image Generation Both the P and Q matrices are stored in 1-6-9 format and have their values scaled by predetermined constants, as discussed in Section D

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T22:26:09.248947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-13T22:26:09.216328Z digest=sha256:336883040cfcf578a1d2ec4000c4081fc5246b9d6e57260c3df745c6d995fcea

Observation 6523559e-ee65-49c8-be55-e09f6b198f63 · outbound

This paper cites This all-reduce is carried out in the 1-6-9 format, using a custom kernel.

Zero-Shot Text-to-Image Generation This all-reduce is carried out in the 1-6-9 format, using a custom kernel

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T22:26:09.251233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-13T22:26:09.216328Z digest=sha256:00d308fcc5dd151fc9e9ffd26493a1bae572f6fd7de67e1361da8f839558d22b

Observation 0cab1972-128e-4201-8802-57967dc05e48 · outbound

This paper cites We use a custom Householder orthogonalization kernel rather than Gram-Schmidt, as we found the latter to be numerically unstable.

Zero-Shot Text-to-Image Generation We use a custom Householder orthogonalization kernel rather than Gram-Schmidt, as we found the latter to be numerically unstable

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T22:26:09.253396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-13T22:26:09.216328Z digest=sha256:211897116c3d0dc822c0559f3deb94cb284baf4579d15f3eb3999b795a7a50d5

Observation 8d255c1e-9a26-4b49-90b7-5b9eb6706e3d · outbound

This paper cites Zero-Shot Text-to-Image Generation.

Zero-Shot Text-to-Image Generation Zero-Shot Text-to-Image Generation

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T22:26:09.255503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-13T22:26:09.216328Z digest=sha256:6f6d92d0314bd69b3fdd11d6f2d3fd902f5c2ae9f07479dc0867a971bbfa1e45

Observation 2d31d7c1-283f-4ac2-bcf7-1054270d789e · outbound

This paper cites As in step (4), we clamp all infinities in the results of the all-reduce to the maximum value of the 1-6-9 format, retaining the sign.

Zero-Shot Text-to-Image Generation As in step (4), we clamp all infinities in the results of the all-reduce to the maximum value of the 1-6-9 format, retaining the sign

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T22:26:09.257566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-13T22:26:09.216328Z digest=sha256:6cfec4eb2f08ab07c7bb1dff800b62f26b357ddbaa10f6a0cc9c7f82f9cfb0d8

Observation 9e9adb57-3835-4f46-a71a-4bb2b3c7c1c4 · outbound

This paper cites Section D explains why we use 32-bit precision for these parameters and their gradients.

Zero-Shot Text-to-Image Generation Section D explains why we use 32-bit precision for these parameters and their gradients

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T22:26:09.259641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-13T22:26:09.216328Z digest=sha256:524de30028059ce32e8a555086b37ed1fa8d80d0b245aed7fa9eac406cd78711

Observation 3dcdd7e8-116b-40b3-8be8-99177b3618f8 · outbound

This paper cites an unresolved cited work.

Zero-Shot Text-to-Image Generation Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-05-13T22:26:09.261650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-13T22:26:09.216328Z digest=sha256:688f97e44a7cfaac7bb91ef48706768044f035b612197b2137569d14f05f7f22

Observation 53a9d960-1f4b-46c3-96fa-8ae917e062e5 · outbound

This paper cites an unresolved cited work.

Zero-Shot Text-to-Image Generation Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-05-13T22:26:09.263358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-13T22:26:09.216328Z digest=sha256:4d79d746770718b32671fda90eb0608ea1ebd3390f376596f309180ce60cfca6

Observation f830d2ff-c5b8-4e9a-a2df-53711d65742e · outbound

This paper cites Like backpropagation, the parameter updates proceed resblock-by-resblock.

Zero-Shot Text-to-Image Generation Like backpropagation, the parameter updates proceed resblock-by-resblock

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T22:26:09.265186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-13T22:26:09.216328Z digest=sha256:23fa8984c5b837e224d5d0b65b095e7d84dd18855fa4736fbaa196972b9c2eb9

Observation 7b04f537-34aa-4a62-9b74-9c063f1b8bd2 · outbound

This paper cites local” gradient averaged over the GPUs on the machine using reduce-scatter, and the “remote.

Zero-Shot Text-to-Image Generation local” gradient averaged over the GPUs on the machine using reduce-scatter, and the “remote

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T22:26:09.267155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-13T22:26:09.216328Z digest=sha256:ec1715bcafac09360dbdc922868e242c9d3627c079f3985e015abdd7aead8f47

Observation 74d2a88f-8c2e-427c-9a5c-b3fb07ededef · outbound

This paper cites We also note the following important optimizations.

Zero-Shot Text-to-Image Generation We also note the following important optimizations

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T22:26:09.268990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-13T22:26:09.216328Z digest=sha256:cbc12ef1b803cf5dfab9fbfd67ad92b3954d6d44208a95abe673efbd1444109b

Observation 445b2588-da2f-44f1-b389-dae00d81d4c0 · outbound

This paper cites For example, while we are running step (2) for resblock i, we can proceed to steps (3)–(8) for all resblocks j > i.

Zero-Shot Text-to-Image Generation For example, while we are running step (2) for resblock i, we can proceed to steps (3)–(8) for all resblocks j > i

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T22:26:09.270883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-13T22:26:09.216328Z digest=sha256:69030c48cf014830cd4de345bfcb536a77e26b2eff9573772ae495e96fb54bfa

Observation 9d4cc561-c45b-4429-baee-51056c1452a7 · outbound

This paper cites For example, we only prefetch the parameters from the preceding resblock when the reduce-scatter operations have finished for the current one.

Zero-Shot Text-to-Image Generation For example, we only prefetch the parameters from the preceding resblock when the reduce-scatter operations have finished for the current one

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T22:26:09.273093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-13T22:26:09.216328Z digest=sha256:d3ac4fa94d4f178856a1c098b88dadc3c02d6e5aebcf06ef96f8f492eeae8c4f

Observation 471f39fe-40ce-4329-be52-d6c76ec1911a · outbound

This paper cites The former influences the bandwidth analysis, which we present in Section E.1.

Zero-Shot Text-to-Image Generation The former influences the bandwidth analysis, which we present in Section E.1

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T22:26:09.275041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-13T22:26:09.216328Z digest=sha256:2a163b78eb61b83669aab1f287fe712903b36f474f84c8e0b807aaeb8b6983f2

Observation ab2b269c-1f6b-43f2-8c18-65a67d7b8d09 · outbound

This paper cites the exact same cat on the top as a sketch on the bottom.

Zero-Shot Text-to-Image Generation the exact same cat on the top as a sketch on the bottom

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T22:26:09.276897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-13T22:26:09.216328Z digest=sha256:2f5169cbf2116baf826a2553049a795107e348d3d46418ffbecb2a5d2468ddc0

Pith citing papers

Observation 922875b4-03ed-4001-85f0-2954776afd30 · inbound

VideoGPT: Video Generation using VQ-VAE and Transformers cites this paper.

VideoGPT: Video Generation using VQ-VAE and Transformers Zero-Shot Text-to-Image Generation

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:26:09.277502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-13T17:24:33.725187Z digest=sha256:017f0e816db347429a100b53857a1090c0c6341f2f10765b763d067938e63d27

Observation 9a92a76e-459a-4ec7-a60e-feb0554eee79 · inbound

GSPMD: General and Scalable Parallelization for ML Computation Graphs cites this paper.

GSPMD: General and Scalable Parallelization for ML Computation Graphs Zero-Shot Text-to-Image Generation

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-05-18T12:36:36.406424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-18T12:36:36.359390Z digest=sha256:4a4a46802f2ba20ded2c4d62e03b1cc384e5dce2579a02ca71443b6ae04351e9

Observation 24bf7e54-9ff9-49f8-bb5e-7a3172ac862f · inbound

Diffusion Models Beat GANs on Image Synthesis cites this paper.

Diffusion Models Beat GANs on Image Synthesis Zero-Shot Text-to-Image Generation

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:26:09.277502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-13T11:16:28.445702Z digest=sha256:663ed432af8f4e5b538ca8e73f41cbed207bc7486d147e6ba0db97e11d07426f

Observation 1ff36d7d-5011-4444-b1d1-a6a74d398e23 · inbound

Decision Transformer: Reinforcement Learning via Sequence Modeling cites this paper.

Decision Transformer: Reinforcement Learning via Sequence Modeling Zero-Shot Text-to-Image Generation

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-05-18T15:11:11.267437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-18T15:11:11.056013Z digest=sha256:d17ee98ee7733010954b5d38c079f2186d7748a3aa245f7ea2f74f3e1890888f

Observation 6376ebb9-6f71-4d83-b400-512f9cc54ad4 · inbound

BEiT: BERT Pre-Training of Image Transformers cites this paper.

BEiT: BERT Pre-Training of Image Transformers Zero-Shot Text-to-Image Generation

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T22:26:09.277502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-13T11:50:11.476015Z digest=sha256:6a51307b40946ab922ddf194bbe690415486c494860641bcbf122a4b2816ff0e

Observation fc5583c4-dda5-4f4f-8bf0-a20117cfbb40 · inbound

LAION-400M: Open Dataset of CLIP-Filtered 400 Million Image-Text Pairs cites this paper.

LAION-400M: Open Dataset of CLIP-Filtered 400 Million Image-Text Pairs Zero-Shot Text-to-Image Generation

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:26:09.277502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-12T10:21:01.062199Z digest=sha256:49a55216a724d49a3fe9f5f6cd79200f5bede7e848b55f744075c8052d84d897

Observation f8fccb61-5695-4904-b2a8-14516ad85837 · inbound

Florence: A New Foundation Model for Computer Vision cites this paper.

Florence: A New Foundation Model for Computer Vision Zero-Shot Text-to-Image Generation

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-05-16T09:38:09.553108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-16T09:38:09.427509Z digest=sha256:a7a3d5b268f4b03d54fe8590920557cfa716cc80a4e57c07b59527aa028d100c

Observation 0c6c1574-9604-4b75-82f7-5e11371a4792 · inbound

GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models cites this paper.

GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models Zero-Shot Text-to-Image Generation

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:26:09.277502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-11T05:56:10.970591Z digest=sha256:79c39d5c0451cc41f5e2362bc00c85842981254713ccf5f6a6bc97ed8507e36a

Observation 2fa11952-b0a5-4e13-8ed9-47c9871090b1 · inbound

High-Resolution Image Synthesis with Latent Diffusion Models cites this paper.

High-Resolution Image Synthesis with Latent Diffusion Models Zero-Shot Text-to-Image Generation

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:26:09.277502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-11T22:02:09.899347Z digest=sha256:737aa4aeb75d879fbce9892228f71b5b7f31a213194ac0a62a452b395c168da2

Observation 0b517447-822b-4ca2-8206-da8b21c1210e · inbound

Text and Code Embeddings by Contrastive Pre-Training cites this paper.

Text and Code Embeddings by Contrastive Pre-Training Zero-Shot Text-to-Image Generation

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-15T19:24:12.045243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-15T19:24:11.907204Z digest=sha256:703f33b376649d81c8749540fcdba5ee71e356f16520509d1a573b073f9d4fb7

Observation af0016f7-f774-4f90-8bf2-504571599d19 · inbound

Hierarchical Text-Conditional Image Generation with CLIP Latents cites this paper.

Hierarchical Text-Conditional Image Generation with CLIP Latents Zero-Shot Text-to-Image Generation

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:26:09.277502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-10T16:55:57.612364Z digest=sha256:0ea90b9b3b2ab8fc8021bf1bababe3fe7cfe58f5990ea4cacaf285a07d3d6fbe

Observation 50945b12-ede7-4cb7-9194-8b2853bd734c · inbound

CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers cites this paper.

CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers Zero-Shot Text-to-Image Generation

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:26:09.277502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-11T12:24:30.071822Z digest=sha256:8c1674a1bfe26dcc2b880bdda6fe72441bae455c2038b7a3c8fc9f5d093a87e2

Observation cf2d3381-bdfe-46ea-9459-2cda4f9010ab · inbound

LLM.int8(): 8-bit Matrix Multiplication for Transformers at Scale cites this paper.

LLM.int8(): 8-bit Matrix Multiplication for Transformers at Scale Zero-Shot Text-to-Image Generation

Reference 66

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T22:26:09.277502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-05-13T13:35:35.972596Z digest=sha256:ea796e545fb697f83b0058b8f641843c95c352fba15eb8a91eb34bff9d86d60c

Observation f1a8f245-3849-4479-9117-cc958aba711f · inbound

DetailCLIP: Injecting Image Details into CLIP's Feature Space cites this paper.

DetailCLIP: Injecting Image Details into CLIP's Feature Space Zero-Shot Text-to-Image Generation

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-05-24T11:09:22.462818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-24T11:08:20.298043Z digest=sha256:57e5e8b56ad742797707d35c40348a694ac508fe40bc905057a8dfedfd292a32

Observation 75ed9063-47eb-4890-a334-3c92310e7798 · inbound

LAION-5B: An open large-scale dataset for training next generation image-text models cites this paper.

LAION-5B: An open large-scale dataset for training next generation image-text models Zero-Shot Text-to-Image Generation

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:26:09.277502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-13T14:22:16.968028Z digest=sha256:68603b4ba012df32b386d97b40994de320dc05a5f4982468557851e0161d8988

Observation 8197d098-652e-44a6-945c-8404abe851bb · inbound

EVA-CLIP: Improved Training Techniques for CLIP at Scale cites this paper.

EVA-CLIP: Improved Training Techniques for CLIP at Scale Zero-Shot Text-to-Image Generation

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:26:09.277502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-13T01:54:21.943160Z digest=sha256:ed88ad43e6f7d8fd2c720a5cc7394623bc1e55d66627362573fff3a5c8647164

Observation 0c4fad37-f6a9-4397-a4cb-ccecf1f5976f · inbound

Shap-E: Generating Conditional 3D Implicit Functions cites this paper.

Shap-E: Generating Conditional 3D Implicit Functions Zero-Shot Text-to-Image Generation

Reference 49

Resolution
verified exact
local_arxiv, observed 2026-05-16T15:32:06.848362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-16T15:32:06.563955Z digest=sha256:b9dada297ad1d67f10d8f2c6c20e1cde7b892c47b03d5e32be402bdbcc678336

Observation fe3a44d7-3951-461e-9de8-f7a6c19d8c3d · inbound

Training Diffusion Models with Reinforcement Learning cites this paper.

Training Diffusion Models with Reinforcement Learning Zero-Shot Text-to-Image Generation

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:26:09.277502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-11T20:16:30.840184Z digest=sha256:2f280e9115d98ceaddff548d02697cece1d49511f014a5f955e49be485443f17

Observation ee2c7628-5660-476a-8c15-f2afeeaa2e90 · inbound

Human Preference Score v2: A Solid Benchmark for Evaluating Human Preferences of Text-to-Image Synthesis cites this paper.

Human Preference Score v2: A Solid Benchmark for Evaluating Human Preferences of Text-to-Image Synthesis Zero-Shot Text-to-Image Generation

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:26:09.277502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-11T08:30:48.564596Z digest=sha256:8034890b10abc1f0f1ed0bfc423c6cde3e6bde317475eae59f50f8148d608e13

Observation 602a8fc1-fbec-47e9-a7e9-4f73de834801 · inbound

Demystifying CLIP Data cites this paper.

Demystifying CLIP Data Zero-Shot Text-to-Image Generation

Reference 104

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T09:20:20.364390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-05-16T09:20:20.143143Z digest=sha256:4f0cc3b85c93fd4b7d28b08b223f45f298cfcd2075a1f7b359fdc19d8b53cd95

Observation 99106562-d136-4f96-a72f-81dee26e93d4 · inbound

VideoPoet: A Large Language Model for Zero-Shot Video Generation cites this paper.

VideoPoet: A Large Language Model for Zero-Shot Video Generation Zero-Shot Text-to-Image Generation

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-05-15T17:51:05.555998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-15T17:51:05.465548Z digest=sha256:e913d779a47218daebea008830f727f82c8f23ec219267e854d71bd63998da99

Observation ac7a1ce9-c5f0-45e5-ad1c-5cd9199600bd · inbound

Chameleon: Mixed-Modal Early-Fusion Foundation Models cites this paper.

Chameleon: Mixed-Modal Early-Fusion Foundation Models Zero-Shot Text-to-Image Generation

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:26:09.277502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-11T10:03:27.919346Z digest=sha256:68cabab2cae454592ecf3bd261decd4ca916e9b94a504950a7757ee414902afe

Observation ab3c2451-2941-4bb4-8ee9-44176c5e732c · inbound

Emu3: Next-Token Prediction is All You Need cites this paper.

Emu3: Next-Token Prediction is All You Need Zero-Shot Text-to-Image Generation

Reference 71

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:26:09.277502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-11T10:56:06.418360Z digest=sha256:b7fcb97bd8b1a28730d19db2e6988dfbeb063c905f142b6943d96e29587e8818

Observation eb22f116-67b5-428a-b15e-c5974a7b51ce · inbound

Mixture-of-Transformers: A Sparse and Scalable Architecture for Multi-Modal Foundation Models cites this paper.

Mixture-of-Transformers: A Sparse and Scalable Architecture for Multi-Modal Foundation Models Zero-Shot Text-to-Image Generation

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-05-18T02:48:45.127956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-18T02:48:44.900467Z digest=sha256:992fb5d9fb414121997065a74880488a44cb85a4ea564e442defce0596be219b

Observation ac0a63f5-98b5-498d-9094-df06117b7f24 · inbound

Architect: Generating Vivid and Interactive 3D Scenes with Hierarchical 2D Inpainting cites this paper.

Architect: Generating Vivid and Interactive 3D Scenes with Hierarchical 2D Inpainting Zero-Shot Text-to-Image Generation

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T20:20:21.407122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:20:21.407122Z digest=sha256:b71281aadf18e688fcebff41aaac3073567b719795a9f03c0d8a706f7333f78d

Observation 17adc1ba-cee6-4d7b-b41b-35329ce75b32 · inbound

Multidimensional Byte Pair Encoding: Shortened Sequences for Improved Visual Data Generation cites this paper.

Multidimensional Byte Pair Encoding: Shortened Sequences for Improved Visual Data Generation Zero-Shot Text-to-Image Generation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-12T19:55:33.894665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:55:33.894665Z digest=sha256:3ad51205aeaf9b2b226c739b92103e6baaaf17dd8f429e8b273fc68cd9783af6

Observation 3329b436-f4f8-4f60-9acb-f93b10840fb2 · inbound

RPN 2: On Interdependence Function Learning Towards Unifying and Advancing CNN, RNN, GNN, and Transformer cites this paper.

RPN 2: On Interdependence Function Learning Towards Unifying and Advancing CNN, RNN, GNN, and Transformer Zero-Shot Text-to-Image Generation

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-12T18:56:48.035436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:56:48.035436Z digest=sha256:1ea70bbd52fca1c79f427b6ee397303b0a7beb0d1efc1f9792facda17ffb72a8

Observation 93aa862c-c758-4532-8ab0-5fb96977256b · inbound

ScImage: How Good Are Multimodal Large Language Models at Scientific Text-to-Image Generation? cites this paper.

ScImage: How Good Are Multimodal Large Language Models at Scientific Text-to-Image Generation? Zero-Shot Text-to-Image Generation

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T23:36:24.751632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T23:36:24.751632Z digest=sha256:f5cc8602d7c9fd74d407b44a5c9a92ed712c9d224aa42b87c344d4c694b7442f

Observation 964dda29-066a-416f-b379-c232fca7c0d1 · inbound

The Efficacy of Transfer-based No-box Attacks on Image Watermarking: A Pragmatic Analysis cites this paper.

The Efficacy of Transfer-based No-box Attacks on Image Watermarking: A Pragmatic Analysis Zero-Shot Text-to-Image Generation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-11T23:23:08.472627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:23:08.472627Z digest=sha256:a43563d79a5404067f340154c1cc51ed78b077651d381fbbdb17829af73631e2

Observation 17fbc4f0-8c30-4718-8741-9644ab4c5291 · inbound

Medical Multimodal Foundation Models in Clinical Diagnosis and Treatment: Applications, Challenges, and Future Directions cites this paper.

Medical Multimodal Foundation Models in Clinical Diagnosis and Treatment: Applications, Challenges, and Future Directions Zero-Shot Text-to-Image Generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T23:17:44.703124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:17:44.703124Z digest=sha256:d67ea473250a675baf226c215786d94e308094f12bd5b44949efa6d3dd28a82b

Observation 20d8dc6f-8bda-4245-951d-cea58b5dd8ec · inbound

Diffusion-based Visual Anagram as Multi-task Learning cites this paper.

Diffusion-based Visual Anagram as Multi-task Learning Zero-Shot Text-to-Image Generation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-11T23:15:26.627187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:15:26.627187Z digest=sha256:8b0d2f1e7402503a4e4e1cee1993253f9b21a3936747b958bf69b02fc18e7d26

Observation 8cd8fda4-7f6b-40fe-9be2-0c73a644c492 · inbound

PyPotteryLens: An Open-Source Deep Learning Framework for Automated Digitisation of Archaeological Pottery Documentation cites this paper.

PyPotteryLens: An Open-Source Deep Learning Framework for Automated Digitisation of Archaeological Pottery Documentation Zero-Shot Text-to-Image Generation

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-11T14:52:42.167477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T14:52:42.167477Z digest=sha256:0aeec5f51c4c93b3833fdf7c8ea142fdb6d8fc58975cf0bcd10131cef8fefe06

Observation 8172c3f8-a721-4b5e-8508-f550d531a49e · inbound

From Specific-MLLMs to Omni-MLLMs: A Survey on MLLMs Aligned with Multi-modalities cites this paper.

From Specific-MLLMs to Omni-MLLMs: A Survey on MLLMs Aligned with Multi-modalities Zero-Shot Text-to-Image Generation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T14:43:41.822155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:43:41.822155Z digest=sha256:573a4ab6a788a7fd0064741ad77fd65109437c774dcf3359e99ae6336fb64278

Observation f9f04929-f78b-4068-af89-a9a8d4f6dbc6 · inbound

Self-control: A Better Conditional Mechanism for Masked Autoregressive Model cites this paper.

Self-control: A Better Conditional Mechanism for Masked Autoregressive Model Zero-Shot Text-to-Image Generation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T12:58:45.076231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:58:45.076231Z digest=sha256:bcddfbac063406136df7169c13265537b4d01e58cc5d488ba128f4a90c0400ca

Observation 7721fd55-6403-4212-aa69-c8188b473bcf · inbound

WikiStyle+: A Multimodal Approach to Content-Style Representation Disentanglement for Artistic Image Stylization cites this paper.

WikiStyle+: A Multimodal Approach to Content-Style Representation Disentanglement for Artistic Image Stylization Zero-Shot Text-to-Image Generation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T12:13:43.931325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:13:43.931325Z digest=sha256:249eb79ba96e13bc5ae22dafd4149b8b0e4e3b0c072bea02e52bb645bf5e203d

Observation 6b601167-6764-4032-90e2-5fff56158bb9 · inbound

A Decade of Deep Learning: A Survey on The Magnificent Seven cites this paper.

A Decade of Deep Learning: A Survey on The Magnificent Seven Zero-Shot Text-to-Image Generation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T16:10:24.504494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:10:24.504494Z digest=sha256:a6c9ce59429bac01b50f2aad0b68819d1b584b738609788e65fa1a157b8f7574

Observation 260a4a83-c0b8-4713-b467-35dcc480bf21 · inbound

Ethics and Technical Aspects of Generative AI Models in Digital Content Creation cites this paper.

Ethics and Technical Aspects of Generative AI Models in Digital Content Creation Zero-Shot Text-to-Image Generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T10:40:07.960038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:40:07.960038Z digest=sha256:9379dd865e1de2c06b55c51869bbae276c140e64a28f76b5d598fa71e96ce9e8

Observation 63760278-8bc4-4dce-9351-2347e91a5e8a · inbound

From Creation to Curriculum: Examining the role of generative AI in Arts Universities cites this paper.

From Creation to Curriculum: Examining the role of generative AI in Arts Universities Zero-Shot Text-to-Image Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T10:32:15.760853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:32:15.760853Z digest=sha256:2534806a93aee7033569930145c916041a4c447d92c14a21a7dfbca703d495a9

Observation 4aa705af-f3c3-46ee-b47e-9865cf5c54f4 · inbound

SubstationAI: Multimodal Large Model-Based Approaches for Analyzing Substation Equipment Faults cites this paper.

SubstationAI: Multimodal Large Model-Based Approaches for Analyzing Substation Equipment Faults Zero-Shot Text-to-Image Generation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T05:51:35.200769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:51:35.200769Z digest=sha256:ea661802e17c1cd5a7a5af4ab65930ed81f29d00fd53d94a83f67217e1745f73

Observation 9578345e-a5e4-4d21-93de-e1410f1aeee4 · inbound

Generative Emergent Communication: Large Language Model is a Collective World Model cites this paper.

Generative Emergent Communication: Large Language Model is a Collective World Model Zero-Shot Text-to-Image Generation

Reference 105

Resolution
unresolved
no resolver link, observed 2026-08-10T23:01:43.140233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T23:01:43.140233Z digest=sha256:3162b03ba0bc001b1e5eb3feec7c5783fc04d219de5e3db495fd62e77fce9e5c

Observation da25f774-4068-44b7-85fe-17d7c4aed2de · inbound

EditAR: Unified Conditional Generation with Autoregressive Models cites this paper.

EditAR: Unified Conditional Generation with Autoregressive Models Zero-Shot Text-to-Image Generation

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-10T21:29:20.056049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:29:20.056049Z digest=sha256:e7fc5f0273399e254881c7c3449fab4d0be7780c8fdbbcfee833abbecce69a57

Observation 1520218f-b7bf-49df-9035-d0f403fdc499 · inbound

Neuro-Symbolic AI in 2024: A Systematic Review cites this paper.

Neuro-Symbolic AI in 2024: A Systematic Review Zero-Shot Text-to-Image Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T21:16:56.168596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:16:56.168596Z digest=sha256:1531d66eba27d15336845aa7483ed7574c48dcea46478a5faea8fd3654e9ab36

Observation ac003d5a-6afe-44cf-9274-1b809f553eb2 · inbound

Averaged Adam accelerates stochastic optimization in the training of deep neural network approximations for partial differential equation and optimal control problems cites this paper.

Averaged Adam accelerates stochastic optimization in the training of deep neural network approximations for partial differential equation and optimal control problems Zero-Shot Text-to-Image Generation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T21:11:06.855735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:11:06.855735Z digest=sha256:08b8e50107d324848c21ca57dc5ffd2b3edf0bb933c247c48860386b994fc0c8

Observation 4987f367-fc3d-46bb-b7ad-78b5f5449137 · inbound

Flow: Modularized Agentic Workflow Automation cites this paper.

Flow: Modularized Agentic Workflow Automation Zero-Shot Text-to-Image Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T20:37:40.418097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:37:40.418097Z digest=sha256:644569bb72e594ea10b76cebdc65c0c2959b8f0dcb546668f39e45e723c6b9ac

Observation df07b8d1-83e3-4e44-b311-ef02de45d24d · inbound

CityLoc: 6DoF Pose Distributional Localization for Text Descriptions in Large-Scale Scenes with Gaussian Representation cites this paper.

CityLoc: 6DoF Pose Distributional Localization for Text Descriptions in Large-Scale Scenes with Gaussian Representation Zero-Shot Text-to-Image Generation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T20:16:33.462899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:16:33.462899Z digest=sha256:16122de347139696d3ad0d79d22802b6210f2f51edbe43e05bfe1c903de5f43e

Observation 0fd74360-c00e-47ac-959d-3f07d4277060 · inbound

Taming Teacher Forcing for Masked Autoregressive Video Generation cites this paper.

Taming Teacher Forcing for Masked Autoregressive Video Generation Zero-Shot Text-to-Image Generation

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-10T17:16:47.533348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:16:47.533348Z digest=sha256:1b567a4539af4f73ac27fcf3e8eee306c279976470ef279f1970333f32bafbab

Observation 55c90488-9a5d-4b72-9d27-26e744e4f188 · inbound

CARING-AI: Towards Authoring Context-aware Augmented Reality INstruction through Generative Artificial Intelligence cites this paper.

CARING-AI: Towards Authoring Context-aware Augmented Reality INstruction through Generative Artificial Intelligence Zero-Shot Text-to-Image Generation

Reference 108

Resolution
unresolved
no resolver link, observed 2026-08-10T12:25:31.481974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T12:25:31.481974Z digest=sha256:819c85d7844a22981ebd4e201f13b73bc62abe585589a1a5043797d9401a550f

Observation 03148481-167a-4e91-b3ed-651a88b27d76 · inbound

CLIP-UP: A Simple and Efficient Mixture-of-Experts CLIP Training Recipe with Sparse Upcycling cites this paper.

CLIP-UP: A Simple and Efficient Mixture-of-Experts CLIP Training Recipe with Sparse Upcycling Zero-Shot Text-to-Image Generation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-09T17:09:26.859300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:09:26.859300Z digest=sha256:9a628f6505483cdcf9c64bdcb161e26aed5b81d4f00ecf96b7d625f9c8c25fd8

Observation 9e2e35ea-cd90-488c-9099-f03f4104af40 · inbound

End-to-end Training for Text-to-Image Synthesis using Dual-Text Embeddings cites this paper.

End-to-end Training for Text-to-Image Synthesis using Dual-Text Embeddings Zero-Shot Text-to-Image Generation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-09T15:09:40.047361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:09:40.047361Z digest=sha256:65123520ec9817a93992b37de78a04cf289f4b6de209867cbc08f5a59db4c417

Observation a399d453-841b-4776-900f-7c8000d4f6a6 · inbound

PyPotteryInk: One-Step Diffusion Model for Sketch to Publication-ready Archaeological Drawings cites this paper.

PyPotteryInk: One-Step Diffusion Model for Sketch to Publication-ready Archaeological Drawings Zero-Shot Text-to-Image Generation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-08T17:31:07.823377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:31:07.823377Z digest=sha256:63d3065c8a29cd2933ab2749a2246dbd666bc530afbb71b6d76c9f4c0741cda5

Observation 10fb234d-9106-4e0f-ae3f-286d9e497118 · inbound

Inferring Questions from Programming Screenshots cites this paper.

Inferring Questions from Programming Screenshots Zero-Shot Text-to-Image Generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T10:10:03.609971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:10:03.609971Z digest=sha256:31f4422d085b77c3f4694ea2b44343d71a0ee4967926f18a73d5a5d88e920757

Observation f9dd7d2b-87b7-44c1-8759-a7eb4f4bf67b · inbound

WILD: a new in-the-Wild Image Linkage Dataset for synthetic image attribution cites this paper.

WILD: a new in-the-Wild Image Linkage Dataset for synthetic image attribution Zero-Shot Text-to-Image Generation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-16T05:52:01.062217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:52:01.062217Z digest=sha256:29eec7295a2572077a479f210d39f02a669d8f527cbbe19dbe1817493812045e

Observation cf02a480-416b-4182-96db-3a4fe7b04987 · inbound

Deepfakes on Demand: the rise of accessible non-consensual deepfake image generators cites this paper.

Deepfakes on Demand: the rise of accessible non-consensual deepfake image generators Zero-Shot Text-to-Image Generation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-15T23:50:44.033955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:50:44.033955Z digest=sha256:e6cd3fc32e5f820b18ad514825bd5e49c8891d95e6c61b1f334ad3ab6ce5260f

Observation 8ef2e3cb-62d4-4fd5-86fb-79808f7ebba0 · inbound

DiffCrysGen: A Score-Based Diffusion Model for Design of Diverse Inorganic Crystalline Materials cites this paper.

DiffCrysGen: A Score-Based Diffusion Model for Design of Diverse Inorganic Crystalline Materials Zero-Shot Text-to-Image Generation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T22:21:29.620026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T22:21:29.620026Z digest=sha256:1e9a1f941acf5bdf5db4a5f503cbc95fba1f230a7a7fceed5049511a4111dfb9

Observation 473dd928-3fa7-4119-8bc4-08fb9121c5f9 · inbound

Multi-modal Synthetic Data Training and Model Collapse: Insights from VLMs and Diffusion Models cites this paper.

Multi-modal Synthetic Data Training and Model Collapse: Insights from VLMs and Diffusion Models Zero-Shot Text-to-Image Generation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T22:37:17.594094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:37:17.594094Z digest=sha256:db79c34e7a0ce4e25c07d1b2c1639920f59292ea8a850252fe2302fe27a66eb1

Observation 4aa9d947-798c-4e0f-a7b6-5bbff39e8e10 · inbound

IMAGE-ALCHEMY: Advancing subject fidelity in personalised text-to-image generation cites this paper.

IMAGE-ALCHEMY: Advancing subject fidelity in personalised text-to-image generation Zero-Shot Text-to-Image Generation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:44.430532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:44.430532Z digest=sha256:a1bf40e48cb0e5a604675c89e15cf91cae5f2123be628ebb6eedde6392abb510

Observation 9e5ad02f-13c7-4bc8-a277-3c4eb8f58e2a · inbound

A collaborative constrained graph diffusion model for the generation of realistic synthetic molecules cites this paper.

A collaborative constrained graph diffusion model for the generation of realistic synthetic molecules Zero-Shot Text-to-Image Generation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:16.889616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:16.889616Z digest=sha256:a665fdbafb22ff1e5779d8b591a702001bdd8756e72612c2964fee17e54f7dfa

Observation e27c383d-4ec4-44e8-8645-9ec6af9eb9ca · inbound

Mitigate One, Skew Another? Tackling Intersectional Biases in Text-to-Image Models cites this paper.

Mitigate One, Skew Another? Tackling Intersectional Biases in Text-to-Image Models Zero-Shot Text-to-Image Generation

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:59.194602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:52:59.194602Z digest=sha256:8e9fea3ce930cc3a44a0234384d01000f9fffb55d337dae95bfb9397cb0c5358

Observation 8b5878c5-cde4-430d-8d15-44266ed48e53 · inbound

Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion cites this paper.

Multimodal LLM-Guided Semantic Correction in Text-to-Image Diffusion Zero-Shot Text-to-Image Generation

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T14:06:52.551359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:06:52.551359Z digest=sha256:61aad63c6fe03688387cf589d4aedef61399f415837b743a70fe9884307a0f00

Observation 2d44ce61-8d6b-448f-b308-199e9a634726 · inbound

EgoZero: Robot Learning from Smart Glasses cites this paper.

EgoZero: Robot Learning from Smart Glasses Zero-Shot Text-to-Image Generation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:56.738485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:56.738485Z digest=sha256:78ffa13141e4f8e9ac9ea30f36b9627013ce5f0e37455d5b25dc439f03c21f2a

Observation 821a4347-7511-40bb-aac1-0a49b15ef796 · inbound

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning cites this paper.

PADAM: Parallel averaged Adam reduces the error for stochastic optimization in scientific machine learning Zero-Shot Text-to-Image Generation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T13:20:48.749797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:20:48.749797Z digest=sha256:ea708a86b96a00f735ddef339cf1709893571d3b9862506135076df89bce3565

Observation 0b42df90-d4f4-4cfd-8661-84c51b32b6e1 · inbound

Fine-Tuning Next-Scale Visual Autoregressive Models with Group Relative Policy Optimization cites this paper.

Fine-Tuning Next-Scale Visual Autoregressive Models with Group Relative Policy Optimization Zero-Shot Text-to-Image Generation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:42.565243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:52:42.565243Z digest=sha256:4ece23de6e6c0d00b1f307c0c02053b9814d9c4a7dd3a47821643b71921f510b

Observation 442a0f88-0949-4fb7-a7d1-8147c8ff6807 · inbound

Humanoid World Models: Open World Foundation Models for Humanoid Robotics cites this paper.

Humanoid World Models: Open World Foundation Models for Humanoid Robotics Zero-Shot Text-to-Image Generation

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-07T11:54:56.809361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:54:56.809361Z digest=sha256:46a971657def3d881fb78c3945b2969fd0fc6fbfca1ae2ec898aa2d690636faa

Observation 96cc0822-947f-497b-bf70-8c2d52e32e5b · inbound

SmartAvatar: Text- and Image-Guided Human Avatar Generation with VLM AI Agents cites this paper.

SmartAvatar: Text- and Image-Guided Human Avatar Generation with VLM AI Agents Zero-Shot Text-to-Image Generation

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-07T10:41:39.779709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:41:39.779709Z digest=sha256:52572fbd46b22cd6a71b40c211ebd828de3c2a5b4980045d6c865acf9fca8f25

Observation 53568e40-68f5-4841-8ef7-5761d669e7ca · inbound

FaSTA$^*$: Fast-Slow Toolpath Agent with Subroutine Mining for Efficient Multi-turn Image Editing cites this paper.

FaSTA$^*$: Fast-Slow Toolpath Agent with Subroutine Mining for Efficient Multi-turn Image Editing Zero-Shot Text-to-Image Generation

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-05-19T08:17:10.821005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-19T08:17:08.223736Z digest=sha256:02d3de3d8172ebd55314759d9c14db6be6891022c57b0336184ce429c304f32a

Observation d57aa31a-e186-49bc-b3a0-8185039cb1fb · inbound

On the Resilience of Underwater Semantic Wireless Communications cites this paper.

On the Resilience of Underwater Semantic Wireless Communications Zero-Shot Text-to-Image Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T21:48:13.474408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:48:13.474408Z digest=sha256:958473b4d6f7d13da8d13da4c57be19bc527b2b130cdf97ef4ca9f49b01689e8

Observation 0f6f9a7e-7e96-4126-97e6-bad372c2e798 · inbound

Transition Matching: Scalable and Flexible Generative Modeling cites this paper.

Transition Matching: Scalable and Flexible Generative Modeling Zero-Shot Text-to-Image Generation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T21:46:28.684612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:46:28.684612Z digest=sha256:edeedfe481ebbcda7941283d4f3963987412e53e00bd6a820675ae289a658078

Observation 3af5aa16-df5c-464d-9d02-6001571e1da0 · inbound

GraphBrep: Learning B-Rep in Graph Structure for Efficient CAD Generation cites this paper.

GraphBrep: Learning B-Rep in Graph Structure for Efficient CAD Generation Zero-Shot Text-to-Image Generation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T19:44:19.898429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:44:19.898429Z digest=sha256:31a2a8bfb26e9f9413eb10585f393636ce7ea2cd479b94572e5e832efc8da345

Observation 3b8c973b-ba5c-4166-9c49-159b4e9bd0fa · inbound

Structured Captions Improve Prompt Adherence in Text-to-Image Models (Re-LAION-Caption 19M) cites this paper.

Structured Captions Improve Prompt Adherence in Text-to-Image Models (Re-LAION-Caption 19M) Zero-Shot Text-to-Image Generation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:21.587329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:21.587329Z digest=sha256:fcd8e13d2cc81cebd4ef21b75673a60a4cf8e96c3d51a5a8b5a0e27632b2432e

Observation d182779b-7dae-4d12-baf7-e4400ab244e8 · inbound

A Review of Generative AI in Aquaculture: Foundations, Applications, and Future Directions for Smart and Sustainable Farming cites this paper.

A Review of Generative AI in Aquaculture: Foundations, Applications, and Future Directions for Smart and Sustainable Farming Zero-Shot Text-to-Image Generation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T17:00:19.177529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:00:19.177529Z digest=sha256:3136ffcbcfbaa573d280ac0fec95630f2e9af494c56630b676030fa404516db8

Observation eb1f0315-49bf-4810-b4ff-4195a773b720 · inbound

Unmasking Synthetic Realities in Generative AI: A Comprehensive Review of Adversarially Robust Deepfake Detection Systems cites this paper.

Unmasking Synthetic Realities in Generative AI: A Comprehensive Review of Adversarially Robust Deepfake Detection Systems Zero-Shot Text-to-Image Generation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T14:34:10.580858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:34:10.580858Z digest=sha256:08ec1272064ed6db1303c523e33205b204ea52e06f24356c1782cb6b261dd3e0

Observation d420a61d-75c9-40f8-8e2d-9502a5ccbcb4 · inbound

Agency Among Agents: Designing with Hypertextual Friction in the Algorithmic Web cites this paper.

Agency Among Agents: Designing with Hypertextual Friction in the Algorithmic Web Zero-Shot Text-to-Image Generation

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T10:37:51.194375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:37:51.194375Z digest=sha256:16c1dbb54c94970c9d1c1ddaf93bc1f2bc756b6192f9e257f5de37957bd12f0e

Observation 725a87a1-2138-4c38-93a8-61406a50411f · inbound

Real-Time 3D Vision-Language Embedding Mapping cites this paper.

Real-Time 3D Vision-Language Embedding Mapping Zero-Shot Text-to-Image Generation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-05T22:52:02.653418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:52:02.653418Z digest=sha256:641ec82eaffd14b792d2753d6d35cbe2c5ac1506cfbb680673b542ec4d0e2cd5

Observation 44de9bf5-2e6d-4854-934b-3df4e357d917 · inbound

Understanding and evaluating computer vision models through the lens of counterfactuals cites this paper.

Understanding and evaluating computer vision models through the lens of counterfactuals Zero-Shot Text-to-Image Generation

Reference 210

Resolution
unresolved
no resolver link, observed 2026-08-05T14:49:31.867345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:49:31.867345Z digest=sha256:58a499cb6975f17adbc6c17ad9922732c726a8f531a8b2075ae68bd1a63cf2db

Observation 87d9060f-42ac-4c74-b9f8-de95ef1ef747 · inbound

Effectively obtaining acoustic, visual and textual data from videos cites this paper.

Effectively obtaining acoustic, visual and textual data from videos Zero-Shot Text-to-Image Generation

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-05T05:01:36.921533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:01:36.921533Z digest=sha256:4d752dc8fc23561261f3421838bfc0e29861f34918bf880d800392e1e21b814e

Observation e4bfc21b-1c53-4578-98e0-b9f4ab7a1ca6 · inbound

Testing chatbots on the creation of encoders for audio conditioned image generation cites this paper.

Testing chatbots on the creation of encoders for audio conditioned image generation Zero-Shot Text-to-Image Generation

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-04T21:25:27.114736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T21:25:27.114736Z digest=sha256:ababaf7ae75cb3a14ebdb926cd159b2c00253b0d7f632707ea593d7cb9e1f7b3

Observation 314eb7db-ce21-4ede-8aab-e3a66df32586 · inbound

Immunizing Images from Text to Image Editing via Adversarial Cross-Attention cites this paper.

Immunizing Images from Text to Image Editing via Adversarial Cross-Attention Zero-Shot Text-to-Image Generation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T17:56:30.917048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:56:30.917048Z digest=sha256:72ce6041d83461b8f8998c82b4f7125821bc4631d85c20228993013bf50b1139

Observation e343fc6e-b8ed-4bc9-a304-3693b1a679ee · inbound

Proto-LeakNet: Towards Signal-Leak Aware Attribution in Synthetic Human Face Imagery cites this paper.

Proto-LeakNet: Towards Signal-Leak Aware Attribution in Synthetic Human Face Imagery Zero-Shot Text-to-Image Generation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-03T23:48:03.476241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:48:03.476241Z digest=sha256:e0bfb9e4cdf09785bf74b1c39ece3ba2b95a86e55e871e4d304823867a6146d2

Observation 5bcfd51e-592b-4b97-859f-b9a5ce9e62df · inbound

Video Deepfake Abuse: How Company Choices Predictably Shape Misuse Patterns cites this paper.

Video Deepfake Abuse: How Company Choices Predictably Shape Misuse Patterns Zero-Shot Text-to-Image Generation

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-03T19:59:09.485773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T19:59:09.485773Z digest=sha256:4b7753286628f10f133ac5140f4304b236c5121fb638fbde05c0d9f57ae898a2

Observation 15ec77e4-7ecd-46b3-81c4-aee1efdfc83c · inbound

Agents of Diffusion: Enhancing Diffusion Language Models with Multi-Agent Reinforcement Learning for Structured Data Generation (Extended Version) cites this paper.

Agents of Diffusion: Enhancing Diffusion Language Models with Multi-Agent Reinforcement Learning for Structured Data Generation (Extended Version) Zero-Shot Text-to-Image Generation

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-03T11:14:51.746511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T11:14:51.746511Z digest=sha256:767d2a8f79ca52da20260deb25f978980fad4f8f1ade0dab0692a1f3761f16a3

Observation f810abbb-7911-41d7-9dbf-037cce92ce3e · inbound

A Marketplace for AI-Generated Adult Content and Deepfakes cites this paper.

A Marketplace for AI-Generated Adult Content and Deepfakes Zero-Shot Text-to-Image Generation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T10:46:23.990908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:46:23.990908Z digest=sha256:e80310e183b2667217459c0998711cadb32a5cb44af1ec51c1e8ce98eaf14e5a

Observation 8c43d029-a5dd-4144-8e51-8b80d202809b · inbound

Predicting integers from continuous parameters cites this paper.

Predicting integers from continuous parameters Zero-Shot Text-to-Image Generation

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-03T01:02:56.438442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:02:56.438442Z digest=sha256:e0db9810a6a327171c567416c8e10314622520fa6ea07a295f0182d160839d92

Observation 5121d0bd-7af6-438b-a861-bc4cea874533 · inbound

SEDGE: Structural Extrapolated Data Generation cites this paper.

SEDGE: Structural Extrapolated Data Generation Zero-Shot Text-to-Image Generation

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T22:26:09.277502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-13T21:04:58.515329Z digest=sha256:d78f250f79d0b17f5c81381c7193e2d9b6094e81b1f5d050967d6c070ae24498

Observation 3704c1c8-779a-4ae3-b6e6-6e9d4745b083 · inbound

SEDGE: Structural Extrapolated Data Generation cites this paper.

SEDGE: Structural Extrapolated Data Generation Zero-Shot Text-to-Image Generation

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-05-15T07:35:13.748025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-15T07:34:02.199098Z digest=sha256:cba3cef50bc9afcc1e97fe3808ed6aeb2640af258425463db46d328ff3020910

Observation 0b63dda0-e888-42a1-bb2d-d8a9b7ac3d47 · inbound

LiveGesture Streamable Co-Speech Gesture Generation Model cites this paper.

LiveGesture Streamable Co-Speech Gesture Generation Model Zero-Shot Text-to-Image Generation

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:26:09.277502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-10T16:46:58.944000Z digest=sha256:16a0e73b1ef6fc8f8ccc7093287ca953892f26c8fc9ec1f71a784f7bef0c839b

Observation dfaf15d9-4d7d-471a-ade1-5346e3df31cf · inbound

BEAT: Tokenizing and Generating Symbolic Music by Uniform Temporal Steps cites this paper.

BEAT: Tokenizing and Generating Symbolic Music by Uniform Temporal Steps Zero-Shot Text-to-Image Generation

Reference 105

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T22:26:09.277502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-05-10T01:18:04.842423Z digest=sha256:4b063aca10ed34578b4a0008907bc66070e28303d65b1d2f83c096114d009f8c

Observation 1f03e6ba-9d84-4376-b053-ae6af8f7a3da · inbound

Who Defines Fairness? Target-Based Prompting for Demographic Representation in Generative Models cites this paper.

Who Defines Fairness? Target-Based Prompting for Demographic Representation in Generative Models Zero-Shot Text-to-Image Generation

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:26:09.277502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-09T23:48:37.948835Z digest=sha256:118a96c08eea51b6c6c20962122b997e20cdbf821fc2bad0fbb07819a6a681ee

Observation 665251b0-9cd5-4a3d-b647-43904baff60d · inbound

Ensemble Distributionally Robust Bayesian Optimisation with Continuous Context cites this paper.

Ensemble Distributionally Robust Bayesian Optimisation with Continuous Context Zero-Shot Text-to-Image Generation

Reference 96

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:26:09.277502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-05-11T03:18:03.316334Z digest=sha256:7fc9064bc14aa5c2c40ddb223ff3be73252db6e768bc35372708b37e7fe467ad

Observation 2e941981-ca77-4278-8fa3-6ee225249b8a · inbound

Rennala MVR: Improved Time Complexity for Parallel Stochastic Optimization via Momentum-Based Variance Reduction cites this paper.

Rennala MVR: Improved Time Complexity for Parallel Stochastic Optimization via Momentum-Based Variance Reduction Zero-Shot Text-to-Image Generation

Reference 139

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T22:26:09.277502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-05-12T01:51:20.003552Z digest=sha256:8328d4a5a86e61b9a875aa03d1f7a3432477571a5e6eb95a1a645052685bc6d5

Observation 4e4c30e0-4369-4246-9cf2-60372997b09c · inbound

Ringmaster LMO: Asynchronous Linear Minimization Oracle Momentum Method cites this paper.

Ringmaster LMO: Asynchronous Linear Minimization Oracle Momentum Method Zero-Shot Text-to-Image Generation

Reference 141

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T13:13:18.606479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-05-20T13:08:52.912250Z digest=sha256:912a7e77c18bf033362d84ac60fff911343f73b833016bee72da0c275772b9ec

Observation a9ff73b0-5904-411b-b95f-fb25d770662b · inbound

LOSCAR-SGD: Local SGD with Communication-Computation Overlap and Delay-Corrected Sparse Model Averaging cites this paper.

LOSCAR-SGD: Local SGD with Communication-Computation Overlap and Delay-Corrected Sparse Model Averaging Zero-Shot Text-to-Image Generation

Reference 143

Resolution
metadata mismatch
local_arxiv, observed 2026-05-21T05:49:40.701655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-05-21T05:49:28.713982Z digest=sha256:5e0d8ff68ecd0f9d2b611ea4dd2c79c5b7c934a15c64353b764004c55df19848

Observation 4a67e0d1-f702-4661-b4a4-534b003071a8 · inbound

Mapping Whisper Representations to Human ECoG Responses with Interpretable Time-Resolved Neural Encoding cites this paper.

Mapping Whisper Representations to Human ECoG Responses with Interpretable Time-Resolved Neural Encoding Zero-Shot Text-to-Image Generation

Reference 126

Resolution
verified exact
local_arxiv, observed 2026-06-28T11:42:04.067501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-06-28T11:39:30.563754Z digest=sha256:813b6ad0591ef4d2b66a94905cb8a140cfc198069a4c971a7deed5c2d83401dc

Observation 9815f686-64c6-466f-98ee-b7ab1cd1aef1 · inbound

ZIPP:Zero-shot Image Personalization from Personas cites this paper.

ZIPP:Zero-shot Image Personalization from Personas Zero-Shot Text-to-Image Generation

Reference 38

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T23:27:27.959436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-06-27T18:11:22.987938Z digest=sha256:fa4102e669446eaa96a859b2a329e089f62329a2a85da734f9354b272692fd19

Observation 97f61ca5-656d-4284-ac76-acc56feb58a1 · inbound

Token-to-Token Alignment of Text Embeddings for Semantic Blending cites this paper.

Token-to-Token Alignment of Text Embeddings for Semantic Blending Zero-Shot Text-to-Image Generation

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-07-04T10:39:46.128417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-26T08:34:23.842302Z digest=sha256:4df0a2bebeee93d018f204a4a5bed37ef641afd1af158d02d1c0aa6bbb6f150c

Observation 27a6e446-bfce-4024-9767-aaef24929011 · inbound

Diff-ID: Identity Consistent Facial Image Generation and Morphing via Diffusion Models cites this paper.

Diff-ID: Identity Consistent Facial Image Generation and Morphing via Diffusion Models Zero-Shot Text-to-Image Generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-31T01:59:33.072872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T01:59:33.072872Z digest=sha256:c4aa74a4db24dfb35816926b1a70d1b5c8e6ac2824f697f282d7d51f87c093dd

Observation 2aa3f75a-318e-4705-8695-8ff9a6a55f00 · inbound

Where Does Generative Difficulty Reside? An Empirical Study of Target Representations cites this paper.

Where Does Generative Difficulty Reside? An Empirical Study of Target Representations Zero-Shot Text-to-Image Generation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T03:02:38.885153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T03:02:38.885153Z digest=sha256:381722ff6ac68f69d999fa69582442d10924c199dbc51adbfda2dfbe072e30dc

Observation 38d58b6f-a98b-4312-8b90-82f00a215da5 · inbound

Evading Chain-of-Thought Monitoring Through Model Poisoning cites this paper.

Evading Chain-of-Thought Monitoring Through Model Poisoning Zero-Shot Text-to-Image Generation

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-15T15:04:48.399633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:04:48.399633Z digest=sha256:35be3c6e50ecf8b4f37c5ae14dd8fe9e2df709659538f991983bc4a6afae75e7

Observation 63e6af98-2245-4982-b6a8-789ff45f2788 · inbound

On MUON optimization: From non-convergence to an error analysis with Polar Express and the Newton-Schulz polynomial from implementations cites this paper.

On MUON optimization: From non-convergence to an error analysis with Polar Express and the Newton-Schulz polynomial from implementations Zero-Shot Text-to-Image Generation

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T21:05:45.963044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:05:45.963044Z digest=sha256:5bc1fcc64f63160686b2e77dbaf0d358cb6b72bd568a93096f0ae6da608a1e3e

Observation fdb59593-1e7d-4656-a9c5-00a9e0f28889 · inbound

NAE: Normalizing AutoEncoder cites this paper.

NAE: Normalizing AutoEncoder Zero-Shot Text-to-Image Generation

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-16T00:22:07.528089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:22:07.528089Z digest=sha256:3a990787b1c83c7fd7223b38c2dadcc631e64fa366c3f5336de4c820c131aaee