Pith. sign in

Paper Citation Record · LEDGER

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation

As of 6 August 2026, this Paper Citation Record lists 85 of 85 outbound references and 1 inbound Pith citation observation for arXiv:2604.18168.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2604.18168 v1

Coverage vector

measured 85 of 85 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-10T05:15:22.907880Z

measured 86 of 86 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-20T18:33:52.933672Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-05-20T18:38:52.917903Z

Reference resolution

85 of 85 outbound references displayed

  • verified exact47
  • verified fuzzy32
  • unresolved1
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch5

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3f069ea7-8e75-419e-a549-12188780915a · outbound

This paper cites Denoising diffu- sion probabilistic models.Advances in neural information processing systems, 33:6840–6851.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Denoising diffu- sion probabilistic models.Advances in neural information processing systems, 33:6840–6851

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T21:44:23.730117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:edbc6184f5fb2298352a08e41da0163c8e7957e89e56109753845de657011e39

Observation 4dbd93fb-70cc-4a87-a478-f15fd048e270 · outbound

This paper cites Denoising Diffusion Implicit Models.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Denoising Diffusion Implicit Models

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-10T09:28:39.625775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:c5080829e3ede134892f14ae9c1caa1ae27a70ca19b748604eced662537982b2

Observation 84666e2e-9a9a-416c-a63c-deca3cafdb49 · outbound

This paper cites Flow matching for generative modeling.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Flow matching for generative modeling

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T21:44:23.732245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:cdb74c40c2e31013256989e459120eb5943b488b6dcae17a1fd60eb07186c51f

Observation 1cddfd30-dd49-4775-a2a2-c12d6db12cd0 · outbound

This paper cites Scaling rectified flowtransformersforhigh-resolutionimagesynthesis.InForty- first international conference on machine learning.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Scaling rectified flowtransformersforhigh-resolutionimagesynthesis.InForty- first international conference on machine learning

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T21:44:23.725387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:23ca8d93aa1638607059e43d7dba766aeb1623628d5d0eb9ec692f68b0fc946c

Observation c4274da3-2d76-428b-81e7-0c8d734bd983 · outbound

This paper cites Consistency models.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Consistency models

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T21:44:23.755134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:ec9371831fea1644e1b9a13a0abd3a82d1c0c7a39533930f5075d3cf5f0d4874

Observation 24eaa755-7c05-4d82-ba2f-06434762e7f7 · outbound

This paper cites Multistep Consistency Models.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Multistep Consistency Models

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:28:39.551210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:3bb38d3f7708226cc45bfbe2cd7b9c56b5e8bd8c4e53885567eaa20049174421

Observation 9f57c5a8-4651-4b9d-951c-ce1580a72fed · outbound

This paper cites Align Your Flow: Scaling Continuous-Time Flow Map Distillation.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Align Your Flow: Scaling Continuous-Time Flow Map Distillation

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:33:41.601618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:c0093e6eb31afa0894686ff48035b846dd3451b8415aeac846fb2e259a259edd

Observation 7e20a4bd-f721-42c0-a456-4ee2efadd77e · outbound

This paper cites Flow map matching with stochastic interpolants: A mathematical framework for consistency models.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Flow map matching with stochastic interpolants: A mathematical framework for consistency models

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T09:33:41.639906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:b7922b1ebe2a95f0e18e86c8041847526bc95097fe31a11f9dd4023d2c91955a

Observation ad82b1cf-94e0-4ffb-ad8b-3a6c5ee9c73a · outbound

This paper cites Mean Flows for One-step Generative Modeling.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Mean Flows for One-step Generative Modeling

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:29:30.500037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:91499f07838689bbd464474e4027bfc09e4831e47b12275a13c9a380784af4c3

Observation 6f52eed5-9cb1-4842-b34f-587929729f21 · outbound

This paper cites Zhang, A.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Zhang, A

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T09:33:41.636415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:bf005078a1f262dc184b1ccbe75ce7e2a95f9c88dc32bf30c9af2a29d57881d5

Observation 9e595388-e436-4c85-9af0-a437fa6a15d4 · outbound

This paper cites SplitMeanFlow: Interval Splitting Consistency in Few-Step Generative Modeling.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation SplitMeanFlow: Interval Splitting Consistency in Few-Step Generative Modeling

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:28:39.588855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:f84545795da520d511e4d1a3208964b145462b745df884d4dd91cc3f55138230

Observation 17af12ea-0515-4483-8499-01b4a9404d26 · outbound

This paper cites Decoupled meanflow: Turning flow models into flow maps for acceler- ated sampling.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Decoupled meanflow: Turning flow models into flow maps for acceler- ated sampling

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:28:39.591490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:1efb5becae7283a7924df1d912aa6a1ccf067368a48740b7c4be9b6a455dc373

Observation 1e62e046-6323-4fb5-b462-afe71a51de37 · outbound

This paper cites Imagenet: A large-scale hierarchical image database.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Imagenet: A large-scale hierarchical image database

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T21:44:23.720593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:5731ccadba61f44cf23482e6201fe2983ae4776be57e1edb5175348f1817cddc

Observation e6b7663d-ebe9-40fd-8242-62a962224a52 · outbound

This paper cites SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformers.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformers

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:56:50.121098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:e1213faa7750ce915a8810a280a51dfa702a2f7e6309afa349b165a37913354d

Observation a93b0b49-62d6-4197-9e08-242ae8bc0634 · outbound

This paper cites Learning trans- ferable visual models from natural language supervision.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Learning trans- ferable visual models from natural language supervision

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T21:44:23.720304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:a04c43ab98ccfdbdc176741e47e6ef0f4a01ed2448f28a6ddf34e72f9cdeb799

Observation 925e388f-5715-4952-9da6-db60f64a966b · outbound

This paper cites Exploring the limits of transfer learning with a unified text-to-text transformer.Journal of machine learning research, 21(140):1–67.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Exploring the limits of transfer learning with a unified text-to-text transformer.Journal of machine learning research, 21(140):1–67

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T21:44:23.704110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:4421a10c96f6f8a4408e79e8fe9d9383f7b4c9be24e5a065f67d9ca6f986e0bc

Observation 06fd4b26-6939-450e-9e7d-10cc1e1e34c1 · outbound

This paper cites Qwen2 Technical Report.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Qwen2 Technical Report

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:09:09.439499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:f14a0fb62facaf0b96efdab3215f658c4d79d6683c615dfd3dc19b440edff9cd

Observation 9f3d9ad2-e725-4af6-8c6f-89c817520d60 · outbound

This paper cites Gemma 2: Improving Open Language Models at a Practical Size.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Gemma 2: Improving Open Language Models at a Practical Size

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-10T12:11:16.746994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:0736dfcc99d8508fc7ebf245f2cd27177daa8a89ec79687f437196338cc8e0d6

Observation b7087298-efb4-45ea-a3e3-146de5a712e2 · outbound

This paper cites 8 Sana-sprint: One-step diffusion with continuous-time consis- tency distillation.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation 8 Sana-sprint: One-step diffusion with continuous-time consis- tency distillation

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:28:39.567279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:4a994d044ea416e704750a9ebd0abaedc0118873c423d29aaadeff84277dbeee

Observation 731fe970-5be0-4ab0-94e3-72c68ccf09b2 · outbound

This paper cites Simplifying, Stabilizing and Scaling Continuous-Time Consistency Models.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Simplifying, Stabilizing and Scaling Continuous-Time Consistency Models

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-13T10:26:23.772771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:f07847e59574caef222bbc30ce16282bd9da44ba491848921061823a34246b13

Observation 9a1f6ecc-891f-4aa5-a9e4-3389e29e495a · outbound

This paper cites Large Scale Diffusion Distillation via Score-Regularized Continuous-Time Consistency.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Large Scale Diffusion Distillation via Score-Regularized Continuous-Time Consistency

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-05-10T09:28:39.639528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:472b06ab2c5baa2e13e738282d5bf3d2112b24a8a02c095f16ba5bad629a10dc

Observation 4d3247bb-44c9-4e2b-a96b-409d69860435 · outbound

This paper cites Consistency Models Made Easy.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Consistency Models Made Easy

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T09:33:41.662547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:116a201f229e9bc46d4fd4797a92de2150b5e94e4233f69d2326d0559820dc25

Observation 5b328d3b-7d01-40e1-9061-6387feb5f2f2 · outbound

This paper cites Improved Techniques for Training Consistency Models.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Improved Techniques for Training Consistency Models

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:04:10.684842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:0c7f8e2f490808bb718675dc29ec7ea0830f6116e68591c9995d2108f985b667

Observation 005ca170-0da8-418e-a4cf-2ed10216f322 · outbound

This paper cites Advancing end- to-end pixel space generative modeling via self-supervised pre-training.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Advancing end- to-end pixel space generative modeling via self-supervised pre-training

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:28:39.612762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:142419c0679370af9254ddd80aab75b67abd1cb9ddbde0055bf01a89e086b042

Observation 86826d24-af32-40d4-88f5-d78d49c3f718 · outbound

This paper cites Diffusion Transformers with Representation Autoencoders.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Diffusion Transformers with Representation Autoencoders

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:34:18.071377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:179d2d2b1e292e8482d844045bc202d6e20169206c843a7f4c6442f7d31ef42d

Observation 57afe4cc-3a5d-4048-883f-4800c8f2cdf1 · outbound

This paper cites Representation Alignment for Generation: Training Diffusion Transformers Is Easier Than You Think.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Representation Alignment for Generation: Training Diffusion Transformers Is Easier Than You Think

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-12T15:09:37.504223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:a9c8416f9cb38adcb415631f5614348694168c5c452e590c74ee7336850c1269

Observation 2ecd1cf9-7c91-493c-901f-faa7fd1f330f · outbound

This paper cites Latent diffusion model without variational autoencoder.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Latent diffusion model without variational autoencoder

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:28:39.548609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:83398a6a24af4fe100674baac354a4285a1c1819a11120d71f1635f7c252adba

Observation 9379771f-a636-438e-9704-b97d8decaeb4 · outbound

This paper cites Representation entanglement for generation: Training diffusion transformers is much easier than you think.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Representation entanglement for generation: Training diffusion transformers is much easier than you think

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T21:44:23.715591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:f57eb7f221c00d3ad851025e441b48370b3a2163fef82212a3c92e8c84efdd2f

Observation c9434ea0-c1c9-489f-a857-17cbdcd8881d · outbound

This paper cites InICCV.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation InICCV

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T21:44:23.699410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:9da348561c09dbb22b7aeb71a2a682d930adfb93d64b9c2082671b05833ba7c3

Observation 18624687-ba25-4118-8323-2cfbc76445ae · outbound

This paper cites Reconstruction Alignment Improves Unified Multimodal Models.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Reconstruction Alignment Improves Unified Multimodal Models

Reference 30

Resolution
metadata mismatch
arxiv_id, observed 2026-06-26T02:15:36.738364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:34e318d1c3cea962e8cddc6b6ea8b604832f2e213425e3bce92d1712d8919e7b

Observation 279719c3-82a0-4c51-9335-39830488559e · outbound

This paper cites Sana 1.5: Efficient scaling of training-time and inference-time compute in linear diffusion transformer.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Sana 1.5: Efficient scaling of training-time and inference-time compute in linear diffusion transformer

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T21:44:23.709567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:40d84c363c9e5d7467d5e22667bee833ed5707cec107f2d6591ef6fa29082249

Observation 47aa68e4-9830-48f4-8b0f-9a521601fa38 · outbound

This paper cites U-net: Convolutional networks for biomedical image segmentation.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation U-net: Convolutional networks for biomedical image segmentation

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T21:44:23.712492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:0dc26b390b5c7405c761602836becd8c298d0709d066b6bee81e90daf18a2612

Observation 0a5a0c6d-5950-4ef3-8a27-04a4b904802d · outbound

This paper cites Scalable diffusion models with transformers.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Scalable diffusion models with transformers

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T21:44:23.718203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:17c7d79854acdfce2bfb735b47cee344be3f6adf9f82999f8d8d9cbb33975de7

Observation 7b5bd5ae-0276-43cb-ab47-0d9e54fd6a31 · outbound

This paper cites S2-guidance: Stochastic self guidance for training-free enhancement of diffusion models.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation S2-guidance: Stochastic self guidance for training-free enhancement of diffusion models

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:33:41.659344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:fa6896f3d5605301576cb6e7c4291bd0607e75a4a69f5695134fa1d8f5700f96

Observation 907c7825-2665-4e8a-8525-c595ff44a7b9 · outbound

This paper cites Taming preference mode collapse via directional decoupling alignment in diffusion reinforcement learning.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Taming preference mode collapse via directional decoupling alignment in diffusion reinforcement learning

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:28:39.634001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:8347237ff83d5a082ac581aaa25eadab66aa74d0f16871bfaae5cb69eb4f2ede

Observation b9562984-82e3-4676-9a48-12165c14b8b5 · outbound

This paper cites Qwen Technical Report.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Qwen Technical Report

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-05-10T09:33:41.642802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:ad52a69dc05fa2fe29d9a05e93b3e1a8b0938977f291ba88da590f2285355d41

Observation 9ec2f277-279c-4a71-8972-e8a26a78daae · outbound

This paper cites Qwen3 Technical Report.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Qwen3 Technical Report

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-05-10T09:28:39.620625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:fc808421c8d641d2ecfe00c3ab7c6da18441a7c7b2446dd3d76a9c4767a24f7c

Observation d7f0a1c9-04c7-4d72-a98e-7ad5651a712f · outbound

This paper cites Gemma: Open Models Based on Gemini Research and Technology.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Gemma: Open Models Based on Gemini Research and Technology

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-10T15:54:09.189864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:67cbf9c620dea5205a526992e74826d4f2988c88d3b602f188ca4260a30102d4

Observation a9db719e-fcfe-49bd-baa4-863eb986598f · outbound

This paper cites Gemma 3 Technical Report.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Gemma 3 Technical Report

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-05-10T09:28:39.617811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:b5ad26ca86d88a911afbfff3f0af31a3f422a83b51f1367c358276bc680ae051

Observation d6fe364b-5a3c-4b71-9499-4cff217e2ffc · outbound

This paper cites High-resolution image synthesis withlatentdiffusionmodels.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation High-resolution image synthesis withlatentdiffusionmodels

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T21:44:23.727838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:abde1c4222577a0032fab0dfecea21cf1f83b429bfacc5e685e55a6ed1d2114a

Observation b0e129c9-4d86-4b9a-8437-5b2c6eca2117 · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-10T15:22:03.751402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:f92b14bca4a55c1efcf8ea0ad07f0ecf4eaf53d1bb780ced2fec08dc038788c8

Observation 7700f5a4-e79f-493e-89ba-dee488f2652b · outbound

This paper cites Pixart-𝛼: Fast training of diffusion transformer for photorealistic text-to-image synthesis.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Pixart-𝛼: Fast training of diffusion transformer for photorealistic text-to-image synthesis

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T21:44:23.737243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:618068d0bf81e17f3acc20b21aeb5c33faff7751b68037765d7530f46716a997

Observation 70510194-090f-412e-a1e4-b72b7d4fe14e · outbound

This paper cites PIXART-{\delta}: Fast and Controllable Image Generation with Latent Consistency Models.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation PIXART-{\delta}: Fast and Controllable Image Generation with Latent Consistency Models

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:28:39.623281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:248248dd688e0b3b6d610d9fdd1947966966dee13847e6b1a0ed7caeefb08ea8

Observation 4be4a9e5-2ab3-4348-b20b-f962c88e7451 · outbound

This paper cites Pixart-𝜎: Weak-to-strong training of diffusion transformer for 4k text-to-image generation.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Pixart-𝜎: Weak-to-strong training of diffusion transformer for 4k text-to-image generation

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T21:44:23.696510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:105dbf039f04201b2ba1434252e5682d201d0b821ec17bbc3794fbf468f97230

Observation 8f4cafac-6826-4145-b41e-161e1732a5a6 · outbound

This paper cites an unresolved cited work.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-05-21T21:44:23.706886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:5b667603ae922b827462c59c18224a98841cfdfe308ea974acb94aacaba93725

Observation 6e216859-95ac-48c4-a414-5aefe50849b0 · outbound

This paper cites Flux-text: A simple and advanced diffusion transformer baseline for scene text editing.arXiv preprint arXiv:2505.03329.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Flux-text: A simple and advanced diffusion transformer baseline for scene text editing.arXiv preprint arXiv:2505.03329

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:33:41.605722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:95d3f31ae4cdad670b2949d66c37f1ad16499d69af442f976af476c0c9e35e8b

Observation 63ba8b72-ade2-480c-b0a9-bfc0a78dd088 · outbound

This paper cites Nano banana.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Nano banana

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T21:44:23.680937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:1b665de331eeb65177debd4dde9cbf0bee12c410e5cf5daf1eb71e2bdd0d345b

Observation 126a2e29-ad8a-48f0-b1d4-c23345e402aa · outbound

This paper cites Qwen-Image Technical Report.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Qwen-Image Technical Report

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:29:07.137498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:164829ed68b58517befe7c64436791be0b1c3ab24312bb116f6ce7a5a0ac0fcc

Observation ce4cba4c-d39b-4163-9ddb-75a2738fb6e5 · outbound

This paper cites HunyuanImage 3.0 Technical Report.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation HunyuanImage 3.0 Technical Report

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-16T02:02:32.945705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:552009e7ce80e7b03cea6b966730e37000e4ea6eaf3c288f7bd3df7a99de5d44

Observation 6cdca740-2263-4059-b665-c36a3b372e93 · outbound

This paper cites Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:28:39.561935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:6efdd1645f570f92992899abfbb9b237d9685e0a540f244a4969560f5c10167b

Observation e982e22b-dde6-42c6-a6b4-63ffc68ace9d · outbound

This paper cites Blip3o-next: Next frontier of native image generation.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Blip3o-next: Next frontier of native image generation

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:28:39.607020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:efd7634cbd3ece14125bab29bdfe434c69cdd3660ea1e60c7ac28d6a7eb03687

Observation 928552b2-61f6-4326-8b44-9b687fad2b42 · outbound

This paper cites Latent Consistency Models: Synthesizing High-Resolution Images with Few-Step Inference.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Latent Consistency Models: Synthesizing High-Resolution Images with Few-Step Inference

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-13T04:15:55.027720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:6ae0200f20fe17f85bf6f575bae59956d9471fc15f9a112ad9948dd14cd80164

Observation bb3cf0db-bcb1-4d7f-81f0-e83f2a3e504c · outbound

This paper cites Consistency Flow Matching: Defining Straight Flows with Velocity Consistency.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Consistency Flow Matching: Defining Straight Flows with Velocity Consistency

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:28:39.583244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:329e2f109ff1027cacd9eb28cc1ad863428eab397714a94e7a117179aa45d9a1

Observation 5780362e-d0f8-4a4c-9ee6-4fd357f51749 · outbound

This paper cites Transition Models: Rethinking the Generative Learning Objective.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Transition Models: Rethinking the Generative Learning Objective

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:28:39.594127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:30ecee16201542aec5e67050c0912980645b92a7265f4433aa0ba089917a0b4e

Observation 13a56e94-eb59-4ce1-a164-21801ea2336d · outbound

This paper cites Microsoft coco: Common objects in context.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Microsoft coco: Common objects in context

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T21:44:23.702133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:dc5455887dd198a9d98659ad79470d9fd40838aefba78159be3f1ba3ea89db51

Observation e6ec3b3e-d11f-4e28-b7c5-891f75cd5523 · outbound

This paper cites Saco loss: Sample-wise affinity consistency for vision-language pre-training.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Saco loss: Sample-wise affinity consistency for vision-language pre-training

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T21:44:23.734872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:d48e14740af97dc775116c748aba1063f9c41280e9fed5650cf7845cd53efcfe

Observation a0232557-c79b-42fc-b4c8-b829612a6ede · outbound

This paper cites DINOv3.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation DINOv3

Reference 57

Resolution
metadata mismatch
local_arxiv, observed 2026-05-10T09:28:39.628421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:a6cf1662bd3e3309f643bded250932f186345d161751dccc4cfe994aa12ca7ee

Observation 00aac801-13ad-4aec-ad65-958dad157038 · outbound

This paper cites ELLA: Equip Diffusion Models with LLM for Enhanced Semantic Alignment.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation ELLA: Equip Diffusion Models with LLM for Enhanced Semantic Alignment

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:43:03.755490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:d438b38a96f360dc62b39b6bdaae9f4af1ba536a7470fd2f20a2dbca5bd5f572

Observation b390162c-fe32-420d-9480-e024d708fc4c · outbound

This paper cites BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:28:39.609965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:594ee0158b5f3b09fd7477542a7b3a688fff06f684ca59bbde24a64d03047da6

Observation 09230663-92aa-4ccf-a58c-26893973f84b · outbound

This paper cites ShareGPT-4o-Image: Aligning Multimodal Models with GPT-4o-Level Image Generation.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation ShareGPT-4o-Image: Aligning Multimodal Models with GPT-4o-Level Image Generation

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:28:39.596721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:67238273b32788883964f3aad932e09d06a7b0e117d2b10db3d2627981737c66

Observation 3d432ef3-d8fb-47df-b223-a5d3b74ead1e · outbound

This paper cites Echo-4o: Harnessing the Power of GPT-4o Synthetic Images for Improved Image Generation.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Echo-4o: Harnessing the Power of GPT-4o Synthetic Images for Improved Image Generation

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:28:39.636695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:b82c96c56fb71d157146ea39fba10a8c3662f9fc930559658e029d64be0e75b8

Observation e65cb7a1-7749-40a6-9ab3-322a91769353 · outbound

This paper cites Geneval: An object-focused framework for evaluating text-to- imagealignment.AdvancesinNeuralInformationProcessing Systems, 36:52132–52152.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Geneval: An object-focused framework for evaluating text-to- imagealignment.AdvancesinNeuralInformationProcessing Systems, 36:52132–52152

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T21:44:23.686655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:ff234f33933e6f66329a8715978bd6efc9c6805604b257281c414545cdfcc433

Observation 434df2f4-984e-46fd-8521-074b35dcca44 · outbound

This paper cites Human Preference Score v2: A Solid Benchmark for Evaluating Human Preferences of Text-to-Image Synthesis.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Human Preference Score v2: A Solid Benchmark for Evaluating Human Preferences of Text-to-Image Synthesis

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:30:48.951760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:42279bbcf283a93c19c2a50cd06a9d5a653c15b90d586c656ad3c65e6cd0a6a1

Observation 1eb8ff30-d822-4b2b-9755-93090830d1da · outbound

This paper cites Cosmos world foundation model platform for physical ai.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Cosmos world foundation model platform for physical ai

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T21:44:23.677327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:6143bc59e9886f81cc5a4e5c4d0de01cc5e2fa8b14015edc85498d14bde84259

Observation 2a99ff9d-d136-4719-b385-a0e61630601d · outbound

This paper cites Lumina-Image 2.0: A Unified and Efficient Image Generative Framework.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Lumina-Image 2.0: A Unified and Efficient Image Generative Framework

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:28:39.604186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:8ce84f472f1d0d4a036aae35e0840f85dd81c6c342e3df136fd0aa5a32f1f3d7

Observation 8d27b40e-ee67-4959-bf6b-d2e7504946e1 · outbound

This paper cites HiDream-I1: A High-Efficient Image Generative Foundation Model with Sparse Diffusion Transformer.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation HiDream-I1: A High-Efficient Image Generative Foundation Model with Sparse Diffusion Transformer

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-05-16T17:13:39.137651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:b9bcc617294ef65859f6fa25b93e970a481db5725471eb74c95f267e93a134ad

Observation c0ffe458-cd3f-4b7b-8de6-d3db4160dcbe · outbound

This paper cites Seedream 3.0 Technical Report.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Seedream 3.0 Technical Report

Reference 67

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:55:38.865721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:708b7c8f1a3d92c1261d79cb759c9632ec9cab34e6ec9466989687a4f4b06d23

Observation 98e17ba2-832a-433d-90f4-01c9902bf7cf · outbound

This paper cites Gpt-image-1.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Gpt-image-1

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T21:44:23.674569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:278740e1beaf1ca7bd9e2f12ebf1ef4ed3e3a1f80aee7876399598d156c78730

Observation 6d0ff3b8-ab01-43ce-a8d0-b03f0767452c · outbound

This paper cites Transfer between Modalities with MetaQueries.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Transfer between Modalities with MetaQueries

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-05-14T22:49:23.311693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:08e1e8b9a279051a8bd8dc4e4d331461a790378400607730355686bac7fe9fc7

Observation c96aa05d-13c8-4ef0-88af-bcbcf5319edf · outbound

This paper cites OpenUni: A Simple Baseline for Unified Multimodal Understanding and Generation.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation OpenUni: A Simple Baseline for Unified Multimodal Understanding and Generation

Reference 70

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:28:39.580506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:24f57f68ab5e0dfd0511324c428bcc16a33a9f0de2820303ad4ef48b29cfc8dd

Observation 3f09bd54-8d22-4e48-9d5f-b6c7e89f3d47 · outbound

This paper cites Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representations.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representations

Reference 71

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:28:39.564612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:eea8eb89d4cc115a3302ad0a069d78c0e063ee38809c4576cfe67ff72ad0e5a3

Observation b0403c62-fbe1-4434-ad2a-a777f90716fa · outbound

This paper cites TBAC-UniImage: Unified Understanding and Generation by Ladder-Side Diffusion Tuning.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation TBAC-UniImage: Unified Understanding and Generation by Ladder-Side Diffusion Tuning

Reference 72

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:28:39.569797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:58126ffa3d4501297977008847b07f3ae2abc4b37322e290f9f592d24937607f

Observation ef46f3d8-b6e0-4506-8923-8125b0f47ab7 · outbound

This paper cites SDXL-Lightning: Progressive Adversarial Diffusion Distillation.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation SDXL-Lightning: Progressive Adversarial Diffusion Distillation

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-05-17T05:08:49.084339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:7acb7d3137a1d7f60a599e837a5d4004f4e1196bca57a60ee3031e14a1e28966

Observation f4495ff5-71c5-4c97-867b-8d6acdbdbb2e · outbound

This paper cites Hyper-sd: Trajectory segmented consistency model for efficient image synthesis.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Hyper-sd: Trajectory segmented consistency model for efficient image synthesis

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T21:44:23.722845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:1081a67f1c1cb2513693739bc4e8096b11ed3554924a15076465c28aedcfea01

Observation f025159d-474c-47ea-94fd-77bce48e82ac · outbound

This paper cites Improved distribution matching distillation for fast image synthesis.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Improved distribution matching distillation for fast image synthesis

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T21:44:23.684809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:6800434b717fcfa87f5161baedb32525ac6c7b95e480060edbe077ee9f7ac2f0

Observation c076c545-6e37-47bf-adaa-3921a74077b5 · outbound

This paper cites Lumina-next: Making lumina-t2x stronger and faster with next-dit.Advances in Neural Information Processing Systems, 37:131278–131315.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Lumina-next: Making lumina-t2x stronger and faster with next-dit.Advances in Neural Information Processing Systems, 37:131278–131315

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T21:44:23.749725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:fc4ff7b329da701e64a817829411559830a3c006a5f15a469f2cec1cd7f006e9

Observation 8b50e645-0356-47d0-8008-17bc23d6465a · outbound

This paper cites Playground v2.5: Three Insights towards Enhancing Aesthetic Quality in Text-to-Image Generation.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Playground v2.5: Three Insights towards Enhancing Aesthetic Quality in Text-to-Image Generation

Reference 77

Resolution
verified exact
arxiv_id, observed 2026-05-15T16:40:06.446550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:87c2d20fe04ede985d1b80100d76fb61d921d3429d17588ce2b38ff0e43ed53a

Observation ce416278-a20e-4dcf-9845-66106728773b · outbound

This paper cites Hunyuan- dit: A powerful multi-resolution diffusion transformer with fine-grained chinese understanding.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Hunyuan- dit: A powerful multi-resolution diffusion transformer with fine-grained chinese understanding

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T21:44:23.752388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:56cc08b3441319ba4656d7526e30b1022e17c00424b9d87cde43df75cecfc3b0

Observation cbec2dbf-7924-4b89-8e6f-4e0e08778e87 · outbound

This paper cites DALL·E 3.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation DALL·E 3

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T21:44:23.701263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:d5ad81e4237b108da4cedec56a5753b8dd55d8132c3dd2082a4b133df8112afb

Observation 341d00f8-d39e-41db-bde1-86461a66d5fa · outbound

This paper cites Fast high- resolution image synthesis with latent adversarial diffusion distillation.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Fast high- resolution image synthesis with latent adversarial diffusion distillation

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T21:44:23.747289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:234550f011c2dce3dfb5aa8807964cbae5846362a43a68ecfdb2612e7eeb943b

Observation 9620d42e-0206-4e69-8043-fa0282e5285b · outbound

This paper cites Text Conditions 𝑡−𝑟𝑢(𝑧,𝑟,𝑡)𝑢(𝑧,𝑟,𝑡)𝑣 𝑡−𝑟𝑢(𝑧,𝑟,𝑡)𝑣𝑢(𝑧,𝑟,𝑡) Figure 7.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Text Conditions 𝑡−𝑟𝑢(𝑧,𝑟,𝑡)𝑢(𝑧,𝑟,𝑡)𝑣 𝑡−𝑟𝑢(𝑧,𝑟,𝑡)𝑣𝑢(𝑧,𝑟,𝑡) Figure 7

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T21:44:23.722527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:8296848e0ac2d71138556e79aba1ce66259621e5ac22a5adebe5a808191f2fd4

Observation 0d375dbb-8fc1-4b4a-bc37-ddcace2b52e7 · outbound

This paper cites We chose OpenUni because it shares the SANA- 1.5 diffusion backbone, but uses a InternVL3–based text encoder.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation We chose OpenUni because it shares the SANA- 1.5 diffusion backbone, but uses a InternVL3–based text encoder

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T21:44:23.712717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:55ea19ceedf5dc73eef8c3903471801a4650bd075d556886ff19d9f34de9f782

Observation 294b7956-d89b-4bac-a097-25d02377dd9c · outbound

This paper cites When generating images from the same prompt and timing diffusion sampling only, BLIP3o-NEXT on H200 takes 1.24 s with 30 steps, while ours takes 0.22/0.12/0.08 s (4/2/1 steps).

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation When generating images from the same prompt and timing diffusion sampling only, BLIP3o-NEXT on H200 takes 1.24 s with 30 steps, while ours takes 0.22/0.12/0.08 s (4/2/1 steps)

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T21:44:23.739780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:e76e45578bf41e88312e73011e4863234f1a4e73337c18e6cf92d4a151790c84

Observation 42b38891-4185-4ebb-ae84-1bf28f468edc · outbound

This paper cites Which result best matches the prompt?.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Which result best matches the prompt?

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T21:44:23.742049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:5477d83cabd85093d70714ce7e0aed5dc9ce9be095a936c4b22e3acfd82535a5

Observation f49fa23d-6c3e-4654-bef9-6cd23e3c56a7 · outbound

This paper cites DPG-Bench evaluation.Generatinghigh-fidelityimages from complex and detail-rich textual prompts in a limited number of denoising iterations is a highly challenging task.

Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation DPG-Bench evaluation.Generatinghigh-fidelityimages from complex and detail-rich textual prompts in a limited number of denoising iterations is a highly challenging task

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T21:44:23.744879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T05:15:22.907880Z digest=sha256:071fcd03be98a212261650c6907ece023273ab822f24d0b6a46c99fd25f05da8

Pith citing papers

Observation 48f3f1e4-b893-4df5-ad80-f151fc77bf54 · inbound

Embedding-perturbed Exploration Preference Optimization for Flow Models cites this paper.

Embedding-perturbed Exploration Preference Optimization for Flow Models Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation

Reference 96

Resolution
verified exact
local_arxiv, observed 2026-05-20T18:38:52.919696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T18:33:52.933672Z digest=sha256:e1aa08b36674d81c4d7413d9a8d9444a76afe0482cf4435bc28f588000b53620