Pith. sign in

Paper Citation Record · LEDGER

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution

As of 19 August 2026, this Paper Citation Record lists 73 of 73 outbound references and 1 inbound Pith citation observation for arXiv:2605.21195.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.21195 v1

Coverage vector

measured 73 of 73 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-21T05:47:21.834145Z

measured 74 of 74 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T00:44:16.917352Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T00:44:18.368421Z

Reference resolution

73 of 73 outbound references displayed

  • verified exact18
  • verified fuzzy54
  • unresolved1
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation db2e2fd7-62d6-46e6-b286-3f0dac047765 · outbound

This paper cites Scheduled sampling for sequence prediction with recurrent neural networks.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Scheduled sampling for sequence prediction with recurrent neural networks

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.377600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:a9a74e4281b6bb3f8e2ed2616dcf3d335865f7aa68f73ba2afcf56fa11ccf8fa

Observation 835dd210-d7db-479f-8fe3-a650e61c5b59 · outbound

This paper cites Estimating or Propagating Gradients Through Stochastic Neurons for Conditional Computation.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Estimating or Propagating Gradients Through Stochastic Neurons for Conditional Computation

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-21T05:49:40.943052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:6da86220fbf53815317e8f0e907b774bf3e459d23b23bfee8e179adf5242ad42

Observation 258e872b-34bb-4814-8ca8-dcca2fb39929 · outbound

This paper cites Improving image generation with better captions.Computer Science.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Improving image generation with better captions.Computer Science

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.418221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:11e570e284da66b5110de7a4b9e770deaf0f413974388f5caf21b9ffc21339ea

Observation a3334b46-bb73-477b-8446-197227deefce · outbound

This paper cites Training diffusion models with reinforcement learning.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Training diffusion models with reinforcement learning

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.465422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:84959364b4c4bf6125a657152dee9e19f06f846d29ef25e65288aa6af608c241

Observation db2ad8f2-f2ab-4ddd-aa5a-714234c05e24 · outbound

This paper cites Chameleon: Mixed-Modal Early-Fusion Foundation Models.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Chameleon: Mixed-Modal Early-Fusion Foundation Models

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-05-21T05:49:40.946283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:ec371e1602497bbb0fb77eb71c41148aea90ad79e5af3c769740ca8dd2e4a93c

Observation 2d1f24a4-fe15-4486-98e5-bac91995117a · outbound

This paper cites MaskGIT: Masked generative image transformer.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution MaskGIT: Masked generative image transformer

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.380493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:671c584bbb9cb3f315dcd966d633f7f65efa54b44c37e3650e1c4a0381da985b

Observation 268762fd-9748-4a32-b68b-9f72aa804967 · outbound

This paper cites Muse: Text-to-image generation via masked generative transformers.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Muse: Text-to-image generation via masked generative transformers

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.389778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:e15f925349be4302557fb4be4dbded99690fe62af31459092fe2bddaa43e12d7

Observation ffa2debe-7263-429d-9f91-d9849d2fb707 · outbound

This paper cites Softvq-vae: Efficient 1-dimensional continuous tokenizer.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Softvq-vae: Efficient 1-dimensional continuous tokenizer

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.485848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:3f85aaa50874e7e88988bbd718cd01afd7f5ff75a6ad598c18189eff07c86066

Observation 79beca6a-5e48-47a3-beaa-3d6c9701db1b · outbound

This paper cites BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-05-21T05:49:40.929209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:71921c460b7a2fe139348194c8b7bebe4c2268c729b163c983b52fef5366ca8a

Observation a7ca868f-348f-4fa3-9325-c9c4796e84cc · outbound

This paper cites Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-05-21T05:49:40.918911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:227ed7cd63947a47d5c7d563a64ed5007bb6bbf291904d833927247f85063e51

Observation 3e356c21-d1fc-4769-9178-26d9a918dd18 · outbound

This paper cites Directly fine-tuning diffusion models on differentiable rewards.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Directly fine-tuning diffusion models on differentiable rewards

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.462571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:305f9202311392ee4563e0d6a74e505cab8de3c257ca6c8105ade93ee1bd2382

Observation 5f675a19-cb74-4e86-96cf-ab269928e8ce · outbound

This paper cites Reward model ensembles help mitigate overoptimization.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Reward model ensembles help mitigate overoptimization

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.361004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:73ea020dbb4755861d00c1607c3d1595a18d37733d80062682b663bc0e15e091

Observation 647b8112-eb94-43c2-9cf4-9a748f4deaa7 · outbound

This paper cites Maximum likelihood from incomplete data via the EM algorithm.Journal of the Royal Statistical Society: Series B.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Maximum likelihood from incomplete data via the EM algorithm.Journal of the Royal Statistical Society: Series B

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.437236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:6b06355d0434dcf4b68c0f404b4f277b040e46c239c2653a02be8ce80d2926eb

Observation da327be3-0b19-4cce-950f-a444ebd5f578 · outbound

This paper cites CogView: Mastering text-to-image generation via transformers.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution CogView: Mastering text-to-image generation via transformers

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.400084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:f25010b8dd2d0e1133ef47575748e6911b17fe8e6688032e52ef10ce6f42f9fa

Observation b23b23b2-edea-48d2-b6ee-0f9bb0139b1e · outbound

This paper cites Taming transformers for high-resolution image synthesis.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Taming transformers for high-resolution image synthesis

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.482532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:eea21dc22beeb9d5f21e16beec289c9b1ee46be0568199c8d92618c82d31f704

Observation 00bcc4aa-984b-4064-973e-01b40deeb51a · outbound

This paper cites Dpok: Reinforcement learning for fine-tuning text-to-image diffusion models.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Dpok: Reinforcement learning for fine-tuning text-to-image diffusion models

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.451507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:bec4f918067cafa98b485c4a115b2363dbaf51e64ee2436d215daa5660380647

Observation 1a3381cf-29c5-4cff-bc00-8ac170d97b99 · outbound

This paper cites Scaling laws for reward model overoptimization.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Scaling laws for reward model overoptimization

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.356071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:6f046b0949c799f6a37e844e4043bab95bc6ce9e74a39983cb6f076f64e60f5a

Observation 9e3f5bf3-cd6e-4d47-8837-c3f80b0b497d · outbound

This paper cites GenEval: An Object-Focused Framework for Evaluating Text-to-Image Alignment.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution GenEval: An Object-Focused Framework for Evaluating Text-to-Image Alignment

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:49:40.887362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:f76d6584062e70ba9b65aac0df4b62d684aedbcbca7c11acf8a9993c3a9132aa

Observation 9047968a-8cf2-4756-be8d-1adcf9e420d0 · outbound

This paper cites Generative adversarial nets.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Generative adversarial nets

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.467945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:38af59fc9c75d702e04cbdf7ddb404d5d4d1b7cb9b2b8076bcd5fd79c496cde8

Observation 1cbc049c-b4d9-4084-ab53-8411ef582ac2 · outbound

This paper cites Bootstrap your own latent: A new approach to self-supervised learning.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Bootstrap your own latent: A new approach to self-supervised learning

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.363459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:620ad72ab1758a01b29dd0ba5a849b859e0dcc5b42744230c952473c4eae8a76

Observation 2369d1a8-69f5-4867-b899-ea405719d91c · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilibrium.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Gans trained by a two time-scale update rule converge to a local nash equilibrium

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.353723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:e1680608c87a8c53d7fa786b68a218f6966938592b1fd632434cb4e78dd4d675

Observation db6e7e4b-2c02-48ef-95bf-5c20baef0421 · outbound

This paper cites Straightening out the straight-through estimator: Overcoming optimization challenges in vector quantized networks.ICML.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Straightening out the straight-through estimator: Overcoming optimization challenges in vector quantized networks.ICML

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.439992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:f599bf60ad690b57eaa720f9f503da16837cf4323ca1e180cc2c844ef5f27389

Observation b058a46d-162c-4847-827b-e9ace4943bcb · outbound

This paper cites Image-to-image translation with conditional adversarial networks.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Image-to-image translation with conditional adversarial networks

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.395854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:0a91a4549b03f1467ed2851cd4e273d8e8e42ffd361ba91284692b65310e79ea

Observation c7c750c3-af60-409c-84e3-3dfbb6901a28 · outbound

This paper cites Categorical reparameterization with Gumbel-softmax.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Categorical reparameterization with Gumbel-softmax

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.374282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:81678bb2c5d25f7981ceabcb62cc1f2720154f2783bb891a89340a11c54ecad8

Observation ca4cf6d2-d980-4039-8557-4a92afa9714f · outbound

This paper cites T2I-R1: Reinforcing Image Generation with Collaborative Semantic-level and Token-level CoT.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution T2I-R1: Reinforcing Image Generation with Collaborative Semantic-level and Token-level CoT

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:49:40.939540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:b130e8c89dc981ed0a8c1e9fa6e5e12f3db3d37592a07040f8830278a6a4d87d

Observation 4d1779f1-4aa7-427e-945c-5220296c0d93 · outbound

This paper cites Fast decoding in sequence models using discrete latent variables.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Fast decoding in sequence models using discrete latent variables

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.454097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:5902c9cc4f1d42de78c18170ad32fff194d9e14ec005a683a7acc1c03e560220

Observation 252948aa-eb6f-4c20-9c65-8e615d8d2006 · outbound

This paper cites Scaling Laws for Neural Language Models.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Scaling Laws for Neural Language Models

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-05-21T05:49:40.895881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:65cf581261a47c58362788b286f24fa10818c325952b1688504edbbfad2f1671

Observation 17e547c3-6cdf-4146-8896-273e98c8b408 · outbound

This paper cites Overcoming catastrophic forgetting in neural networks.Proceedings of the national academy of sciences.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Overcoming catastrophic forgetting in neural networks.Proceedings of the national academy of sciences

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.402512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:81ddaf5b1ba18aa232e304ce8d715acd77489898f145bece5d2d3474ccfb633f

Observation be44dd44-0f06-4900-b493-a8e8e15be454 · outbound

This paper cites Rl with kl penalties is better viewed as bayesian inference.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Rl with kl penalties is better viewed as bayesian inference

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.457034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:a714d7e2c8b3f275f1d57b85b930e45159e5b36a0c114d5e74eb0f9520cdad18

Observation cc197f03-4d9c-4fef-a091-90bca4f42eeb · outbound

This paper cites Repa-e: Unlocking vae for end-to-end tuning of latent diffusion transformers.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Repa-e: Unlocking vae for end-to-end tuning of latent diffusion transformers

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.434350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:def1eefa76e5b66eb1498d4de62b07a123687fbd390fe8054709f0784f185dca

Observation 6e20c38a-dfc5-4474-8c14-540a4a85e694 · outbound

This paper cites Reinforcement Learning and Control as Probabilistic Inference: Tutorial and Review.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Reinforcement Learning and Control as Probabilistic Inference: Tutorial and Review

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-05-21T05:49:40.922660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:f2490794697ec5afa6235bc91fd2868f08c1c78b5b04b3c57068d390774b2d72

Observation 3ed9c092-213e-4874-b0c0-a54b1fd9485b · outbound

This paper cites Mergevq: A unified framework for visual generation and representation with disentangled token merging and quantization.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Mergevq: A unified framework for visual generation and representation with disentangled token merging and quantization

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.476791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:23abd54e62e1578b15e752062f1895995cadf1c41e169687e61ae390d24c1b80

Observation a00f1cdf-ef07-4fbc-970b-5e75ae6aa320 · outbound

This paper cites Va-π: Variational policy alignment for pixel-aware autoregressive generation.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Va-π: Variational policy alignment for pixel-aware autoregressive generation

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.383433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:52023c248540ecd2e757a6a1ff1d0fbf74710903b6806951574474bc8ae0f402

Observation 2895b185-b4ac-4d28-87ac-3f08c4ed41e7 · outbound

This paper cites Microsoft COCO: Common objects in context.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Microsoft COCO: Common objects in context

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.470854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:e5b926000ee2c1c251d081411fb9f5b72e1c41181186e9bae92940fcd3df78ac

Observation 5270a4d4-ee6b-4e3a-998e-d70a5bcb3ecd · outbound

This paper cites Flow matching for generative modeling.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Flow matching for generative modeling

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.415475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:25ee99f61b067d21b6f673907695dec10826e4e22dae06aef2d147effdfc2923

Observation 3124c93e-500a-4ac3-b566-d83a7d20d87f · outbound

This paper cites Flow-grpo: Training flow matching models via online rl.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Flow-grpo: Training flow matching models via online rl

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.504245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:dab17ba0ee4cde96a042e9ab5bdfd9e1e8d09ecef288f825305e8c3622f2ca8d

Observation 9b72ce2e-b051-42bb-b5d7-9f3e6428fc2a · outbound

This paper cites Decoupled weight decay regularization.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Decoupled weight decay regularization

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.459808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:bc8a537c3926a87c8d1ac09fba370a0219a26908ef23f39b59d0f0cad9a2d2c3

Observation 9b86cf57-c2e3-45f2-85f3-543a1f838036 · outbound

This paper cites Open-MAGVIT2: An Open-Source Project Toward Democratizing Auto-regressive Visual Generation.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Open-MAGVIT2: An Open-Source Project Toward Democratizing Auto-regressive Visual Generation

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:49:40.890541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:5ce3dc35c728842eac42c8eb70ba0b3e15e66a560e89f8842aa89acb37ad21f8

Observation a3e3c970-89a2-477a-b93c-d088cf3168e6 · outbound

This paper cites A view of the em algorithm that justifies incremental, sparse, and other variants.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution A view of the em algorithm that justifies incremental, sparse, and other variants

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.366121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:bd21bdc983194e1ccb45d52210636bd4e89b1ce9102af2315dd876e0841add92

Observation fe5c188d-28d5-4cc6-90ef-9a49c3d79606 · outbound

This paper cites Training language models to follow instructions with human feedback.NeurIPS.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Training language models to follow instructions with human feedback.NeurIPS

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.371457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:a5c3fdf8a91bd89315b2737b48d45a6c1d58eda9ea3d42dd31b8fab27d081fe4

Observation 9137680d-1a3e-44de-a9cd-1450bda5abaf · outbound

This paper cites Freeman, and Yu-Xiong Wang.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Freeman, and Yu-Xiong Wang

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.496257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:c4212bfeb529f3c2999eacf7a199ec0de8fe916554f6640160d689c32da99540

Observation 89afd83d-ad0e-4974-a251-bb568dd56ed6 · outbound

This paper cites Scalable diffusion models with transformers.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Scalable diffusion models with transformers

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.493788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:2fe3fa1a97823b14e51fcb1042d68ebac8fc6500795fcb79112a67538f738f5a

Observation 3feb2a8d-1b1b-4bd4-94a7-b51667d0fded · outbound

This paper cites Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-05-21T05:49:40.904468Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:62169c2c27fb965159cd123bfc4dd0fa410a951b2060dfbfd00638c668ebeebc

Observation 320f6548-1bc8-47ea-b981-e48324685b76 · outbound

This paper cites Reinforcement learning by reward-weighted regression for operational space control.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Reinforcement learning by reward-weighted regression for operational space control

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.488818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:54cbc3bed49bc6cac435ac80eca2d5949be10aad25187917e2196a7cca5ad86c

Observation 22391721-289a-46c4-81af-87c7c340988f · outbound

This paper cites Sdxl: Improving latent diffusion models for high-resolution image synthesis.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Sdxl: Improving latent diffusion models for high-resolution image synthesis

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.473560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:983e8f592e348f07a79608beebdcbc524322bbd42e3fa0ad78be19453f5bf5da

Observation 998adf9c-7bc3-406d-abe3-4598238b86b9 · outbound

This paper cites Aligning Text-to-Image Diffusion Models with Reward Backpropagation.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:49:40.925892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:586d8645bc7c150bb90820888f079ce6716da9b18ef70f965b63b0146c0996d0

Observation b0bd8bd7-27c3-4ad2-8331-05e46cb215a1 · outbound

This paper cites Qwen2.5 technical report.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Qwen2.5 technical report

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.501501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:2cc6f6b66d5a15620bee349d292ed6f5489cf355cbe2472ec810dd093e4a9c0b

Observation 1d74986a-d75c-47be-b3d0-ca03e3af65d0 · outbound

This paper cites Learning transferable visual models from natural language supervision.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Learning transferable visual models from natural language supervision

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.392998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:ceb2bc5b13bc14b28f446dbcfa7c23e803694eb1a10740044a9668e581b49283

Observation 12671343-6041-4760-8183-cf512836a71a · outbound

This paper cites Zero-shot text-to-image generation.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Zero-shot text-to-image generation

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.386057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:dac0170d83ed86c63daa458d7e52d3bf6cfd2b3c36bb0a866d58053713018b98

Observation 06bb285b-4ebf-4cba-92e7-c744690086ed · outbound

This paper cites Sequence level training with recurrent neural networks.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Sequence level training with recurrent neural networks

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.358541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:ec863ea3654b9c17793a01fb4386b6aa6aa01fd0daff89a1fd2b6955bf2d7937

Observation 6e2ec4e9-7129-4cff-9701-a7565a47cd5f · outbound

This paper cites Generating diverse high-fidelity images with VQ-V AE-2.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Generating diverse high-fidelity images with VQ-V AE-2

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.368817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:d48152e2b318980ceb14f6922464e0bc15d1151c8fe5cede39e236fa459fc72d

Observation eba44567-e1b8-47c8-931b-e91dfdafb10b · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution High-resolution image synthesis with latent diffusion models

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.491385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:3fef11c4c3053595be6a08d0c363940ab9a1dca0a2b0ad3de1ed42adc3d1a5c2

Observation fc73c602-8a6e-4197-923c-b468de688c15 · outbound

This paper cites Proximal Policy Optimization Algorithms.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Proximal Policy Optimization Algorithms

Reference 53

Resolution
verified exact
local_arxiv, observed 2026-05-21T05:49:40.893219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:4c7ac8857355b87862e093b06f9cc0c312ad5845832f5b5ce57216bbbf018586

Observation 166f7554-a7ba-4629-a787-355201b2604b · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 54

Resolution
verified exact
local_arxiv, observed 2026-05-21T05:49:40.907757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:b7ba19ebe7db62438bd1d7de0c925657d6e035c91f85057822e072727890c4e6

Observation 4764e018-30cd-4cae-bcdd-00c4cae8f206 · outbound

This paper cites Scalable image tokenization with index backpropagation quantization.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Scalable image tokenization with index backpropagation quantization

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.423558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:8ca3ca02706e874a5e80114bd677ea580342499eeea770c7ea7fdfaff84b783c

Observation e7b2041d-67f8-46d0-948d-c1b2a788d58d · outbound

This paper cites Journeydb: A benchmark for generative image understanding.NeurIPS.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Journeydb: A benchmark for generative image understanding.NeurIPS

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.410605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:fe4f0cb59892d12e4514727cea35aa79cf71a491fe3cc5d86f15fb22de056c60

Observation 04043bd1-4302-4c54-b35f-8a2cc443b8a5 · outbound

This paper cites Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-05-21T05:49:40.898442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:6f1658d29cf086118ea8a06d40ef4d8299c4cc85605e173342833511024d6d21

Observation 4cde9640-29e7-4918-88de-4f5ff581f8b7 · outbound

This paper cites Mean teachers are better role models: Weight-averaged consistency targets improve semi-supervised learning results.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Mean teachers are better role models: Weight-averaged consistency targets improve semi-supervised learning results

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.431547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:69ce5be9f903eb9d75662abea8056929de32bd43a91689154a0a530bccb64522

Observation 8ae6a361-1703-47a9-a7ff-add8d84f01ff · outbound

This paper cites Visual autoregressive modeling: Scalable image generation via next-scale prediction.NeurIPS.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Visual autoregressive modeling: Scalable image generation via next-scale prediction.NeurIPS

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.498961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:6dacfc815022f01095b2d1e59f0be65ca44b40480dcafd2af8ece8eec9c6d2ea

Observation 010c0cc4-a98a-4697-b5e0-13a97fd71478 · outbound

This paper cites Neural discrete representation learning.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Neural discrete representation learning

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.405231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:5f909ce806fa78922646d07478fca6a6295bc81bf9451310b45cf1699cc37a41

Observation 829e91de-b4a4-449f-aa06-2a4fe8cfe1c2 · outbound

This paper cites Diffusion model alignment using direct preference optimization.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Diffusion model alignment using direct preference optimization

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.479405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:0c663b5244ac74c6efb4038cde9d191d0bbafc0e9af2bdff5649d7dd9f5d7127

Observation 9eb7b858-8c79-4dae-a80a-29698f4c93e2 · outbound

This paper cites SimpleAR: Pushing the Frontier of Autoregressive Visual Generation through Pretraining, SFT, and RL.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution SimpleAR: Pushing the Frontier of Autoregressive Visual Generation through Pretraining, SFT, and RL

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:49:40.884259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:56508d2a88ef3b63608bfdeb81296b38bff375df1c54b8cc397aab3d7fa5d03c

Observation 758f3bb4-ea34-4db1-9cd1-55e07836234e · outbound

This paper cites Emu3: Next-Token Prediction is All You Need.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Emu3: Next-Token Prediction is All You Need

Reference 63

Resolution
verified exact
local_arxiv, observed 2026-05-21T05:49:40.914523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:e54bf87c1d0f18caa188290f0768151ef4fecc71d837b9b55e8ac222ff6da0b6

Observation bacc4c14-0fc6-4a36-8d32-9b0fda4aba92 · outbound

This paper cites Williams.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Williams

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.445121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:f7498d2a8db6c6b3597614f57fa45a532ad0d3a3dad6e6395dbf34b2b57ea79f

Observation 4a06f67c-2df2-43d9-9776-81527c0560e0 · outbound

This paper cites an unresolved cited work.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Unresolved cited work

Reference 65

Resolution
unresolved
raw_fallback, observed 2026-05-21T05:54:41.442460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:9a4a39864e635c5af88ca31cadf8ae4c4ea99663e57219f0da34b1d7a3ce2780

Observation a3f5fbd8-24b7-477b-93df-3184d7ceaa5c · outbound

This paper cites Human Preference Score v2: A Solid Benchmark for Evaluating Human Preferences of Text-to-Image Synthesis.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Human Preference Score v2: A Solid Benchmark for Evaluating Human Preferences of Text-to-Image Synthesis

Reference 66

Resolution
verified exact
local_arxiv, observed 2026-05-21T05:49:40.901207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:fca44f729ba3093f949555a091568b8cb7c7b74188b6a082e790966b7d207422

Observation 406f16d3-5e8a-4213-b879-18d379e0a399 · outbound

This paper cites Gigatok: Scaling visual tokenizers to 3 billion parameters for autoregressive image generation.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Gigatok: Scaling visual tokenizers to 3 billion parameters for autoregressive image generation

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.426179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:dcbb8dadce1213b3f515fb983b8bd4a5893cc5b2ad4ed8f8f93deb1d8fc978da

Observation b7ef1236-a161-4797-b15e-4484ab5a9a1b · outbound

This paper cites Imagereward: Learning and evaluating human preferences for text-to-image generation.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Imagereward: Learning and evaluating human preferences for text-to-image generation

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.420804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:249305dfdc37f91c8c2369baf28ef1d2343facef38ac94357aa7df5320e22b63

Observation e6d1e4ef-a56c-445f-b189-a6f2f08b37fc · outbound

This paper cites Scaling autoregressive models for content-rich text-to-image generation.TMLR.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Scaling autoregressive models for content-rich text-to-image generation.TMLR

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.412931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:047be3f47bef219b6051e8819e854d3287074d0b17cd93e8e63b6cb36b6482f7

Observation 49c00af3-c48c-44bb-81a9-23746d6ca8a2 · outbound

This paper cites An image is worth 32 tokens for reconstruction and generation.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution An image is worth 32 tokens for reconstruction and generation

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.428780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:f51188738d927c0f8ead366f28790712171e271258de86f498199d50a7241813

Observation e18ad42a-e5b5-44fd-8d5f-1fed50eb0348 · outbound

This paper cites Group critical-token policy optimization for autoregressive image generation.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Group critical-token policy optimization for autoregressive image generation

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.408153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:a29ee589d7d2f3a5a2951082e66bb14c10d028229c0f164b6d8c9e6bdd421bb4

Observation c7578e74-1dcf-4652-8874-a5b4cf6ffc0e · outbound

This paper cites The unreasonable effectiveness of deep features as a perceptual metric.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution The unreasonable effectiveness of deep features as a perceptual metric

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:54:41.448413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:8bf8366349cb1ca0aa132adf01767815c3e67fdac60d3c3ab298ac607ad1140e

Observation e0583df5-fc43-4f66-be56-590770960393 · outbound

This paper cites Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model.

RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model

Reference 73

Resolution
verified exact
local_arxiv, observed 2026-05-21T05:49:40.949420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T05:47:21.834145Z digest=sha256:fc2f90a5aa1234ffb101b4ac2910d2caed4e10ed7d8bea23aaa6c74b758cc8a8

Pith citing papers

Observation 2a3574f6-5c24-470c-b556-17fca20f6ffa · inbound

Beyond Token-Level Cross-Entropy: Fr\'echet Distributional Post-Training for Autoregressive Image Generation cites this paper.

Beyond Token-Level Cross-Entropy: Fr\'echet Distributional Post-Training for Autoregressive Image Generation RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution

Reference 36

Resolution
metadata mismatch
local_arxiv, observed 2026-08-05T00:44:18.374263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T00:44:16.917352Z digest=sha256:dbed6356020174c5aa833a7c27b067eb228ecb54dfcd99ad0d4afab93d591acf