Pith. sign in

Paper Citation Record · LEDGER

Native-Resolution Image Synthesis

As of 7 August 2026, this Paper Citation Record lists 87 of 87 outbound references and 5 inbound Pith citation observations for arXiv:2506.03131.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.03131 v1

Coverage vector

measured 87 of 87 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:15:27.939105Z

measured 92 of 92 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T10:59:53.959732Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-17T05:49:08.291505Z

Reference resolution

87 of 87 outbound references displayed

  • verified exact1
  • verified fuzzy14
  • unresolved72
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f161e984-723d-47a5-ad79-9f907838ce09 · outbound

This paper cites GPT-4 Technical Report.

Native-Resolution Image Synthesis GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:20.987565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:20.987565Z digest=sha256:36e5dc89e9121528b7d7b6d0be971411c9b8cdbd2582f6fadb4f50bcd0dded87

Observation 2f3135aa-8ffd-44f2-86a9-b3dc8ff00d74 · outbound

This paper cites Stochastic Interpolants: A Unifying Framework for Flows and Diffusions.

Native-Resolution Image Synthesis Stochastic Interpolants: A Unifying Framework for Flows and Diffusions

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:21.066092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:21.066092Z digest=sha256:f9631a1462acd956725df9051da7c79cec6d034f870914ba4875512cb52e1922

Observation 25b1c9d2-6bca-4aae-99ae-4fea246ec421 · outbound

This paper cites Building Normalizing Flows with Stochastic Interpolants.

Native-Resolution Image Synthesis Building Normalizing Flows with Stochastic Interpolants

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:21.162998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:21.162998Z digest=sha256:206fada526d2969eda7ecdbf04167249febfb0a5a4c2f56966190d47687e8cfc

Observation 0a9ecd36-6c81-4671-bc92-07b176fd83f5 · outbound

This paper cites Qwen2.5-VL Technical Report.

Native-Resolution Image Synthesis Qwen2.5-VL Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:21.344876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:21.344876Z digest=sha256:9f7f0d5d5a7a6eb32b19d1af3dc2e673df8436384b727735347143f3a77b6fad

Observation 98f66a24-4a99-4702-b379-d7a5dff555e7 · outbound

This paper cites eDiff-I: Text-to-Image Diffusion Models with an Ensemble of Expert Denoisers.

Native-Resolution Image Synthesis eDiff-I: Text-to-Image Diffusion Models with an Ensemble of Expert Denoisers

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:21.483651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:21.483651Z digest=sha256:94e8be1699de612dabda9fa5a6719894276d6d194875bdb9822d88ce57b9b987

Observation 482644b1-f70c-4076-9850-9e8fd82dd5a2 · outbound

This paper cites Language models are few-shot learners.

Native-Resolution Image Synthesis Language models are few-shot learners

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:21.596345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:21.596345Z digest=sha256:85452f968a4cc1c232746f826658ff2b11bb3f101a2e7e68fd6ee65e5a137e66

Observation 36887104-8b08-42f6-9869-9f24112971e3 · outbound

This paper cites Muse: Text-To-Image Generation via Masked Generative Transformers.

Native-Resolution Image Synthesis Muse: Text-To-Image Generation via Masked Generative Transformers

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:21.740312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:21.740312Z digest=sha256:151f519ebd9f3fc878e7fbffe181685f4b212ac3f9658c5cc2be67d94348ac0f

Observation 24041d9d-d77c-4be2-8835-cd52dd56f01a · outbound

This paper cites Maskgit: Masked generative image transformer.

Native-Resolution Image Synthesis Maskgit: Masked generative image transformer

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:21.900913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:21.900913Z digest=sha256:bcbc562c2dcada5b906ae1c27afce98388ec5893c9821052e03a54fcdea81910

Observation 7bd3b75c-c651-403f-9c05-5ed7b935108d · outbound

This paper cites CLEX: Continuous Length Extrapolation for Large Language Models.

Native-Resolution Image Synthesis CLEX: Continuous Length Extrapolation for Large Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:22.062379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:22.062379Z digest=sha256:75d724a69614a7ec97d9b6ea80bb7e434c4a3a110efe109499b64f22dcfbc58f

Observation bf88a2d5-a92a-43bc-b731-4b7e242a8bc7 · outbound

This paper cites Pixart-sigma: Weak-to-strong training of diffusion transformer for 4k text-to-image generation.

Native-Resolution Image Synthesis Pixart-sigma: Weak-to-strong training of diffusion transformer for 4k text-to-image generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:22.223665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:22.223665Z digest=sha256:006f20ed9c43741a8f3087afd572e369083d845d81abd97615911f6b10839af3

Observation cf4dee61-1d22-41b8-974f-d9699b0d446e · outbound

This paper cites PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis.

Native-Resolution Image Synthesis PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:22.301847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:22.301847Z digest=sha256:55a340a4e9eb5af78c5980e8e05a7f02bfbeb4347efdd7c65a11377c0248ff2e

Observation a12c5aa2-2178-4d83-a23b-245e5e273836 · outbound

This paper cites Deep Compression Autoencoder for Efficient High-Resolution Diffusion Models.

Native-Resolution Image Synthesis Deep Compression Autoencoder for Efficient High-Resolution Diffusion Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:22.346398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:22.346398Z digest=sha256:52bc8a2c98c862aec33085b0849e26d8341e3762947d99030872c17b4b970f09

Observation 37a638ae-47dd-4725-b78c-fbb8d9a149ae · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

Native-Resolution Image Synthesis Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:22.444972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:22.444972Z digest=sha256:16bcdcf5e8b095a85747a69b2153655cb30420286a5baa2acf6bd03888a72e26

Observation 30a6ae75-d721-4469-978d-ae9d04b7e3c4 · outbound

This paper cites Ramadge, and Alexander Rudnicky.

Native-Resolution Image Synthesis Ramadge, and Alexander Rudnicky

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:22.559084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:22.559084Z digest=sha256:1c96c7795b359797548616038edbc663d604494055405ecda6cc5db6db2eab2f

Observation 4e167c0e-2123-43cc-8301-97552952d1c9 · outbound

This paper cites FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning.

Native-Resolution Image Synthesis FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:22.663379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:22.663379Z digest=sha256:8186ffa5e46cfbb2d1689ce0ecd444117a53f6e580f7feeb26e97455e3abc697

Observation a0eef25d-31f6-4a4f-ac30-1290f02b085f · outbound

This paper cites Patch n' Pack: NaViT, a Vision Transformer for any Aspect Ratio and Resolution.

Native-Resolution Image Synthesis Patch n' Pack: NaViT, a Vision Transformer for any Aspect Ratio and Resolution

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:22.753088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:22.753088Z digest=sha256:e84dd18947927f23bb95fe7e79995ce623a7b181eec69ff1e0199bb6a5ba1500

Observation 3d12a58b-b8ce-424e-b677-f71b6e6545a2 · outbound

This paper cites Imagenet: A large-scale hierarchical image database.

Native-Resolution Image Synthesis Imagenet: A large-scale hierarchical image database

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:22.857667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:22.857667Z digest=sha256:f9d14e099c593f64206891fb170a144685dd9c0268c4cf3339888f0cd226ec52

Observation 075d9e44-70a8-4a9f-8aed-baa97876e823 · outbound

This paper cites Diffusion models beat gans on image synthesis.NeurIPS, 2021.

Native-Resolution Image Synthesis Diffusion models beat gans on image synthesis.NeurIPS, 2021

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:22.952406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:22.952406Z digest=sha256:d33b68459075616f9961bf324bc24512e064f6618e1d5f0fb713e99725ad02d3

Observation 21b77b61-a5a6-43a9-aa68-e0b466eb3bfa · outbound

This paper cites Cogview2: Faster and better text-to-image generation via hierarchical transformers.Advances in Neural Information Processing Systems, 35:16890–16902, 2022.

Native-Resolution Image Synthesis Cogview2: Faster and better text-to-image generation via hierarchical transformers.Advances in Neural Information Processing Systems, 35:16890–16902, 2022

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:23.035803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:23.035803Z digest=sha256:ba47aa84c08dddf568363938e4bdbe8b2afd99a969a04670f72f281257355406

Observation 7f65b2fc-1edc-4fbf-a366-a0a6211fc287 · outbound

This paper cites LongRoPE: Extending LLM Context Window Beyond 2 Million Tokens.

Native-Resolution Image Synthesis LongRoPE: Extending LLM Context Window Beyond 2 Million Tokens

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:23.081198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:23.081198Z digest=sha256:7bd97639c30e78f914ec03b9c3852e9af29c9e21c71fa6ff9240e284ad8cadc7

Observation b6634200-3a29-45aa-bf3b-019f36625017 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Native-Resolution Image Synthesis An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:23.137904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:23.137904Z digest=sha256:8eaa119efb855121b4f7533b171fedf3d86880e8816390af6eaff6faaae7eb45

Observation 9266811a-34a5-4f0d-a98c-7a7c7ff95be3 · outbound

This paper cites The Llama 3 Herd of Models.

Native-Resolution Image Synthesis The Llama 3 Herd of Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:23.254388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:23.254388Z digest=sha256:c725bd02ac1d131c90245f09a4ea053f21c943b938383a8db62ecc1ee5c4a094

Observation a75da351-980a-499c-96e2-3b060461a921 · outbound

This paper cites Scaling rectified flow transformers for high-resolution image synthesis.

Native-Resolution Image Synthesis Scaling rectified flow transformers for high-resolution image synthesis

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:15:28.880751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:15:23.341339Z digest=sha256:50350d1fdd9735bd27d59361e18fb7355fcf460f47475bcbc0f3c1d11f39802b

Observation 2735251d-14fc-475c-b0af-faa2d07ed9c6 · outbound

This paper cites Make-a-scene: Scene-based text-to-image generation with human priors.

Native-Resolution Image Synthesis Make-a-scene: Scene-based text-to-image generation with human priors

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:15:28.868042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:15:23.502549Z digest=sha256:53413cc28efb51f491a049c9057816a8ff9936a9de842b98b4dbd0895e37c2d2

Observation fca8e983-a16a-4156-8cd2-08fe6dec3114 · outbound

This paper cites MDTv2: Masked Diffusion Transformer is a Strong Image Synthesizer.

Native-Resolution Image Synthesis MDTv2: Masked Diffusion Transformer is a Strong Image Synthesizer

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:23.667405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:23.667405Z digest=sha256:82a5e28169d2750d2ff285961da8629239f6e4a46779054db3964c3fc429abdf

Observation 690b89dd-e414-4967-ac73-ece6bb013ac7 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Native-Resolution Image Synthesis DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:23.833220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:23.833220Z digest=sha256:cda7c39c4ec6a290b5adf80df1028b7b85ad1c036d62d5391a7bef2c949dcd0e

Observation 59c4f0f8-6bc4-4735-a902-d56b017cf2cc · outbound

This paper cites LM-Infinite: Zero-Shot Extreme Length Generalization for Large Language Models.

Native-Resolution Image Synthesis LM-Infinite: Zero-Shot Extreme Length Generalization for Large Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:23.945547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:23.945547Z digest=sha256:20fa272336b3c3b6be3f9950bdb77e9863dd20b6b3de26caed399d1eddb3573b

Observation 97a0a116-776d-47a9-b3eb-f5fb76a55301 · outbound

This paper cites an unresolved cited work.

Native-Resolution Image Synthesis Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:15:28.854999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:15:23.972830Z digest=sha256:d285fa5903d8b91aadb5e13a5114f44359658633a1ace2780889a08961bb9e00

Observation 8227c911-c8cc-4169-a446-eaf113c41080 · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilibrium.NeurIPS, 2017.

Native-Resolution Image Synthesis Gans trained by a two time-scale update rule converge to a local nash equilibrium.NeurIPS, 2017

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:24.033223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:24.033223Z digest=sha256:d9501dfa97d5bede47229801ebbe31239e371acb7cff805ccb6fb583ff4bff1d

Observation 9e245ee0-ed18-44f3-80af-a4c50841404c · outbound

This paper cites Denoising diffusion probabilistic models.NeurIPS, 2020.

Native-Resolution Image Synthesis Denoising diffusion probabilistic models.NeurIPS, 2020

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:24.115501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:24.115501Z digest=sha256:e57ee3fa2e97acbaf23c2fa7f10385d74b6b948d2c90947efca0aa2506de45d5

Observation cd4ca75d-a2af-4e68-a9ec-0461fbeceaa5 · outbound

This paper cites Lora: Low-rank adaptation of large language models.

Native-Resolution Image Synthesis Lora: Low-rank adaptation of large language models

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:15:28.826312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:15:24.240951Z digest=sha256:bac28936411982af9206c31fdfe7dade51dde33bdb619d26656f2bf1788125fd

Observation 24e79e3f-64d7-4472-a757-c37db8b6d5c8 · outbound

This paper cites Estimation of non-normalized statistical models by score matching.

Native-Resolution Image Synthesis Estimation of non-normalized statistical models by score matching

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:15:28.813464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:15:24.313090Z digest=sha256:49ade3613fdb2d6193e9a46e032d7fd032553e303722905a9eef28982a950d71

Observation d7feca47-4f14-4ee1-8e51-1838f68d4963 · outbound

This paper cites Springer Science & Business Media, 2005.

Native-Resolution Image Synthesis Springer Science & Business Media, 2005

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:15:28.800801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:15:24.402250Z digest=sha256:46241066befd3998e0c842cdc44bf80ee1e59cc6ae21fd57af90670d5c97f7e5

Observation f8eb2c10-6433-47ed-9611-8d1e20dcf930 · outbound

This paper cites Scaling Laws for Neural Language Models.

Native-Resolution Image Synthesis Scaling Laws for Neural Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:24.469159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:24.469159Z digest=sha256:6715e8a915b73bf37ce684495b24884b1d62e8581de335744b6e5179bfac6bc4

Observation e89498af-fbf2-4efd-9657-89d87486a412 · outbound

This paper cites Elucidating the design space of diffusion-based generative models.NeurIPS, 2022.

Native-Resolution Image Synthesis Elucidating the design space of diffusion-based generative models.NeurIPS, 2022

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:15:28.788083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:15:24.558757Z digest=sha256:4106c17bd6d78af9ffabbb0a0850162a08cdb2d93c63142781f6db6a49e98d0d

Observation 0cdcc470-883f-4204-a57c-48b908bba491 · outbound

This paper cites Analyzing and improving the training dynamics of diffusion models.

Native-Resolution Image Synthesis Analyzing and improving the training dynamics of diffusion models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:24.651388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:24.651388Z digest=sha256:49b22dd84e40f57c9e4bc334f1d50286566595bee1d808f86edc8fdc8e705643

Observation 40cc2cb7-e7c6-486d-a911-b22ce04c4bfc · outbound

This paper cites The impact of positional encoding on length generalization in transformers.Advances in Neural Information Processing Systems, 36:24892–24928, 2023.

Native-Resolution Image Synthesis The impact of positional encoding on length generalization in transformers.Advances in Neural Information Processing Systems, 36:24892–24928, 2023

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:24.746027Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:24.746027Z digest=sha256:aae979321b9882f8583ea3398d24d78e992a44a95cb7631e7dfc786cc49c6a59

Observation 3093cf34-c5da-4be4-8bc6-d61fc33f4068 · outbound

This paper cites Segment anything.

Native-Resolution Image Synthesis Segment anything

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:24.811862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:24.811862Z digest=sha256:df9d759125b1dca868a42bfe9aaebbd0ac65b2db5392a2a85c0896abf405e71a

Observation 14fe30ba-120e-49c9-858d-5e0e414a31ad · outbound

This paper cites Efficient Sequence Packing without Cross-contamination: Accelerating Large Language Models without Impacting Performance.

Native-Resolution Image Synthesis Efficient Sequence Packing without Cross-contamination: Accelerating Large Language Models without Impacting Performance

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:24.901809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:24.901809Z digest=sha256:5bd3483de670b8cad3d60e2688f847018e2d29b92a4d45d799975ebeaf46c149

Observation 59142b54-efb8-48f4-b716-dd3ad09df58a · outbound

This paper cites Kynkäänniemi, T.

Native-Resolution Image Synthesis Kynkäänniemi, T

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:15:28.752721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:15:24.993696Z digest=sha256:7af75a88f05121da0fedf8d9a12f072785b163ec8eee854e466776c991e1626d

Observation d0934f07-f05c-41d8-ae65-faedd23d444a · outbound

This paper cites Microsoft coco: Common objects in context.

Native-Resolution Image Synthesis Microsoft coco: Common objects in context

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:25.083512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:25.083512Z digest=sha256:5d69a10631eab2f4841f8a3bca8dc465ca2330f7fa0a17cc507f7c6e35c359f8

Observation 56bab1c0-cbb0-42d6-89b9-beea2051c08e · outbound

This paper cites Flow Matching for Generative Modeling.

Native-Resolution Image Synthesis Flow Matching for Generative Modeling

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:25.150326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:25.150326Z digest=sha256:4cbe8c315a0392559d3d6562ca35b17f32deaa3445151164261ab8f9959a12d7

Observation 7f2c05bc-dbca-441d-ad90-ed26bc2b99c3 · outbound

This paper cites DeepSeek-V3 Technical Report.

Native-Resolution Image Synthesis DeepSeek-V3 Technical Report

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:25.242734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:25.242734Z digest=sha256:2c22fb79e3b27f358ac91a6568e9fc7af6e5129f48d1546ead6c652f98c73777

Observation 8eb9cfa1-38f8-4c25-92cd-1445e60c001e · outbound

This paper cites Flow straight and fast: Learning to generate and transfer data with rectified flow.

Native-Resolution Image Synthesis Flow straight and fast: Learning to generate and transfer data with rectified flow

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:25.343513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:25.343513Z digest=sha256:7b90574ce65cc72cd0293f16acde6ef67f6f57eba8ffa1a26042755b99717f84

Observation 737abe9f-7f54-4bd9-9676-6e4f7176a03c · outbound

This paper cites Ntk-aware scaled rope allows llama models to have extended (8k+) context size with- out any fine-tuning and minimal perplexity degradation.

Native-Resolution Image Synthesis Ntk-aware scaled rope allows llama models to have extended (8k+) context size with- out any fine-tuning and minimal perplexity degradation

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:15:28.724807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:15:25.437263Z digest=sha256:8fbc7390f0e90f0cf9d32cff2f9bf6cad97732f378b1d9541d45b0a8861d184f

Observation 1949db0f-b8ea-4485-9e50-8ad4295c7220 · outbound

This paper cites Fit: Flexible vision transformer for diffusion model.

Native-Resolution Image Synthesis Fit: Flexible vision transformer for diffusion model

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:15:28.712428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:15:25.532035Z digest=sha256:f04b6974d1046b1c18d708ba4ded31fc47156a1a912a04d9a3ad5379210a21a7

Observation 9c7ee035-06ae-454e-b0ec-dc99ee5045b9 · outbound

This paper cites SiT: Exploring Flow and Diffusion-based Generative Models with Scalable Interpolant Transformers.

Native-Resolution Image Synthesis SiT: Exploring Flow and Diffusion-based Generative Models with Scalable Interpolant Transformers

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:25.626014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:25.626014Z digest=sha256:7f8135f0b340c77592a84901e9e77540e78aa834eb49ac05f0434f316326dd03

Observation 7d6d7170-d0b0-4583-97be-6fb4415c56dc · outbound

This paper cites SDEdit: Guided Image Synthesis and Editing with Stochastic Differential Equations.

Native-Resolution Image Synthesis SDEdit: Guided Image Synthesis and Editing with Stochastic Differential Equations

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:25.685719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:25.685719Z digest=sha256:ffafdfc150415579613dc0d7795610b346490545db566354ab3b16e8b05a5566

Observation 9924311c-de1d-4408-8f2f-f291691e1ec4 · outbound

This paper cites The llama 4 herd: The beginning of a new era of natively multimodal ai innovation, april 2025, 2025.

Native-Resolution Image Synthesis The llama 4 herd: The beginning of a new era of natively multimodal ai innovation, april 2025, 2025

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:25.754276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:25.754276Z digest=sha256:69d6d992223efa31b8a941c0ec51869743d80d34acd3a87d7ee1ff85dbb309e8

Observation 176abfe9-1a39-4f1a-9b10-41cea373b490 · outbound

This paper cites Generating Images with Sparse Representations.

Native-Resolution Image Synthesis Generating Images with Sparse Representations

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:25.847259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:25.847259Z digest=sha256:c6de87baa1c654e68b73a76986a46f355e0d1fe74f656feae0cf3a4541ec9d3e

Observation 50145533-0526-4762-8c89-8f8e77d0c46b · outbound

This paper cites GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models.

Native-Resolution Image Synthesis GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:25.908440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:25.908440Z digest=sha256:721ec0267b706009419b7ff53d61881e66f43fb063952463284cbf8c15dd4ea5

Observation 1be92052-8696-4bc8-bb13-94b934de05ee · outbound

This paper cites Improved denoising diffusion probabilistic models.

Native-Resolution Image Synthesis Improved denoising diffusion probabilistic models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:26.003489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:26.003489Z digest=sha256:21a7f340f80bac657a8548131c67e9f6a218bbffc087f8d7c24a5f91ac7c88ac

Observation ee5d10d4-65d7-47d6-8d57-c7b7d402f663 · outbound

This paper cites Pearson Educación, 1997.

Native-Resolution Image Synthesis Pearson Educación, 1997

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:15:28.684253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:15:26.101969Z digest=sha256:1b3ee9c69cceb6c3c98fc737d0c4a58b2c930aaa6da3972cf978958110f3ab27

Observation dac0b850-91bb-4a63-ae5b-d2a21564678b · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

Native-Resolution Image Synthesis DINOv2: Learning Robust Visual Features without Supervision

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:26.216276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:26.216276Z digest=sha256:ebdbbcb9befe9c6a58fc295e7d9865dd997e2eb076b091f252a840403c7ac206

Observation e035af05-5512-49fd-9f08-8178163eca4f · outbound

This paper cites Scalable diffusion models with transformers.

Native-Resolution Image Synthesis Scalable diffusion models with transformers

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:26.240491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:26.240491Z digest=sha256:957f4fec56d66e1a4d7ccf4fbce8abc0ce2d2a4a7e5d18b59c9b5c6258667741

Observation c6229fcc-5086-40ef-b099-856cb9f7ca87 · outbound

This paper cites YaRN: Efficient Context Window Extension of Large Language Models.

Native-Resolution Image Synthesis YaRN: Efficient Context Window Extension of Large Language Models

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:26.393194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:26.393194Z digest=sha256:52202bd74d63021af748b5d247dae1bf7b18a5f276987dda6633136a214a54b4

Observation 316110c1-aa34-4e6a-be20-971ffabbd9d2 · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

Native-Resolution Image Synthesis SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:26.611490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:26.611490Z digest=sha256:0785f23023e0d3cf26a869a775ec585c49cd9d91785a9bebf01c07da8082fac8

Observation 5eb6cec8-4910-4f10-9042-b9a2d792f6e9 · outbound

This paper cites Train short, test long: Attention with linear biases enables input length extrapolation.

Native-Resolution Image Synthesis Train short, test long: Attention with linear biases enables input length extrapolation

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:15:28.664013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:15:26.733572Z digest=sha256:ba25f54d2e8ce962bcf66ca03b6a570f43c5c037dcfdfe987df3184c46a059a4

Observation 5a6bdc94-a183-4f9d-bf8a-d1d796de94ed · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

Native-Resolution Image Synthesis Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:26.885811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:26.885811Z digest=sha256:fb02134caebf8ce416bd6140c17a4b6bfe0ed49b67918843d634cbeedcd53957

Observation 8ffbd902-cd6a-4e03-a83b-f6f1c0ac3e0c · outbound

This paper cites Zero-shot text-to-image generation.

Native-Resolution Image Synthesis Zero-shot text-to-image generation

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:15:28.651576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:15:27.054517Z digest=sha256:2a1644e75679df25b59111d42a55a661e52c1fd03050915d413205bda15581be

Observation 0bd08008-c6e1-48c2-b0be-83378756de88 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Native-Resolution Image Synthesis High-resolution image synthesis with latent diffusion models

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:27.070558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:27.070558Z digest=sha256:ad0e608c25a476538fa8048499f04331d773bc728aefa227a4c0698b9ea6152a

Observation c56deb2a-9f32-42ed-9608-37e81cf4df64 · outbound

This paper cites Salimans, I.

Native-Resolution Image Synthesis Salimans, I

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:15:28.631694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:15:27.151937Z digest=sha256:c5edc73653fc3c640466be68fc8d0fb3522149051e5f70c074408adf75baf52e

Observation 44ae72e7-1e62-4d6c-aeed-948c21a8ffa2 · outbound

This paper cites Adversarial diffusion distillation.

Native-Resolution Image Synthesis Adversarial diffusion distillation

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:27.248286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:27.248286Z digest=sha256:b0a470b5659f44080379303ef2ed0ca925ede418bdfa16a9cdbb820ee14f0d32

Observation 2578c484-905e-4f72-8bfa-528f3bc0bc88 · outbound

This paper cites Generative modeling by estimating gradients of the data distribution.

Native-Resolution Image Synthesis Generative modeling by estimating gradients of the data distribution

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:27.252346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:27.252346Z digest=sha256:9ce532f15955e8577d732ef42f5c7ebe55fac54bf4391666039d77961f55ac29

Observation 0f6a762a-4f27-4862-8e9c-0f80710cc57a · outbound

This paper cites Improved techniques for training score-based generative models.NeurIPS, 2020.

Native-Resolution Image Synthesis Improved techniques for training score-based generative models.NeurIPS, 2020

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:27.280745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:27.280745Z digest=sha256:73faff10627e5e16fa09882caba965e11f61509106ce50b11d4b1c1b54336c48

Observation 9fc7b962-7f85-4147-aac0-8b289c67f08d · outbound

This paper cites Score-Based Generative Modeling through Stochastic Differential Equations.

Native-Resolution Image Synthesis Score-Based Generative Modeling through Stochastic Differential Equations

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:27.370087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:27.370087Z digest=sha256:0330899dc563e5d5238eb427316cbadcce70b199daf888b74fef8789181742f7

Observation a50211fc-31df-4d42-880a-43456de3815d · outbound

This paper cites Springer, 2013.

Native-Resolution Image Synthesis Springer, 2013

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:15:28.594798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:15:27.447708Z digest=sha256:5758e8f00550a88850fc9a6d7b518b7808d78144c7ca996d7de93ea54adbf3c2

Observation 89af2ae6-5db2-4524-b0e6-10fc1ffcdc73 · outbound

This paper cites Roformer: Enhanced transformer with rotary position embedding.Neurocomputing, 2024.

Native-Resolution Image Synthesis Roformer: Enhanced transformer with rotary position embedding.Neurocomputing, 2024

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:27.526473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:27.526473Z digest=sha256:7d3e53692563711d95cf799ac08585032fd4e7911e2b045cf689131b6b237d8f

Observation 8a5d7fff-8a42-40cb-957c-aa7367634621 · outbound

This paper cites Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation.

Native-Resolution Image Synthesis Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:27.613456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:27.613456Z digest=sha256:f6eec69f42e8005f213d19b06759f06e7ffaed7b64183f065463fb75a88b7c79

Observation b8dae6e2-b953-4ce0-8233-84253194770e · outbound

This paper cites Seed1.5-VL Technical Report.

Native-Resolution Image Synthesis Seed1.5-VL Technical Report

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:27.691230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:27.691230Z digest=sha256:c7106b777c4f47909169bb0f33b738ef25d42fce811918ad59e205aab85ba919

Observation 4f4b7648-4cc5-42d5-9820-a4a588dfc9c6 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Native-Resolution Image Synthesis Gemini: A Family of Highly Capable Multimodal Models

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:27.762445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:27.762445Z digest=sha256:f8afe68c93edd6ccc2aefb35ec52157f5fe9004b1f0bc2c9c8153c3a53af35b7

Observation 983512aa-42be-4090-a934-f5d5f5aa5f3f · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

Native-Resolution Image Synthesis Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:27.846475Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:27.846475Z digest=sha256:f1c6dc1ed260a1f6198b80096733d422e3adc9de6e1cd6ee9b26de7bf9efdedf

Observation c2b9a487-b8ab-48e9-9476-b94bf5e51ded · outbound

This paper cites Gemma 3 Technical Report.

Native-Resolution Image Synthesis Gemma 3 Technical Report

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:27.879113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:27.879113Z digest=sha256:0cd5812c7e99cf9080cf4531abe1cef65a71272207bbadc5d11f449b334d0475

Observation ca6ce34a-977b-44cb-b5e1-2db8299ff808 · outbound

This paper cites Kimi k1.5: Scaling Reinforcement Learning with LLMs.

Native-Resolution Image Synthesis Kimi k1.5: Scaling Reinforcement Learning with LLMs

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:27.883333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:27.883333Z digest=sha256:e064a6cd6047e828bc7fa4a45a9e21804d98b5a9a5a8b54b459f94204b80ec8e

Observation ddef2aa0-1b0b-4acd-b6ca-d818d7237e27 · outbound

This paper cites Kimi-VL Technical Report.

Native-Resolution Image Synthesis Kimi-VL Technical Report

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:27.894108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:27.894108Z digest=sha256:753c36a69310a63c35fa0b6744a6f41b3555609c640e409492827ce61e542060

Observation bf6fd23a-01d3-413b-a43b-f0cfa40d0d82 · outbound

This paper cites Visual autoregressive modeling: Scalable image generation via next-scale prediction.Advances in neural information processing systems, 37:84839–84865, 2024.

Native-Resolution Image Synthesis Visual autoregressive modeling: Scalable image generation via next-scale prediction.Advances in neural information processing systems, 37:84839–84865, 2024

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:27.897881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:27.897881Z digest=sha256:13e7363abe3beaba6b8e92e6acee17e741474f28a1b09f50b3bee45d2f79ee74

Observation 2717444f-37c8-4a18-b2ae-3112867016f3 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Native-Resolution Image Synthesis LLaMA: Open and Efficient Foundation Language Models

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:27.901577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:27.901577Z digest=sha256:80c743a232fc660f02530e10992d76ed7fc61b80a170838f0bf181b6619b7197

Observation 8d8f8cfe-f9e0-4bdd-8cae-ecbeeee76172 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Native-Resolution Image Synthesis Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:27.905511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:27.905511Z digest=sha256:8b5a4f6098ae2eeffc20467bf9051e3d595bbc3ee77e460a4061adc111603eda

Observation bf02c922-3487-4c70-8f08-d8fbfc4217d2 · outbound

This paper cites FlowDCN: Exploring DCN-like Architectures for Fast Image Generation with Arbitrary Resolution.

Native-Resolution Image Synthesis FlowDCN: Exploring DCN-like Architectures for Fast Image Generation with Arbitrary Resolution

Reference 79

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:15:28.075956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:15:27.909683Z digest=sha256:062983e01e8b68d7dfd51edffb08d65badc0ddcea0af6c9802bbfdf108cc6507

Observation 835a150f-1d8a-44a9-8172-b9ec0f5a1bc0 · outbound

This paper cites FiTv2: Scalable and Improved Flexible Vision Transformer for Diffusion Model.

Native-Resolution Image Synthesis FiTv2: Scalable and Improved Flexible Vision Transformer for Diffusion Model

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:27.913353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:27.913353Z digest=sha256:5c0830c29731fd565013b683b7a883b4149e902c69a8d2d32a665cd483ceaf10

Observation b84bce03-8b41-4f15-861f-92659f6b1a16 · outbound

This paper cites SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformers.

Native-Resolution Image Synthesis SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformers

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:27.917023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:27.917023Z digest=sha256:c290e417ac6fecc50d05ee2ef482667bf42e80ca36097c81bf1e84be28c86763

Observation 9d157082-e74f-4f26-9d4f-e2affd1a205a · outbound

This paper cites Qwen2.5 Technical Report.

Native-Resolution Image Synthesis Qwen2.5 Technical Report

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:27.920953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:27.920953Z digest=sha256:3d03c5dea593d657b28a90a1c905ad858d1ed766a13434220200046170a06b6d

Observation d4e4bee8-5a77-40fe-8ac7-355eaf138879 · outbound

This paper cites CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer.

Native-Resolution Image Synthesis CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:27.924274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:27.924274Z digest=sha256:4f05a97db41bed7a914a953e39c3064fd5d86ef042801574dfca25a46e8dfa4b

Observation 0947f8ca-f047-4568-9340-f074930b1d8f · outbound

This paper cites MiniCPM-V: A GPT-4V Level MLLM on Your Phone.

Native-Resolution Image Synthesis MiniCPM-V: A GPT-4V Level MLLM on Your Phone

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:27.927710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:27.927710Z digest=sha256:444c2969016b8308e036aa5b3e3e471883eeb79c9d0058a9314e3f12d5d0f58d

Observation aa682b6b-9d61-4183-9d58-2f8001db09e3 · outbound

This paper cites Scaling Autoregressive Models for Content-Rich Text-to-Image Generation.

Native-Resolution Image Synthesis Scaling Autoregressive Models for Content-Rich Text-to-Image Generation

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:27.931555Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:27.931555Z digest=sha256:cb327e33f203159e2ea2496056a915c43e8dbb4c9428b565ee2f1b1a665d574d

Observation 96980f9b-1acf-47de-821d-41f3b86bbe48 · outbound

This paper cites Representation Alignment for Generation: Training Diffusion Transformers Is Easier Than You Think.

Native-Resolution Image Synthesis Representation Alignment for Generation: Training Diffusion Transformers Is Easier Than You Think

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:27.935094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:27.935094Z digest=sha256:c4283664151caab2ae5ffa5dd25422ff7c48447327269d82cb7b31731db9fc91

Observation cb2c27da-bdac-40a4-9dc7-a7ede0d331e0 · outbound

This paper cites PoSE: Efficient Context Window Extension of LLMs via Positional Skip-wise Training.

Native-Resolution Image Synthesis PoSE: Efficient Context Window Extension of LLMs via Positional Skip-wise Training

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:27.939105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:27.939105Z digest=sha256:ac2389198ab5f5ff53bc4ac4d7271002707359bef9f50f46c6b29090d6be5d6a

Pith citing papers

Observation 27e6125d-c332-49bb-83cc-97af61f0e6bc · inbound

PixNerd: Pixel Neural Field Diffusion cites this paper.

PixNerd: Pixel Neural Field Diffusion Native-Resolution Image Synthesis

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-06T10:59:53.959732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:59:53.959732Z digest=sha256:6e13e3d7c06cfe4b9224658763b4ffa02b3ef26b81f3354e5eca74cbc9a29903

Observation 9359812e-9cde-4758-a93a-988074429599 · inbound

Transition Models: Rethinking the Generative Learning Objective cites this paper.

Transition Models: Rethinking the Generative Learning Objective Native-Resolution Image Synthesis

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-05T10:19:54.506238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:19:54.506238Z digest=sha256:4e3816b2a247bd8b889667d754c5b4ff5f020782e7038849d76719d7c7cf93d7

Observation fcc0f4d4-3bec-46a9-afc1-1ea50c3a58ef · inbound

DeCo: Frequency-Decoupled Pixel Diffusion for End-to-End Image Generation cites this paper.

DeCo: Frequency-Decoupled Pixel Diffusion for End-to-End Image Generation Native-Resolution Image Synthesis

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-17T05:49:08.294010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-17T05:47:24.669763Z digest=sha256:6d7d17dee1fee8d4151659fece2c3c6d92151794823e6f553c658ffcd9a9feff

Observation cdf67ea2-5bff-46e8-b1f9-b09fb03a980a · inbound

VibeToken: Scaling 1D Image Tokenizers and Autoregressive Models for Dynamic Resolution Generations cites this paper.

VibeToken: Scaling 1D Image Tokenizers and Autoregressive Models for Dynamic Resolution Generations Native-Resolution Image Synthesis

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:51:11.672181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T04:12:58.100610Z digest=sha256:7858b7601f62fa246d67830d923c57ae56df9da2b03e470ba8c792943c22b7ec

Observation de2cb3be-e99b-4213-9930-2e68296213a9 · inbound

Mage-Flow: An Efficient Native-Resolution Foundation Model for Image Generation and Editing cites this paper.

Mage-Flow: An Efficient Native-Resolution Foundation Model for Image Generation and Editing Native-Resolution Image Synthesis

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T13:38:57.401548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:38:57.401548Z digest=sha256:142d3f8b2112dfad040b4814114880c3f51b860a04b941b10f97cdec9c1204ff