Pith. sign in

Paper Citation Record · LEDGER

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions

As of 17 August 2026, this Paper Citation Record lists 46 of 46 outbound references and 0 inbound Pith citation observations for arXiv:2504.21292.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.21292 v1

Coverage vector

measured 46 of 46 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T05:13:23.656049Z

measured 46 of 46 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

46 of 46 outbound references displayed

  • verified exact0
  • verified fuzzy31
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 43f91412-12f4-402c-b694-21aad15dec5f · outbound

This paper cites Butterworth.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions Butterworth

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:13:25.966795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T05:13:23.175746Z digest=sha256:c2c02b1ac04763e52d0ab16c808fad7571b199e46c607cf913a4ef37428e0914

Observation e7d4d3cd-aa4a-4d6e-911a-ec62119f8453 · outbound

This paper cites Flash Diffusion: Accelerating Any Conditional Diffusion Model for Few Steps Image Generation.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions Flash Diffusion: Accelerating Any Conditional Diffusion Model for Few Steps Image Generation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-16T05:13:23.186771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:13:23.186771Z digest=sha256:52dc4b8d459b6992fbb7d26aa9a46508f92789f6e5b57852740d2da0c2521c00

Observation 220a54da-d045-4b39-bcaa-13504f7f157c · outbound

This paper cites Pixart-Σ: Weak-to-strong training of diffusion transformer for 4k text-to-image generation.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions Pixart-Σ: Weak-to-strong training of diffusion transformer for 4k text-to-image generation

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:13:25.893786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T05:13:23.191134Z digest=sha256:2970c7f4d6b25260d35fc7b82e771c96afd23525942999e689e0d555c8dc6293

Observation 169f7757-9882-407c-b96c-0bd968dab5b0 · outbound

This paper cites Simple baselines for image restoration.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions Simple baselines for image restoration

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:13:25.881518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T05:13:23.195286Z digest=sha256:e8d5fdbddd19cc0ba66fe67ad42eef38ef1b9e5d294ab254193406240728d8e7

Observation 440b074c-1963-43e2-8844-b4f6b1ab1a7f · outbound

This paper cites How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-16T05:13:23.201201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:13:23.201201Z digest=sha256:cdab0882bd64609d8db1ca8d104ac651633b3d2e96d5dc3923c7a281290ffb6e

Observation 8ecd76a3-1904-4344-9473-7ae4d76bd0b5 · outbound

This paper cites Rifegan: Rich feature generation for text-to-image synthesis from prior knowledge.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions Rifegan: Rich feature generation for text-to-image synthesis from prior knowledge

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:13:25.826809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T05:13:23.205603Z digest=sha256:abbca9904a5d4a87cf4c3249ea047037ccd05d28a5ce80aa616735011df8c906

Observation 9e648e70-1a8d-44ae-82f3-57e171ad2ce1 · outbound

This paper cites Kaplan, and Enrico Shippole.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions Kaplan, and Enrico Shippole

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:13:25.766066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T05:13:23.209601Z digest=sha256:ec68e71d701b5e3fffec867893372de2cdf8d08a5d0355abf364d70d5f3ed183

Observation e65840dc-93dc-4a1b-9366-1da507335c6a · outbound

This paper cites Transformers are ssms: Generalized models and efficient algorithms through structured state space duality.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions Transformers are ssms: Generalized models and efficient algorithms through structured state space duality

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:13:25.753001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T05:13:23.214499Z digest=sha256:a71cd4d3745b4a69298228b9857dd97a88c8adf65d3fc72c13d7af6a83568f4c

Observation 3477e814-f06d-4785-829d-e2b3f3615f3c · outbound

This paper cites Taming transformers for high-resolution image synthesis.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions Taming transformers for high-resolution image synthesis

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:13:25.488215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T05:13:23.218081Z digest=sha256:03c555943c4f753a5bead7976640f38955d53d2b22c649068f8ff3d3ee139a31

Observation 7f555d8b-8d37-4321-95c9-af2e28228302 · outbound

This paper cites Mamba: Linear-Time Sequence Modeling with Selective State Spaces.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions Mamba: Linear-Time Sequence Modeling with Selective State Spaces

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T05:13:23.222435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:13:23.222435Z digest=sha256:f128d91e216f966bc4eff472cbe6568dcbee0aa54181101874f17e11db20555c

Observation dc921f8e-2e7f-4f71-96c6-c1639c4f77c4 · outbound

This paper cites Efficient diffu- sion training via min-snr weighting strategy.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions Efficient diffu- sion training via min-snr weighting strategy

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:13:25.476244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T05:13:23.253133Z digest=sha256:d07360c0ab9c989aa44485ac3929f4ff7ceefe1a237fc13d7e4a783360f13746

Observation 9128077e-87c4-4fbd-84c8-1b3d9b7bab7d · outbound

This paper cites Neighborhood attention transformer.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions Neighborhood attention transformer

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:13:25.357697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T05:13:23.295471Z digest=sha256:193591538ab1ff54e24710156b8212a835c3badb40b4501b657ba117e7f2cc92

Observation 0e58f682-61d2-46eb-86ce-b6238bac75e0 · outbound

This paper cites Classifier-free diffusion guidance.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions Classifier-free diffusion guidance

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-16T05:13:23.316049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:13:23.316049Z digest=sha256:d5f39846ed08ceb5ce8dd7c1f8648796278fbcbc327a3e8663866552863a29cd

Observation b893b62c-f6a4-4858-a5d3-6ea4f831a3bd · outbound

This paper cites Denoising diffu- sion probabilistic models.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions Denoising diffu- sion probabilistic models

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:13:25.292563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T05:13:23.319834Z digest=sha256:c6c111ecb7e9ee8c3c48873242d5781f9a98137cbf8366e24a02c9a4a32cc9ec

Observation 276fecd5-2493-48d6-8fef-42d232e4358d · outbound

This paper cites Alias-free generative adversarial networks.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions Alias-free generative adversarial networks

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:13:25.234747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T05:13:23.323490Z digest=sha256:b8e9244e1f95f4210bb63cb857d93eae220888393189ffa41968ddfad8d1a7c6

Observation dc756dc6-b5a3-493c-b60f-4c5aff047ddb · outbound

This paper cites Elucidating the design space of diffusion-based generative models.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions Elucidating the design space of diffusion-based generative models

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:13:25.188853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T05:13:23.327206Z digest=sha256:be77161ebeadb766af0b3c4c30d1e377826122d1d01a8b551600bcdf65db2cc0

Observation 0df50b9b-cae6-40b6-abb7-9f8afa36dd61 · outbound

This paper cites Analyzing and improving the training dynamics of diffusion models.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions Analyzing and improving the training dynamics of diffusion models

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:13:25.177483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T05:13:23.331790Z digest=sha256:2960ed7336316276ed839f7a51bfee03cc19534ebd181ff1a3730526f76735d9

Observation c1d3fd5d-3da1-48e3-b490-3f3eba1f5e13 · outbound

This paper cites an unresolved cited work.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-16T05:13:25.165990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T05:13:23.335657Z digest=sha256:81725a99723fe09bec484bc8cfe9e21a98ebc6c33e010a905bf7d8936df92424

Observation ed7de0cd-bf6a-4282-b846-6cf3b8ffe322 · outbound

This paper cites an unresolved cited work.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-16T05:13:24.990278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T05:13:23.338982Z digest=sha256:17d6416e4a6a2ae9ab9cac095fa874bf9afdd108c845bf4ef0a180eb2112149f

Observation 2a2049d4-7da3-4801-a954-26c1e83f79c0 · outbound

This paper cites Faster Diffusion via Temporal Attention Decomposition.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions Faster Diffusion via Temporal Attention Decomposition

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-16T05:13:23.342919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:13:23.342919Z digest=sha256:973ed7ffeb0de89bbcaa591d907723771a51ae96200157b0331a0b1e29146cad

Observation 1ce73f91-59b4-402d-b201-ec65c663fcb5 · outbound

This paper cites LinFusion: 1 GPU, 1 Minute, 16K Image.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions LinFusion: 1 GPU, 1 Minute, 16K Image

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-16T05:13:23.346402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:13:23.346402Z digest=sha256:4a15ffdcaf8fa9ed951f3a43519e27c1a18fc235fb1f1c670f7ef98839971607

Observation 01f18ee4-3953-4363-bca3-0481593350fc · outbound

This paper cites More Control for Free! Image Synthesis with Semantic Diffusion Guidance.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions More Control for Free! Image Synthesis with Semantic Diffusion Guidance

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-16T05:13:23.400547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:13:23.400547Z digest=sha256:e23e50439784eddbcf0c27bcd9f117debcc9e29170b19e92239a1e47c864e920

Observation 92f33a93-d365-4541-8bdb-b352ccbbbc7f · outbound

This paper cites Swin transformer: Hierarchical vision transformer using shifted windows.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions Swin transformer: Hierarchical vision transformer using shifted windows

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:13:24.977269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T05:13:23.443728Z digest=sha256:14f10be5679ca694abe78da55cc5329978b13f4fb877e5921588bbce7b3d7ebe

Observation 802b7057-62c9-4fdc-9369-eaf4b656c76c · outbound

This paper cites Token caching for diffusion transformer acceleration.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions Token caching for diffusion transformer acceleration

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-16T05:13:23.467486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:13:23.467486Z digest=sha256:769f3e25510f30657e90dc5f1d4882b26d9d96d69ffcf6df06cde67655ae59ed

Observation 7197f8be-b6cb-47d8-b49d-0f910cb3ef38 · outbound

This paper cites DPM-Solver++: Fast Solver for Guided Sampling of Diffusion Probabilistic Models.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions DPM-Solver++: Fast Solver for Guided Sampling of Diffusion Probabilistic Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-16T05:13:23.478064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:13:23.478064Z digest=sha256:046a20318566df101f8c4c5e5018771325b0afcbd9f7113345ecd492b0170c50

Observation 313f2e72-1254-4b13-b288-9f7b931c9127 · outbound

This paper cites Understanding the effective receptive field in deep convolu- tional neural networks.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions Understanding the effective receptive field in deep convolu- tional neural networks

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:13:24.909520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T05:13:23.483044Z digest=sha256:0a86d681a1fd0530c6b658018c0b08446cd30b18e96e19c8489e0429ce84c572

Observation 3f35b6bf-8c54-48d5-a74c-a2a3273b5b27 · outbound

This paper cites Kingma, Stefano Ermon, Jonathan Ho, and Tim Salimans.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions Kingma, Stefano Ermon, Jonathan Ho, and Tim Salimans

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:13:24.796516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T05:13:23.486050Z digest=sha256:9c968806fed3169203140ad924af0bb21f431fe260b16745d8b6230aa98fa4c1

Observation 0c64192b-3673-4ed4-b1b7-c66a021043c2 · outbound

This paper cites GLIDE: towards photorealistic image gen- eration and editing with text-guided diffusion models.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions GLIDE: towards photorealistic image gen- eration and editing with text-guided diffusion models

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:13:24.783837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T05:13:23.489064Z digest=sha256:788c6469bff8084318fabe0eb90dbe346a90d9a0c34db4e5760b025f14d85fc6

Observation e388da6e-fc73-4bbc-ab5e-6d73dff89b34 · outbound

This paper cites an unresolved cited work.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-16T05:13:24.698053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T05:13:23.492387Z digest=sha256:32f4a56d7704243cbf0403a6b17d78b5e8075104c6c48c9f792d330cdd08c596

Observation 040c90cb-ab5f-4c65-adbb-043919e320f1 · outbound

This paper cites Scalable diffusion models with transformers.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions Scalable diffusion models with transformers

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:13:24.610380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T05:13:23.495465Z digest=sha256:4edaba6f332f019e3c43fc0f5b7f669ef8d7c994512efb6ad3c6f4ccdbfa91ee

Observation 1f5361bb-6d0f-4b23-9c71-7a5fe37ed309 · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-16T05:13:23.498577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:13:23.498577Z digest=sha256:c3206a751918ef5d136f98f927bdf64c915f8716a503c80691c415aecbcdf623

Observation e177a634-6f2c-4dd4-91ea-a04dbae4060b · outbound

This paper cites Proakis and Dimitris G.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions Proakis and Dimitris G

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:13:24.599543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T05:13:23.501853Z digest=sha256:6d6fd1fc6afe439dc9401a77cfebd208379768d55b4c5331f8617e9b277fe0be

Observation 4ada7950-00d7-4029-86d5-61379ccc04e1 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions High-resolution image synthesis with latent diffusion models

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:13:24.471349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T05:13:23.505803Z digest=sha256:e21fefaa787d07ce41e6bc485fe08e7acb110314ff597e61e880add3c3df4c32

Observation 7d6c679d-5c39-44c9-90d8-a29cd4d030c9 · outbound

This paper cites U-net: Convolutional networks for biomedical image segmentation.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions U-net: Convolutional networks for biomedical image segmentation

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:13:24.460101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T05:13:23.510755Z digest=sha256:4235592fe2349202c05fc73f8482c9fd347428c56ea3ea221a408cfd74a9ee94

Observation cb1efbf7-1a4d-4add-8d42-1bac940b4e41 · outbound

This paper cites Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-16T05:13:23.516270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:13:23.516270Z digest=sha256:d81041fe5a7eac26638243428d9a00dffb0a8964ab083ad64f6932ab8dc055b0

Observation 0402eedd-cd4d-4700-bfb4-7ab1194bf52a · outbound

This paper cites Stylegan-t: Unlocking the power of gans for fast large-scale text-to-image synthesis.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions Stylegan-t: Unlocking the power of gans for fast large-scale text-to-image synthesis

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:13:24.448438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T05:13:23.536075Z digest=sha256:5d31504ac14dae53db2f4700f6917a97e465dec08a9cc6b54dc2eca15c9196b5

Observation 5316c712-2f78-461b-a5b7-49061dc37d3e · outbound

This paper cites Laion-5b: A large-scale dataset for training ai models.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions Laion-5b: A large-scale dataset for training ai models

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:13:24.409166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T05:13:23.569300Z digest=sha256:4c8d545d1ab58039ec49854d9de7fd3e3affb3fa8420d031ce3e2b7741768985

Observation aceb1237-0934-4799-880e-361589e470fe · outbound

This paper cites FORA: Fast-Forward Caching in Diffusion Transformer Acceleration.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions FORA: Fast-Forward Caching in Diffusion Transformer Acceleration

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-16T05:13:23.602018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:13:23.602018Z digest=sha256:c39d85117f9c2c7e89f739de36abb83646626b2a29076cec5edc3f93c29d9976

Observation 06d5847f-c57a-4677-a417-f2782a359fe2 · outbound

This paper cites Denoising diffusion implicit models.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions Denoising diffusion implicit models

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:13:24.383639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T05:13:23.623244Z digest=sha256:e4b2bb493c6070eb25b67b1679c71c12a9edc4c5b15e577ab4ef059c06825a91

Observation 4d3724bb-943a-4a00-a707-cbcc5c8072b7 · outbound

This paper cites Kingma, Ab- hishek Kumar, Stefano Ermon, and Ben Poole.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions Kingma, Ab- hishek Kumar, Stefano Ermon, and Ben Poole

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:13:24.373164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T05:13:23.627787Z digest=sha256:036c5d040baf375fc01be7b157a1f2fe431eb170d78bf4462d2e99f88fbf9bd2

Observation fadae87b-05bb-47df-b0bc-1a559cbd99cf · outbound

This paper cites Expos- ing flaws of generative model evaluation metrics and their unfair treatment of diffusion models.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions Expos- ing flaws of generative model evaluation metrics and their unfair treatment of diffusion models

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:13:24.360969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T05:13:23.632240Z digest=sha256:83a8e3452991a65daff9a387186db5b0b9018916aa7df0113b7bc088711321ab

Observation 28217a4f-e2fc-4163-a21f-bae3f5197bd8 · outbound

This paper cites Rethinking the incep- tion architecture for computer vision.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions Rethinking the incep- tion architecture for computer vision

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:13:24.297407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T05:13:23.636363Z digest=sha256:66b8f578226ed9d4b6e5f55306f25e0f4f35fac294c634821b46fc3c9af92677

Observation 8a4571b2-d12c-44a4-a198-c91a29bdb008 · outbound

This paper cites text2image-multi-prompt: A multi-prompt dataset for text-to-image generation.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions text2image-multi-prompt: A multi-prompt dataset for text-to-image generation

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:13:24.274251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T05:13:23.640562Z digest=sha256:71a4bd345a2bff7d3bbe41d311892e6d7aeafb275633f46d933e1a7e18136db8

Observation b0c7b602-70d8-4380-82dd-ab6acfa35708 · outbound

This paper cites Attention is all you need.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions Attention is all you need

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:13:24.192012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T05:13:23.644724Z digest=sha256:ffc9ab6bb0511575e9dec7cc720e052befed1df57803e217b951e71e9d90c94a

Observation dfeb0648-83c3-4146-9757-7b50c9b77d56 · outbound

This paper cites midjourney-v5-202304-clean.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions midjourney-v5-202304-clean

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:13:24.103327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T05:13:23.651842Z digest=sha256:2d30516ee871ac2394a579260546e0a26cc43bed1bc8f7419eb541de09a8acf0

Observation c6cdd0fa-bf29-46d7-8663-a87e6a4ae62c · outbound

This paper cites Ditfastattn: Attention compression for diffusion transformer models.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions Ditfastattn: Attention compression for diffusion transformer models

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:13:24.091649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T05:13:23.656049Z digest=sha256:e7c6b738baf6e71253a6453b449f33940849e323f68b84738dd2cac4cedb9071

Pith citing papers

No inbound Pith citation observations are available.