Pith. sign in

Paper Citation Record · LEDGER

LinFusion: 1 GPU, 1 Minute, 16K Image

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:2409.02097.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2409.02097 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 19 of 19 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:52:25.338444Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T18:13:48.401465Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation bf69749e-acf1-484b-9ce6-84628edf7e92 · inbound

EfficientViM: Efficient Vision Mamba with Hidden State Mixer based State Space Duality cites this paper.

EfficientViM: Efficient Vision Mamba with Hidden State Mixer based State Space Duality LinFusion: 1 GPU, 1 Minute, 16K Image

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T15:07:15.935114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:07:15.935114Z digest=sha256:bfbe797923951b3ca88a052eb2add2656e2b190caa5159155117d9ad349a4b67

Observation 19d1ea95-5c5d-4bba-a4f9-829a63a55faf · inbound

Text-to-Image Synthesis: A Decade Survey cites this paper.

Text-to-Image Synthesis: A Decade Survey LinFusion: 1 GPU, 1 Minute, 16K Image

Reference 226

Resolution
unresolved
no resolver link, observed 2026-08-12T13:33:39.794240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:33:39.794240Z digest=sha256:f0e6d0cc0b902fdf1b76e0bde9dca36c8aa501ec5e6c5dffeeba330a0de1b774

Observation 7417f335-81de-4992-9dbe-44141fff43b6 · inbound

SnapGen: Taming High-Resolution Text-to-Image Models for Mobile Devices with Efficient Architectures and Training cites this paper.

SnapGen: Taming High-Resolution Text-to-Image Models for Mobile Devices with Efficient Architectures and Training LinFusion: 1 GPU, 1 Minute, 16K Image

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-11T16:56:01.715656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:56:01.715656Z digest=sha256:eeb5e664688377ae1e6e236aa920f4d215119bc963a0171f5c7b636a2ffb4cd1

Observation b457ad1b-99c7-44ec-b189-1aa15756176a · inbound

CLEAR: Conv-Like Linearization Revs Pre-Trained Diffusion Transformers Up cites this paper.

CLEAR: Conv-Like Linearization Revs Pre-Trained Diffusion Transformers Up LinFusion: 1 GPU, 1 Minute, 16K Image

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T10:54:30.888959Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:54:30.888959Z digest=sha256:fdf7eb4cf8a1f6aabde70604e10545ed4d6c9fc089a298e715b10ab05aa1a0e9

Observation fa2c8bcd-df98-43d8-a1c8-066f1b21ff69 · inbound

Parallel Sequence Modeling via Generalized Spatial Propagation Network cites this paper.

Parallel Sequence Modeling via Generalized Spatial Propagation Network LinFusion: 1 GPU, 1 Minute, 16K Image

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-10T17:16:43.672571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:16:43.672571Z digest=sha256:005ba96c2140da2509e9030fb0c12da27f82d7a17d9572b2145fbabf84cbed9d

Observation 175fa8c8-19da-46bc-96bb-62737264b31c · inbound

Pushing the Boundaries of State Space Models for Image and Video Generation cites this paper.

Pushing the Boundaries of State Space Models for Image and Video Generation LinFusion: 1 GPU, 1 Minute, 16K Image

Reference 755

Resolution
unresolved
no resolver link, observed 2026-08-09T17:07:38.320756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:07:38.320756Z digest=sha256:dba66a3c4d009e1ec67140b0f55e9c5f6597ec904a11f2e4de9c572359b2311c

Observation 0a5f403d-ccf1-4369-8efd-ea393381b59f · inbound

Turbo2K: Towards Ultra-Efficient and High-Quality 2K Video Synthesis cites this paper.

Turbo2K: Towards Ultra-Efficient and High-Quality 2K Video Synthesis LinFusion: 1 GPU, 1 Minute, 16K Image

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-16T11:52:25.338444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:52:25.338444Z digest=sha256:c3c17e3237d6bcd7af8967784e56c64bde27b62877a527a5946ed5236360e5f1

Observation 1ce73f91-59b4-402d-b201-ec65c663fcb5 · inbound

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions cites this paper.

Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions LinFusion: 1 GPU, 1 Minute, 16K Image

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-16T05:13:23.346402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:13:23.346402Z digest=sha256:f891d9bc313864eafd3133928589eb5ca518f6eac0d1a3c3a8449ca8b5b1aa84

Observation 07f81108-c6cd-481c-b297-ffa44504f9ba · inbound

Long-Context State-Space Video World Models cites this paper.

Long-Context State-Space Video World Models LinFusion: 1 GPU, 1 Minute, 16K Image

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:03:17.907050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:03:17.907050Z digest=sha256:205d18c064b0e0241739b09d54753fe05e6474369b11c09e5add1a7098f3d2c8

Observation d98c8544-43ef-4d71-ba2b-244e82ba604d · inbound

Video World Models with Long-term Spatial Memory cites this paper.

Video World Models with Long-term Spatial Memory LinFusion: 1 GPU, 1 Minute, 16K Image

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:39.512049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:39.512049Z digest=sha256:093c5feb52f355f23be84a6f8f6c21592257a8c63a22312227185bdde5665986

Observation cd191463-0216-4f84-963f-05e85a2ca15b · inbound

Exploring Diffusion Transformer Designs via Grafting cites this paper.

Exploring Diffusion Transformer Designs via Grafting LinFusion: 1 GPU, 1 Minute, 16K Image

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T10:29:04.695484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:29:04.695484Z digest=sha256:01f2fcaebb621849c5580adabc9a88ef7443da381194804cc2af37862677bec5

Observation cf12ec93-5e9e-4d3a-b3e8-8c43e14b608e · inbound

Diffusion Transformer-to-Mamba Distillation for High-Resolution Image Generation cites this paper.

Diffusion Transformer-to-Mamba Distillation for High-Resolution Image Generation LinFusion: 1 GPU, 1 Minute, 16K Image

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T18:44:38.453093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:44:38.453093Z digest=sha256:c2ac024cdb78f006c3d5572d1a5cfc90e66ebcf0654aaec3dee20f5a337cb23b

Observation 19a70a53-fd62-493b-b926-1a63816c9bf1 · inbound

Consistent and Controllable Image Animation with Motion Linear Diffusion Transformers cites this paper.

Consistent and Controllable Image Animation with Motion Linear Diffusion Transformers LinFusion: 1 GPU, 1 Minute, 16K Image

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-05T22:21:05.627730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:21:05.627730Z digest=sha256:962860ea93a7052be5747ee8345222b423da05615557b168b944ebcccbcebe63

Observation 6bbafdf3-6dff-4da1-ac68-89a60b371c21 · inbound

UltraImageGen: Efficient Ultra-High-Resolution Image Generation with Hierarchical Local Attention cites this paper.

UltraImageGen: Efficient Ultra-High-Resolution Image Generation with Hierarchical Local Attention LinFusion: 1 GPU, 1 Minute, 16K Image

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-04T09:20:06.875247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:20:06.875247Z digest=sha256:94bfe4982a0274e6b324d117fd269e85f9f30fbc1a73567172c5c85cfacae0be

Observation 3924b6fe-0d69-44da-9a0e-4d67809dc8fa · inbound

SnapGen++: Unleashing Diffusion Transformers for Efficient High-Fidelity Image Generation on Edge Devices cites this paper.

SnapGen++: Unleashing Diffusion Transformers for Efficient High-Fidelity Image Generation on Edge Devices LinFusion: 1 GPU, 1 Minute, 16K Image

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-03T10:56:13.971631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:56:13.971631Z digest=sha256:eae8becd45325c8f907fc3021c38d1b8ec6c5ec471c036d5c09daa65f3dc8af8

Observation 4147bfd2-a1d1-43ee-a93c-b309a0d6b2b8 · inbound

HierEdit: Region-Aware Hierarchical Diffusion for Efficient High-Resolution Editing cites this paper.

HierEdit: Region-Aware Hierarchical Diffusion for Efficient High-Resolution Editing LinFusion: 1 GPU, 1 Minute, 16K Image

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:43:22.561733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-20T14:40:20.517450Z digest=sha256:20a7c8d2a2e8c2e4c27ff53853e13bb99b9c422140a72c09b830cffba5121dbc

Observation 15d1ca9e-256c-4ec1-9c1d-c0a458e771e5 · inbound

PixVerve: Advancing Native UHR Image Generation to 100MP with a Large-Scale High-Quality Dataset cites this paper.

PixVerve: Advancing Native UHR Image Generation to 100MP with a Large-Scale High-Quality Dataset LinFusion: 1 GPU, 1 Minute, 16K Image

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-20T05:23:03.904990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-20T05:19:30.372528Z digest=sha256:e8663e5a568653e4d6fede71373c7b278964d2d863571eeaf3d28032ceb678c2

Observation 626281fc-71c4-4dd3-8924-4904de66788d · inbound

JetViT: Efficient High-Resolution Vision Transformer with Post-Training Attention Search cites this paper.

JetViT: Efficient High-Resolution Vision Transformer with Post-Training Attention Search LinFusion: 1 GPU, 1 Minute, 16K Image

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:13:48.402865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-29T18:12:21.124678Z digest=sha256:5b26146852add94299528de5081762ec1801f9cc684d48593e5bd2ac3aac5716

Observation 38cd8e5f-3b5d-40f3-8dd3-02c8dd248124 · inbound

Scaling Parallel Sequence Models to Foundation-Scale Vision Encoders cites this paper.

Scaling Parallel Sequence Models to Foundation-Scale Vision Encoders LinFusion: 1 GPU, 1 Minute, 16K Image

Reference 123

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T19:32:35.228393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-28T19:23:08.100056Z digest=sha256:831e9b26fb9f011ac3d376a9199b446e32d5614ace82ede3757ded131758bec6