Pith. sign in

Paper Citation Record · LEDGER

Learning-to-Cache: Accelerating Diffusion Transformer via Layer Caching

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2406.01733.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.01733 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T20:53:59.400434Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 68593e84-379b-40e8-b2fd-bbb5509dca58 · inbound

PipeFusion: Patch-level Pipeline Parallelism for Diffusion Transformers Inference cites this paper.

PipeFusion: Patch-level Pipeline Parallelism for Diffusion Transformers Inference Learning-to-Cache: Accelerating Diffusion Transformer via Layer Caching

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-24T01:18:42.398669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-24T01:17:11.261301Z digest=sha256:c6ee1aee49625da34cd2a4145b8ef402d46c2e93d595ba1550c61b0c23f351f1

Observation 61b3cd58-0392-422d-9aca-7bbffb729098 · inbound

Accelerating Diffusion Transformer via Error-Optimized Cache cites this paper.

Accelerating Diffusion Transformer via Error-Optimized Cache Learning-to-Cache: Accelerating Diffusion Transformer via Layer Caching

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-09T20:53:59.400434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T20:53:59.400434Z digest=sha256:33fe9a4329df12c6e507783b4dbd891af41bf50d6010e1cf033b56190e268a6f

Observation 829f5af7-79e7-464a-9ae0-971a6cb21b3b · inbound

Efficient Diffusion Models: A Survey cites this paper.

Efficient Diffusion Models: A Survey Learning-to-Cache: Accelerating Diffusion Transformer via Layer Caching

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-09T16:13:36.755965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T16:13:36.755965Z digest=sha256:5758349a5b1913cac65f87b8095909e6d52f0f49219075a5b6c642962c543ecd

Observation d8c3a417-8360-489f-b328-292dce66eb14 · inbound

RainFusion: Adaptive Video Generation Acceleration via Multi-Dimensional Visual Redundancy cites this paper.

RainFusion: Adaptive Video Generation Acceleration via Multi-Dimensional Visual Redundancy Learning-to-Cache: Accelerating Diffusion Transformer via Layer Caching

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T13:45:09.410733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:45:09.410733Z digest=sha256:880c596902d359d7a37ce85622e66b787e6a8fe255b20bb621ef91430a659c49

Observation c31556c7-3364-4fe8-a5c9-86f377709fc5 · inbound

SkipVAR: Accelerating Visual Autoregressive Modeling via Adaptive Frequency-Aware Skipping cites this paper.

SkipVAR: Accelerating Visual Autoregressive Modeling via Adaptive Frequency-Aware Skipping Learning-to-Cache: Accelerating Diffusion Transformer via Layer Caching

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T05:05:00.152790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:05:00.152790Z digest=sha256:5423d53d600d5c240bdf6aadcf641acce8bac76a7adac911de759249d063147c

Observation d02a5606-e915-4bf8-892a-ba68de491c37 · inbound

DisCa: Accelerating Video Diffusion Transformers with Distillation-Compatible Learnable Feature Caching cites this paper.

DisCa: Accelerating Video Diffusion Transformers with Distillation-Compatible Learnable Feature Caching Learning-to-Cache: Accelerating Diffusion Transformer via Layer Caching

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-16T07:30:44.403100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-16T07:28:08.873654Z digest=sha256:8ee4f85505989bcc32ea9dbd3f8f436b77a2a916383d41fbf8c29ab3ce0a4b0b

Observation 4625368b-bcb5-4845-8a89-9cebe998ecd0 · inbound

AdaCorrection: Adaptive Offset Cache Correction for Accurate Diffusion Transformers cites this paper.

AdaCorrection: Adaptive Offset Cache Correction for Accurate Diffusion Transformers Learning-to-Cache: Accelerating Diffusion Transformer via Layer Caching

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-15T22:56:50.255035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T22:55:15.572867Z digest=sha256:dd4ed64cb74721767b3718288b6106736caad69949837898973359ab66bf51e2

Observation e1433d7d-a275-4bc1-968f-53b6163b80d6 · inbound

LESA: Learnable Stage-Aware Predictors for Diffusion Model Acceleration cites this paper.

LESA: Learnable Stage-Aware Predictors for Diffusion Model Acceleration Learning-to-Cache: Accelerating Diffusion Transformer via Layer Caching

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T21:23:43.299853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:23:43.299853Z digest=sha256:779405d8efdb60165dfda5a1f498dd5d251f6881d3a525c101d0017e4067bb71

Observation e4195a67-9b26-4346-8864-d8dc672cffc6 · inbound

S2O: Early Stopping for Sparse Attention via Online Permutation cites this paper.

S2O: Early Stopping for Sparse Attention via Online Permutation Learning-to-Cache: Accelerating Diffusion Transformer via Layer Caching

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-15T19:36:32.944227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T19:32:52.948154Z digest=sha256:f70c05eb382c512db5cbee614e5dcba2f452fb5c8b27599db81b3dc07880495f

Observation 4923d8da-09eb-4f6b-a655-f3c9fadb43cc · inbound

Not All Denoising Steps Are Equal: Model Scheduling for Faster Masked Diffusion Language Models cites this paper.

Not All Denoising Steps Are Equal: Model Scheduling for Faster Masked Diffusion Language Models Learning-to-Cache: Accelerating Diffusion Transformer via Layer Caching

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-16T07:47:33.002955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-16T07:43:45.415485Z digest=sha256:997102088e9ca5332a7c46f9a189b97b03a77c2edc7d4dddf79a4baa9504babf

Observation e3c1db61-0a81-487f-bf8c-4f61f74457fd · inbound

CoCoDiff: Optimizing Collective Communications for Distributed Diffusion Transformer Inference Under Ulysses Sequence Parallelism cites this paper.

CoCoDiff: Optimizing Collective Communications for Distributed Diffusion Transformer Inference Under Ulysses Sequence Parallelism Learning-to-Cache: Accelerating Diffusion Transformer via Layer Caching

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-10T10:24:20.640324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T10:23:19.236803Z digest=sha256:79efbd7dd7e886bcc2c5e2c6c9a575a09b2b884c30a57ccf160ba6765ed888a3

Observation 17d2dd36-afec-4d65-9b70-5e6a87ecde36 · inbound

DiTango: Cost-Effective Parallel Diffusion Generation with Selective Attention State Reuse cites this paper.

DiTango: Cost-Effective Parallel Diffusion Generation with Selective Attention State Reuse Learning-to-Cache: Accelerating Diffusion Transformer via Layer Caching

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T22:45:27.256030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T22:45:27.256030Z digest=sha256:a469f90c1b4a6a74c6a85ec295a0b5bd9971d5479df26522be48fd8382ba171d