Pith. sign in

Paper Citation Record · LEDGER

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling

As of 23 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 4 inbound Pith citation observations for arXiv:2411.15381.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.15381 v2

Coverage vector

measured 52 of 52 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T14:27:01.856685Z

measured 56 of 56 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-14T06:49:04.349040Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

52 of 52 outbound references displayed

  • verified exact2
  • verified fuzzy28
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 6b339838-4b1f-4ca5-8aef-35986f752893 · outbound

This paper cites https://archive.org/details/archiveteam-twitter-stream-2018-04, 2018.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling https://archive.org/details/archiveteam-twitter-stream-2018-04, 2018

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:27:02.893254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T14:27:01.595608Z digest=sha256:6a7e296f927181177996e8744fc488b74b3e148908fa4150fe4549809c43a784

Observation 3f516ddc-f2b9-4a5d-bab4-6bcc75ed5b41 · outbound

This paper cites build, train, and deploy machine learning models at scale.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling build, train, and deploy machine learning models at scale

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:27:02.876943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T14:27:01.601213Z digest=sha256:9e85d2688b0be64229295f5a2c5c321478b4d01ecfd910f8b4969ee44ae2d6e4

Observation 8f690e63-0b8b-462d-bf61-ea9fd31dfbec · outbound

This paper cites https://developer.nvidia.com/nvidia-triton-inference-server, 2022.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling https://developer.nvidia.com/nvidia-triton-inference-server, 2022

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:27:02.859926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T14:27:01.606258Z digest=sha256:1dc2e15e8a336dbd496a9a5ddc8e814834b63b4b50be882baf3003980459fa8c

Observation 86fedf91-a8dc-47e3-aea6-4273b14aee71 · outbound

This paper cites Adobe firefly.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling Adobe firefly

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:27:02.843476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T14:27:01.611958Z digest=sha256:027932a5177b4095baf60d9541d96f54a954116907419c547b829866d9875ad8

Observation 58999957-9aff-4a58-880e-3deefce50e6a · outbound

This paper cites D., Williams, T., Sitaraman, R.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling D., Williams, T., Sitaraman, R

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:27:02.826968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T14:27:01.617052Z digest=sha256:23ccf29eaf5a20e294e7265c5194246f443cc0201274d318d6c43959f1d8ae68

Observation ec542047-654e-44bc-b8c1-dbe7c2be5ffe · outbound

This paper cites Batch: Machine learning inference serving on serverless platforms with adaptive batching.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling Batch: Machine learning inference serving on serverless platforms with adaptive batching

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:27:02.810385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T14:27:01.622178Z digest=sha256:3f750f5330ab94ec44bd470d69fbce077a5b125cf4fa09a91acecc3b231f949f

Observation c8ff1979-f004-4dfa-bb97-b99e5955f732 · outbound

This paper cites Adaptive neural networks for efficient inference.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling Adaptive neural networks for efficient inference

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:27:02.792984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T14:27:01.627824Z digest=sha256:36e4f8c4dee75fdde40ba1c5f85939973033ae01d0ae87b7125e970f492b5f75

Observation 93b30f8b-798e-404c-b557-c20d29f92779 · outbound

This paper cites FrugalGPT: How to Use Large Language Models While Reducing Cost and Improving Performance.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling FrugalGPT: How to Use Large Language Models While Reducing Cost and Improving Performance

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T14:27:01.632654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:27:01.632654Z digest=sha256:83ffed93acbdf3963956150f79ae2384c6eefe6bef24f0fa8c29105f2fd19dfa

Observation d2c6f702-a993-49e4-8ad1-dd4d94f6d50c · outbound

This paper cites J., Gonzalez, J.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling J., Gonzalez, J

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:27:02.775787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T14:27:01.637742Z digest=sha256:4e835df61cb3106f5c57549f85d4456a9d2e317b1f5aa0862ce04d7e21eaa3eb

Observation 1af2c845-98d8-4cb8-bd55-9faa881b24c6 · outbound

This paper cites J., Gonzalez, J.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling J., Gonzalez, J

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:27:02.758588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T14:27:01.642574Z digest=sha256:74a5a6a331a927430fbf86bb722c3c924708515c2fec42c59bd96c277e3b78f3

Observation 5511b38f-1a35-4125-8dab-0aeb9a02a198 · outbound

This paper cites Inferline: latency-aware provisioning and scaling for prediction serving pipelines.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling Inferline: latency-aware provisioning and scaling for prediction serving pipelines

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:27:02.742137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T14:27:01.647120Z digest=sha256:d0030b2d2c88e54bd93075c40cd90c58b6600f4de2a0cbf98fb6ac34595847ea

Observation e5d89cde-ccfc-4b19-bac1-99bf2d7ce8c4 · outbound

This paper cites Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T14:27:01.652242Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:27:01.652242Z digest=sha256:ca074fb51735dd31714c5b9c7a5833c97248f079dfd1bc20f518823f219722f9

Observation 2b3756cf-e1f7-4930-b950-1b63b1470ed1 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T14:27:01.657683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:27:01.657683Z digest=sha256:d110e07fcc84da240bdff17bad7d10388b24688ebe715853ca1c82f6084c5bf7

Observation 8d171270-cf19-41ae-9fa3-1077eb2cbab0 · outbound

This paper cites Serving \ DNNs \ like clockwork: Performance predictability from the bottom up.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling Serving \ DNNs \ like clockwork: Performance predictability from the bottom up

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:27:02.726100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T14:27:01.663000Z digest=sha256:ea97608401f557f1fd72d889c5432293f526a977299d5640f3992424b3775a68

Observation c6e34ff0-4642-4fa4-a7b2-cb0676c77479 · outbound

This paper cites R., Mishra, C.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling R., Mishra, C

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:27:02.708691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T14:27:01.667892Z digest=sha256:92528ae90a1b03921a0ec1ef7413ad23399bdf97a5424b86e7f3e3c14f8bd238

Observation 8bead135-69f7-4fc0-bca4-b0912a73c18e · outbound

This paper cites Sommelier: Curating dnn models for the masses.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling Sommelier: Curating dnn models for the masses

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:27:02.690982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T14:27:01.672853Z digest=sha256:6ac12068b8be64bf36d76997a45c8280826687b23644f0d64e8e7d05f0355b66

Observation afa21dde-a470-43fe-bd5e-44e1adb90d51 · outbound

This paper cites Language Model Cascades: Token-level uncertainty and beyond.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling Language Model Cascades: Token-level uncertainty and beyond

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T14:27:01.678122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:27:01.678122Z digest=sha256:0f5d881a83468acdbba17540e9f246f64d8ce86cc12eb42fd3ffd63169abbc65

Observation fb15b216-d4a8-4b73-8447-d9d70d83ae09 · outbound

This paper cites URL https://www.gurobi.com/.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling URL https://www.gurobi.com/

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:27:02.674398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T14:27:01.683265Z digest=sha256:42ab31a65c7e65510e635fb8e7ce8bb24989106a7ffe02c2f5c63963f06f2914

Observation 831dba85-a24b-47af-90bc-7754b85066ce · outbound

This paper cites Deep Residual Learning for Image Recognition.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling Deep Residual Learning for Image Recognition

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T14:27:01.688497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:27:01.688497Z digest=sha256:c3b49953760bf2fde2aaf0eb641e925cd9765bfbb115968b22f79cfd8494d9e4

Observation 596d9d61-41ed-4515-be45-4f5c66affc04 · outbound

This paper cites CLIPS core: A reference-free evaluation metric for image captioning.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling CLIPS core: A reference-free evaluation metric for image captioning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T14:27:01.693864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:27:01.693864Z digest=sha256:0fe10cb896d2318196d244f96a03cd797719bad39ab1356e768517762be43b3d

Observation 221f98b3-6b47-4258-b385-46660cb19b42 · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilibrium.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling Gans trained by a two time-scale update rule converge to a local nash equilibrium

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:27:02.657412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T14:27:01.699078Z digest=sha256:9c66ac386f3691ba351fa045f1afd114833c05b1bd57bf23ca6c942febaeb470

Observation 675856fc-7aa9-4be7-a41c-f922385c0ffc · outbound

This paper cites Scrooge: A cost-effective deep learning inference system.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling Scrooge: A cost-effective deep learning inference system

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:27:02.640758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T14:27:01.703949Z digest=sha256:3279e64ccd37c4f044cc4d2fb0441fdb3a55f1e12daf1780c830a05eda10722b

Observation 8e738f5b-25ea-4398-8501-25fd38922ff3 · outbound

This paper cites Hugging face models hub, 2024.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling Hugging face models hub, 2024

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:27:02.624099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T14:27:01.708704Z digest=sha256:3599ca9c269ae18b6b045ccd4df02d63011a0deb03640d099c2180d4cc94bf8e

Observation ccba5706-5591-4be1-ad01-5bbfc39aed41 · outbound

This paper cites Pick-a-Pic: An Open Dataset of User Preferences for Text-to-Image Generation.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling Pick-a-Pic: An Open Dataset of User Preferences for Text-to-Image Generation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T14:27:01.713566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:27:01.713566Z digest=sha256:3f5869305efe68dca893c713938a345e9a114fbb1b92c990ea78636312a7a590

Observation 6073e157-e86f-4063-9833-cfbb4dcdb6c7 · outbound

This paper cites Willump: A Statistically-Aware End-to-end Optimizer for Machine Learning Inference.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling Willump: A Statistically-Aware End-to-end Optimizer for Machine Learning Inference

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-12T14:27:02.224872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T14:27:01.718943Z digest=sha256:06f0d9c511a0404fdd690cf758bd5360dbdcae06b429450d2e30767ed2e9eee5

Observation e6bee921-81a2-4d50-9fd9-b0c18ccd2fab · outbound

This paper cites CascadeBERT: Accelerating Inference of Pre-trained Language Models via Calibrated Complete Models Cascade.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling CascadeBERT: Accelerating Inference of Pre-trained Language Models via Calibrated Complete Models Cascade

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T14:27:01.724169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:27:01.724169Z digest=sha256:ffa0677ae0198cfa5e7cdf132a1ce7f051caf37c470ccb45bc53d6aaae3ac5a9

Observation ebc4d86d-2161-4947-a409-a2d6f5d06e81 · outbound

This paper cites SDXL-Lightning: Progressive Adversarial Diffusion Distillation.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling SDXL-Lightning: Progressive Adversarial Diffusion Distillation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T14:27:01.729111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:27:01.729111Z digest=sha256:ef9b86160f6855f440cb5db962998402d695727dee16c626d2a7eba6566b2a52

Observation 071f4e84-85a7-4aa1-9acc-20a0a828fe2a · outbound

This paper cites an unresolved cited work.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T14:27:01.734396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:27:01.734396Z digest=sha256:cacdaf080de930674a62c03cd3d5c886fca63ee0bcb627d693c226113a4b8490

Observation 1a7ed603-b3e1-4db8-8b5e-54a276106107 · outbound

This paper cites Midjourney.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling Midjourney

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:27:02.595379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T14:27:01.739242Z digest=sha256:22bafcd86a26afb4733236611a3d462bf887756fd036ff9382fbc545fe2c8344

Observation 61a01d18-ab89-45a8-b837-01d7674aabc9 · outbound

This paper cites Tensorflow-serving: Flexible, high-performance ml serving.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling Tensorflow-serving: Flexible, high-performance ml serving

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:27:02.576033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T14:27:01.743995Z digest=sha256:debb24ee6da5678f32bdb2774567be395b348dbbcb376da4fbf92c753513df90

Observation 408dae85-b0fd-44d7-ae8f-c8947e41f859 · outbound

This paper cites PyTorch: An Imperative Style, High-Performance Deep Learning Library.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling PyTorch: An Imperative Style, High-Performance Deep Learning Library

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-12T14:27:01.749149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:27:01.749149Z digest=sha256:5dc98d41ebf61ec9226eee646b938cf70fdf8936b7d422d9e7aaa8b99022363c

Observation 3743c3a7-a30f-4d40-a3b8-f1801f5fa9f3 · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-12T14:27:01.754594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:27:01.754594Z digest=sha256:c9eec0389a8e29e4e0321b7fcace4fc36e7702640a28819f039ae4d107dbb4e1

Observation d3e8aab3-6934-4a2f-a9cd-e867cf47761f · outbound

This paper cites Torchserve.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling Torchserve

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:27:02.559073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T14:27:01.760153Z digest=sha256:b7f5e3a2e4a0b9cf9a9e8f06a695b63994ede33fe64593ed5cb7d64a970dc4d4

Observation 1f1eb2e6-3714-460f-9e35-e3e656c2b8d2 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling High-resolution image synthesis with latent diffusion models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T14:27:01.764989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:27:01.764989Z digest=sha256:d01711fc2338dcbcb66939a70b5ce1c84b5c8b03c9877b9dba4e3bd694c93215

Observation e847266e-ed6b-4b24-96ac-0a05db6a3a9e · outbound

This paper cites J., and Kozyrakis, C.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling J., and Kozyrakis, C

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:27:02.530876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T14:27:01.769792Z digest=sha256:6e538bd127a59e3dd01fbc9e57793956c5dcaeaa464b3c0bd5e9cf3160a02b8c

Observation 9387b0e1-0e28-4c43-9032-838476b8398c · outbound

This paper cites J., and Kozyrakis, C.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling J., and Kozyrakis, C

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T14:27:01.774915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:27:01.774915Z digest=sha256:214bb90239de78fb636d0f40f928dfdf001be750f5da25a084b7afc53bdf274e

Observation 5b040390-ec84-4503-bc81-f202af14a8af · outbound

This paper cites Adversarial Diffusion Distillation.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling Adversarial Diffusion Distillation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-12T14:27:01.779693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:27:01.779693Z digest=sha256:adee432eaea517dd48b7d012d11da2eee9447989ed3a0dbe3ce5562494dd6f5a

Observation c59f4356-5ad6-4298-b962-a5860d120513 · outbound

This paper cites Serverless in the wild: Characterizing and optimizing the serverless workload at a large cloud provider.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling Serverless in the wild: Characterizing and optimizing the serverless workload at a large cloud provider

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:27:02.512873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T14:27:01.784847Z digest=sha256:2415a0a7c76da9ddebc38a6c8b590e9d301efa80448d20c711bbb15171140235

Observation d815e354-5d72-437f-8a28-1efbf8d16643 · outbound

This paper cites Fast video classification via adaptive cascading of deep models.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling Fast video classification via adaptive cascading of deep models

Reference 39

Resolution
verified exact
doi, observed 2026-08-12T14:27:01.899001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T14:27:01.789531Z digest=sha256:f415e1d3da065ae2759b9cba222208a39b637843840ca27f92ce6a9e469ee376

Observation 19ee7a8c-2775-49ec-a92d-3270d01c7919 · outbound

This paper cites Nexus: A gpu cluster engine for accelerating dnn-based video analysis.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling Nexus: A gpu cluster engine for accelerating dnn-based video analysis

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:27:02.495517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T14:27:01.794294Z digest=sha256:cbb0bcb792423b02683add813542f48d1e2ba53e482dd50fa126ec3545bd6c0d

Observation ff5a7a63-9d26-40fb-b520-776126789cc3 · outbound

This paper cites F., Thompson, J.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling F., Thompson, J

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:27:02.476846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T14:27:01.799021Z digest=sha256:8fb280e267dec79b0f32ec2b680ef14acadd2961534ea9be397c240834ed12d2

Observation c13e7ee5-d1f4-47af-ad5f-386596fee70a · outbound

This paper cites SDXS: Real-Time One-Step Latent Diffusion Models with Image Conditions.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling SDXS: Real-Time One-Step Latent Diffusion Models with Image Conditions

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T14:27:01.804062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:27:01.804062Z digest=sha256:e56f11e4353a326353b980fc3438cf9422c84fe0feb613de609293f223d28165

Observation 13a5051f-25bc-4043-a85b-7fef7d6c38b6 · outbound

This paper cites Tensorflow serving.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling Tensorflow serving

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:27:02.458559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T14:27:01.809350Z digest=sha256:8c1ed85fe5023b2054b490982aa1c80da324d5fbef998498d1c9e41128f09381

Observation 3c9d8c47-a515-4899-88e0-8d96b6993664 · outbound

This paper cites and Jones, M.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling and Jones, M

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:27:02.439690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T14:27:01.814699Z digest=sha256:44cb09983ec822b70fe92dfbf254e3e7a8860d1a99489a0764d4b84052b85b53

Observation 59fbc0a5-2a44-49ed-819a-3e18267f3e40 · outbound

This paper cites Diffusers: State-of-the-art diffusion models.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling Diffusers: State-of-the-art diffusion models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T14:27:01.819689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:27:01.819689Z digest=sha256:8989c8e7c5eb58c79ec0d725582f6480e2961c72cf362b2dde8f034ed2d69b1d

Observation adb7d2c2-4ffe-401a-b5ef-2a1263f5640a · outbound

This paper cites Rafiki: Machine Learning as an Analytics Service System.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling Rafiki: Machine Learning as an Analytics Service System

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-12T14:27:01.824439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:27:01.824439Z digest=sha256:22e9c9b938a77ebe45c9c6e8e90d77ca1cca630004beb96489547753246c2298

Observation 89663e8d-cda5-4c9d-9623-e969ec33819b · outbound

This paper cites Tabi: An efficient multi-level inference system for large language models.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling Tabi: An efficient multi-level inference system for large language models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-12T14:27:01.830067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:27:01.830067Z digest=sha256:e391a6c29f3cf8a52b5f99be1ffaf87934412a55b18d5c52bb5bc920690e436e

Observation 95183afe-e1c2-4a92-8924-015001f28317 · outbound

This paper cites DiffusionDB: A Large-scale Prompt Gallery Dataset for Text-to-Image Generative Models.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling DiffusionDB: A Large-scale Prompt Gallery Dataset for Text-to-Image Generative Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-12T14:27:01.835347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:27:01.835347Z digest=sha256:cd5bb4db30d640da2e5b8e10f5e379a8bc36a8b08af826b3d23dfb8d9dedfa77

Observation ef582d90-0749-4755-ae06-615f15cc131c · outbound

This paper cites Bartscore: Evaluating generated text as text generation.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling Bartscore: Evaluating generated text as text generation

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:27:02.410087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T14:27:01.841104Z digest=sha256:1a5c26696739f4d526ab247146842895f7115fa3f1815124bf0205ae900332e0

Observation 009eb8c5-4744-4047-8ab1-5ddf6b08bfd3 · outbound

This paper cites an unresolved cited work.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:27:02.392540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T14:27:01.846613Z digest=sha256:72aac226b3b2b979b5289daa29e94016d33987ad8d88542a66f3d2c83d989f8a

Observation 9e38a922-a2a5-46c9-bd5c-f1ce786729f0 · outbound

This paper cites \ Model-Switching \ : Dealing with fluctuating workloads in \ Machine-Learning-as-a-Service \ systems.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling \ Model-Switching \ : Dealing with fluctuating workloads in \ Machine-Learning-as-a-Service \ systems

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:27:02.374533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T14:27:01.851437Z digest=sha256:e3678c5900d58817f9a7d43ae711573a1249892ea8e74cac63a514b55cc72d2f

Observation 7de3969a-5eb3-4168-9beb-f383f9ca6902 · outbound

This paper cites write newline.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling write newline

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-12T14:27:01.856685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:27:01.856685Z digest=sha256:2d2c79dd4e68575acb9273ffb89ab2340182ea3c4d2ecb4799a0e8264df1567c

Pith citing papers

Observation f91bb7ea-c123-4962-a7a1-6a9b52d1bebf · inbound

DisagFusion: Asynchronous Pipeline Parallelism and Elastic Scheduling for Disaggregated Diffusion Serving cites this paper.

DisagFusion: Asynchronous Pipeline Parallelism and Elastic Scheduling for Disaggregated Diffusion Serving DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-06-29T20:53:57.613373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-29T20:46:51.896596Z digest=sha256:603d0814273b90d35c69f5d884f405c7278ca630dd6d152e18751a8faa29a4b5

Observation 245aba40-2b5f-4cbe-849f-fc9060b86073 · inbound

GF-DiT: Scheduling Parallelism for Diffusion Transformer Serving cites this paper.

GF-DiT: Scheduling Parallelism for Diffusion Transformer Serving DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-03T16:28:39.323307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-27T05:35:01.339479Z digest=sha256:84fa95134ca1637c1c53744403c180265f51dd45d438f449eb774af64dbae9c4

Observation 40dac2ac-8156-42d6-8f8f-9a1b85da25d0 · inbound

GF-DiT: Scheduling Parallelism for Diffusion Transformer Serving cites this paper.

GF-DiT: Scheduling Parallelism for Diffusion Transformer Serving DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-03T23:59:06.300197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-07-03T23:50:39.241880Z digest=sha256:c9798d0526038c84cda5117223a7cd90f47e228db58ede06823e0a2438c0b523

Observation 9f45b247-6672-4d76-81e7-15703380acc9 · inbound

Xema: Efficient Diffusion Serving through Fine-Grained Memory Management and Auto-Configuration cites this paper.

Xema: Efficient Diffusion Serving through Fine-Grained Memory Management and Auto-Configuration DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-14T06:49:04.349040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:49:04.349040Z digest=sha256:20374bd63b9fa254a1013edcf380fa1205eb53fc7a5a815c07692286aa2f371e