Pith. sign in

Paper Citation Record · LEDGER

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud

As of 12 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 1 inbound Pith citation observation for arXiv:2411.15664.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.15664 v1

Coverage vector

measured 37 of 37 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T14:07:52.471138Z

measured 38 of 38 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T20:34:34.676288Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-10T20:34:35.675801Z

Reference resolution

37 of 37 outbound references displayed

  • verified exact1
  • verified fuzzy16
  • unresolved17
  • parse uncertain1
  • malformed identifier2
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 364f3b2d-7051-41f6-93c8-dedeace87310 · outbound

This paper cites slideshare.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud slideshare

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.205858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:07:52.342326Z digest=sha256:500dcdaa0bb7ba006fe87a3b84ad231fa823773b79b288953238876979c5687f

Observation 5df39438-f0dd-40b9-b554-337378de84b6 · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:07:53.196280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:07:52.346076Z digest=sha256:08fea12cb4dc37cd0106b306bd054cfa371f0ab4dbc2dd9551606c0083ccfca3

Observation 1023118d-9954-4aca-9d91-080b1315d27e · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 3

Resolution
parse uncertain
raw_fallback, observed 2026-08-12T14:07:53.185421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:07:52.349499Z digest=sha256:50d9c85a096ff6788f2b067504c0eb7e3435088ab2cbbddb092f221b60d6b717

Observation feb893f8-64be-4fa3-b183-850deaffcb9b · outbound

This paper cites com/ jeremydaly/ lambda-warmer.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud com/ jeremydaly/ lambda-warmer

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.175555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:07:52.353297Z digest=sha256:96bc1ee95f96c2d699fd9cdc6b3f0fff490be00c36b6db0fcb416066c0767fb9

Observation d3177d5b-5d1d-4900-a5f4-aa86aa6436e9 · outbound

This paper cites com/ google-cloud/ 3-solutions-to-mitigate-the-cold-starts-on-cloud-run-8c60f0ae7894.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud com/ google-cloud/ 3-solutions-to-mitigate-the-cold-starts-on-cloud-run-8c60f0ae7894

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.164567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:07:52.356999Z digest=sha256:239df4488c906308c3680fda9cfc5864c362f1699cb09a5088929fa8c6ff2add

Observation 2f795c3d-195c-42f5-8547-405954d828c8 · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:07:53.154936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:07:52.360480Z digest=sha256:7b028bf37b2fa623d0443d579bcfd20788337a92a06e06a017a69731b1b04694

Observation 3d03005d-8e8d-46a6-8ae3-291f4f5dbaba · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:07:53.145956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:07:52.363413Z digest=sha256:70f76656ec4ac2351156987ff90258890cac831ff238044fd0f484ec23f909d3

Observation 202f8e13-e3c3-4812-bcdc-c0c01875a679 · outbound

This paper cites microsoft.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud microsoft

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.137112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:07:52.366346Z digest=sha256:478b681742057b685e97ea0849bda4797f78813607215b00390137e763f88848

Observation de4aaac9-7e90-4eba-a5dc-089d7fca0403 · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:07:53.127094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:07:52.368997Z digest=sha256:8e5ddb826a73c1ded17d7daec5e671a557161d42c4fa051b7782491423311b0c

Observation 2fcba9bd-9c35-43a4-a845-12c0b7fa35f5 · outbound

This paper cites https: // www.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud https: // www

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.117528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:07:52.371989Z digest=sha256:24045165b5327d8ed2a04f669213d13a78de36536be2381d8e030b0015b44857

Observation 82eec83a-4d1c-40dd-b542-37ad77f34f47 · outbound

This paper cites SedAI ( https: // www.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud SedAI ( https: // www

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.107679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:07:52.375096Z digest=sha256:0b2e4442148fdf7daa08af3b0292914b9f99ab6686cbb1e91d8d391caf86b6f3

Observation a4c88b13-02c7-44e8-a963-1a4f7d05878c · outbound

This paper cites Snowflake ( https: // www.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Snowflake ( https: // www

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.097763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:07:52.379176Z digest=sha256:7757ea8eb3edefc6048feacdf0caf61c8441b3f307b7622883f6c9644a0992c1

Observation 4324c903-1289-4ab5-a36b-8b9f2f5b0cde · outbound

This paper cites ( https: // en.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud ( https: // en

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.088216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:07:52.382791Z digest=sha256:7d3a8880821e9df83ffc9b7e10cbc8fccd0c344fd73682bdf647bc7698e7825e

Observation 45d9e220-c714-4936-8a6c-48e75dbbadc6 · outbound

This paper cites Fast Distributed Inference Serving for Large Language Models.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Fast Distributed Inference Serving for Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T14:07:52.386091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:07:52.386091Z digest=sha256:bca6477286f08cbb2914eaf86ee992c10737c2c43aa8bfef148f242253e318f7

Observation de6462a3-8c62-4370-b450-504444aa0743 · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:07:53.077833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:07:52.390158Z digest=sha256:0672a9e3d7db65d2c7b3f8a858cae50e9f25aa2fe885b887fe565f97df00104f

Observation 68f846e0-6afe-439f-9d34-8d3fd4386f23 · outbound

This paper cites Catalyzer: Sub-millisecond startup for serverless computing with initialization-less booting.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Catalyzer: Sub-millisecond startup for serverless computing with initialization-less booting

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T14:07:52.398677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:07:52.398677Z digest=sha256:74074194cb5ea3de8311a7416aee9c853e57751da68d3792b8fbe648c16a8657

Observation 4e46f8db-aba4-4517-846a-6e2e7015b7ef · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:07:53.059267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:07:52.402455Z digest=sha256:ed2e5130847b8c9c1b7aa2cc435a244747577894347a8b7eeede0d964deffc77

Observation 12f96f92-095c-409e-95ce-7c044a2ed2ec · outbound

This paper cites Centralized core-granular scheduling for serverless functions.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Centralized core-granular scheduling for serverless functions

Reference 18

Resolution
malformed identifier
no resolver link, observed 2026-08-12T14:07:52.406118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:07:52.406118Z digest=sha256:663290bd26550706218a0b8fd3cc660d0a33e3084ca7432fbb5b2744880df1c4

Observation 0f092b08-7059-4c5c-ae16-0930f2ba1aff · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:07:53.069167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:07:52.394148Z digest=sha256:bd632f7c807c9c9e67409ffb56bd4faac7aa31f6b51fb9ff6c203bf1070f34c7

Observation db06b321-02d0-434d-8d5e-5a639575bc15 · outbound

This paper cites Faaslight: General application-level cold-start latency optimization for function-as-a-service in serverless comput- ing.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Faaslight: General application-level cold-start latency optimization for function-as-a-service in serverless comput- ing

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.049715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:07:52.409303Z digest=sha256:02916cabda3c4acb05761ab79352b6fd882221a23870b1167d00f55e344d0ef6

Observation 5a4c2f0e-8182-4dd1-8406-34b232294aaf · outbound

This paper cites Rapid task provision- ing with Serverless-Optimized containers.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Rapid task provision- ing with Serverless-Optimized containers

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.039443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:07:52.413546Z digest=sha256:5174dc5a31d3bca30f7caf68b6fa43ad0b2b761242fa6eb106ca5c536701836d

Observation b4ba1b40-1a9f-4c4a-b1e0-f48fe07ae907 · outbound

This paper cites Ghobaei-Arani.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Ghobaei-Arani

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.029020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:07:52.416575Z digest=sha256:56c68b22d142e194f1f7b8d1b1989a0eff60f752fa9f174e02c4b526ef4c5e69

Observation 79557d58-7223-4ac3-8d6a-aa67466ddae0 · outbound

This paper cites Cold start latency in serverless com- puting: A systematic review, taxonomy, and future directions.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Cold start latency in serverless com- puting: A systematic review, taxonomy, and future directions

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.019552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:07:52.419557Z digest=sha256:9970c78ba10b28b9f048f18bef86f7d030bb5722e60b5c721859f93d44dddda4

Observation b542313f-fb5d-4c01-84dd-f9165069a116 · outbound

This paper cites What is serverless computing? IBM ( https: // www.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud What is serverless computing? IBM ( https: // www

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.009677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:07:52.422766Z digest=sha256:a48e52f854a6fd18be110564c807ce28b255c6c5cecb732b2eb89d4ad82b25cc

Observation 33d95291-1641-4918-a0e1-733e3aaf5e5f · outbound

This paper cites Mitigating Cold Starts in Serverless Platforms: A Pool-Based Approach.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Mitigating Cold Starts in Serverless Platforms: A Pool-Based Approach

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T14:07:52.425584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:07:52.425584Z digest=sha256:f5f60088df22ba9e558132b615003d69cdb0ffd954a5f4163b4c1f9e54fc30d4

Observation 11398b45-2b56-40d3-9991-a9412e7f1d4f · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:07:52.998521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:07:52.429465Z digest=sha256:e664d36f9fcf32d6907dbf8c40ed7bd9def751cb8b1b8cdde12258c5d50f660d

Observation 0e5868b2-1ac2-456a-bfa0-1064a03ebf44 · outbound

This paper cites Persson and W.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Persson and W

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:52.987551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:07:52.433660Z digest=sha256:6d287c54c904d7411cb680151bfec009b88a539eb61e6bfeb86cd727f98c8791

Observation b4ad41ec-044c-48f0-96e5-3288dc7dee95 · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T14:07:52.437985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:07:52.437985Z digest=sha256:20dfaa3fe4365c426a2ef8824e61ffe80489c27f78057fd29e169c82ddc5dfe0

Observation 14f82162-c99a-4d65-b6f0-640b1c60b349 · outbound

This paper cites Pietzuch https: //api.semanticscholar.org/CorpusID: 11 51997872.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Pietzuch https: //api.semanticscholar.org/CorpusID: 11 51997872

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:52.975524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:07:52.441604Z digest=sha256:6404800298df8e0be6546e3f382db0a4f8780e79ef848a63f695fc53dae0853b

Observation f63de849-60e9-4273-bc54-25fbc1fda6b5 · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:07:52.964010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:07:52.445502Z digest=sha256:6c3ec8ec635d0ced49f03d37ed3e01d9595811e534a83306ed3aae853ea90d82

Observation eb9f7d2c-ddeb-4acc-95b0-dd7eff3418a4 · outbound

This paper cites Rise of the Planet of Serverless Computing: A Systematic Review.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Rise of the Planet of Serverless Computing: A Systematic Review

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-08-12T14:07:52.674384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:07:52.448829Z digest=sha256:981c634c00e54822239f8fa266d3c2b725f1b5e774d7aacabf9a9b67f99744c5

Observation 727ca945-9b4f-450d-a8cd-e42ab8a30ace · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:07:52.952132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:07:52.452351Z digest=sha256:817eafcca6e3421f9572ae1769207554421f85927b8af5ff3d8322e3334f594e

Observation 3aeb15bc-302b-4493-a87b-30bbfd068181 · outbound

This paper cites FlexGen: High-Throughput Generative Inference of Large Language Models with a Single GPU.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud FlexGen: High-Throughput Generative Inference of Large Language Models with a Single GPU

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-12T14:07:52.463915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:07:52.463915Z digest=sha256:45cda0ad7c16371f9121988e0efe8e6b4c21955208f742ab57deb1fd646811c3

Observation 0daaa7ee-668f-4558-bc16-d7a6f5ed40da · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T14:07:52.460464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:07:52.460464Z digest=sha256:9d32f2f8092e7f28782fcdc8639d39f99bd337d110d20d4a5019ed3c3d857a8c

Observation ade5bb26-b9e9-4a20-97a5-b1e61e5ec008 · outbound

This paper cites Taming serverless cold start of cloud model inference with edge comput- ing.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Taming serverless cold start of cloud model inference with edge comput- ing

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:52.942131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T14:07:52.471138Z digest=sha256:0c59cee1b52c967bfb8f9e7a6101a8ab51a7fcd6d109b3f7d08d055ea98b220a

Observation 11983ed9-bda9-45a0-bd46-fcb59feb7f7c · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 36

Resolution
malformed identifier
no resolver link, observed 2026-08-12T14:07:52.467684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:07:52.467684Z digest=sha256:53628173926852ff2b3bcf1e86894ddcf0e268465ace78e4c96b46d609c4a739

Observation ae7ea405-8a5d-4248-bc6f-6f03e4a0382b · outbound

This paper cites ServerlessLLM: Low-Latency Serverless Inference for Large Language Models.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud ServerlessLLM: Low-Latency Serverless Inference for Large Language Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-12T14:07:52.456185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:07:52.456185Z digest=sha256:00ff1d066a401a646f3bcef34b88197e7f486394c1324a33f5c52ac345012d0c

Pith citing papers

Observation 36daebe2-8f27-449a-814d-24b40e8aae63 · inbound

Addressing the sustainable AI trilemma: a case study on LLM agents and RAG cites this paper.

Addressing the sustainable AI trilemma: a case study on LLM agents and RAG Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud

Reference 2024

Resolution
verified exact
local_arxiv, observed 2026-08-10T20:34:35.680096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T20:34:34.676288Z digest=sha256:9930c50644183cb3f0c035d2098625c1f84670f19ac24e879f1b58364a478480