Pith. sign in

Paper Citation Record · LEDGER

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud

As of 13 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 1 inbound Pith citation observation for arXiv:2411.15664.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.15664 v1

Coverage vector

measured 37 of 37 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T14:07:52.471138Z

measured 38 of 38 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T20:34:34.676288Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-10T20:34:35.675801Z

Reference resolution

37 of 37 outbound references displayed

  • verified exact1
  • verified fuzzy16
  • unresolved17
  • parse uncertain1
  • malformed identifier2
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 364f3b2d-7051-41f6-93c8-dedeace87310 · outbound

This paper cites slideshare.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud slideshare

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.205858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:07:52.342326Z digest=sha256:d86b8c936f38a7ec08dfe700f48ec8a846dfacf71ed2df4f5af2c1490699cb9d

Observation 5df39438-f0dd-40b9-b554-337378de84b6 · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:07:53.196280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:07:52.346076Z digest=sha256:db7c635983f6af159f95de96dcc88a6fd312b1cfd0436bbcbdc87a4320ffaae9

Observation 1023118d-9954-4aca-9d91-080b1315d27e · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 3

Resolution
parse uncertain
raw_fallback, observed 2026-08-12T14:07:53.185421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:07:52.349499Z digest=sha256:8d2151bd72be05d4e812e2877f8aed4030eceef1eafa70d4dc64562ac0fdc56f

Observation feb893f8-64be-4fa3-b183-850deaffcb9b · outbound

This paper cites com/ jeremydaly/ lambda-warmer.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud com/ jeremydaly/ lambda-warmer

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.175555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:07:52.353297Z digest=sha256:8dcf83eff7d20a2596d7b2861c88831b599d27c2e0ee8ffcb3429fe987e5d64e

Observation d3177d5b-5d1d-4900-a5f4-aa86aa6436e9 · outbound

This paper cites com/ google-cloud/ 3-solutions-to-mitigate-the-cold-starts-on-cloud-run-8c60f0ae7894.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud com/ google-cloud/ 3-solutions-to-mitigate-the-cold-starts-on-cloud-run-8c60f0ae7894

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.164567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:07:52.356999Z digest=sha256:5012dc6c3ef2c52c0e40d3248beb211b5b7359e4c0b8687594d868eeaced77e8

Observation 2f795c3d-195c-42f5-8547-405954d828c8 · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:07:53.154936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:07:52.360480Z digest=sha256:2ce7f22dbe5905f0289cbe3a6f3b8a078703c3a70e77c547444d082800580de0

Observation 3d03005d-8e8d-46a6-8ae3-291f4f5dbaba · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:07:53.145956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:07:52.363413Z digest=sha256:f4c9bf9c1d5ec8bbe52c4dc3ecafcf8748cd462661d0449ba8434c6889d522e1

Observation 202f8e13-e3c3-4812-bcdc-c0c01875a679 · outbound

This paper cites microsoft.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud microsoft

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.137112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:07:52.366346Z digest=sha256:400df9a43b020951a3db6ac1209c58fc0989cda4cd87d6eb1dbb578b5e06b4eb

Observation de4aaac9-7e90-4eba-a5dc-089d7fca0403 · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:07:53.127094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:07:52.368997Z digest=sha256:1ba2ddce7361e175887ced1fc4a1a1c2d8e55d5865534cc336beb1d80b0fb3b2

Observation 2fcba9bd-9c35-43a4-a845-12c0b7fa35f5 · outbound

This paper cites https: // www.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud https: // www

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.117528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:07:52.371989Z digest=sha256:1f13689f003c42ca7768838e6f464ce2c3e3b06aaa768d0a49ba816438e0d1a3

Observation 82eec83a-4d1c-40dd-b542-37ad77f34f47 · outbound

This paper cites SedAI ( https: // www.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud SedAI ( https: // www

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.107679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:07:52.375096Z digest=sha256:b12ec73c8ccb75adfca4995f99c38c3817fbd2d8281684817d0c42d413ef707c

Observation a4c88b13-02c7-44e8-a963-1a4f7d05878c · outbound

This paper cites Snowflake ( https: // www.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Snowflake ( https: // www

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.097763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:07:52.379176Z digest=sha256:e2c400d44744907dc3245b2c0ff71e0f56c34845139a19937033d11c25c8c016

Observation 4324c903-1289-4ab5-a36b-8b9f2f5b0cde · outbound

This paper cites ( https: // en.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud ( https: // en

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.088216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:07:52.382791Z digest=sha256:961341ca9c81f8c861a69ed322955ef5069a923d9a038e79d4c25872cdf0f7a0

Observation 45d9e220-c714-4936-8a6c-48e75dbbadc6 · outbound

This paper cites Fast Distributed Inference Serving for Large Language Models.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Fast Distributed Inference Serving for Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T14:07:52.386091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:07:52.386091Z digest=sha256:e2964303279e4e4811a687368b9ccd93014df83736010d89fa62aa3bc4a7b271

Observation de6462a3-8c62-4370-b450-504444aa0743 · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:07:53.077833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:07:52.390158Z digest=sha256:531fdb9fe3f7ff580c03bd28560056ddd9b307272c659ec148886736034d4c99

Observation 68f846e0-6afe-439f-9d34-8d3fd4386f23 · outbound

This paper cites Catalyzer: Sub-millisecond startup for serverless computing with initialization-less booting.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Catalyzer: Sub-millisecond startup for serverless computing with initialization-less booting

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T14:07:52.398677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:07:52.398677Z digest=sha256:44f514a66585c88a728d9ffc15f98a1d01108b4eb947d58470e68313be451ad7

Observation 4e46f8db-aba4-4517-846a-6e2e7015b7ef · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:07:53.059267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:07:52.402455Z digest=sha256:58da860e71afd3ba63ebb7f9cbe122ebab3544f8b3f8bbd9ffec73e2fb2fe05a

Observation 12f96f92-095c-409e-95ce-7c044a2ed2ec · outbound

This paper cites Centralized core-granular scheduling for serverless functions.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Centralized core-granular scheduling for serverless functions

Reference 18

Resolution
malformed identifier
no resolver link, observed 2026-08-12T14:07:52.406118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:07:52.406118Z digest=sha256:cf9021d1035ea5923c9342b2bfa3271f49ea2f4644a75908765b83e36493eb6a

Observation 0f092b08-7059-4c5c-ae16-0930f2ba1aff · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:07:53.069167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:07:52.394148Z digest=sha256:10cabd5d5df55c70e031c2f72255879a131587f06480398cce13819535239dac

Observation db06b321-02d0-434d-8d5e-5a639575bc15 · outbound

This paper cites Faaslight: General application-level cold-start latency optimization for function-as-a-service in serverless comput- ing.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Faaslight: General application-level cold-start latency optimization for function-as-a-service in serverless comput- ing

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.049715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:07:52.409303Z digest=sha256:18b6556fbe8b67990d9091e976134cee23921f5130c4ba3c56ee91e34c19d74f

Observation 5a4c2f0e-8182-4dd1-8406-34b232294aaf · outbound

This paper cites Rapid task provision- ing with Serverless-Optimized containers.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Rapid task provision- ing with Serverless-Optimized containers

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.039443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:07:52.413546Z digest=sha256:657282b13567d220fe9a227259d7fd05d4017d2e254c157d4d59aac5d2b1504e

Observation b4ba1b40-1a9f-4c4a-b1e0-f48fe07ae907 · outbound

This paper cites Ghobaei-Arani.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Ghobaei-Arani

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.029020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:07:52.416575Z digest=sha256:2378ad0e9108bca4db31d293056e47a9a265911d5cd9f01e343efd0c041b77a6

Observation 79557d58-7223-4ac3-8d6a-aa67466ddae0 · outbound

This paper cites Cold start latency in serverless com- puting: A systematic review, taxonomy, and future directions.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Cold start latency in serverless com- puting: A systematic review, taxonomy, and future directions

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.019552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:07:52.419557Z digest=sha256:a901ab3543e807a650019aa14090bd9644a70e0966cd7d3f86cddb7e5793d681

Observation b542313f-fb5d-4c01-84dd-f9165069a116 · outbound

This paper cites What is serverless computing? IBM ( https: // www.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud What is serverless computing? IBM ( https: // www

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.009677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:07:52.422766Z digest=sha256:7e712e406ce8f5dcc658b4306c53212edfe287f92796f465b3c18ecc4dd0872d

Observation 33d95291-1641-4918-a0e1-733e3aaf5e5f · outbound

This paper cites Mitigating Cold Starts in Serverless Platforms: A Pool-Based Approach.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Mitigating Cold Starts in Serverless Platforms: A Pool-Based Approach

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T14:07:52.425584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:07:52.425584Z digest=sha256:4c672e7bbb9c639d41008c8b369d41586d6b0adc0cedcd372017f2ead2e888cd

Observation 11398b45-2b56-40d3-9991-a9412e7f1d4f · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:07:52.998521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:07:52.429465Z digest=sha256:aea262833ce728acfc0284cdc6b2d96cc17e1b7305ef894cde8de1fb97ac216e

Observation 0e5868b2-1ac2-456a-bfa0-1064a03ebf44 · outbound

This paper cites Persson and W.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Persson and W

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:52.987551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:07:52.433660Z digest=sha256:a9035c884c516445cd607c9453414a8763568df97e43d8355e646bf2324a407d

Observation b4ad41ec-044c-48f0-96e5-3288dc7dee95 · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T14:07:52.437985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:07:52.437985Z digest=sha256:6b9ff22d96cca6ad4bf3d47ca5d87d95e05509d3701e31f3d19af85a0988ce63

Observation 14f82162-c99a-4d65-b6f0-640b1c60b349 · outbound

This paper cites Pietzuch https: //api.semanticscholar.org/CorpusID: 11 51997872.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Pietzuch https: //api.semanticscholar.org/CorpusID: 11 51997872

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:52.975524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:07:52.441604Z digest=sha256:17ac5d650dbe3183a7b82bfde16e79981c756da8d98f5dfe3a08ab4941ccafbc

Observation f63de849-60e9-4273-bc54-25fbc1fda6b5 · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:07:52.964010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:07:52.445502Z digest=sha256:ee3cbc422d1be9e8bd0ec64852dce4946d09dc86808913022da3dabf033a0bc9

Observation eb9f7d2c-ddeb-4acc-95b0-dd7eff3418a4 · outbound

This paper cites Rise of the Planet of Serverless Computing: A Systematic Review.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Rise of the Planet of Serverless Computing: A Systematic Review

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-08-12T14:07:52.674384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:07:52.448829Z digest=sha256:a90666f71e2bb1e2ec2731fa62318e4afb232e0578a727b2f906dd697564df86

Observation 727ca945-9b4f-450d-a8cd-e42ab8a30ace · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:07:52.952132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:07:52.452351Z digest=sha256:6943824f87a1a4c0e21848b4518314ae158643d8aca9a043f8183f4e66ed9906

Observation 3aeb15bc-302b-4493-a87b-30bbfd068181 · outbound

This paper cites FlexGen: High-Throughput Generative Inference of Large Language Models with a Single GPU.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud FlexGen: High-Throughput Generative Inference of Large Language Models with a Single GPU

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-12T14:07:52.463915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:07:52.463915Z digest=sha256:53e387d5d021d879489a5c062b0dd319c4dc5af2aeeb07eb6c05cdbb26905f4d

Observation 0daaa7ee-668f-4558-bc16-d7a6f5ed40da · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T14:07:52.460464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:07:52.460464Z digest=sha256:2f1045de4292f1b94c784c977938f43405b2d666253f2c8177fc24f4aee8963f

Observation ade5bb26-b9e9-4a20-97a5-b1e61e5ec008 · outbound

This paper cites Taming serverless cold start of cloud model inference with edge comput- ing.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Taming serverless cold start of cloud model inference with edge comput- ing

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:52.942131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T14:07:52.471138Z digest=sha256:3e7e911b20ba417b957e2cbf954d183acfa2b828a8385bc2ec8747cff1cb2fdf

Observation 11983ed9-bda9-45a0-bd46-fcb59feb7f7c · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 36

Resolution
malformed identifier
no resolver link, observed 2026-08-12T14:07:52.467684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:07:52.467684Z digest=sha256:28e1c24c03082d86701d7d3fdbb10cc08f0a4380d3733929146f13e4a4ebdd2d

Observation ae7ea405-8a5d-4248-bc6f-6f03e4a0382b · outbound

This paper cites ServerlessLLM: Low-Latency Serverless Inference for Large Language Models.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud ServerlessLLM: Low-Latency Serverless Inference for Large Language Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-12T14:07:52.456185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:07:52.456185Z digest=sha256:ec005b49fd12667c26fe4ff06303c7c9b31edd1d0ae26f3f018fa7429aba34ee

Pith citing papers

Observation 36daebe2-8f27-449a-814d-24b40e8aae63 · inbound

Addressing the sustainable AI trilemma: a case study on LLM agents and RAG cites this paper.

Addressing the sustainable AI trilemma: a case study on LLM agents and RAG Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud

Reference 2024

Resolution
verified exact
local_arxiv, observed 2026-08-10T20:34:35.680096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T20:34:34.676288Z digest=sha256:5cb02ac21cf97b2991606b6f67dcd4b3275b798dd5e7e68fd3453f534a8cee68