Pith. sign in

Paper Citation Record · LEDGER

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud

As of 19 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 2 inbound Pith citation observations for arXiv:2411.15664.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.15664 v1

Coverage vector

measured 37 of 37 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T14:07:52.471138Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:07:59.365233Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-10T20:34:35.675801Z

Reference resolution

37 of 37 outbound references displayed

  • verified exact1
  • verified fuzzy16
  • unresolved17
  • parse uncertain1
  • malformed identifier2
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 364f3b2d-7051-41f6-93c8-dedeace87310 · outbound

This paper cites slideshare.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud slideshare

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.205858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T14:07:52.342326Z digest=sha256:aa2293456b989b34aaa53dc037de89862c2c746199cc5e5bbb7a9624a159f84b

Observation 5df39438-f0dd-40b9-b554-337378de84b6 · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:07:53.196280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T14:07:52.346076Z digest=sha256:b43d65d37e47d83472644deb10df226927422510af53f9263fefa250963b2d61

Observation 1023118d-9954-4aca-9d91-080b1315d27e · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 3

Resolution
parse uncertain
raw_fallback, observed 2026-08-12T14:07:53.185421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T14:07:52.349499Z digest=sha256:bab6c27372c9547e2c7a116d14a2f8ce4e9599a89d54f70e7473c459efdce1cb

Observation feb893f8-64be-4fa3-b183-850deaffcb9b · outbound

This paper cites com/ jeremydaly/ lambda-warmer.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud com/ jeremydaly/ lambda-warmer

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.175555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T14:07:52.353297Z digest=sha256:0e8c5a997c566fe56c08f2f03fe4ebe60fc89039a7d7b18b94af3c8bfb66d8a5

Observation d3177d5b-5d1d-4900-a5f4-aa86aa6436e9 · outbound

This paper cites com/ google-cloud/ 3-solutions-to-mitigate-the-cold-starts-on-cloud-run-8c60f0ae7894.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud com/ google-cloud/ 3-solutions-to-mitigate-the-cold-starts-on-cloud-run-8c60f0ae7894

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.164567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T14:07:52.356999Z digest=sha256:334073f4f4d3e610ed7b48eedfc1517b680dbb6d9dbac0a0b5461573d8b17d09

Observation 2f795c3d-195c-42f5-8547-405954d828c8 · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:07:53.154936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T14:07:52.360480Z digest=sha256:972fcb63feaac3eeba027a25c4ebd6fc0d48fcc8769f9179635e069ee1ada490

Observation 3d03005d-8e8d-46a6-8ae3-291f4f5dbaba · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:07:53.145956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T14:07:52.363413Z digest=sha256:aa7efb531ee4244bec30c4a9568a0ffe63f31063aca8607fef1c8c169b775f45

Observation 202f8e13-e3c3-4812-bcdc-c0c01875a679 · outbound

This paper cites microsoft.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud microsoft

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.137112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T14:07:52.366346Z digest=sha256:5ccb256b0dd9bda2c29384e5b6c624f0e424e2358f035c098b7705bec1a0c73f

Observation de4aaac9-7e90-4eba-a5dc-089d7fca0403 · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:07:53.127094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T14:07:52.368997Z digest=sha256:624b6a0571881d6b8be83ef5d0afdfd2f84586715ac9edb2f437aa242f029b51

Observation 2fcba9bd-9c35-43a4-a845-12c0b7fa35f5 · outbound

This paper cites https: // www.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud https: // www

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.117528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T14:07:52.371989Z digest=sha256:64e8cf39763b2906a1837af07c1e258dc03dd926ee2ee654d6ff64def583613a

Observation 82eec83a-4d1c-40dd-b542-37ad77f34f47 · outbound

This paper cites SedAI ( https: // www.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud SedAI ( https: // www

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.107679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T14:07:52.375096Z digest=sha256:7dc022c0b961b86cfb73d5f370e87b40efc64051adf51327f004dfed36d4846c

Observation a4c88b13-02c7-44e8-a963-1a4f7d05878c · outbound

This paper cites Snowflake ( https: // www.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Snowflake ( https: // www

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.097763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T14:07:52.379176Z digest=sha256:fe6e1ebc6c8911e9535126cb4931410151b491648f790c13fad9f22e451788df

Observation 4324c903-1289-4ab5-a36b-8b9f2f5b0cde · outbound

This paper cites ( https: // en.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud ( https: // en

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.088216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T14:07:52.382791Z digest=sha256:ed49bcc172ebc2c3442b4b5ff1cede5aa65f3b90102f5cf50e45286bffa1acfa

Observation 45d9e220-c714-4936-8a6c-48e75dbbadc6 · outbound

This paper cites Fast Distributed Inference Serving for Large Language Models.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Fast Distributed Inference Serving for Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T14:07:52.386091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:07:52.386091Z digest=sha256:e3a629774e69c9532c3593c61496d40ae850b50e3be831cebc36c65358fe28ba

Observation de6462a3-8c62-4370-b450-504444aa0743 · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:07:53.077833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T14:07:52.390158Z digest=sha256:58854c8094078cea20c117f3f2152800fe3bd311e52fce816fb8ca10439fc5a0

Observation 68f846e0-6afe-439f-9d34-8d3fd4386f23 · outbound

This paper cites Catalyzer: Sub-millisecond startup for serverless computing with initialization-less booting.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Catalyzer: Sub-millisecond startup for serverless computing with initialization-less booting

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T14:07:52.398677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:07:52.398677Z digest=sha256:f43d39e0b6aa1a39321b2dc00ea9cb592335531c88cadc833031c36a1ab351ae

Observation 4e46f8db-aba4-4517-846a-6e2e7015b7ef · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:07:53.059267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T14:07:52.402455Z digest=sha256:0701db634fdaa763a92697a7814c11dcf3967b296013d3c01215a050bcb6d3ae

Observation 12f96f92-095c-409e-95ce-7c044a2ed2ec · outbound

This paper cites Centralized core-granular scheduling for serverless functions.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Centralized core-granular scheduling for serverless functions

Reference 18

Resolution
malformed identifier
no resolver link, observed 2026-08-12T14:07:52.406118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:07:52.406118Z digest=sha256:c8d2c842b2e10b40cf214783b6710496aa57ade05acb9dc4c7730e4e4493f213

Observation 0f092b08-7059-4c5c-ae16-0930f2ba1aff · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:07:53.069167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T14:07:52.394148Z digest=sha256:aeafb7317869cd83d1c124c250e2079012b166723cacfad3a98c03ab7fb3cded

Observation db06b321-02d0-434d-8d5e-5a639575bc15 · outbound

This paper cites Faaslight: General application-level cold-start latency optimization for function-as-a-service in serverless comput- ing.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Faaslight: General application-level cold-start latency optimization for function-as-a-service in serverless comput- ing

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.049715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T14:07:52.409303Z digest=sha256:fee3edea0019c946e0f56a2dcf865ea06f33ed7e52cf58a678281513ecd32aa8

Observation 5a4c2f0e-8182-4dd1-8406-34b232294aaf · outbound

This paper cites Rapid task provision- ing with Serverless-Optimized containers.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Rapid task provision- ing with Serverless-Optimized containers

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.039443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T14:07:52.413546Z digest=sha256:204b0531a229466b9f16b61209f9737d2755772a9fe21541900a8d178d87b6a0

Observation b4ba1b40-1a9f-4c4a-b1e0-f48fe07ae907 · outbound

This paper cites Ghobaei-Arani.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Ghobaei-Arani

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.029020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T14:07:52.416575Z digest=sha256:e8ac35a0603763e6e91414bada8daecba76d6ddfe0a1bede75841ec0eb815ce2

Observation 79557d58-7223-4ac3-8d6a-aa67466ddae0 · outbound

This paper cites Cold start latency in serverless com- puting: A systematic review, taxonomy, and future directions.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Cold start latency in serverless com- puting: A systematic review, taxonomy, and future directions

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.019552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T14:07:52.419557Z digest=sha256:91410cd738de47a313e55bce2a36e633b4b743011bc655588d265b8131c3c24f

Observation b542313f-fb5d-4c01-84dd-f9165069a116 · outbound

This paper cites What is serverless computing? IBM ( https: // www.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud What is serverless computing? IBM ( https: // www

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:53.009677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T14:07:52.422766Z digest=sha256:49cfbfa170fa78ba7495235e9dcf322d938466495430646d50598cb5c3dea0b8

Observation 33d95291-1641-4918-a0e1-733e3aaf5e5f · outbound

This paper cites Mitigating Cold Starts in Serverless Platforms: A Pool-Based Approach.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Mitigating Cold Starts in Serverless Platforms: A Pool-Based Approach

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T14:07:52.425584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:07:52.425584Z digest=sha256:4071b1be19600385b28f0b295c566b4ed3e146fc9085230ae1ea1a408351a11e

Observation 11398b45-2b56-40d3-9991-a9412e7f1d4f · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:07:52.998521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T14:07:52.429465Z digest=sha256:38becb816bd45168000f8aaf1a96d3a738db7de8326efa4dab0a4a5365df0369

Observation 0e5868b2-1ac2-456a-bfa0-1064a03ebf44 · outbound

This paper cites Persson and W.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Persson and W

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:52.987551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T14:07:52.433660Z digest=sha256:886233761018d8903bb440bc4beb3b8297c7b6bb3c3f4040c9bb02454abd8776

Observation b4ad41ec-044c-48f0-96e5-3288dc7dee95 · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T14:07:52.437985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:07:52.437985Z digest=sha256:5e27e13622fc26e41d3e014fd1b8bcf13b4cb809b51728a4d9ad8f7a866c0fae

Observation 14f82162-c99a-4d65-b6f0-640b1c60b349 · outbound

This paper cites Pietzuch https: //api.semanticscholar.org/CorpusID: 11 51997872.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Pietzuch https: //api.semanticscholar.org/CorpusID: 11 51997872

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:52.975524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T14:07:52.441604Z digest=sha256:952be3e028ce59f00c8990dd851c890f0747bc1cd9a293539b28c8147592fd82

Observation f63de849-60e9-4273-bc54-25fbc1fda6b5 · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:07:52.964010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T14:07:52.445502Z digest=sha256:a2c330d7fac30a01e7d8db638fd2ccc9839507dd2a9ef1ae0d3210f06d17b05a

Observation eb9f7d2c-ddeb-4acc-95b0-dd7eff3418a4 · outbound

This paper cites Rise of the Planet of Serverless Computing: A Systematic Review.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Rise of the Planet of Serverless Computing: A Systematic Review

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-08-12T14:07:52.674384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T14:07:52.448829Z digest=sha256:715eccd2f1ada0bcf5cad1fb078edfd89f34dc2f77ed503e3c12c0d09b1da3d3

Observation 727ca945-9b4f-450d-a8cd-e42ab8a30ace · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-12T14:07:52.952132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T14:07:52.452351Z digest=sha256:d090d867c82a9c706e39f5f24ec1173a9298facec1a9cadafaa42645038d1320

Observation 3aeb15bc-302b-4493-a87b-30bbfd068181 · outbound

This paper cites FlexGen: High-Throughput Generative Inference of Large Language Models with a Single GPU.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud FlexGen: High-Throughput Generative Inference of Large Language Models with a Single GPU

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-12T14:07:52.463915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:07:52.463915Z digest=sha256:1d7c39d957ffd6a72561f6d64a8bad13d76538a01b6c44d92837b2084545a672

Observation 0daaa7ee-668f-4558-bc16-d7a6f5ed40da · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T14:07:52.460464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:07:52.460464Z digest=sha256:599a4533b46173284fb0ec85ebe32decf8aeb2976536c58fcbb97f63cc15c601

Observation ade5bb26-b9e9-4a20-97a5-b1e61e5ec008 · outbound

This paper cites Taming serverless cold start of cloud model inference with edge comput- ing.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Taming serverless cold start of cloud model inference with edge comput- ing

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T14:07:52.942131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-12T14:07:52.471138Z digest=sha256:75007d1860920218791119a1f01ee224ac169af389ce3467a4e6f13a80b14a62

Observation 11983ed9-bda9-45a0-bd46-fcb59feb7f7c · outbound

This paper cites an unresolved cited work.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud Unresolved cited work

Reference 36

Resolution
malformed identifier
no resolver link, observed 2026-08-12T14:07:52.467684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:07:52.467684Z digest=sha256:6bd9c18759f2473635cdd543d7f0cf84b30a17b32186fbefbeff02f27c1a5676

Observation ae7ea405-8a5d-4248-bc6f-6f03e4a0382b · outbound

This paper cites ServerlessLLM: Low-Latency Serverless Inference for Large Language Models.

Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud ServerlessLLM: Low-Latency Serverless Inference for Large Language Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-12T14:07:52.456185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:07:52.456185Z digest=sha256:7696f4a0f4b9e3e41882b1721f9ed6a45b98ef7f82bb43749d773e0382c4285b

Pith citing papers

Observation 36daebe2-8f27-449a-814d-24b40e8aae63 · inbound

Addressing the sustainable AI trilemma: a case study on LLM agents and RAG cites this paper.

Addressing the sustainable AI trilemma: a case study on LLM agents and RAG Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud

Reference 2024

Resolution
verified exact
local_arxiv, observed 2026-08-10T20:34:35.680096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T20:34:34.676288Z digest=sha256:4b3e9c6e5883d979452cba41ea49005e8c81d4057ee3523b81653553a83f5aff

Observation d8e55695-8058-4494-be69-13e34c0f8129 · inbound

A Survey of Foundation Model-Powered Recommender Systems: From Feature-Based, Generative to Agentic Paradigms cites this paper.

A Survey of Foundation Model-Powered Recommender Systems: From Feature-Based, Generative to Agentic Paradigms Enabling Efficient Serverless Inference Serving for LLM (Large Language Model) in the Cloud

Reference 246

Resolution
unresolved
no resolver link, observed 2026-08-16T11:07:59.365233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:07:59.365233Z digest=sha256:0f36438bc3edd00f47e53631112f0f9b483a7b8dd064af2aee59e9f685994e5c