Pith. sign in

Paper Citation Record · LEDGER

Towards Sustainable Large Language Model Serving

As of 15 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 0 inbound Pith citation observations for arXiv:2501.01990.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.01990 v1

Coverage vector

measured 40 of 40 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T22:59:41.651892Z

measured 40 of 40 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

40 of 40 outbound references displayed

  • verified exact2
  • verified fuzzy27
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation bf68ce9e-ca05-4a27-a76d-5e784ee23243 · outbound

This paper cites AWS Trainium.

Towards Sustainable Large Language Model Serving AWS Trainium

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:59:42.618385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:59:41.429822Z digest=sha256:d502e6b2462651de1de127e4d02f6aac4ee092570a4c0044448d3181af20ed17

Observation 13a151b6-dd24-4d8f-ae37-d3a3c715e157 · outbound

This paper cites Reducing the carbon impact of generative AI inference (today and in 2035).

Towards Sustainable Large Language Model Serving Reducing the carbon impact of generative AI inference (today and in 2035)

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:59:42.601568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:59:41.435938Z digest=sha256:7d363b64542650d497746cad25fed038dca53c2dde5275779a8b2ce56c48e426

Observation a3c9fd5d-e1c6-479a-aba0-72b07a5cf7ed · outbound

This paper cites PaLM: Scaling Language Modeling with Pathways.

Towards Sustainable Large Language Model Serving PaLM: Scaling Language Modeling with Pathways

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T22:59:41.442520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:59:41.442520Z digest=sha256:1c5371648ee4ae7ab9684d8d626e713e3355105b6103b13830a6058cf2ae8ba3

Observation 2c591211-a6c6-4fb4-81ac-d4d3a4b5837f · outbound

This paper cites Electricity maps.

Towards Sustainable Large Language Model Serving Electricity maps

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:59:42.586450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:59:41.448735Z digest=sha256:75f188bd91db4aacc8cd7b0c5296cff1f4c887ed22acc778f7dae9920863b18e

Observation 329323fe-7efe-4c62-952c-0dd3d4a92101 · outbound

This paper cites Fine- tuning giant neural networks on commodity hardware with automatic pipeline model parallelism.

Towards Sustainable Large Language Model Serving Fine- tuning giant neural networks on commodity hardware with automatic pipeline model parallelism

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:59:42.570198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:59:41.454609Z digest=sha256:56256f41b958294601a2adfb5f0ff02d045657000daa9f6cb421739f49697201

Observation b6ea0d18-0423-4e23-a6d0-f45b5cfc6b56 · outbound

This paper cites LLMCarbon: Modeling the end-to-end carbon footprint of large language models.

Towards Sustainable Large Language Model Serving LLMCarbon: Modeling the end-to-end carbon footprint of large language models

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:59:42.552869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:59:41.460255Z digest=sha256:9527a557863a13ae28425ed28cf76b0fbbae87c744fc59e4e0218ce2784c8d77

Observation 8cc2e6eb-1f51-40c2-8680-d3e3c2076fd0 · outbound

This paper cites Cost-Efficient Large Language Model Serving for Multi-turn Conversations with CachedAttention.

Towards Sustainable Large Language Model Serving Cost-Efficient Large Language Model Serving for Multi-turn Conversations with CachedAttention

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T22:59:41.466114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:59:41.466114Z digest=sha256:a57c0d88894686c8490f67530e2e6e23869cba3a19664be718031333bd27808e

Observation 4c6a22a2-27f0-4d3d-a75a-c2a82677db15 · outbound

This paper cites Accelerate AI development with Google cloud TPUs.

Towards Sustainable Large Language Model Serving Accelerate AI development with Google cloud TPUs

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:59:42.536143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:59:41.472529Z digest=sha256:d3c3e20d9c38add1b74f3ff04f5c80fe132dab94228a8b1c3ce69f7fcd498042

Observation 96fee8bd-69e2-4de8-8eea-afdb59706d0a · outbound

This paper cites Why your internet habits are not as clean as you think.

Towards Sustainable Large Language Model Serving Why your internet habits are not as clean as you think

Reference 9

Resolution
verified exact
raw_fallback, observed 2026-08-10T22:59:42.008197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:59:41.477761Z digest=sha256:59b784bf4817c6dbf9af94444906be098fb5479a47eddbc3dcff9776ecb4d00f

Observation b05c6086-3116-4798-8d37-638a1b306e88 · outbound

This paper cites Lee, David Brooks, and Carole-Jean Wu.

Towards Sustainable Large Language Model Serving Lee, David Brooks, and Carole-Jean Wu

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:59:42.517160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:59:41.484547Z digest=sha256:023dc035e4a340c9f6a43d7bcbc41c536cfd227a9392f7b2c3118270d9db469e

Observation 77fdce02-bfd6-459d-bce6-8230f3785780 · outbound

This paper cites Small Language Models Fine-tuned to Coordinate Larger Language Models improve Complex Reasoning.

Towards Sustainable Large Language Model Serving Small Language Models Fine-tuned to Coordinate Larger Language Models improve Complex Reasoning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T22:59:41.490618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:59:41.490618Z digest=sha256:19b9427f54d9e60b58d4b582305d316be2bf75538ce7030c4fac6d8373926689

Observation 0f53dac2-aca2-4afe-b0b8-395965e139e7 · outbound

This paper cites Fast inference from transform- ers via speculative decoding.

Towards Sustainable Large Language Model Serving Fast inference from transform- ers via speculative decoding

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:59:42.500433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:59:41.497085Z digest=sha256:4faf7d49800d3bdf8d4d840c4f2aea6ce7b035af63c07658073d7fdb94b71f4c

Observation 69e53815-ecc2-4b6a-8650-8ea4c21dae10 · outbound

This paper cites Toward sustainable HPC: Carbon footprint estimation and environmental implications of HPC systems.

Towards Sustainable Large Language Model Serving Toward sustainable HPC: Carbon footprint estimation and environmental implications of HPC systems

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:59:42.476795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:59:41.503036Z digest=sha256:3fe48ff8221312b395397f01d6ddf054e512e14f5ea4f1c0d8b3e975e2f8076f

Observation 92569acb-41ac-4e53-9c05-0fae7cf8d56d · outbound

This paper cites Is your code generated by ChatGPT really correct? rigorous evaluation of large language models for code generation.

Towards Sustainable Large Language Model Serving Is your code generated by ChatGPT really correct? rigorous evaluation of large language models for code generation

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:59:42.459306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:59:41.508613Z digest=sha256:b96ebf47ea37b672130f9025f95371fdeb0212d2748c457e41524d10affafc04

Observation 20024992-6927-462c-ac1e-05abb80d4657 · outbound

This paper cites Counting Carbon: A Survey of Factors Influencing the Emissions of Machine Learning.

Towards Sustainable Large Language Model Serving Counting Carbon: A Survey of Factors Influencing the Emissions of Machine Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T22:59:41.514240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:59:41.514240Z digest=sha256:e7b6133c3b21f506b19a579ae51e275284a46c3d3149b9f2c3ead98fea906624

Observation d0010800-0cf1-4e04-a5b7-63e112816355 · outbound

This paper cites Estimating the carbon footprint of BLOOM, A 176B parameter language model.

Towards Sustainable Large Language Model Serving Estimating the carbon footprint of BLOOM, A 176B parameter language model

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:59:42.442749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:59:41.520109Z digest=sha256:07f138b28b0abe0061c409226b1ab1aa9503908c70239024a5c22d42c96326a0

Observation 76366ff8-20a0-4b66-b317-3c606d82517f · outbound

This paper cites Bringing carbon awareness to multi-cloud application delivery.

Towards Sustainable Large Language Model Serving Bringing carbon awareness to multi-cloud application delivery

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:59:42.425374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:59:41.525636Z digest=sha256:42159d547fc752a295e848888e9e26ffad7fb663f79a366dd1e1063e4eafdc2b

Observation 503f2a08-f7b3-4897-a370-b1fbfa013d60 · outbound

This paper cites Sitaraman.

Towards Sustainable Large Language Model Serving Sitaraman

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:59:42.407842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:59:41.531525Z digest=sha256:821ca211dcebdf13a3c09eb80772f2ee122e89819ea2431d62a64d44771921a0

Observation d6f3b6d0-3dc0-4fab-9a9b-84bd779a5a5a · outbound

This paper cites Sitaraman, and Prashant Shenoy.

Towards Sustainable Large Language Model Serving Sitaraman, and Prashant Shenoy

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:59:42.282128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:59:41.537901Z digest=sha256:53353ac675b0d12f13d7ed0c229f9ec614dd5c9111e91a0ce6475773527c8e1d

Observation abc497d3-ef3d-4821-993d-21f3772bb373 · outbound

This paper cites MTIA v1: Meta’s first-generation AI inference accelerator.

Towards Sustainable Large Language Model Serving MTIA v1: Meta’s first-generation AI inference accelerator

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:59:42.265653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:59:41.543487Z digest=sha256:4e7bad92ad6c3a94407d076cc19a6392f80da182d08d9d4eab0ee5f2949de965

Observation 6a3138b1-8208-441f-be0c-2e524f2bbaf1 · outbound

This paper cites NVIDIA HGX AI Supercomputer.

Towards Sustainable Large Language Model Serving NVIDIA HGX AI Supercomputer

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:59:42.248337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:59:41.548196Z digest=sha256:f0f94c9f6b3e1d7e8b3f0777a73c2a559c75a4921ee5b1bf7560a3257a95ea05

Observation 9002b9f7-abea-4745-ad59-16353fe32a26 · outbound

This paper cites NVIDIA management library (NVML).

Towards Sustainable Large Language Model Serving NVIDIA management library (NVML)

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:59:42.233086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:59:41.552771Z digest=sha256:f235d6fb4e486297d89ed38b2655a46dce817ac27eeebdf9d5b14a6714c6c6df

Observation 54bb7531-422b-48af-ba70-7b522918f5a7 · outbound

This paper cites Ashraf, Christian Engelmann, Mallikarjun Shankar, and James H.

Towards Sustainable Large Language Model Serving Ashraf, Christian Engelmann, Mallikarjun Shankar, and James H

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:59:42.216849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:59:41.557390Z digest=sha256:ceba385ff183d4219bd65c988e0eeec12f622b440daa5e8515128092f84bb6b1

Observation 5de76dd5-9c77-48c6-95d1-dc14369130eb · outbound

This paper cites Splitwise improves GPU usage by splitting LLM inference phases.

Towards Sustainable Large Language Model Serving Splitwise improves GPU usage by splitting LLM inference phases

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:59:42.200609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:59:41.562712Z digest=sha256:01ee4a344064ab337e6029caacbb23a5cfde743f50836cc4a51bb026ba056d16

Observation a6aa644a-c1f2-4126-ab03-997972818433 · outbound

This paper cites So, Maud Texier, and Jeff Dean.

Towards Sustainable Large Language Model Serving So, Maud Texier, and Jeff Dean

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:59:42.184194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:59:41.568407Z digest=sha256:1bbfdd88ac3bfe40b6739a91a47f0834a5090286538ba9f3ed56caddb0d81a7a

Observation be7924a2-302d-4d43-8cef-03592f7b667d · outbound

This paper cites HuggingGPT: Solving AI tasks with ChatGPT and its friends in hugging face.

Towards Sustainable Large Language Model Serving HuggingGPT: Solving AI tasks with ChatGPT and its friends in hugging face

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:59:42.167871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:59:41.574133Z digest=sha256:000a6cf19f65c01a63b902eb3cac50e132361865b7e3788526fede53ac017490

Observation eb8f7b95-e03c-49a2-8f4a-3ca3c00adc85 · outbound

This paper cites PowerInfer: Fast Large Language Model Serving with a Consumer-grade GPU.

Towards Sustainable Large Language Model Serving PowerInfer: Fast Large Language Model Serving with a Consumer-grade GPU

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T22:59:41.579934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:59:41.579934Z digest=sha256:f2811bff54864b318aa98c370cfc63c02e8d394a62646e70b1243f5e8e7317d3

Observation fba87fc4-3f3e-48c3-8199-7be765e74fe7 · outbound

This paper cites Hashimoto.

Towards Sustainable Large Language Model Serving Hashimoto

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:59:42.151010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:59:41.585451Z digest=sha256:c9fcc9865af5ac2691121e459bbd12a8db03fc536a00ba9a91f53dcb5cf391c3

Observation f6ffe0fd-31f4-4bfe-8192-98fcac1e5c2e · outbound

This paper cites NVIDIA Tesla T4.

Towards Sustainable Large Language Model Serving NVIDIA Tesla T4

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:59:42.131793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:59:41.591163Z digest=sha256:3127d70108340b38997fb468331ef1d7f3e356b65f5eecc3584fa65396b13c1d

Observation 63a7e315-6af4-47a2-8c1d-02942f7f7adc · outbound

This paper cites NVIDIA RTX 6000 Ada Generation.

Towards Sustainable Large Language Model Serving NVIDIA RTX 6000 Ada Generation

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:59:42.113678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:59:41.596627Z digest=sha256:7d1ef1c141a4e2071db1abe990bc85d2e093e95c805059be6584ec933fd80128

Observation d449df61-6ba9-4d72-99ec-3c392337f535 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Towards Sustainable Large Language Model Serving LLaMA: Open and Efficient Foundation Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T22:59:41.602195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:59:41.602195Z digest=sha256:2568e5b2539f505a7d41c39908b0b84b3131f29f2c9500440472c8763154f841

Observation 02d3299f-b931-4491-bc3f-442959e8924d · outbound

This paper cites FreshLLMs: Refreshing Large Language Models with Search Engine Augmentation.

Towards Sustainable Large Language Model Serving FreshLLMs: Refreshing Large Language Models with Search Engine Augmentation

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T22:59:41.607643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:59:41.607643Z digest=sha256:df84c9b86e1c300eec844d5166f54a65908a5ea7fb4ee1434ef47333c7c51e08

Observation 87d7fde2-a4ed-4953-be19-6b080b4fdaad · outbound

This paper cites Peeling back the carbon curtain: Carbon optimization challenges in cloud computing.

Towards Sustainable Large Language Model Serving Peeling back the carbon curtain: Carbon optimization challenges in cloud computing

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:59:42.097126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:59:41.613273Z digest=sha256:7a6452113478dc8d4b46a1e53909a55c581bc0dd8c57cadc2e9ba284d5edd40a

Observation 14800bd2-dc19-4eb9-837b-cc0a9272dc63 · outbound

This paper cites Gen AI’s environmental ledger: A closer look at the carbon footprint of ChatGPT.

Towards Sustainable Large Language Model Serving Gen AI’s environmental ledger: A closer look at the carbon footprint of ChatGPT

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:59:42.080318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:59:41.618823Z digest=sha256:8ec5084e4b6e0dddc4fea6535e7a6a83fa0e54975cd362b045e3cfcce3952182

Observation b7ac0204-d0ff-4dce-a4bb-e06d7a34dc26 · outbound

This paper cites Small Models are Valuable Plug-ins for Large Language Models.

Towards Sustainable Large Language Model Serving Small Models are Valuable Plug-ins for Large Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T22:59:41.624539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:59:41.624539Z digest=sha256:0a13002c72bb9c18eca84eb4951b7bf300e951b309cb05cd6f01e52b8ef6dfe9

Observation 55516410-530a-45c1-90e1-374fafe0503e · outbound

This paper cites mLoRA: Fine-Tuning LoRA Adapters via Highly-Efficient Pipeline Parallelism in Multiple GPUs.

Towards Sustainable Large Language Model Serving mLoRA: Fine-Tuning LoRA Adapters via Highly-Efficient Pipeline Parallelism in Multiple GPUs

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-10T22:59:41.630016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:59:41.630016Z digest=sha256:990cc58050f88474394dbd51bbb77efcb4f30498f113d8d1960ea0088d733b3a

Observation c5182538-f188-48fe-9e35-5c0c2f12a564 · outbound

This paper cites RAFT: Adapting Language Model to Domain Specific RAG.

Towards Sustainable Large Language Model Serving RAFT: Adapting Language Model to Domain Specific RAG

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T22:59:41.636057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:59:41.636057Z digest=sha256:5c86bd4aa809d36587770bade2fc48104c83bd708e82598f0d318b450c8c6380

Observation 0c4fd526-13e3-44a6-b712-ba3d5b1ed46d · outbound

This paper cites A GNN-based day ahead carbon intensity forecasting model for cross-border power grids.

Towards Sustainable Large Language Model Serving A GNN-based day ahead carbon intensity forecasting model for cross-border power grids

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:59:42.062685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:59:41.641223Z digest=sha256:ad94b2e0aed6acc85d444477211e7f69f05cd2b4805140fa18fa76c941ff0dfe

Observation 7ad3930f-8210-48ee-837b-669f6b1b7e1c · outbound

This paper cites Embodied Carbon Accounting through Spatial-Temporal Embodied Carbon Models.

Towards Sustainable Large Language Model Serving Embodied Carbon Accounting through Spatial-Temporal Embodied Carbon Models

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-08-10T22:59:41.715700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-10T22:59:41.646483Z digest=sha256:8fe4acaa9d1fc692788a10bd5995636e03474faeee46dd83550cdf582e0633b2

Observation 0f926cef-6847-43f8-bfe5-259bb2b16ab8 · outbound

This paper cites DistServe: Disaggregating Prefill and Decoding for Goodput-optimized Large Language Model Serving.

Towards Sustainable Large Language Model Serving DistServe: Disaggregating Prefill and Decoding for Goodput-optimized Large Language Model Serving

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T22:59:41.651892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:59:41.651892Z digest=sha256:3c44afb5dac6838a0976cbdf5863313e33c4bf8fd720b8c9371b0de8528db154

Pith citing papers

No inbound Pith citation observations are available.